Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

117 results about "Structured content" patented technology

Structured content is information or content that is organized in a predictable way and is usually classified with metadata. XML is a common storage format, but structured content can also be stored in other standard or proprietary formats.

Financial document automatic auditing method and device and medium

The invention discloses a financial document automatic auditing method and device and a medium, and relates to the technical field of financial reimbursement auditing. The method comprises the following steps: receiving a financial document to be audited and an associated attachment document, respectively extracting a first entity set, and extracting a second entity set from unstructured content; constructing a dynamic knowledge graph state space based on the entity set, wherein the dynamic knowledge graph state space comprises an entity vector generated by an entity embedding algorithm and a relation vector generated by a relation coding algorithm; defining a reinforcement learning action space, wherein the reinforcement learning action space comprises three types of atomic operations of newly adding and deleting a triple and adjusting confidence; in combination with the real-time document flow, the historical case library and the audit result data, dynamically evolving the knowledge graph through atomic operation, and calculating a value return value of each operation; pre-judging accumulated return values of different operation sequences by utilizing a Monte Carlo tree search algorithm, and pruning a low return sequence; executing the optimized operation sequence to update the knowledge graph; and finally, based on the updated atlas, triggering a logic verification rule to generate an auditing result.
Owner:INSPUR GENERSOFT CO LTD

Semantic and situational knowledge collaborative modeling declarative knowledge construction method and device, computer equipment and readable storage medium

The invention discloses a declarative knowledge construction method and device for semantic and situational knowledge collaborative modeling, computer equipment and a readable storage medium, and relates to the field of data processing.The method comprises the steps that firstly, a multi-modal document is analyzed, a chapter abstract is extracted, and structured content is obtained; entities, events and multi-modal knowledge points are extracted from the structured content, and cross-modal fusion is carried out on the entities, the events and the multi-modal knowledge points; carrying out anaphora resolution based on the fused knowledge points, and constructing a double atlas containing a knowledge atlas, a affair atlas and a four-dimensional relation triple; clustering the double maps to obtain a theme community, and performing association mapping on the community, the triple and the entity event, the chapter abstract and the multi-modal knowledge point to form association knowledge; vectorizing the associated knowledge and establishing a vector knowledge index; and carrying out compression ratio and accuracy evaluation on the knowledge through an evaluation system, and feeding back and optimizing the whole knowledge construction process. According to the method, multi-modal knowledge deep fusion and semantic scene collaborative modeling are realized, and the knowledge structuring degree and the application reliability are improved.
Owner:DARK MATTER ARTIFICIAL INTELLIGENT (BEIJING) TECHNOLOGY CO LTD

Report generation method and system based on multi-agent architecture

The invention provides a report generation method and system based on a multi-agent architecture, and the method comprises the steps: firstly receiving and analyzing a report generation instruction, obtaining a theme demand, a framework specification and an initial reference material, then starting a multi-agent cooperation framework, generating an agent task distribution table, and generating an agent task distribution table; and the resource retrieval agent executes network resource directional retrieval according to the task allocation table to generate an associated resource set, and the content extraction agent performs structured conversion on the associated resource set and the initial reference material to obtain a structured content unit with chapter codes. The method comprises the following steps: performing module classification and logic series connection on a structured content unit by a report integration agent in combination with a framework specification to generate a report first draft, and finally performing content verification and optimization on the report first draft by a multi-agent collaborative framework to generate a final report text conforming to the framework specification, thereby realizing automation, intellectualization and high efficiency of report generation. And report quality is improved.
Owner:JIEHELIX (SHANGHAI) MEDICAL TECH CO LTD

Method for creating and correcting large model fine tuning data set

The invention discloses a method for creating and correcting a large model fine tuning data set, and relates to the technical field of large models. The method comprises the following steps: converting received structured and unstructured texts into structured content units in a unified format; generating and distilling an initial question and answer pair by using a teacher model taking a pre-trained large language model as a core; screening out high-quality question and answer pairs according to a quantitative evaluation system containing correlation, accuracy and integrity dimensions; performing diversified optimization on high-quality question and answer pairs by applying a generalization strategy and multi-round distillation, and performing dynamic regulation and control based on language feature difference to avoid content homogenization; and finally, guiding the teacher model to generate ternary structure data containing a thinking chain, and executing a verification and self-correction algorithm so as to generate a final fine tuning data set containing a verified reasoning process.
Owner:YGSOFT INC

Regional wind field multi-source heterogeneous work order data rapid structuring system fusing large language model

The invention provides a regional wind field multi-source heterogeneous work order data fast structuring system fused with a large language model, which comprises a plurality of modules, and is characterized in that a data input module receives various work order images; the OCR layout recognition module comprehensively recognizes the data, extracts key information, determines a spatial position and outputs text fragments with coordinates and layout anchor point information, and the rule field extraction module recognizes format category fields from the text fragments and calculates rule matching confidence; the semantic field recognition module recognizes descriptive fields from unstructured content and outputs related confidence information, the fusion decision module calculates final confidence of field candidates according to a predetermined function and judges output or complementation according to a threshold value, and the data consistency verification module is connected with an asset database to verify key information. The data complementation module generates complementation candidates and confidence coefficients, the man-machine collaborative closed-loop module pushes suggestions and receives feedback, and the result output module outputs standardized data. According to the method, the multi-source heterogeneous work order data can be effectively processed.
Owner:SHAANXI HUADIAN NEW ENERGY POWER GENERATION CO LTD +1

Cross-platform content generation and distribution method based on multi-modal AI

The invention discloses a cross-platform content generation and distribution method based on a multi-modal AI, and belongs to the technical field of cross-platform content generation and distribution, and the method comprises the steps: carrying out the content analysis and feature extraction of an original material based on a multi-modal AI model, and generating a structured content label and a semantic vector. By combining a vector matching degree formula of target portrait features and platform features, an adaptation strategy is dynamically generated, it is ensured that content not only conforms to platform rules, but also can accurately reach a target group, the conversion rate and user viscosity are finally improved, visual, text and semantic vectors are aligned through a Transform multi-modal fusion model, a cross-modal joint representation vector is generated, and the user experience is improved. And in combination with an adversarial generative network, differentiated variants of the same theme are generated in batches, and a content diversity score mechanism ensures that generated contents are balanced between creativity and compliance.
Owner:QUZHOU TIMES ENGINE NETWORK TECHNOLOGY CO LTD

Structured auxiliary writing system based on multi-agent cooperation

The invention relates to the technical field of auxiliary writing systems, in particular to a structured auxiliary writing system based on multi-agent collaboration. The system specifically comprises a writing task demand analysis and modeling module, a multi-agent system architecture construction module, a writing task disassembly and subtask distribution module, a multi-source knowledge retrieval and integration module, a structured content generation module, a format specification and quality evaluation module, a logic verification and content adjustment module and a user feedback iteration and model optimization module. The writing task demand analysis and modeling module firstly performs surface analysis on a writing task of a user, determines the type, theme, target audience, core demand and quality requirement of the writing task, and converts a fuzzy writing demand into a quantifiable and executable structured task model; according to the invention, the limitation of a traditional writing mode and an existing single-function auxiliary tool is solved.
Owner:ZHILONG INNOVATION (BEIJING) TECHNOLOGY CO LTD

PPT exporting method and device based on HTML page, equipment and medium

The invention discloses a PPT exporting method, device and equipment based on an HTML page and a medium, and is applied to the technical field of computers. The method comprises the steps that the content type of a front-end page constructed based on the HTML is recognized; calling a native interface of the slide generation tool to export the structured content of the front-end page to the PPT; using a snapshot technology to render the three-dimensional model of the front-end page after the view angle is selected to obtain a rendering result, converting the rendering result into an image and embedding the image into the PPT; and rendering decorative chart elements of the front-end page by using a delay rendering technology, intercepting an image of a corresponding area after rendering, and exporting the intercepted image to the PPT. According to the method, high-fidelity, high-automation, editable and visually consistent multi-source heterogeneous content PPT export is realized, and a high-quality migration path from Web visual content to a PowerPoint presentation document is effectively opened.
Owner:SHANDONG CVICSE MIDDLEWARE CO LTD

Generative services for content search and chat interfaces in a collaboration platform

Embodiments described herein relate to systems and methods for automatically generating content, generating API requests and / or request bodies, structuring user-generated content, and / or generating structured content in collaboration platforms, such as documentation systems, issue tracking systems, project management platforms, and other platforms. The systems and methods described use a network architecture that includes a search and chat generative interfaces, which may be used to produce generative answers, identify relevant content, and take actions within a respective content collaboration platform.
Owner:ATLASSIAN PTY LTD

Multifunctional integrated digital network broadcasting system based on intelligent algorithm

The invention relates to the technical field of digital network broadcasting, in particular to a multifunctional integrated digital network broadcasting system based on an intelligent algorithm. Comprising a multi-modal data acquisition module, a broadcast content generation module, a broadcast scheduling decision module, an abnormal interference suppression module and a terminal collaborative feedback module. The multi-modal data acquisition module collects multi-source information of a coverage area in real time. And the broadcast content generation module performs intelligent clustering analysis based on the information to generate structured content. And the broadcast scheduling decision module adopts a graph attention scheduling algorithm to optimize a scheduling strategy. And the abnormal interference suppression module identifies an interference source by using a residual convolutional network, and repairs an interfered signal through a recurrent neural network. And the terminal collaborative feedback module executes playing and data feedback, and dynamically adjusts receiving parameters and pushing strategies by means of federal reinforcement learning. The intelligent level, the adaptability, the user satisfaction and the overall service efficiency of the digital network broadcasting system are remarkably improved.
Owner:GUANGDONG RUIZHAO AUDIO EQUIPMENT CO LTD

Artificial intelligence-based video content understanding method and system

The present application relates to the technical field of artificial intelligence, in particular to a video content understanding method and system based on artificial intelligence, which first acquires a video scene unit set of a video to be understood arranged in time sequence and containing scene element and object behavior information; then establishes a semantic behavior association graph of scene element information and object behavior information for each scene unit; further generates a hierarchical understanding process containing scene element semantic association analysis and object behavior logical association analysis based on the semantic behavior association graph; dynamically optimizes and adjusts the hierarchical understanding process according to the analysis result; and finally outputs a structured content understanding result containing scene semantic description and object behavior logical description based on the optimized process, so as to comprehensively, deeply and dynamically understand the video content.
Owner:LIANGSHAN BRANCH OF SICHUAN TOBACCO

Automated content creation and content services for collaboration platforms

Embodiments described herein relate to systems and methods for automatically generating content, generating API requests and / or request bodies, structuring user-generated content, and / or generating structured content in collaboration platforms, such as documentation systems, issue tracking systems, project management platforms, and other platforms. The systems and methods described use a network architecture that includes a prompt generation service and a set of one or more purpose-configured large language model instances (LLMs) and / or other trained classifiers or natural language processors used to provide generative responses for content collaboration platforms.
Owner:ATLASSIAN PTY LTD +1

Special effect video generation method and system

The invention relates to a special effect video generation method and system, and the method comprises the following steps: S1, carrying out video content understanding based on a Vision-Language Model, inputting an original video frame sequence to the Vision-Language Model, and outputting a structured content label and a special effect content candidate list to the Vision-Language Model; s2, performing special effect style injection on the video generation model by adopting a LoRA fine tuning technology so as to realize personalized and content-driven special effect animation generation; and S3, fusing the special effect layer output by the video generation model with the original video frame to generate a final special effect video with visual impact, natural transition sense and dynamic consistency. According to the invention, the video content can be automatically identified and the special effect video conforming to the video content is generated.
Owner:GIANT MOBILE TECH CO LTD

Standard draft auditing and correcting method and device, electronic equipment and storage medium

The invention provides a standard draft auditing and correcting method and device, electronic equipment and a storage medium, and relates to the technical field of data processing, and the method comprises the steps: obtaining a to-be-audited standard draft, and preprocessing the to-be-audited standard draft to obtain structured content data and structured format data of the to-be-audited standard draft; inputting the structured content data into a pre-trained content correction model to obtain a standard draft marked with to-be-corrected content; wherein the content correction model is obtained through reinforcement learning training based on a large language model; comparing the structured format data with a pre-constructed format knowledge database to obtain a standard draft marked with a to-be-corrected format problem; and re-checking the to-be-corrected content and the to-be-corrected format problem, carrying out content correction on the to-be-corrected content after re-checking, and carrying out format correction on the to-be-corrected format problem to obtain a correct standard draft. The auditing efficiency and accuracy of the standard draft are improved.
Owner:BEIJING MUNICIPAL PUBLIC SECURITY BUREAU ARTIFICIAL INTELLIGENCE SECURITY RESEARCH CENTER

Consistency comparison method and system for multi-mode electronic signed files

The invention relates to the technical field of data processing, and particularly discloses a consistency comparison method and system for multi-modal electronic signed files, and the method comprises the following steps: S1, carrying out the multi-dimensional Hash calculation of a to-be-compared file, and generating a comprehensive Hash consistency score; s2, judging whether the comprehensive hash consistency score reaches a preset threshold value or not: if yes, judging that the comparison files are consistent; if not, cross-modal conjoint analysis is carried out on the text, the image and the structured content of the comparison file through a multi-modal deep learning model, and a consistency prediction score is generated; s3, judging whether the comparison files are consistent or not according to the consistency prediction score and a preset threshold value, and generating a judgment result; by comprehensively applying multi-dimensional hash code analysis and combining and using multi-modal model comparison and covering data of different formats and types, subtle changes of various contents such as texts, images and tables can be detected, so that the high efficiency and accuracy of the electronic signed file are improved.
Owner:SICHUAN TUOTUO DI SCI & TECH CO LTD

Multi-party cross-platform query and content creation service and interface for collaboration platforms

Embodiments described herein relate to systems and methods for automatically generating content, generating API requests and / or request bodies, structuring user-generated content, and / or generating structured content in collaboration platforms, such as documentation systems, issue tracking systems, project management platforms, and other platforms. The systems and methods described use a network architecture that includes a generative interface panel having multiple automated assistant services. Each assistant service may access a prompt generation service and a set of one or more purpose-configured large language model instances (LLMs) and / or other trained classifiers or natural language processors used to provide generative responses for content collaboration platforms.
Owner:ATLASSIAN PTY LTD

Document classification splitting method, device, system and equipment and storage medium

The invention relates to the technical field of document processing, and provides a document classification splitting method, device, system and equipment and a storage medium, and the method comprises the following steps: obtaining a PDF format document, and judging whether the PDF format document is in a scanning PDF format; if the document in the PDF format is in a scanning PDF format, analyzing each structured content of the document in the PDF format on the basis of a multi-mode OCR (Optical Character Recognition) engine; if the document in the PDF format is not in the scanning PDF format, analyzing each structured content of the document in the PDF format based on a structured document analysis engine; sorting each structured content into a plurality of document logic units according to context semantics based on a reinforcement learning model; and for each document logic unit, fusing the text feature, the visual feature and the structural feature of the document logic unit, and carrying out classified marking on the document logic unit based on the classification model. The PDF format document can be intelligently and automatically classified and split, the splitting accuracy is high, and the splitting efficiency is high. And the splitting effect is good.
Owner:CHINA CITIC BANK CO LTD

Systems and methods for generating structured conversational ai content from unstructured and structured data sources

A system and method are disclosed for generating conversational content from human-readable documents. The method includes receiving a document comprising unstructured or semi-structured content and extracting linguistic and layout features using a language model and layout analysis techniques. The document is segmented into atomic content blocks representing discrete semantic units. For at least one content block, a natural language question is generated using a neural model, and a corresponding answer is extracted or synthesized. An optional rephrasing step modifies the surface form of the question or answer while preserving semantic meaning. Each question-answer pair is reviewed using automated or human-in-the-loop mechanisms for accuracy and alignment. Approved content is stored in a structured repository along with metadata supporting traceability and deployment. The system supports enterprise-scale generation of high-quality conversational data for downstream applications such as chatbots, virtual assistants, and retrieval-based AI systems.
Owner:KNOWBL LLC

Video content understanding method and system based on artificial intelligence

The invention relates to the technical field of artificial intelligence, in particular to a video content understanding method and system based on artificial intelligence, and the method comprises the steps: firstly obtaining a video scene unit set which is arranged according to a time sequence of a to-be-understood video and contains scene elements and object behavior information; establishing a semantic behavior association graph of scene element information and object behavior information for each scene unit; generating a hierarchical understanding process containing scene element semantic association analysis and object behavior logic association analysis based on the semantic behavior association graph; dynamically optimizing and adjusting the hierarchical understanding process according to an analysis result; and finally, outputting a structured content understanding result containing scene semantic description and object behavior logic description based on the optimized process, thereby comprehensively, deeply and dynamically understanding the video content.
Owner:LIANGSHAN BRANCH OF SICHUAN TOBACCO

File structured information extraction method and device, equipment, medium and product

The invention discloses a file structured information extraction method and device, equipment, a medium and a product, and relates to the technical field of data processing, and the method comprises the following steps: determining a file content type of a to-be-processed file; under the condition of determining that the file content type is the image content file, performing text recognition on the to-be-processed file, and determining a to-be-processed text contained in the to-be-processed file and a text region coordinate corresponding to the to-be-processed text in the to-be-processed file; carrying out structured content entity recognition on the to-be-processed text, and determining structured content entities contained in the to-be-processed text and content entity coordinates respectively corresponding to the structured content entities in the text region coordinates; and constructing content entity relation data among the structured content entities according to the content entity coordinates, and performing structured information extraction on the to-be-processed text according to the content entity relation data to obtain target structured information contained in the to-be-processed file. According to the invention, the accuracy and integrity of structured information extraction can be improved.
Owner:HANGZHOU HAOLINK INTELLIGENT TECHNOLOGY CO LTD +1

File structured information extraction method, device, equipment, medium and product

The application discloses a kind of extraction methods, device, equipment, medium and product of file structured information, it is related to data processing technical field, comprising: determining the file content type of to-be-processed file;In the case where it is determined that file content type is image content file, to-be-processed file is carried out text recognition, and the text area coordinates of to-be-processed text contained in to-be-processed file and to-be-processed text in to-be-processed file are determined;To-be-processed text is carried out structured content entity identification, and the content entity coordinates of each structured content entity in text area coordinates respectively corresponding to structured content entity contained in to-be-processed text are determined;According to each content entity coordinates, the content entity relationship data between each structured content entity is constructed, and according to content entity relationship data, to-be-processed text is carried out structured information extraction, and the target structured information contained in to-be-processed file is obtained.The application can improve the accuracy and integrity of structured information extraction.
Owner:HANGZHOU HAOLINK INTELLIGENT TECHNOLOGY CO LTD +1

Vehicle repair file processing method and related equipment

The invention provides a vehicle repair file processing method and related equipment, and the method comprises the steps: receiving a vehicle repair file uploaded by a user, and recognizing the format of the vehicle repair file; converting the automobile repair file into a Markdown format by using a conversion method matched with the format of the automobile repair file to obtain a target automobile repair file; inputting the target automobile repair file into a preset language large model to obtain file content extracted by the language large model; and exporting the file content according to a preset business data format. In the scheme, the automobile repair file format is recognized and uniformly converted into the Markdown format, the problem of complex format analysis of multi-column typesetting, table nesting and image-text mixing is solved, the content is standardized, the standardized file is input into a large language model, the natural language understanding ability of the standardized file is utilized, professional terms are precisely recognized, and the structured content is extracted; the extracted content is exported according to the preset service format, so that unified management of the automobile repair file content is realized, and the data utilization rate is improved.
Owner:LAUNCH TECH CO LTD

Method and apparatus for the intelligent capture and preprocessing of unstructured content for upstream or downstream workflow integration

This technology relates to field of Intelligent Document Processing (IDP) and more particularly to the methods, apparatus, and systems that enables the capture and handling of unstructured content; enabling the pre-processing of such captured content to extract, classify, convert, and / or summarize the captured content; providing at least one of the extracted content, the original captured content, and / or any generated ancillary information about the captured content (metadata); potentially allowing the efficient triage of such content, extracted content, and / or metadata by automated or manual processes; and determining whether to forward at least one of such captured content, extracted information, and / or metadata to upstream or downstream workflow processes or request corrections or additions to such content from the source of the provided content.
Owner:ETHERFAX LLC

Dynamic ai system for context-aware, domain-specific workflow management

A system is disclosed for generating context-aware, domain-specific responses using a pre-trained language model in combination with a semantic search engine and specialized processing modules. The system includes a classification model to identify user intent, an extraction model to determine parameters, and a plurality of parameter functions that generate outputs such as sentiment classifications, semantically similar content, and dynamically constructed prompts. By integrating retrieved structured and unstructured content into tailored prompts, the system provides accurate, domain-specific query responses without requiring retraining of the underlying language model. The architecture supports efficient and scalable workflow management in dynamic environments while reducing computational overhead compared to conventional approaches.
Owner:CELLIGENCE INTERNATIONAL LLC

Retrieval recall method and device based on table analysis, equipment and storage medium

The invention provides a retrieval recall method and device based on table analysis, equipment and a storage medium, and is applied to the fields of finance and medical treatment. The method provided by the invention comprises the following steps: obtaining a table to be analyzed, if the layout structure of the table to be analyzed comprises unstructured content, segmenting the unstructured content of the table to obtain a subdivided table, and analyzing the subdivided table to obtain a sliced table; enabling the large language model to generate first characteristic information corresponding to the slice table according to the slice table, wherein the first characteristic information comprises an abstract, a question and answer pair and label information; obtaining a user question, and generating second characteristic information corresponding to the user question according to the user question; a retrieval resource object is determined according to the similarity between the second characteristic information corresponding to the user question and the first characteristic information corresponding to the slice table, a retrieval recall result is generated according to the retrieval resource object, and the retrieval resource object comprises the slice table. According to the method, the accuracy of the retrieval recall result can be improved.
Owner:PING AN TECH (SHENZHEN) CO LTD

Self-evolution demonstration document generation method and system based on multi-modal and knowledge graph

PendingCN122262361AAvoid Distortion of Detailsavoid inconsistent styleSemantic analysisSpecial data processing applicationsEngineeringContinual learning
The application provides a self-evolution presentation document generation method and system based on multi-modal and knowledge graph, wherein the self-evolution presentation document generation method comprises the following steps: constructing an enterprise-level picture meta-database containing a vector index; parsing a source document to obtain structured content and generating a presentation document outline; searching for matching candidate pictures in the enterprise-level picture meta-database based on the target page core content in the outline; querying a design decision knowledge graph to generate layout and color matching planning; assembling a presentation document according to the planning and synchronously generating design decision metadata; driving knowledge graph updating by displaying the design decision metadata and obtaining user modification feedback on the design; and the self-evolution presentation document generation system comprises functional modules for realizing the above method. The application solves the technical problems of poor picture quality, inaccurate image-text matching, rigid design, high copyright risk and difficulty in continuous learning and evolution from use in the existing automatic presentation document generation technology.
Owner:CHINA RAILWAY TUNNEL GROUP CO LTD +1

Curation And Provision Of Digital Content

A method includes accessing a structured content item from a first database and event data from a second database, the event data including sets of event attributes in a multi-dimensional namespace and associated with a respective point in time; determining a relevancy profile characterizing a metric of relevancy of the structured content item over a respective time interval, the metric of relevancy including a distance in the multi-dimensional namespace between attributes associated with the structured content and the sets of event attributes; generating, using the relevancy profile, second digital content including a subset of the structured content item; and providing the second digital content for rendering on a device. Related apparatus, systems, techniques and articles are also described.
Owner:NANT HOLDINGS IP LLC

An electronic bidding system and method based on XML structured data

The application relates to the technical field of electronic bidding systems, in particular to an electronic bidding system and method based on XML structured data, which comprises the following units: a structured content division unit which divides bidding content information into multiple segments and converts the segments into corresponding XML segments; a data binding unit which establishes a mapping relationship between structured label variables and associated data in a system database or user input project information; a template storage unit which combines the screened XML segments into complete XML bidding templates and stores the templates according to industry categories; a bidding file generation unit which calls the bidding templates matched with the user input project information, automatically fills the associated data into the label variables, and generates complete bidding XML files; a bidding file generation unit which classifies the project data filled by bidders according to requirements and generates bidding XML file data; and an evaluation assisting unit which extracts the project data bound by the structured label variables and generates a horizontal comparison analysis table.
Owner:BEIJING JINGNENG TENDERING & COLLECTIVE PROCUREMENT CENT CO LTD

Automated content creation and content services for collaboration platforms

Embodiments described herein relate to systems and methods for automatically generating content, generating API requests and / or request bodies, structuring user-generated content, and / or generating structured content in collaboration platforms, such as documentation systems, issue tracking systems, project management platforms, and other platforms. The systems and methods described use a network architecture that includes a prompt generation service and a set of one or more purpose-configured large language model instances (LLMs) and / or other trained classifiers or natural language processors used to provide generative responses for content collaboration platforms.
Owner:ATLASSIAN PTY LTD +1

Smart identification of indicator text with full-text search or optimized document analysis

Several aspects for optimizing unstructured document analysis comprise operating a document system, where the document system comprises a plurality of documents comprising unstructured content and a full-text index; receiving a request to identify documents comprising a type of data elements; selecting a sample out of the plurality of documents; determining data elements of the type in the sample of documents; determining an indicator context expression for the type of data elements out of the determined data elements of the type; determining a query for searching, using a search engine, the full-text index using the indicator context expression; and determining the documents in the document system being compliant to the determined query.
Owner:INTERNATIONAL BUSINESS MACHINE CORPORATION