Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

111 results about "Text searching" patented technology

In text retrieval, full text search refers to techniques for searching a single computer-stored document or a collection in a full text database.

Full-text retrieval method and system fusing various types of documents

The invention provides a full-text retrieval method and system fusing various types of documents, and relates to the technical field of information retrieval, and the method comprises the following steps: obtaining document representation through document content extraction and structure recognition, generating a cross-modal semantic vector by using word embedding and nonlinear transformation, constructing a hierarchical index and a cross-document association graph, and obtaining a full-text retrieval result; the basic correlation score is calculated after the query request is received, and the comprehensive score of the candidate content segments is calculated based on the association graph to determine the optimal retrieval result, so that unified representation and retrieval of heterogeneous documents are realized, the cross-document retrieval precision and relevance are improved, and the processing capability of a retrieval system on complex queries is enhanced.
Owner:BEIJING CHANGFA TECH CO LTD

Relative fuzziness for fast reduction of false positives and false negatives in computational text searches

A computer-implemented method for computational textual search to find and display search identified information in documents. Search queries are processed over one or more documents either at a server or user device using match schemes that produce both binary (match or no match) and non-binary (i.e. multiple matching values) results and have relative fuzziness relationships. A fuzzier match scheme implies more results and fewer false negatives and a less fuzzy match scheme implies fewer results and fewer false positives. This fuzziness relationship allows, without changing a search string, for users to quickly change from a match scheme to a fuzzier or less fuzzy match scheme depending on user evaluation of results as containing too many false negatives or too many false positives—without becoming a programmer of complex search string metadata or an expert user of advanced search capabilities.
Owner:DENNINGHOFF KARL LOUIS

Text processing method, text processing apparatus, electronic device, and computer-readable storage medium

A text processing method includes obtaining a query text, invoking a search engine interface based on the query text to obtain a plurality of text search results corresponding to the query text, obtaining, from the plurality of text search results, a plurality of answer text segments matching the query text, determining a relevance between the query text and each of the plurality of answer text segments, determining one of the plurality of answer text segments that corresponds to a maximum relevance as a reference text of the query text, and invoking a language model based on the query text and the reference text to obtain a reply text of the query text.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

System and methods for audio data analysis and tagging

A system for automated processing and analysis of audio files for large data sets in a cloud environment. A unified analytic environment can integrate audio machine learning models for processing and analysis with a knowledge management system, including graph presentations of tracked entities, linked to audio files and / or associated translations and transcripts. Entities within such data can be searched or filtered and proposed for tracking, or identified as tracked objects. These features can allow triage and prioritization of audio files for analysis. User interfaces can facilitate feedback on transcription and translation outputs, thereby improving present outputs and future inputs and outputs. Entities speaking or referred to can be found, tagged, and distinguished in audio files (e.g., using speaker identification in audio files, text searching in transcripts, etc.) Users can provide feedback and input on various aspects of a system, to enhance or adjust initial automated or other machine learning outputs.
Owner:PALANTIR TECHNOLOGIES INC

Large model-based credit review knowledge question and answer method, equipment and medium

The invention discloses a question answering method and device for credit review knowledge based on a large model and a medium, and the method comprises the steps: obtaining a to-be-processed credit review question, extracting a core keyword from the question through a natural language processing algorithm, and converting the question into a deep semantic vector through a semantic vector model; inputting the questions into an intention classification model, generating filtering labels, inputting the questions into a distributed full-text search engine, screening a target knowledge base according to the filtering labels, and performing retrieval in the target knowledge base based on the core keywords and the deep semantic vectors to obtain a preliminary retrieval result set; calculating the score of each preliminary retrieval result in the result set according to the core keyword, the deep semantic vector and the weight factor, and screening out final reference content in a descending order; and on the basis of credit and loan business compliance requirements, risk control rules and question and answer output specifications, the credit and loan prompt words are constructed, and the final reference content is input into the large model to obtain structured credit and loan review answers, so that the question and answer efficiency and accuracy are improved.
Owner:INDUSTRIAL AND COMMERCIAL BANK OF CHINA

Systems and Methods for a Context-Aware Retrieval System for FPGA Design Implementation and Closure Processes

PendingUS20260057184A1Semantic analysisRelational databasesFull text searchTechnical specifications
Systems or methods of the present disclosure may provide a design tool for adjusting designs implemented on programmable logic devices. The present disclosure includes receiving documentation and receiving a full text search. The documentation may include user guides, technical specifications, and design files such as HDL code, constraints, timing reports, and / or design assistant / rule violation (DRC) reports. The present disclosure also includes determining semantic search for embedding vectors based on the documentation and the full text search. Furthermore, the present disclosure includes providing the semantic search to a large language model (LLM).
Owner:KOTIYAL SAURABH +2

BEV-based scene-level text search point cloud retrieval method and apparatus, and electronic device

The invention discloses a BEV-based scene-level text search point cloud retrieval method and device and electronic equipment, and the method comprises the steps: projecting a three-dimensional point cloud into an aerial view, and generating a BEV image; encoding the image by using a pre-trained vision-language model to obtain BEV feature vectors, storing the BEV feature vectors in a vector database, and constructing a point cloud feature library; meanwhile, encoding a user query text to obtain a text feature vector; searching a matching scene from the feature library by calculating the cross-modal similarity between the text and the BEV feature vector; and finally, mapping a retrieval result to original point cloud data and outputting the original point cloud data. By means of the method, the feature alignment problem of the point cloud and the natural language is solved, and the retrieval method which directly and efficiently utilizes essential features of the point cloud data to be in seamless joint with the natural language is achieved.
Owner:MOLAR INTELLIGENCE INFORMATION TECHNOLOGY (HANGZHOU) CO LTD

Method and system for text search capability of live or recorded video content streamed over a distributed communication network

A server receives and rebroadcasts live streaming video content from a video capture device, such as a mobile phone or unmanned surveillance vehicle. The server includes a media server configured to stream selected video content to a client device, a video analysis system configured to analyze the live video content and generate object detection data, a storage system configured to store the generated object detection data and an identifier of the associated live video content, and a search engine configured to receive a text-based search request, search the object detection data stored in the storage system for relevant search results, and generate a list of live and stored video content associated with the relevant search results.
Owner:AERYON LABS

Urban rail vehicle electronic history data processing method

The present application provides a kind of urban rail vehicle electronic history data processing method, the method establishes vehicle operation configuration in combination with vehicle structure design, and vehicle configuration data structure is structured using database technology;Vehicle history data granularity refined to the smallest maintainable unit is used, the correctness and timeliness of data are guaranteed through data analysis and integration technology based on business process, the whole life cycle stage document of vehicle is stored using data storage technology, and preliminary vehicle knowledge framework is established in combination with full-text search technology, to provide basic data support for later subway vehicle procurement, technical reform, operation.In history data management, block chain technology is added to ensure the confidentiality and integrity of data flow, to prevent sensitive data from being leaked and tampered with during transmission, and to enhance the security and correctness of vehicle history data, to provide reliable data security guarantee for subsequent history sharing.
Owner:BEIJING MASS TRANSIT RAILWAY OPERATION CORPORATION LIMITED

Intelligent customer service question and answer processing method, server and storage medium

The invention belongs to the field of artificial intelligence, and particularly relates to an intelligent customer service question and answer processing method, a server and a storage medium. The intelligent customer service question and answer processing method comprises the following steps: acquiring consultation questions input by a user, and intelligently classifying the consultation questions to obtain category labels of the consultation questions; searching the consultation question through a vector search channel and a full-text search channel to obtain two search result sets; adjusting the initial weight based on the similarity score of each candidate answer in the two retrieval result sets to obtain a fusion weight of the consultation question in the two retrieval channels; calculating a comprehensive score of each candidate answer based on the fusion weight, and merging and sorting the candidate answers according to the comprehensive score to obtain a candidate answer set; and generating a final answer to the consultation question based on the candidate answer set. According to the method provided by the invention, the retrieval precision and recall ratio are improved, the sorting and correlation of retrieval results are optimized, and the solution rate of user questions is improved.
Owner:SHENZHEN POWEROAK NEWENER CO LTD

OA cloud service platform based on data sharing

The invention relates to an OA cloud service platform based on data sharing, and the platform comprises a data time sequence analysis module which is used for carrying out the seasonal decomposition and trend prediction of standardized historical sales data through employing a Prophet algorithm, and generating a sales prediction data set containing periodic features and a trend curve; the fusion module is used for receiving the structured product database and the sales prediction data set, performing feature alignment and dimension association with business data of an enterprise ERP system, and outputting a multi-source fusion data cube with a unified space-time label; and the spatio-temporal data warehouse architecture module is used for constructing a distributed data lake based on a Hadoop ecological system, implementing column type storage optimization on the multi-source fusion data cube, and integrating an Elasticsearch full-text retrieval engine to form a data warehouse supporting spatio-temporal dimension online analysis processing. According to the invention, the enterprise business process management efficiency, the data security and the decision intelligent level are improved.
Owner:TIANJIN HAIANDA NAVIGATION ENGINEERING SERVICE CO LTD

RAG document splitting optimization method and system in operator field

The invention discloses an RAG document splitting optimization method and system in the operator field, and belongs to the technical field of large model optimizing.The method comprises the steps that a document is uploaded, and source files and file information are stored through minio and paradeDB respectively; constructing a document loader, automatically selecting a corresponding loader according to a file type to carry out file processing, converting the file into a uniform Markdown text, analyzing a picture in the document, and converting the picture into a base64 format; constructing a picture processor, extracting a picture base64 character string in the text, processing a bitmap and a vector diagram, and converting the bitmap and the vector diagram into a Markdown picture reference format; constructing a document divider; vector conversion; and recalling the text. According to the method, the problem that picture extraction, conversion and processing are complex in the current RAG is solved, vector retrieval, full-text retrieval and mixed retrieval functions can be achieved through one database, and the operation and maintenance cost is reduced.
Owner:INSPUR TIANYUAN COMM INFORMATION SYST CO LTD

A data processing method, apparatus, device, and medium

This application provides a data processing method, apparatus, device, and medium. The method includes: acquiring a first text containing business text data; performing a risk assessment on the first text to obtain a risk category result corresponding to the first text; if the risk category result is a first risk category, acquiring keywords of the business text data and searching for the keywords in a standard database; if a request text matching the keywords is found in the standard database, determining the feedback text corresponding to the request text as the business processing result corresponding to the business text data; if no request text matching the keywords is found in the standard database, performing text search processing on the business text data in a target knowledge graph to obtain a business processing result matching the business text data. Implementing this application embodiment can improve the security of text data.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Text search method and system, electronic equipment, storage medium and product

The invention discloses a text search method and system, electronic equipment, a storage medium and a product, and relates to the field of large model technology and artificial intelligence. The method is applied to a server and comprises the steps that a first text sent by a client side and context information associated with the first text are received, the first text is used for representing a text selected in operation pages displayed on the client side, and the context information at least comprises page information of the operation pages and historical operation information of the operation pages; performing semantic analysis on the first text to obtain a semantic analysis result of the first text; searching based on the semantic analysis result and the context information to obtain a search result; and sending the search result to the client. According to the method and the device, the technical problem of relatively low accuracy of a search result during intelligent word segmentation search is solved.
Owner:HANGZHOU ALICLOUD FEITIAN INFORMATION TECH CO LTD

Sentence generation device, control method thereof, information processing system, and program

To provide a sentence generation device, a control method of the same, an information processing system, and a program which simply determines reliability of a generated reply sentence.SOLUTION: A sentence generation device includes: a quoted sentence retrieval unit which acquires a sentence related to a question sentence as a quoted sentence; an answer generation unit which generates an answer sentence to the question sentence by using the quoted sentence; a reliability calculation unit which calculates the reliability of the answer sentence on the basis of the quoted sentence and the answer sentence; and a hallucination determination unit which verifies consistency of the answer sentence and the quoted sentence, and calculates the reliability with the reliability calculation unit.SELECTED DRAWING: Figure 1
Owner:CANON DENSHI KK

Full-text retrieval method and system, computer equipment and computer readable storage medium

The invention belongs to the field of information retrieval, particularly relates to a full-text retrieval method and system, computer equipment and a computer readable storage medium, and aims to solve the problem of improving the full-text retrieval accuracy. The method comprises the following steps: segmenting a document entering a corpus; calculating paragraph weights of words contained in each segmented document in the segmented document; calculating the document weight of the word in the document according to the paragraph weight of the word; calculating the query weight of the query word, wherein the calculation method of the query weight is the same as the calculation method of the paragraph weight; determining a corresponding target word in a corpus according to the query word; respectively calculating one or more query relevancy according to the query weight and the document weight of the target word; and taking the document corresponding to the document weight corresponding to the maximum n query relevancy as a query result. Semantic features are introduced into weight calculation, and document segmentation processing is combined, so that the full-text retrieval accuracy is effectively improved.
Owner:TONGFANG KNOWLEDGE DIGITAL PUBLISHING TECH CO LTD +1

Multi-line text searching and replacing method and system

The invention relates to a multi-line text searching and replacing method and system, and the method comprises the steps: carrying out the preprocessing of an original multi-line text, a multi-line searching content and a multi-line replacing content, removing the head and tail spaces of each line, and obtaining a corresponding logic text, a logic searching content and a logic replacing content, judging whether the logic search content contains a preset wildcard character or not, if the logic search content does not contain the preset wildcard character, carrying out simple character string matching on the logic text to obtain a logic index search result, converting the logic index search result into an original character index by utilizing a preset index mapping relation for positioning, and executing simple replacement operation, and if the logic search content contains the preset wildcard character, carrying out simple character string matching on the logic text to obtain a logic index search result; and if the logic text is matched with the fuzzy item list, regarding the logic search content as an integral matching block, performing sliding comparison according to rows in a row array corresponding to the logic text, if the matching succeeds, extracting the fuzzy item list and the starting and stopping row indexes of the matching block, performing positioning by utilizing an index mapping relation, and executing fuzzy replacement operation of inserting the fuzzy item list into a placeholder.
Owner:SHENZHEN POWER SUPPLY BUREAU

Smart identification of indicator text with full-text search or optimized document analysis

Several aspects for optimizing unstructured document analysis comprise operating a document system, where the document system comprises a plurality of documents comprising unstructured content and a full-text index; receiving a request to identify documents comprising a type of data elements; selecting a sample out of the plurality of documents; determining data elements of the type in the sample of documents; determining an indicator context expression for the type of data elements out of the determined data elements of the type; determining a query for searching, using a search engine, the full-text index using the indicator context expression; and determining the documents in the document system being compliant to the determined query.
Owner:INTERNATIONAL BUSINESS MACHINE CORPORATION

Method and device for generating token by computing power acceleration of intelligent computing cloud platform

PendingCN122364446AComputing centerData set
This invention provides a method and apparatus for generating tokens using computing power acceleration on an intelligent computing cloud platform. It relates to the fields of intelligent computing centers, smart computing centers, computing infrastructure, and smart computing cloud technologies. The method for generating tokens using computing power acceleration on an intelligent computing cloud platform includes: Step S1, establishing a full-text index of a target data set; Step S2, responding to a received data query request, determining whether to enable document index acceleration; Step S3, if document index acceleration is enabled, extracting scalar filtering conditions and vector filtering conditions; Step S4, performing a full-text search based on the scalar filtering conditions to obtain a first data set; Step S5, performing a vector search on the vector index of the first data set based on the vector filtering conditions to obtain a matching second data set, and returning it to the user. This invention can improve the speed of vector database processing with scalar constraints on intelligent computing cloud platforms, reduce resource consumption, and increase token generation speed.
Owner:DATACANVAS LTD

An efficient multi-modal contrastive deep hashing retrieval method for medical big data

The application discloses a kind of efficient multi-modal contrast depth hash retrieval method for medical big data, it is related to artificial intelligence technical field, "search by image" function can let researcher search all similar image cases in database, not be influenced by previous expert diagnosis conclusion, provide possibility for further enrich and correct database.The "text search" function of cross-modal is more suitable for primary doctors or researchers.The method occupies less storage space, can realize fast search across modalities, by utilizing the potential correlation between medical report and its corresponding x-ray image, an efficient multi-modal medical data retrieval model is developed, the storage space is reduced, the efficiency of medical big data retrieval is improved, so that doctors can better learn, research and clinical diagnosis.
Owner:LIAONING UNIVERSITY OF TECHNOLOGY

Full-text search processor

[Problem to solve]To provide a hardware accelerator processor for full-text searches.[Solution]There is provided a full-text search processor, comprising: character storage elements for assigning and temporarily storing therein, search target text data to be searched through to a first address to an Nth address byte by byte; character detection circuits for receiving coded characters included in the search keyword byte by byte as comparison data, and sequentially detecting storage positions, on the character storage elements, of all of coded characters included in a search keyword; character string detection circuits for sequentially detecting positions, on the character storage elements, of coded characters which match a sequence of all of the coded characters included in the search keyword; and result output circuits for receiving search results of the character string detection circuits and outputting a position of the beginning or a position of the end of the character string that matches the search keyword.
Owner:INOUE KATSUMI

Intelligent search method and system and storage medium

The invention provides an intelligent search method and system and a storage medium, and relates to the technical field of information search and artificial intelligence. The method comprises the following steps: performing structured analysis processing on user query content to obtain a semantic feature vector corresponding to the user query content; for financial term semantic vectors contained in a preset financial term semantic library, determining the similarity between the semantic feature vector and the financial term semantic vectors in the financial term semantic library; according to the similarity, a target search mode used when the user query content is matched is determined in preset multi-mode search modes, and the multi-mode search modes comprise full-text search and / or pure semantic search; matching the user query content in a financial term semantic library by adopting a target search mode to obtain a target financial term corresponding to the user query content; and performing intelligent search on the target financial terms in the knowledge base to obtain and output a search result corresponding to the query content of the user. The method improves the accuracy of search results.
Owner:AGRICULTURAL BANK OF CHINA

Text search method and device, electronic equipment and storage medium

Embodiments of the present application provide a text search method and device, electronic equipment and storage medium, and relate to the technical field of retrieval. First, according to a plurality of word data, an index library and preset search information, all texts containing the preset search information are filtered to obtain at least one first text; then, the at least one first text is sorted to obtain at least one second text; and according to the category and content of the second text, label information corresponding to each second text is generated; finally, the identification and label information corresponding to each second text are taken as search results and displayed. According to the category and content of the filtered texts, the corresponding label information is generated, and the texts and the label information are taken as the search results for display, thereby saving storage space, and text search can be realized without large equipment.
Owner:BEIJING PIXEL SOFTWARE TECH

Information publishing method, device and equipment of cross-border matching system, and storage medium

The present specification relates to the technical field of computer network communication, and provides an information publishing method and device of a cross-border matching system, equipment and a storage medium. The method comprises the following steps: receiving supply and demand information; acquiring a search formula of each supply and demand information partition in the cross-border matching system; using each search formula to respectively perform full-text search on a partition keyword in the supply and demand information, and obtaining a corresponding partition keyword search result; calculating an evaluation score of each partition keyword search result; determining a supply and demand information partition to which the supply and demand information belongs according to the evaluation score; and publishing the supply and demand information to the supply and demand information partition to which the supply and demand information belongs. The embodiment of the present specification can improve the publishing efficiency and publishing accuracy of the supply and demand information in the cross-border matching system.
Owner:BANK OF CHINA

Book cover recognition method and device, electronic equipment and storage medium

The present disclosure provides a book cover recognition method and device, electronic equipment and storage medium, the method comprising: performing text recognition on a book cover image to be recognized to obtain text content contained in the book cover image; inputting the book cover image into a pre-trained press classification model to obtain a target press contained in the book cover image; based on the text content, the target press and the book cover image, performing feature extraction using a pre-trained feature extraction model to obtain a feature vector corresponding to the book cover image; performing text search based on the text content to obtain a text search result, and performing vector search based on the feature vector to obtain a vector search result; and determining target book information corresponding to the book cover image according to the text search result and the vector search result. The present scheme can search for more reliable vector search results, thereby improving the accuracy of book cover recognition.
Owner:深圳市星桐科技有限公司

A method for batch browsing and structured reading of DWG files based on lightweight design engine

The application discloses a kind of DWG file batch browsing and structured reading method based on lightweight design engine.The application constructs the complete process of "batch preprocessing-multithreading analysis-structured extraction-cross-platform browsing" by integrating open source DWG analysis technology and lightweight design architecture.The method realizes bottom DWG format analysis based on LibreDWG, generates lightweight intermediate files through redundant data elimination and geometry simplification, realizes efficient processing of batch files by combining multi-thread task scheduling;At the same time, structured data such as layer, entity geometry parameter and text attribute can be automatically extracted and standardized stored;Finally, cross-platform batch browsing of desktop and Web is supported, with interactive functions such as layer control, text search and measurement.The application is independent of commercial CAD software, and the batch processing efficiency is improved by more than 300% compared with traditional solutions, the lightweight file size is reduced by 60%-80%, and can be widely used in engineering design collaboration, BIM data integration, drawing automatic review and other scenes.
Owner:TIANHE INTELLIGENT MFG BEIJING TECH CO LTD

Article generation method and device, electronic equipment and storage medium

The invention discloses an article generation method and device, electronic equipment and a storage medium, and belongs to the technical field of content generation. The method comprises the steps of obtaining multiple document contents based on a specified article theme and / or keyword; generating an initial article based on the article theme, the keyword and the plurality of document contents by utilizing a first large language model; determining positions where illustrations need to be added in the initial article and illustration description information corresponding to the positions by utilizing a second large language model; searching the illustration in an image-text search material library based on the illustration description information to obtain a first search result, and determining a target illustration based on the first search result; adding a target illustration at the position of the initial article; and generating a target article based on the initial article added with the illustration. Therefore, technologies such as a large language model and image-text search can be combined, and the SEO article with high quality and rich illustrations can be quickly produced.
Owner:GUANGZHOU BOGUAN TELECOMM TECH LTD

A file label implementation method based on a distributed file system

This invention discloses a method for implementing file tags based on a distributed file system, comprising: building a distributed file system on a multi-node server; establishing a file dimension tag information group storage facility, building a database to store tag group data, and periodically synchronizing the data to a full-text search engine to improve the retrieval speed of tag group information; constructing tag library rules, creating corresponding custom tag features on the distributed file system management platform by writing relevant programs for custom tag dimension management, and realizing the definition and management of tag dimensions for distributed file system files; and building an automatic tag identification capability for distributed system files, automatically identifying tags for distributed file system files based on pre-set word segmentation rules by writing relevant code for file tag setting logic, thereby meeting users' needs for classifying, managing, retrieving, collecting, and sharing distributed file system files.
Owner:THE 28TH RES INST OF CHINA ELECTRONICS TECH GROUP CORP