Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

66 results about "Text database" patented technology

A text database is a system that maintains a (usually large) text collection and provides fast and accurate access to it. These two goals are relatively orthogonal, and both are critical to profit from the text collection. Traditional database technologies are not well suited to handle text databases.

Multi-modal agricultural question and answer method and system for generating RAG (Retrieval Enhanced Generation) based on retrieval

The invention discloses a multi-modal agricultural question and answer method and system for generating RAG based on retrieval enhancement, and the method comprises the steps: collecting and constructing crop disease image data containing a farmland complex background and corresponding text description, and forming an agricultural image-text knowledge database; based on the agricultural image-text knowledge database, multi-modal features of an input image and a query text are extracted, image-text joint similarity retrieval is carried out, and knowledge image-text candidates are obtained; inputting the knowledge image-text candidates into an image-text rearrangement module for rearrangement; taking the reordered image-text knowledge as condition input, accessing a large language model, and generating diagnosis description and prevention and treatment suggestions for the current crop diseases; and outputting a multi-modal question and answer result according to the image and question input by the user. According to the method, the agricultural image-text database is constructed, and multi-modal retrieval, rearrangement and large language model generation technologies are combined, so that accurate and efficient crop disease multi-modal question answering is realized, and the reliability of diagnosis and prevention suggestions is improved.
Owner:HEFEI INSTITUTE OF PHYSICAL SCIENCE CHINESE ACADEMY OF SCIENCES

Methods, systems, and storage media for constructing knowledge graphs in the civil aviation service sector

This invention discloses a method, system, and storage medium for constructing a knowledge graph in the civil aviation service field. The method includes: S1, using a BERT-BiLSTM-CRF algorithm model to extract entities and obtain interrelated entity vector sequences, feature vector sequences, and annotation sequences; S2, using a convolutional neural network model to extract sentence vectors and their contained entity vectors, and employing n filters to identify and extract a database of entity-relationship-entity triples; S3, using a conditional random field entity node integration model to integrate the annotation information and store it in the corresponding entities as entity attribute values; S4, using the triple database and the integrated entity attribute values ​​to link and fuse them to construct a civil aviation knowledge graph. This invention can obtain a comprehensive and accurate civil aviation knowledge graph of entity relationships based on a civil aviation knowledge text database. It can not only satisfy passenger knowledge question answering and querying needs but also serve as a training and educational resource, improving the overall service level.
Owner:CHINA ACAD OF CIVIL AVIATION SCI & TECH +1

Blue-green algae image recognition method and system based on hierarchical self-adaption and domain driving

The invention provides a blue-green algae image recognition method and system based on hierarchical self-adaption and domain driving, and the method comprises the steps: enhancing an image through employing an improved dark channel algorithm; constructing a blue-green algae biological attribute text database, and performing synonym replacement and sentence pattern recombination; multi-scale visual features are extracted through a hierarchical adaptive Swin Transform model, and key region characterization is enhanced in combination with dynamic spectrum attention; the text is input into a Bio-ALBERT model, and semantic embedding of field optimization is generated through term mask prediction and attribute relation pre-training; constructing a two-layer heterogeneous graph by using a graph attention interaction network GAIN, calculating a cross-modal association weight through a bidirectional graph attention mechanism, and outputting a cross-modal graph feature; multi-scale cross-modal association is modeled through a hierarchical graph attention fusion mechanism, and a comparison alignment loss optimization model is combined; and high-precision blue-green algae identification is realized. According to the method, through multi-scale perception, domain semantic adaptation and graph structure fusion, the accuracy of blue-green algae detection in a complex environment is improved.
Owner:ANHUI AGRICULTURAL UNIVERSITY

Knowledge question and answer method and device, processor and storage medium

The invention provides a knowledge question-answering method and device, a processor and a storage medium, and belongs to the field of information retrieval, and the method comprises the following steps: constructing a professional knowledge base for a specific field; wherein the professional knowledge base at least comprises a full-text database and a vector database; receiving a user query, and preprocessing the user query to obtain a standardized query; executing multi-path knowledge recall from the plurality of heterogeneous data sources in parallel based on the standardized query to generate a fusion recall result; wherein the plurality of heterogeneous data sources comprise but are not limited to a full-text database, a vector database, an internal retrieval API (Application Program Interface) plug-in and an internet retrieval interface; the historical interaction information, the fusion recall result and the standardized query are assembled into cue words through the instruction template; and inputting the cue word into the large language model to generate a reply text, and outputting the reply text in multiple modes. Through the method provided by the invention, the accuracy and coverage rate of knowledge recall can be improved, the answer is closer to real consultation, and the user experience is improved.
Owner:CHINA CONSTRUCTION BANK +1

Methods, devices, storage media, and processors for querying similar text

This application provides a method, apparatus, storage medium, and processor for querying similar text. The method includes: when the text type of the text to be queried is a first type of vocabulary, decomposing the text to be queried into multiple short words; determining a first target word vector for each short word; determining a second target word vector for the text to be queried based on the first target word vectors of all short words, thereby determining a first similarity between each word in the text database and the text to be queried; determining multiple first words in the text database that semantically match the text to be queried based on the first similarity, and a target similarity between each word in the text database and each first word; determining words in the text database whose target similarity is greater than a preset similarity threshold as target words matching the text to be queried. This greatly reduces the situation where text vectors cannot be recognized or have large errors, improving the success rate and accuracy of text recognition.
Owner:ZHONGKE YUNGU TECH

Audio evaluation method, computer device and storage medium

The application relates to an audio evaluation method, a computer device and a storage medium. By obtaining a plurality of emission probabilities corresponding to the audio to be evaluated, obtaining first transition probabilities between each template character in a template audio text corresponding to the audio to be evaluated, and second transition probabilities between each preset character in a preset audio text database, determining a first decoding path and a second decoding path based on the decoding of the emission probabilities, the first transition probabilities and the second transition probabilities, comparing a template decoding path corresponding to the template audio with the first decoding path and the second decoding path, and determining an evaluation result of the audio to be evaluated based on the first evaluation value and the second evaluation value determined thereby. Compared with the traditional multi-dimensional evaluation of audio, the scheme decodes two evaluation values based on the emission probabilities, the first transition probabilities and the second transition probabilities, and then determines the evaluation result, solves the misjudgment problem in multi-dimensional evaluation, and improves the accuracy of audio evaluation.
Owner:TENCENT MUSIC ENTERTAINMENT TECH (SHENZHEN) CO LTD

Civil aviation safety report topic modeling method based on large language model and semantic enhancement

The application discloses a kind of civil aviation safety report theme modeling methods based on large language model and semantic enhancement, its method includes: civil aviation safety theme semantic feature model extracts the text semantic vector and event structure vector of safety report text, and obtains the fusion feature vector of safety report text by fusion;Safety report text database is carried out cluster clustering processing and obtains initial cluster set and noise sample set;Select representative front p% as the representative sample set of cluster;Construct noise sample evaluation repair mechanism module, and select the candidate cluster to which noise sample belongs using noise sample evaluation repair mechanism module;Comprehensive gain function is constructed using representative sample set in cluster, and representative sample iteration screening processing of candidate sample is carried out in cluster.The application realizes the theme modeling goal of semantic accuracy, comprehensive coverage and stable result by multi-module collaborative innovation, and provides reliable technical support for management decision.
Owner:CHINA ACAD OF CIVIL AVIATION SCI & TECH

Multi-view clustering method and system based on self-paced learning and view weighting

The application discloses a multi-view clustering method and system based on self-step learning and view weighting, and belongs to the technical field of multi-view data processing. The method comprises the following steps: normalizing a multi-view data set, splicing the multi-view data, initializing each view clustering kernel and distribution matrix by using a kmeans algorithm, and calculating each view weight; each view sample weight matrix, each view clustering kernel and distribution matrix are iteratively updated in sequence through a target function; and when an iteration end condition is met, a final clustering kernel and distribution matrix are output. The clustering system comprises an acquisition module, a preprocessing module, a construction module, an optimization module and a clustering output module. The self-step learning model is used to sequentially learn clustering data and finally obtain a clustering result. Through view weighting, the model can selectively learn information of different views, thereby effectively improving clustering accuracy. The application can be applied to retrieval of an image database, a text database and the like.
Owner:INST OF ELECTRONICS & INFORMATION ENG OF UESTC IN GUANGDONG

Automatic international disease classification code coding method based on multi-synonym matching network

The invention provides an automatic international disease classification code coding method based on a multi-synonym matching network. The method comprises the following steps: S1, preparing a medical text database; s2, a main structure layer of a multi-synonym matching network algorithm model is constructed, a main calculation structure is designed, and the main structure layer of the multi-synonym matching network algorithm model comprises an encoder, a label encoder and a decoder; s3, training by using the prepared data set for training to obtain a model weight for reasoning, and evaluating a training result; s4, setting of various parameters of the multi-modal large language model is adjusted, the multi-modal large language model is designed to be combined with a multi-synonym matching network algorithm, the function of the multi-modal large language model is matched with a method target task, and a multi-modal ICD coding model based on a multi-synonym matching network is obtained; and S5, testing the multi-mode automatic international disease classification code coding model based on the multi-synonym matching network to enable the model to meet task requirements. According to the method disclosed by the invention, the ICD coding accuracy and convenience are effectively improved.
Owner:GUANGDONG BOHUA UHD INNOVATION CENT CO LTD

Data query method and related device

The invention discloses a data query method and a related device, and relates to the field of computers, and the method comprises the following steps: obtaining a query text, a database mode and a description text of the database mode; based on the query text, the database mode and the description text, calling a large language model, and generating an initial structured query language corresponding to the query text; calling a large language model, extracting a mode element from the initial structured query language to obtain a first mode element having an association relationship with the query text, and establishing a mapping relationship between a target mode element and the query text; based on the last structured query language and the mapping relation, calling a large language model, and generating a current structured query language corresponding to the query text; and under the condition that the current structured query language is the same as the last structured query language, performing data query based on the current structured query language to obtain query data indicated by the query text. The data query difficulty can be reduced.
Owner:AGRICULTURAL BANK OF CHINA

Caching of text analytics based on topic demand and memory constraints

An embodiment includes analyzing text content of a user query to identify via natural language processing (NLP) a query topic. The embodiment maps the query topic to a topic cluster at a node of a hierarchical model of a text database. The embodiment generates query demand data indicative of demand for the topic cluster based on user queries. The embodiment identifies the topic cluster as a topic-cache candidate based on the query demand data. The embodiment compares an amount of memory required for storing text associated with the first topic cluster to available cache memory. The embodiment caches the text of the topic cluster candidate upon determining that there is sufficient available cache memory space.
Owner:INTERNATIONAL BUSINESS MACHINE CORPORATION

Text processing method and apparatus, device, and medium

A text processing method includes the steps described below. Multiple texts in a preset text database are acquired, and at least one keyword is extracted from each text. Based on each keyword which is extracted, a keyword table is established. A mapping relationship between each keyword and a phrase element in a phrase set corresponding to a text in which each keyword is located is the same as a mapping relationship between the text in which each keyword is located and the phrase element in the corresponding phrase set. A keyword pair having an association relationship is determined in the keyword table. A phrase element in a phrase set corresponding to each keyword in the keyword pair are updated. According to a mapping relationship between each keyword and phrase elements in an updated phrase set, a relationship chart between a keyword and a phrase is established.
Owner:BEIJING YOUZHUJU NETWORK TECH CO LTD

Information recommendation method and device, equipment, storage medium and program product

The invention relates to an information recommendation method and device, equipment, a storage medium and a program product. According to one embodiment of the invention, the method comprises the following steps: in response to an information recommendation condition meeting a user at present, obtaining a target copywriting template from a pre-constructed recommended copywriting database, the target copywriting template being presented in a natural language form; pre-determined target recommendation content is filled in the target copywriting template to obtain target recommendation information, the target recommendation content is determined based on equipment operation habits of the user in a historical time period, and the target recommendation content comprises at least one of equipment identification, a room where equipment is located and equipment operation; and pushing the target recommendation information to the user. According to the information recommendation method and device, the user can quickly and clearly understand the equipment control content recommended based on the previous operation habits of the user, so that whether recommendation is accepted or not is better determined based on own requirements, the intelligence and individuation level of information recommendation can be improved, and the user experience is improved.
Owner:BEIJING XIAOMI MOBILE SOFTWARE CO LTD

Large model rapid question-answering method and system based on traffic engineering image-text database

The invention belongs to the technical field of model questioning and answering, and discloses a large-model rapid questioning and answering method and system based on a traffic engineering image-text database. Comprising the following steps: establishing a database, and clustering the database to construct a traffic engineering problem type set in each business field; obtaining a question language input by a user, and matching a user question type; analyzing the data in the question type of the user and a question language input by the user to generate a plurality of question answers, calculating a confidence score of each question answer, selecting the question answers based on the confidence scores, and recommending the question answers to the user; obtaining user feedback data according to the push result, analyzing the user feedback data, generating an update judgment coefficient, and executing an update instruction based on the update judgment coefficient; according to the method, whether updating is needed or not is judged according to the quantitative score of the feedback item type, the fluctuation coefficient and the updating judgment coefficient generated by the accumulation coefficient, so that the updating of the business domain module is more targeted, and resource waste caused by blind updating can be effectively avoided.
Owner:BEIJING TEXIDA TRAFFIC INVESTIGATION DESIGNING INST CO LTD

Blue-green algae image recognition method and system based on hierarchical self-adaptation and domain driving

The application provides a cyanobacteria image recognition method and system based on hierarchical adaptation and field driving. In the recognition method, an improved dark channel algorithm is used to enhance the image; a cyanobacteria biological attribute text database is constructed and synonym replacement and sentence restructuring are performed; multi-scale visual features are extracted through a hierarchical adaptive Swin Transformer model, and key region representation is enhanced in combination with dynamic spectrum attention; text is input into a Bio-ALBERT model, and field-optimized semantic embedding is generated through term mask prediction and attribute relationship pre-training; a two-layer heterogeneous graph is constructed using a graph attention interaction network (GAIN), cross-modal correlation weights are calculated through a bidirectional graph attention mechanism, and cross-modal graph features are output; multi-scale cross-modal correlations are modeled through a hierarchical graph attention fusion mechanism, the model is optimized in combination with a contrast alignment loss, and high-precision cyanobacteria recognition is achieved. Through multi-scale perception, field semantic adaptation and graph structure fusion, the application improves the accuracy of cyanobacteria detection in complex environments.
Owner:ANHUI AGRICULTURAL UNIVERSITY

Text mining-based ship fire risk factor identification method

The invention discloses a ship fire risk factor identification method based on text mining. The ship fire risk factor identification method comprises the following steps: 1) obtaining ship fire risk text data and establishing a text database; 2) performing word segmentation on the ship fire risk text data to obtain a ship fire risk word segmentation feature item list; 3) obtaining a ship fire risk keyword list; 4) extracting related phrases, and constructing a keyword related phrase set; 5) clustering the keyword related phrase set, and performing semantic analysis on a clustering result to obtain a ship fire risk factor list; 6) calculating the occurrence probability of the risk factors, and obtaining a ship fire risk factor probability table; according to the ship fire hazard risk factor risk assessment method, the importance of the risk factors and the fire hazard scene is quantitatively assessed through the text mining technology and the semantic analysis model.
Owner:WUHAN UNIV OF TECH

Method, device, equipment and product for automatically transcribing voice into HIS (hospital information system) of outpatient medical record

The invention discloses a method, a device, equipment and a product for automatically transcribing an HIS (Hospital Information System) through voice of an outpatient medical record, and relates to the technical field of automatic data entry. The method comprises the following steps: firstly, converting doctor inquiry voice into a text by applying a medical ASR technology, then carrying out entity recognition, term standardization and relation extraction through a medical NLP technology, generating structured JSON data with medical record field names as keys and entity information as values, and then fusing an Accession API and computer vision template matching technology to obtain a text database; according to the method, all the control capable of being filled in the target HIS interface is automatically detected, the mapping relation between the control and the medical record field name is established, finally, the corresponding entity information is extracted from the JSON data according to the mapping relation, and the corresponding control is automatically filled through the RPA or Computer Use technology, so that the medical term recognition accuracy can be remarkably improved, and the medical term recognition efficiency is improved. And seamless integration with a multi-brand HIS system is realized, full-automatic and cross-platform input of voice medical records is achieved, and practical application and popularization are facilitated.
Owner:CHENGDU ZIJIELIU TECH CO LTD

Auxiliary diagnosis result and reference prescription generation method and device, medium and computer equipment

This invention discloses a method and system for generating auxiliary diagnostic results and reference prescriptions. First, a medical record database and a treatment strategy database are constructed and stored in a vector database and a text retrieval database, respectively. For a target medical record, vectorization and textualization are performed. Then, candidate medical records and treatment strategies are retrieved from the medical record vector database, medical record text database, treatment strategy vector database, and treatment strategy text database via four channels. Subsequently, the RRF algorithm is used to fuse and sort the vector and text channels, and the candidate results are further refined based on a pre-trained semantic ranking model to obtain a small number of reference medical records and treatment strategies highly relevant to the current condition. Finally, prompt words are constructed by combining the target medical record, reference medical records, and treatment strategies to generate candidate results such as TCM diagnosis, TCM syndrome type, Western medicine diagnosis, and TCM prescriptions, serving as auxiliary references for doctors' clinical decision-making. This significantly improves the accuracy, interpretability, and efficiency of intelligent assisted diagnosis and treatment.
Owner:ZHEJIANG GUSHENG INTELLIGENT TECHNOLOGY CO LTD

Intelligent quantitative evaluation method and system for community fire-fighting toughness

The invention discloses an intelligent quantitative evaluation method and system for community fire-fighting toughness, and the method comprises the steps: collecting community fire accident cases, building a community fire accident text database in combination with a community fire-fighting toughness key stage, and recognizing community fire-fighting toughness related entities and mutual relationships; determining a community fire accident knowledge graph mode layer, establishing a knowledge graph, and concluding and extracting community fire-fighting toughness influence factors; a TOPSIS method is improved based on community fire-fighting toughness characteristics, key influence factors are screened, a community fire-fighting toughness evaluation index system is established, and a toughness network is constructed; and in combination with parameter learning and a Delphi method, establishing a Bayesian network evaluation model, carrying out causal reasoning and sensitivity analysis, quantitatively evaluating the community fire-fighting toughness level, identifying weak links and key influence indexes of the community fire-fighting toughness, and pointedly providing a promotion strategy. According to the method and the system, a practical tool is provided for a decision department to strengthen fire safety management and promote flexible sustainable community construction.
Owner:CHINA UNIV OF MINING & TECH (BEIJING)

Electric power inspection target detection false alarm elimination method and device, computer equipment, readable storage medium and program product

The invention relates to an electric power inspection target detection false alarm elimination method and device, computer equipment, a computer readable storage medium and a computer program product. The method comprises the following steps: inputting an electric power inspection image into an image multi-modal large model for visual feature extraction to obtain visual feature information of a suspected defect target; performing analysis based on the visual feature information to obtain visual attribute parameters of the suspected defect target; searching in a power inspection business text database by using the target power transmission line and the suspected defect target to obtain business text data of power equipment to which the suspected defect target belongs; wherein the business text data comprises first feature information and second feature information; processing is carried out based on the visual attribute parameters, the first feature information and the second feature information, and a feature matching result and a similarity comparison value are obtained; and false alarm detection elimination is carried out on the suspected defect target based on the processing result. By adopting the method, false alarms can be eliminated, and the accuracy and stability of a power system are guaranteed.
Owner:SHENZHEN POWER SUPPLY BUREAU

A large model hallucination mitigation method based on multi-agent cooperation and dynamic state tracking

This invention discloses a method for mitigating hallucinations in large-scale narratives based on multi-agent collaboration and dynamic state tracking. The method includes: generating an event graph composed of multiple atomic events using an outline agent; mapping these atomic events from an event occurrence time series to a text narrative time series using an orchestration agent; constructing a non-linear narrative outline; retrieving historical fragments related to the semantics and sentiment of the current chapter from the generated historical text; generating the text content of the current chapter using a writing agent; performing real-time logical consistency checks on the content during generation; triggering a correction mechanism when entity state changes in the generated content violate preset state transition rules; and parsing the entity state changes after the current chapter is finalized, updating the global state table and historical text database. This invention effectively mitigates hallucinations in long narratives, significantly improving the logical rigor and long-term coherence of the plot development while maintaining the literary quality of the text.
Owner:HANGZHOU JUNTONG FUTURE TECHNOLOGY CO LTD

An event-driven incremental knowledge extraction and fusion method and apparatus

This invention discloses an event-driven incremental knowledge extraction and fusion method and apparatus. The method includes: retrieving a set of factual text information from a text database; preprocessing the set of factual text information to obtain a set of factual text information; and performing knowledge fusion processing on the set of factual text information to obtain a set of factual knowledge ontology information.
Owner:CHINESE PEOPLES LIBERATION ARMY UNIT 61618

Intelligent responsibility judgment system and method for express customer service work order based on multi-modal large model

The invention discloses an express customer service work order intelligent responsibility judgment system and method based on a multi-modal large model, and relates to the technical field of artificial intelligence, and the system comprises a work order registration module which obtains multi-modal work order data based on an express business background system, and builds a structured complaint work order text database; the work order classification module is used for driving a Qwen-Plus large language model based on the structured complaint work order text database and executing complaint structured hierarchical classification to obtain an accurate complaint problem label; the work order responsibility judgment module is used for dynamically triggering ASR analysis, sign-in graph identification and transfer graph verification processes based on an accurate complaint problem label to obtain a multi-modal evidence analysis result; and the result output module fuses the complaint problem label and the multi-modal evidence analysis result, intelligently judges the responsibility branch in combination with a business rule engine, generates a responsibility branch responsibility judgment result, and pushes the responsibility branch responsibility judgment result to a related business system through an API interface to complete the whole-process intelligent responsibility judgment of express from work order input to responsibility judgment. The labor cost is reduced.
Owner:SHANGHAI YUANQING INFORMATION TECH CO LTD

Compliance risk avoidance methods, systems, devices, and media for generative artificial intelligence

This invention provides a method, system, device, and medium for compliance risk avoidance in generative artificial intelligence. The method includes: acquiring user input text; preprocessing the input text and calculating its compliance potential value, wherein the compliance potential value is positively correlated with the compliance relevance, discourse empowerment, and group influence of each character in the input text; based on the calculated compliance potential value, using a normalization algorithm to calculate an empirical threshold for the input text, and dividing the input text into three levels according to the empirical threshold; constructing a three-level text database; generating a final question-and-answer result by calling the three-level text database according to the level of the empirical threshold corresponding to the input text; and displaying the final question-and-answer result to the user. This addresses how to ensure that the generated content does not have content compliance issues while providing users with more authoritative, comprehensive, and reliable answers when applying generative artificial intelligence in areas involving compliance expression related to content compliance.
Owner:NORTH CHINA ELECTRIC POWER UNIV

Flight status recognition system and method

The patent discloses a flight state recognition system and method, and relates to the cross field of computer software and avionics system.The main technical scheme of the present application is as follows: constructing a text database;collecting pilot voice data and performing voice recognition;inputting the voice data;matching the input voice data with the constructed text database, calculating the matching degree, and recognizing the flight state of the airplane by using the matching degree.The pilot voice recognition flight state process can quickly infer the real flight state of the airplane, recognize the dangerous conditions that the airplane may encounter during the flight process, improve the response speed of the abnormal flight state recognition, and implement fault disposal according to the abnormal recognition result.
Owner:BEIJING AERONAUTIC SCI & TECH RES INST OF COMAC +1

Method for generating improved question-answer pair quality by using retrieval enhancement

According to the field of natural language processing annotation data acquisition and quality prompting in text travel, in the field of traditional natural language processing, when a question and answer scene contains complex questions with rich background information, the challenge that the quality of generated answers is insufficient is often faced, and the ability of a natural language model trained through question and answer pairs in a special field is limited; the invention provides a method for generating improved question and answer pair quality by utilizing retrieval enhancement, which comprises the following steps of: selecting a language model through a multi-dimensional evaluation mechanism, acquiring text data related to a theme by adopting a mode of combining a web crawler and field investigation, cleaning and sorting the data, constructing a comprehensive text database, screening seed question and answer pairs, and obtaining the quality of the improved question and answer pairs. It is ensured that the generated question and answer pairs are sufficient in number and credible in content, and the question and answer pairs are further optimized through expert feedback and annotation so as to provide more accurate and high-quality answers; according to the method, the response efficiency and the answer accuracy of the question answering system in processing complex questions are remarkably improved.
Owner:TAIYUAN UNIVERSITY OF TECHNOLOGY +1

Systems and methods for automating conversion of drawings to indoor maps and plans

Automating conversion of drawings to indoor maps and plans. One example is a computer-implemented method comprising: preprocessing a CAD drawing to create a text database containing text from the CAD drawing and associations of the text with locations within the CAD drawing; determining a floor depicted in the CAD drawing, the determining results in a floor-level outline; identifying a plurality of room-level outlines within the floor-level outline, the plurality of room-level outlines corresponds to a respective plurality of rooms; selecting a name of a first room from the plurality of rooms, the selecting based on text within the text database; and creating an indoor map including the name of the first room, the name of the first room associated with a location of the first room within the floor-level outline.
Owner:POINTR LTD

Document duplicate checking method based on multi-index mechanism, storage medium and equipment

The invention discloses a document duplicate checking method based on a multi-index mechanism. The document duplicate checking method comprises the following steps: storing an original file of a document into a file database MinIO; extracting structured text information of the document according to the identified document type, cutting the structured text information into text blocks, storing the text blocks into a text database Elasticsearch, and allocating a text block index; generating a semantic vector code of each text block, storing the semantic vector code in a vector database Milvus, and distributing a unique identifier for each text block; storing the metadata information of the document and the text block index and the unique identifier corresponding to the document into a relational database Mysql; the method comprises the following steps: performing full-text duplicate checking on a document to be subjected to duplicate checking by utilizing a text database Elasticsearch to obtain a most similar matched document, calling a relational database Mysql, searching a unique identifier corresponding to the matched document, and performing semantic vector coding similarity duplicate checking in a vector database Milvus. According to the method, more accurate duplicate checking evaluation and feedback of semantic hierarchy and structural hierarchy can be provided.
Owner:CHINA TELECOM DIGITAL INTELLIGENCE TECH CO LTD