Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

384results about "Metadata multimedia retrieval" patented technology

Personalized teaching content generation method and system based on digital portraits

The invention discloses a personalized teaching content generation method and system based on a digital portrait, and the method comprises the steps: collecting the multi-dimensional learning data of a student in real time, analyzing the data keyword of the student, constructing the portrait of the student, generating a teaching target based on the portrait of the student, and enabling the teaching target to comprise a teaching theme, exercises related to the theme, and knowledge points related to the theme. Converting the teaching target into an executable instruction of a large language model through a structured prompt engineering technology; the large language model generates initial teaching content according to the teaching instruction; verifying and optimizing the generated initial teaching content to ensure that the generated teaching content conforms to a teaching target and can adapt to the current cognitive level and learning requirements of students; and outputting the verified and optimized teaching content to the students and automatically adapting to the content presentation form according to the learning style preference of the students.
Owner:SHAANXI NORMAL UNIV +1

Meeting content intelligent generation processing method and system based on multi-modal large model

The invention discloses a conference content intelligent generation processing method and system based on a multi-modal large model, and the method comprises the steps: collecting the original data of a conference, and completing the standardization preprocessing; inputting a multi-modal large model, extracting multi-modal features and carrying out semantic alignment; executing cross-modal hash coding, generating binary codes and establishing an index database; performing hash retrieval on the related fragments, and constructing a conference content directed graph; based on a conference content directed graph structure, searching an optimized path by adopting a Monte Carlo tree; and generating structured conference content, and outputting a summary, an abstract and an action item. According to the method, efficient extraction, accurate retrieval and structured intelligent generation of the conference content are realized by fusing a multi-modal large model, cross-modal Hash coding and Monte Carlo tree search.
Owner:NANJING WEITEXI NETWORK SCI & TECH

Photo content clustering for digital picture frame display and automated frame storytelling

A method and system for automated routing of pictures taken on mobile electronic devices to a digital picture frame including a camera, microphone, and speaker integrated with the frame, and a network connection module allowing the frame for direct contact and upload of photos from electronic devices or from photo collections of community members. Clustering photos by content is used to improve display and to respond to photo viewer desires. Trends or patterns can be detected from the photo collections and that information used for various purposes beyond photo display. The frame includes a conversational intelligence that provides a verbal communication with a viewer, such as for determining an identity or preferences of the frame viewer, determining photos to display for the viewer, discussing displayed photos with the viewer, or telling stories or life histories to the viewer based upon photo content.
Owner:PUSHD INC

Electronic government enterprise service intelligent response method based on AI cue word engineering

The invention relates to the crossing field of artificial intelligence technology and e-government, and discloses an e-government enterprise service intelligent response method based on AI cue word engineering, comprising: constructing and maintaining a government knowledge graph and a scenarized cue word template library; receiving and analyzing a service request input by an enterprise, and extracting an enterprise feature tag and a business intention; selecting a target cue word template based on the matching degree of the enterprise feature tag and the scene tag of the template in the scene cue word template library, and generating an adapted cue word template; retrieving associated knowledge data from the government affair knowledge graph according to the service request, and filling the adapted cue word template with the retrieved associated knowledge data to generate a structured cue word; inputting the structured cue word into an AI model to drive the AI model to generate a government affair service response; and collecting multi-dimensional evaluation feedback of government affair service response, and carrying out iterative optimization. The demand understanding precision is improved, and the risk of core information misjudgment is reduced.
Owner:SICHUAN ENRISING INFORMATION TECH CO LTD

Tobacco marketing hotspot event real-time analysis method based on big data and AI

The invention provides a tobacco marketing hotspot event real-time analysis method based on big data and AI, and relates to the technical field of big data and artificial intelligence, and the method comprises the steps: collecting multi-source heterogeneous data through a distributed crawler, and achieving the structured processing of unstructured data through the semantic analysis and multi-modal fusion technology; hot event identification and early warning are carried out by combining deep learning and a propagation dynamics model; further fusing the knowledge graph, NLP and space-time analysis to generate brand specification popularity ranking and trend prediction; and finally, an evaluation model is constructed based on historical and real-time data, new product research and development, brand promotion and supply chain optimization strategies are output, and full-process intelligent decision support is realized.
Owner:SHANDONG INSPUR DIGITAL BUSINESS TECHNOLOGY CO LTD

Intelligent traffic accident liability affirmation method and system based on multi-agent cooperation mechanism

The invention relates to the technical field of artificial intelligence and intelligent traffic, in particular to a traffic accident liability intelligent affirmation method and system based on a multi-agent cooperation mechanism and a large language model. The whole process of credible information verification, responsibility affirmation reasoning and standard document generation is simulated. Wherein the legal expert agent adopts a competing mechanism, and the fairness and the accuracy of affirmation are improved through preliminary affirmation, double-party defense and final judgment. The reasoning ability of a large language model and accurate knowledge retrieval of a retrieval enhancement generation technology are fused, the real-time performance and accuracy of legal clause quotation are ensured, and a road traffic accident identification document conforming to specifications is automatically generated. The method effectively solves the problems that traditional manual identification is low in efficiency and high in subjectivity, and an existing intelligent method lacks interpretability and legal accuracy.
Owner:SICHUAN POLICE COLLEGE +1

Advertisement creativity matching method based on multi-modal content generation

The invention discloses an advertisement creativity matching method based on multi-modal content generation, and relates to the technical field of digital media content generation, and the method comprises the following steps: building a cross-modal time anchoring belt facing advertisement creativity matching, carrying out metaphor level decomposition on input text information, marking a symbol axis for image information, and carrying out data processing on the image information; obtaining an initial semantic boundary list; and constructing a culture fingerprint database according to the initial semantic boundary list, and mapping the territory taboo information and the brand symbol information into constraint tags to obtain a semantic guardrail set. According to the method, through cross-modal time anchoring and semantic boundary control, accurate correspondence of the text and the image in time and semantic levels is achieved, and it is ensured that generated content is clear in semantic meaning and adaptive in culture. In combination with breathing type phase traction and cultural fingerprint dynamic adjustment, multi-modal content rhythm and emotion are coordinated and unified, brand expression is kept stable, and the overall consistency and propagation effect of advertisement creativity are improved.
Owner:大根控股股份有限公司

Multi-modal data retrieval, generation and synthesis method and system based on artificial intelligence driving

The invention relates to the technical field of artificial intelligence, in particular to a multi-modal data retrieval, generation and synthesis method and system based on artificial intelligence driving, and the method comprises the steps of multi-modal feature extraction, cross-modal alignment, feature fusion, multi-modal retrieval and sorting and multi-modal generation. Different deep learning models are adopted to carry out feature extraction on multi-modal data and convert the multi-modal data into vectors, through cross-modal alignment, different modal feature vectors are mapped to the same vector space, through feature fusion, correlation weights are calculated through an attention mechanism, weighted summation is carried out on fusion features, and through multi-modal retrieval and sorting, multi-modal data are obtained. According to the method, the similarity between a query vector and candidate data is calculated and sorted, and finally, through multi-modal generation, retrieval knowledge is used as external knowledge to be fused into a generation model, and a multi-modal synthesis answer is generated, so that information in multi-modal data can be effectively integrated and utilized.
Owner:BEIJING SGITG ACCENTURE INFORMATION TECH CO LTD +1

Task processing method and intelligent device with body

The invention discloses a task processing method and an intelligent device, and relates to the technical field of intelligent devices, and the method comprises the steps: obtaining multi-modal data which comprises task information of a to-be-processed task in a real-time interaction scene of a physical entity and an environment; performing feature extraction and feature alignment processing on different modal data in the multi-modal data to obtain a first feature after feature alignment; performing cross-modal data retrieval on a knowledge base based on the first feature to obtain target retrieval data; performing feature fusion on different modal data features in the first features to obtain second features; and making a decision based on the target retrieval data and the second feature by using a decision model to obtain an action instruction sequence, and executing the to-be-processed task based on the action instruction sequence.
Owner:LENOVO (BEIJING) LTD

Learning resource recommendation method and system based on knowledge tracking and retrieval enhancement generation

The invention provides a learning resource recommendation method based on knowledge tracking and retrieval enhancement generation, and belongs to the technical field of education. The method comprises the steps of obtaining data of a user learning platform and / or learning content uploaded by a user, and constructing a background knowledge base related to a current learning task of the user; acquiring learning interaction behavior data of the user, constructing a knowledge state tracking model, and forming a current knowledge mastering state of the user; wherein the learning interaction behavior data of the user comprises knowledge points, exercises, videos and texts; and outputting personalized learning resource recommendation and question and answer response by utilizing a retrieval enhancement generation model in combination with the background knowledge base and the knowledge mastering state. Therefore, dynamic, personalized and knowledge-accurate learning content recommendation can be realized.
Owner:GUANGZHOU PANYU POLYTECHNIC

Dynamic multimedia data hash retrieval method and system based on extensible increment

The invention discloses a dynamic multimedia data hash retrieval method and system based on extensible increment, and relates to the technical field of multimedia data retrieval. The method comprises the following steps: acquiring dynamic multimedia data to be retrieved; a pre-trained Hash retrieval large model is constructed, the Hash retrieval large model is trained by taking a bit extensible Hash center as global supervision information and taking tag cosine similarity as local supervision information, and specifically, generalization feature representation of new multimedia data maintaining new and old class discrimination is obtained through forward propagation; constructing a linear mapping relationship between the generalization feature representation and the hash code, and introducing an auxiliary variable to continuously update the hash function without playback; and utilizing the trained Hash retrieval large model to generate a query Hash code for the dynamic multimedia data to be retrieved, and utilizing the query Hash code to retrieve. According to the method, low memory occupation, high updating efficiency and non-forgetting retrieval of the dynamic multimedia data stream in the open environment are realized.
Owner:SHANDONG JIANZHU UNIV

Method and apparatus for the collection and management of quantitative data on unusual aerial phenomena via a citizen network of personal devices

The present invention relates to a method and apparatus to quantify unusual aerial phenomena via a network of citizen operated personal devices. A feature of the present invention is an app that runs on popular personal devices. A further feature of the invention is a server system in communication with said app. A further feature of the invention is a means by which a user having spotted a potential event can quickly engage the app to begin a data collection mode. A further feature of the invention is a data collection mode that simultaneously records data including but not limited to video, audio, 9-axis IMU data, GPS coordinates, and time. A further feature of the invention is a method of tagging the collected data with a cryptographic signature 800 to ensure integrity. A further feature of the invention is real-time app communication to a centralized server. A further feature of the invention is an alert to other users indicating something of interest is happening nearby. A further feature of the invention is the a means by which the server can distinguish interesting events from non-interesting events. A further feature of the inventions is a means by which the server can analyze the data to obtain scientifically useful quantitative information.
Owner:RANDALL MITCH

Keyword filtering for digital content recommendation

Methods, systems, and apparatus, including computer-readable storage media, for keyword list filtering as part of identifying digital content responsive or relevant to a search query or request for content. A user, such as a content provider, may generate a keyword list associated with digital content of the content provider. Keyword lists, however, may be built over the course of years and can grow to include millions of keywords. Further, these keyword lists are often not maintained in line with changes in a content provider's digital content delivery strategy or context. An artificial intelligence (AI) model may be trained to generate a summary of the digital content associated with the content provider. That summary, along with the keyword list of the content provider, is provided as input into the AI model, which is trained to provide, as output, a recommendation to keep or remove a keyword from the keyword list.
Owner:GOOGLE LLC

Broadcast content real-time analysis interaction method and system based on AI multi-mode large model

The invention relates to the technical field of AI intelligent calculation, in particular to a broadcast content real-time analysis interaction method and system based on an AI multi-modal large model, and the method comprises the steps: achieving the precise event recognition and structural description through the efficient modeling of single-modal real-time data, and providing the standardized input for the subsequent content generation; then, broadcast texts meeting object features and scene requirements are generated based on event information and global static audience feature configuration, and quantified semantic priorities are calculated to support strategy decision; on the basis, performing deep intention recognition on the generated content by utilizing multi-modal large model reasoning, introducing targeted parameters such as historical broadcast diversity punishment and interference sensitivity adjustment, and generating an executable scheduling strategy; and finally, through area matching and terminal capability verification, mapping the strategy parameter into a specific equipment control instruction, and issuing and executing the specific equipment control instruction to form a closed loop from content generation to precise broadcasting.
Owner:GUANGZHOU ANSPER TECH CO LTD

Methods and apparatus to credit media presentations for online media distributions

Methods and apparatus to credit media presentations for online media distributions are disclosed. Example methods and apparatus determine a presenter of a media session based on a user agent identifier extracted from a proxy record associated with the media session, and in response to determining that the user agent identifier does not identify a publisher, identify a first domain referenced by a URL associated with the media session, and, in response to determining that the first domain matches a domain pattern associated with a publisher and does not match a domain in a list of hosting domains, classify the media of the media session as being published by the publisher associated with the matching domain pattern, the publisher being different from the presenter.
Owner:THE NIELSEN CO (US) LLC

Method and System for Real-Time Collaboration, Task Linking, and Code Design and Maintenance in Software Development

A method for automated document processing and task assignment including receiving an input document from a document source, performing a content analysis on the input document, extracting metadata from the input document, identifying an identified task type to be performed, generating standardized metadata by converting the metadata into a standardized JSON format, storing the standardized metadata in a database, determining a user assignment for the identified task type, the including an assigned user, generating an action item including the identified task type, the user assignment, and the standardized metadata, and adding the action item to a task management system for processing by the assigned user.
Owner:MADISETTI VIJAY

Visual compression and retrieval method and device of document, equipment and storage medium

The invention discloses a document visual compression and retrieval method and device, equipment and a storage medium, and relates to the technical field of computers, the method comprises the following steps: obtaining a to-be-processed document page image, segmenting the image into a plurality of image blocks, and determining the structure category and the structure importance score of each image block; obtaining a plurality of structure regions based on structure category and spatial position aggregation, and distributing a preset number of compression tokens for each region in combination with structure category weights and importance scores; generating structure anchor point tokens corresponding to the regions by the compressed tokens to form a set; receiving a query request, converting the query request into a query vector, and performing retrieval in the anchor point token set to obtain a target structure region; and performing local decoding reconstruction based on the compressed token of the target region, and outputting a region image or a structure mask. According to the method, through structure-guided self-adaptive compression and fine-grained retrieval, the long document processing efficiency is greatly improved, and the compression effect and the retrieval accuracy are both considered.
Owner:BEIJING DIGITAL CHINA CLOUD COMPUTING CO LTD

Teaching courseware generation method, device and system and computer storage medium

The invention provides a teaching courseware generation method and device, and relates to the technical fields of document processing, artificial intelligence, natural language processing and the like. According to the specific scheme, the method comprises the following steps: generating a teaching design general outline comprising a unit identifier of a teaching material unit and module information of at least two teaching modules based on teaching elements of the teaching material unit input by a user; based on the teaching design general outline, teaching resources are generated for each teaching module; based on the teaching design general outline, detecting whether the learning targets of the teaching resources among different teaching modules are consistent and / or whether the teaching resources of each teaching module deviate from respective teaching targets; in response to detection that the learning targets of the teaching resources are inconsistent and / or deviate from the teaching targets, regenerating the teaching resources of the corresponding teaching modules, and continuing detection until the learning targets of the teaching resources of different teaching modules are consistent and the teaching resources of all the teaching modules do not deviate from the respective teaching targets; and generating a target teaching courseware of the teaching material unit based on the teaching resources.
Owner:HUA CHUAN INTERNATIONAL HOLDINGS GROUP CO LTD

Systems and methods for dynamic media asset modification

The present disclosure provides systems and methods for transforming media assets using data retrieved from external sources. A system can identify a request to update one or more media assets maintained in a database of a media asset system. The system can retrieve, from a remote data system identified in the request, data corresponding to object metadata of each media asset of the one or more media assets. The system can generate, for each media asset of the one or more media assets, an updated media asset to include the data retrieved from the remote data system. The system can modify the object metadata of each of the one or more media assets based on the data. The system can update, responsive to the request, the database with each updated media asset. The updated media assets can be transmitted to client devices for display in information resources.
Owner:ZARTECH HOLDINGS LLC

Systems and methods for calculating a predicted time when a user will be exposed to a spoiler of a media asset

Methods and systems are described herein for calculating a predicted time when a first user will be exposed to a spoiler of a first media asset. A media guidance application may retrieve an initial transmission time of the first media asset and a transmission time of a second media asset that precedes the transmission time of the first media asset. The media guidance application may retrieve electronic communications from a second user whose electronic communications the first user has viewed, and determine a subset of the electronic communications that correspond to the second media asset. The media guidance application may calculate a length of time from the transmission time of the second media asset to the earliest availability time of an electronic communication in the subset. The media guidance application may calculate the predicted time based on the transmission time of the first media asset and the length of time.
Owner:ADEIA GUIDES INC

System and Method for Automated Integration of Contextual Information with a Series of Digital Images Displayed in a Display Space

A user interface application displays, in the user interface application, an image, or the portion thereof, in a display space. While the user interface application continues to display the image, or portion thereof, a messaging platform application searches in one or more digital data sources for, and retrieves, contextual information based on the displayed image, or portion thereof, without receiving user input to request searching in the one or more digital data sources for contextual information based on the displayed image, or portion thereof. The messaging platform application detects, one or more user interactions with one or more of the user interface application, the display space, or the image or the portion thereof, and displays a portion of the retrieved contextual information as related digital data content in a location within a field of view of the display space, based in part on the detected one or more user interactions.
Owner:TECTONIQ INC

Intellectual property big data fusion analysis and service platform

The invention relates to the technical field of computers, and discloses an intellectual property big data fusion analysis and service platform which comprises a data acquisition preprocessing module, a multi-modal feature coding module, a knowledge graph construction and reasoning module, a multi-modal fusion and correlation analysis module and an intelligent service and application module. According to the technical scheme, retrieval accuracy and efficiency can be improved, traditional retrieval limitation is overcome through deep semantic understanding, association discovery and multi-dimensional intelligent analysis, intellectual property work efficiency is greatly improved, and the method is particularly suitable for complex technical patents.
Owner:JIANGSU BAITENG TECH CO LTD

Intelligent exhibition hall multi-mode interactive digital human system and implementation method

The invention relates to the technical field of computer data processing, and discloses an intelligent exhibition hall multi-modal interaction digital human system and an implementation method, which are used for solving the problem that multi-source data of multi-modal interaction lacks a unified clock and verifiable timestamp alignment mechanism in a traditional method. According to the method, a unified time domain and a time version are established on an edge side, terminal access is restrained, and an acquisition time mark, an access time mark and a serial number are written in data or a state; performing gating shunting according to a time version, performing de-duplication and out-of-order rearrangement based on a serial number and double time marks, and generating and solidifying a session window evidence index; locking a transaction window boundary according to the evidence index, generating a participation source list and a transaction number, establishing a fragment reference relationship and generating an alignment voucher; in the linkage stage, phase division issuing is carried out according to a preparation phase, an execution phase and a confirmation phase, backward reading verification is carried out according to an action sequence number, a transaction log is archived, and alignment, rechecking and playback of an interaction link are achieved.
Owner:SUZHOU CHUANGJIE MEDIA EXHIBITION CO LTD

Simmary generation method and related device

The invention discloses a summary generation method and a related device, and relates to the technical field of artificial intelligence, and the method comprises the steps: obtaining multi-source information of a target conference, obtaining a summary generation prompt instruction according to the multi-source information, the summary generation prompt instruction is used for instructing a summary generation large model to generate a summary according to the content of the multi-source information, and generating a structured conference summary by referring to the hierarchical structure in the conference presentation, and inputting a summary generation prompt instruction into a pre-trained summary generation large model to obtain the conference summary of the target conference. According to the method, the hierarchical structure of the conference summary can be reconstructed based on the hierarchical structure of the conference presentation, so that the content of the multi-source information can be presented as a more logical expression in the conference summary, and the readability of the conference summary is improved.
Owner:HKUST IFLYTEK (SHANGHAI) TECH CO LTD

Methods for personalized search and recommendation on smart TVs

This invention relates to a personalized search and recommendation method for smart TVs, specifically in the field of film and television recommendation. It utilizes a deep neural network model to generate embedding feature vectors for each film and television program in a media resource library. User profiles are obtained based on MAC addresses or voiceprint IDs, and embedding feature vectors for user-preferred films and television programs are generated using the same deep neural network model. A film and television knowledge graph is constructed, with the relationships between knowledge graph nodes serving as the reasoning for recommendations to the user. The method obtains the first- and second-degree neighbor film and television IDs for any given film and television program. When a user inputs a search term, the first- and second-degree neighbor film and television IDs are obtained using the corresponding film and television ID. The distance between the embedding feature vector of each film and television ID in the first- and second-degree neighbor programs and the embedding feature vector of the user-preferred films and television programs is calculated. These distances are then sorted from smallest to largest, and the top N films and television programs are selected for recommendation to the user. This invention solves the problem of weak correlation between search recommendation results and specific user preferences in existing technologies. This invention is applicable to personalized film and television recommendation.
Owner:SICHUAN CHANGHONG ELECTRIC CO LTD

Audio and video retrieval method and device, electronic equipment and storage medium

This invention provides an audio / video retrieval method, apparatus, electronic device, and storage medium. The method includes: obtaining current search conditions; retrieving each secondary data body based on a primary search table; each secondary data body corresponds to a secondary search table, and each secondary data body includes at least one data block; each data block includes at least one audio / video data, vehicle information corresponding to each audio / video data, and statistical information corresponding to the vehicle information; retrieving statistical information from each data block in the corresponding secondary data body based on each secondary search table; and when target statistical information that satisfies the current search conditions is retrieved, and it is determined that the target data block corresponding to the target statistical information contains target vehicle information corresponding to the current search conditions, the audio / video data corresponding to the target vehicle information is determined as the target audio / video data corresponding to the current search conditions. This invention eliminates the need to traverse all audio / video data, reducing retrieval time and improving retrieval speed.
Owner:HANGZHOU HOPECHART

System and method for automated integration of contextual information with a series of digital images displayed in a display space

A user interface application displays a series of digital images in a display space. A user interface receives input to select one of the series of digital images, or portion thereof. The user-selected digital image, or portion thereof, is displayed in a location with the field of view of the displayed series of digital images or display space. While the user interface application continues to display the series of digital images in the display space, an application, such as a chatbot, searches one or more digital data sources for, and retrieves, contextual information based on the displayed user-selected digital images or portion thereof, without receiving user input to perform the searching.
Owner:TECTONIQ INC

Methods, devices, equipment and storage media for searching multimedia data

This application provides a method, apparatus, device, and computer-readable storage medium for searching multimedia data. The method includes: based on an acquired search request, determining the multimedia data and text data to be processed, and acquiring a trained cross-modal representation model, wherein the trained cross-modal representation model is obtained through two training phases using multimedia training data, the identifier information of the multimedia training data, and the search text corresponding to the multimedia training data; inputting the multimedia data and the text data into the trained cross-modal representation model to obtain the semantic similarity of each element of the multimedia data and the text data; and determining and outputting search results based on the semantic similarity and each element of the multimedia data. This application can improve the relevance between the search results and the search content itself.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Digital twinborn-based academic publishing knowledge service intelligent management system and method

The invention discloses an academic publication knowledge service intelligent management system and method based on digital twinning, and relates to the field of knowledge service and technology, and the system comprises a data storage module, a dimension weight quantification module, a weight distribution difference analysis module, an effective trigger distribution model construction module and a response priority analysis module. According to the method, the initial weight of each dimension is quantified through six dimensions including multi-dimensional intelligent retrieval, covering knowledge elements, semantic understanding and user portraits, the weight can be dynamically allocated according to user retrieval content features such as whether keywords contain limits or not, new and old attributes of the users and keyword expansion capacity, the weight is updated based on subjective operation contents of the users, and the user experience is improved. The retrieval strategy is deeply matched with the individual requirements of the user, the problem of low retrieval efficiency is greatly reduced, and the operation cost of repeatedly modifying retrieval conditions by the user is reduced.
Owner:QINGDAO UNIV

Media content memory retrieval

Aspects of the subject disclosure may include, for example, a media consumption database that stores data elements describing conditions under which electronic media content is consumed by a user on an electronic device. A search of the media consumption database based on at least a portion of the conditions may result in at least a portion of the electronic media content to be re-presented to an electronic device of the user Other embodiments are disclosed.
Owner:AT&T INTELLECTUAL PROPERTY I L P