Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

86 results about "Object description" patented technology

Image object detection method, system and apparatus, and storage medium

Embodiments of the present description provide an image object detection method. The method comprises: on the basis of an image to be retrieved, an object description text, and an object retrieval condition, determining, by means of an object detection model, a target position of an object to be retrieved in said image, wherein the object description text is used for describing said object, and the object retrieval condition comprises at least one of a mask image, a pose, and a texture corresponding to said object.
Owner:ZHEJIANG DAHUA TECH CO LTD

Mandatory access control method and device based on process function context

The invention discloses a mandatory access control method and device based on a process function context, and the method comprises the steps: collecting a security context associated with a system call initiated by a target process, so as to generate a standardized object description; mapping the object description into a target function classification identifier, so as to obtain a process function context view of the target process according to the target function classification identifier; constructing a target decision key for access decision based on the current policy era, the qualifier, the function classification identifier, the view identifier of the process function context view and the isolation domain abstract; and querying the multi-level cache according to the target decision key to determine a matched target access decision. Therefore, context-sensitive judgment and cross-component consistency taking the functional context as the center are realized.
Owner:BEIJING METRO INFORMATION DEV CO LTD

Service processing

Object description information that is transmitted by a biometric recognition apparatus is received, the object description information includes a biometric feature of a target object and location information of the target object. Identity information of the target object is obtained according to the biometric feature. A service information set in association with the identity information is obtained. From the service information set, one or more pieces of candidate service information are selected. The one or more pieces of candidate service information are transmitted to a terminal device associated with the identity information. At least a first piece of target service information returned by the terminal device is received. At least a first service corresponding to the first piece of target service information is processed. Apparatus and non-transitory computer-readable storage medium counterpart embodiments are also contemplated.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Upside down reinforcement learning for text-to-image generation

A method, apparatus, non-transitory computer readable medium, and system for image processing include obtaining an input prompt including an image quality level and a description of an object, generating an image embedding based on the input prompt, where the image embedding represents the object and the image quality level in a vector space, and generating a synthetic image based on the image embedding, where the synthetic image depicts the object and has the image quality level.
Owner:ADOBE INC

Industrial robot self-adaptive grabbing method and system based on multi-mode perception

The invention provides an industrial robot self-adaptive grabbing method based on multi-modal sensing, which comprises the following steps of: 1, acquiring multi-modal data of an object and an environment through a multi-modal sensing module; the multi-mode sensing module comprises at least two sensing modules of a visual sensing unit, a touch sensing unit, a force sensing unit and an auditory sensing unit; 2, preprocessing and feature extraction are carried out on the multi-modal data, feature information of different modals is fused through a multi-modal fusion algorithm, and comprehensive object description information is generated; 3, generating an optimal grabbing strategy based on the comprehensive object description information; wherein the grabbing strategy comprises the position and posture of a grabbing point, a grabbing path and grabbing force; and fourthly, the industrial robot executes the optimal grabbing strategy, the first step to the third step are repeated, and the grabbing strategy is adjusted in a self-adaptive mode till grabbing is completed. Stable and efficient grabbing can be achieved.
Owner:HUNAN INST OF INFORMATION TECH +2

Distributed operation-oriented multi-agent collaborative recommendation method and system

The invention discloses a distributed operation-oriented multi-agent collaborative recommendation method and system, and relates to the technical field of intelligent recommendation, and the method comprises the steps: configuring a private domain operation agent in each independent private domain platform, deploying a big language model-based construction demand analysis agent at a user side to receive a natural language demand description input by a user, and constructing a multi-agent collaborative recommendation system; generating recommendation task instructions of different private domain platforms; the method comprises the following steps: receiving user preference description and recommendation object description, sending the description to a corresponding private domain operation agent, calling a recommendation model to generate a recommendation result list based on the received user preference description and recommendation object description, collecting recommendation result lists returned by all private domain operation agents by a demand analysis agent, and integrating and classifying the recommendation result lists to generate a comprehensive recommendation list. According to the method, an efficient, safe and extensible distributed recommendation architecture is constructed by introducing a user demand analysis agent, a private domain operation agent and a federal learning mechanism.
Owner:广东省华南技术转移中心有限公司 +1

Cross-platform page code generation method based on large model

The invention discloses a cross-platform page code generation method based on a large model, and relates to the technical field of front-end development, and the method comprises the following steps: cooperatively analyzing a UI design drawing through a plurality of special large models, and carrying out real-time cross validation and dynamic compensation on an analysis result by adopting a multi-model mutual verification mechanism; based on a predefined standard specification, dynamic adjustment is carried out through a platform adaptation rule base; generating a page object description tree embedded with the input / output processing function; cross-platform page codes are generated through a code generation engine, and sandbox testing and automatic correction are executed; outputting a target platform code by utilizing a unified compiling engine; an adaptive optimization loop is constructed based on test feedback. According to the method, through integration of multi-model collaborative analysis, platform rule dynamic adaptation, sandbox verification and closed-loop optimization, the problems of large analysis deviation, poor cross-platform compatibility, uncontrollable code quality and the like are solved, and end-to-end high-quality automatic generation from a design drawing to multi-platform codes is realized.
Owner:CHENG DU ZHONG KE JI YUN RUAN JIAN YOU XIAN GONG SI

An intelligent SQL generation system based on multi-level intent recognition and a generation method thereof

The application discloses a kind of intelligent SQL generation method and system based on multistage intention recognition, comprising: receiving user natural language query;Vector matching step: vector matching is carried out based on knowledge base, obtain Top-K candidate query object, and knowledge base contains multi-layer structure description information;Large language model intention understanding step: determine final query object in combination with candidate query object description;Load branch workflow, parse field and table structure, dynamically configure database table and field information;SQL generation step: generate SQL sentence in accordance with syntax rule, including automatically adding field, time condition conversion, field priority matching operation;Execute SQL sentence and return result, and record log if it fails to prompt correction.The application provides an efficient, accurate and user-friendly database query solution by natural language processing technology combined with database knowledge base.
Owner:SICHUAN ZHONGLI JIAHUA INFORMATION TECH CO LTD

Print text content display method and device based on cross-platform consistency

PendingCN121722337ADigital output to print unitsImage resolutionFont rasterization
The invention discloses a printed text content display method and device based on cross-platform consistency, and the method comprises the steps: obtaining a to-be-printed character string and a defined text object description protocol, and obtaining a text attribute structure according to the defined text object description protocol; obtaining the resolution ratio of the ink-jet printer nozzle, and enabling the resolution ratio of the PC upper computer to be consistent with the resolution ratio of the ink-jet printer nozzle according to the resolution ratio of the ink-jet printer nozzle; adopting a text shaping engine to obtain corresponding font list information according to the text attribute structure; adopting a font rasterization engine to obtain corresponding font contour information according to the font list information and the text attribute structure; obtaining the pixel width and height of the whole line of text according to the font list information and the font contour information; according to the obtained font list information, font contour information and resolution parameters, drawing to a display screen by adopting a Skia rendering engine; therefore, cross-platform measurement and drawing consistency of the text content of the ink-jet printer is realized.
Owner:SOJET MARKING TECH (XIAMEN) CO LTD

Cochlea object description method and system, equipment, storage medium and program product

The embodiment of the invention provides a cochlea object description method and system, equipment, a storage medium and a program product, and the method comprises the steps: extracting a region-of-interest image of a cochlea object from a medical image comprising the cochlea object, carrying out the segmentation of the cochlea object, obtaining a segmentation result of the cochlea object, and carrying out the segmentation of the cochlea object according to a plurality of preset radiomics feature types, respectively extracting corresponding radiomics features from the region-of-interest image, obtaining voxels of the cochlea object from the segmentation result, calculating respectively corresponding morphological parameter values based on voxel coordinates of the voxels of the cochlea object according to a plurality of preset morphological parameters, and determining the cochlea object according to the plurality of preset morphological parameters and the plurality of morphological parameter values based on the plurality of radiomics features and the plurality of morphological parameter values. A description data set of the cochlea object is constructed to describe the cochlea object using the plurality of radiomics features and the plurality of morphological parameter values. According to the scheme provided by the embodiment of the invention, the comprehensiveness and accuracy of description aiming at the cochlea are improved.
Owner:BEIJING FRIENDSHIP HOSPITAL CAPITAL MEDICAL UNIV +1

Symptom information determination method and apparatus, electronic device, and storage medium

The application relates to a symptom information determination method and device, electronic equipment and a storage medium. The method comprises the following steps: acquiring object description information to be processed; using a target symptom recognition network to obtain corresponding target symptom information by taking the object description information as input, wherein the target symptom recognition network is obtained by machine learning training of multiple sample pairs and adjusting parameters of a preset network during the training process, each sample pair indicates a pair of object description information samples and symptom information samples, the training process comprises learning a first type correlation degree and a second type correlation degree, the first type correlation degree represents a correlation degree between two heterogeneous samples in the sample pair, and the second type correlation degree represents a correlation degree between two heterogeneous samples from different sample pairs. The application improves the accuracy and efficiency of symptom recognition. The embodiments of the application can be applied to various scenes such as cloud technology, artificial intelligence, intelligent transportation and auxiliary driving.
Owner:腾讯医疗健康(深圳)有限公司

Pixel based with object based decision making approach for driving

A method of a pixel based with object based decision making for driving, the method includes receiving, at a first machine learning process of an artificial intelligence agent, a sensed information unit; receiving, at a second machine learning process of the artificial intelligence agent, object descriptive information regarding an object captured in the sensed information unit; generating, by the first machine learning process, a pixel-based path planning output related to a suggested pixel-based path segment of a vehicle; generating, by the second machine learning process, an object-based path planning output related to a suggested object-based path segment of the vehicle; and generating, by at least in part processing the pixel-based path planning output in correspondence with the object-based path planning output, a driving related output with respect to the vehicle.
Owner:AUTOBRAINS TECH LTD

Image generation method, training method of reward model

Embodiments of the present application provide an image generation method, a reward model training method, an electronic device, a storage medium and a computer program product. The method comprises: obtaining object description text of a target object, original object image and scene information matched with a to-be-generated image carried by an image generation instruction; taking the scene information, the object description text and the original object image as input information, calling a pre-tuned visual language model to generate background image design description text adapted to the scene information, wherein when the scene information is a weak signal in the input information, the background image design description text can reflect scene preference characteristics matched with the scene information; calling a preset image generation model to generate an image of the target object matching the scene information under the constraint of the background image design description text, the image of the target object in the original object image and layout information. The image generated by the method is more matched with the scene preference corresponding to the scene information.
Owner:HANGZHOU ALIBABA INT NETWORK TECH CO LTD

Task execution method, device, apparatus, computer storage medium and product

The application discloses a task execution method, device, equipment, computer storage medium and product. The method comprises the following steps: receiving a voice text of a user, processing the voice text into a semantic vector; based on the semantic vector, searching a corresponding object description vector in an environment vector database, and obtaining an object description text of the object description vector; based on the voice text and the object description text, generating an intent recognition prompt information, inputting the intent recognition prompt information into an intent recognition model, and obtaining an object information set recognized by the intent recognition model. The adaptability of the robot to the environment in the home service scene can be improved, the recognition accuracy of the user's intent can be improved, and the autonomy and flexibility of the robot task execution can be realized.
Owner:CHINA MOBILEHANGZHOUINFORMATION TECH CO LTD +1

Training artificial intelligence (AI) engines for custom object description generation

An artificial intelligence (AI) engine assists in the creation of product descriptions. For example, a system that includes the AI engine receives raw data including a description of a product and generates a search query based on the raw data to search the internet for possible product descriptions that satisfy the query. The system can parse the product descriptions into categories of content that map to sections of a template and creating the custom product description based on a selection and combination of the categories of content of the product descriptions matched to sections of the template. An AI engine is trained based on the selection and combination of the categories of content and ranks the content based on a frequency of being included in custom product descriptions such that the system can later recommend content in accordance with the ranking for creating a custom product description.
Owner:TRUSTCLARITY INC

Audio-video player control method based on voice instruction

The application relates to the technical field of audio and video control, and discloses an audio and video player control method based on a voice instruction. The method comprises the following steps: collecting original voice instruction streams of a user, the instruction streams containing a time domain audio signal sequence, environmental noise spectrum and user pronunciation characteristic parameters, so that voice information can be comprehensively captured; performing multi-modal instruction analysis processing on the original voice instruction streams, generating a structured control instruction set containing an acoustic control intention identifier, a semantic operation object description and context association parameters, and improving analysis accuracy; then performing player state adaptation based on the set, generating a dynamic control response sequence containing device state adjustment commands, media content positioning parameters and interface interaction logic identifiers, driving the player to perform multi-dimensional control operations and generating real-time playing control effect feedback data; and finally optimizing multi-modal analysis parameters according to the feedback data, generating an adaptive instruction analysis strategy, and optimizing the control experience of the user on the audio and video player.
Owner:ONWAY TECH LTD

Video processing methods, apparatus, computer equipment and storage media

This disclosure proposes a video processing method, apparatus, computer device, and storage medium. The method includes: decomposing a video to obtain multiple video segments; determining multiple action event information corresponding to each of the multiple video segments; determining a target video segment from the multiple video segments based on the action event information; identifying object description information from the target video segment; and determining whether a target action event has occurred in the scene described by the video based on the object description information. Because the video is first decomposed, and a target video segment is determined based on the decomposed video segments, and target action events in the scene are identified and judged around the target video segment, the method achieves rapid identification and judgment of target action events around the target video segment, effectively reducing the computational resources consumed in target action event identification and judgment, and improving the identification and judgment effect of target action events.
Owner:JD DIGITS HAIYI INFORMATION TECHNOLOGY CO LTD

Session message processing method and apparatus, computer device, and storage medium

ActiveCN116644040BExpand archiving methodsImplement automatic archivingFile metadata searchingFile/folder operationsThumbnailMediaFLO
The application relates to a conversation message processing method and device, computer equipment and a storage medium. The method comprises the following steps: in a first social application, a first conversation window of a conversation between a first conversation object and a second conversation object is displayed; in the first conversation window, a message card of a media conversation message generated by the second conversation object forwarding a target media object belonging to a media social platform into the conversation through a second social application is displayed; the message card displays thumbnail description information of the target media object; in the case that the media conversation message is archived after authorization of the second conversation object, in response to an archiving query instruction for the conversation, text format conversation archive information of the conversation is displayed; the conversation archive information comprises message archive information of the media conversation message, and the message archive information comprises object description information matched with the thumbnail description information and used for describing the target media object. The method can widen the archiving mode of the conversation message.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Robot control based on natural language (NL) input and based on descriptor(s) of object(s) that are present in environment with robot and relevant to the nl input

PendingUS20260186488A1Robot environmentMap Location
Some implementations relate to generating, based on processing captured vision data instances throughout an environment: regions of interest, and an estimated map location and region embedding(s) for each region of interest. Some implementations additionally or alternatively relate to determining, based on (1) a free form (FF) natural language (NL) instruction for a robot to perform a task and (2) generated region embedding(s) for identified regions of interest in an environment: object descriptors that describe objects that are relevant to performing the task and that are likely present in the environment. Some implementations additionally or alternatively relate to utilizing a subset of object descriptor(s), determined to be descriptive of object(s) that are relevant to performing the task of an FF NL instruction and likely included in the environment, in determining robotic skill(s) for robot(s) to implement in performing the task specified in the FF NL instruction.
Owner:GDM HOLDING LLC

Cascade CT data target identification method and device and ray scanning detection system

The invention provides a cascaded CT data target identification method. The method comprises the following steps: acquiring three-dimensional CT data; performing initial identification on the three-dimensional CT data to obtain a first object description set of a target; according to a first object description set of the target, a second object description set is generated, the first object description set is an object description set for three-dimensional data, and the second object description set is an object description set for two-dimensional data; and performing fine recognition on the target by using a cascaded fine recognition method to obtain a fine recognition result of the target, in the cascaded fine recognition method, taking the second object description set as an input of a first deep learning network for two-dimensional data, and taking the second object description set as an output of a second deep learning network for the two-dimensional data. And processing the second object description set by using the first deep learning network.
Owner:TSINGHUA UNIVERSITY +1

Image description method based on 3D spatial relationship and multi-agent debate

The invention discloses an image description method based on a 3D spatial relationship and multi-agent debate. Respectively inputting the image into an object understanding model and a depth estimation model, and obtaining object description and a relative / absolute position; inputting the acquired data into a space diagram module for processing to obtain a position relation text description; constructing an entity relationship library, extracting a semantic relationship between entities from the library, inputting the extracted data and the text description into a large language model for fusion to obtain a new question text, retrieving support text paragraphs related to the new question text from a knowledge corpus by utilizing a retriever, and performing iterative refining to obtain a text paragraph set; and processing the image, the new question text and the text paragraph set to obtain detailed and accurate long text description of the final image. According to the method, the object understanding model and the grammar analysis technology are combined, natural language description of the extended object is refined, diversity and distinction are enhanced, ambiguity between the objects is reduced, and image description accuracy is improved.
Owner:JIANGXI NORMAL UNIV

Cascade CT data target identification method and device and ray scanning detection system

The invention provides a cascaded CT data target identification method. The method comprises the following steps: acquiring three-dimensional CT data; performing initial identification on the three-dimensional CT data to obtain an object description set of a target; and performing fine identification on the target by using a cascaded fine identification method to obtain a fine identification result of the target, in the cascaded fine identification method, taking the object description set as an input of a first deep learning network for three-dimensional CT data, and taking the object description set as an output of the first deep learning network for the three-dimensional CT data. And processing the object description set by using the first deep learning network.
Owner:TSINGHUA UNIVERSITY +1

Package similarity search for lost item identification

A first object out of a plurality of objects that have passed an object handling system is identified. To that end, image embeddings of images of the plurality of objects are generated by a multimodal object finder model that includes at least one neural network. A first object description text describing the first object is passed to the multimodal object finder model to generate a text embedding. Using a similarity search, a most similar image embedding among the image embeddings that is most similar to the text embedding is found, and a corresponding image and / or a text is output.
Owner:SICK PRODUCT & COMPETENCE CENTER AMERICAS LLC

Session message processing method and apparatus, computer device, and storage medium

The application relates to a conversation message processing method and device, computer equipment and a storage medium. The method comprises the following steps: in a first social application, a first conversation window of a conversation between a first conversation object and a second conversation object is displayed; in the first conversation window, a message card of a media conversation message generated by the second conversation object forwarding a target media object belonging to a media social platform into the conversation through a second social application is displayed; the message card displays thumbnail description information of the target media object; in the case that the media conversation message is archived after authorization of the second conversation object, in response to an archiving query instruction for the conversation, text format conversation archive information of the conversation is displayed; the conversation archive information comprises message archive information of the media conversation message, and the message archive information comprises object description information matched with the thumbnail description information and used for describing the target media object. The method can widen the archiving mode of the conversation message.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Method and apparatus for enforcing access control based on process functional context

The application discloses a method and device for enforcing access control based on process function context, wherein the method comprises: collecting a security context associated with a system call initiated by a target process to generate a standardized object description; mapping the object description to a target function classification identifier to obtain a process function context view of the target process according to the target function classification identifier; constructing a target decision key for an access decision based on a current policy epoch, a qualifier, a function classification identifier, a view identifier of the process function context view and an isolation domain digest; and querying a multi-level cache according to the target decision key to determine a matching target access decision. Thus, context-sensitive decision based on a function context is realized, and cross-component consistency is achieved.
Owner:BEIJING METRO INFORMATION DEV CO LTD

Index information generation method and device, video retrieval method and device and electronic equipment

The invention provides an index information generation method, a video retrieval method and device and electronic equipment, and can be applied to the technical field of video information extraction and retrieval. The index information generation method comprises the following steps: dynamically sampling a detection video according to the video content of the detection video to obtain a plurality of sampled video frames, the sampled video frames comprising a target object; screening a plurality of key video frames from the plurality of sampling video frames according to object information for the target object in the plurality of sampling video frames; analyzing the plurality of key video frames by using a multi-modal large model to obtain object description information and object associated text information corresponding to each key video frame; and generating index information for the detection video according to a multi-modal information set consisting of the plurality of key video frames, the object description information corresponding to each key video frame and the object associated text information, so as to determine an information set by using the index information when retrieval information for the detection video is received.
Owner:NUCTECH CO LTD

Cross-platform mobile terminal automatic testing method and device based on natural language driving

The invention discloses a cross-platform mobile terminal automatic testing method and device based on natural language driving and a medium. The method comprises the following steps: analyzing a natural language instruction in a preset format file, and extracting an operation semantic feature and a target object description; matching an underlying control protocol instruction corresponding to the system type of the current tested equipment from a preset cross-platform instruction abstraction library according to the operation semantic features; acquiring a current interface image of the equipment, and retrieving a target coordinate area matched with the target object description by using a hybrid positioning mechanism; and finally, generating and executing an automatic test action. According to the method, the problem of inaccurate element positioning in a complex UI scene is solved through weighted fusion and intersection-union correlation of text and icon areas in a hybrid positioning mechanism; and meanwhile, the abstract library is utilized to shield the underlying protocol difference between the Android and the iOS, so that the cross-platform test driven by the natural language is realized, the script maintenance cost is remarkably reduced, and the test efficiency is improved.
Owner:DONGGUAN QIAOAN ZHILIAN TECHNOLOGY CO LTD

Object descriptor tokens with object tokens for object detection

A device for object detection includes one or more memories configured to store image data; and processing circuitry connected to the one or more memories, the processing circuitry configured to: generate bird's-eye-view (BEV) object feature data from the image data, including BEV object tokens, the BEV object tokens being indicative of a first set of information used for object detection in the image data; generate an input for a transformer encoder based on at least some of the BEV object feature data or the image data; generate object descriptor tokens based on applying the transformer encoder to the input, the object description tokens being indicative of a second set of information used for object detection in the image data, the second set of information being usable for classifying real objects; and output object detection information based on the BEV object tokens and the object descriptor tokens.
Owner:QUALCOMM INC

Method for controlling discrete event systems

The invention relates to a method, implemented by a computer infrastructure, for controlling the sequential change in a resource of a discrete event system, said change being modelled by a state machine having a plurality of states and at least one transition for changing the current state of the resource from a first state to a second state on the occurrence of a predefined event, the method comprising the following steps: reading a first computer object representing the resource, said first computer object describing the current state of the resource and comprising a function that is intended to be executed automatically by the computer infrastructure in the event of a request to change the current state of the resource from the first state to the second state; executing the function associated with the requested state transition from the first state to the second state.
Owner:OVOCHAIN

Object management method and device, equipment and storage medium

Embodiments of the invention disclose an object management method and apparatus, a device and a storage medium. The method comprises the steps of determining a to-be-processed object matched with an obtained object identifier; determining an enhancement processing mode of the to-be-processed object and object description information of the to-be-processed object; determining enhancement mode coding information of the enhancement processing mode according to the enhancement processing mode; according to the to-be-processed object, the object description information and the enhancement mode coding information, an object enhancement evaluation result of the to-be-processed object is predicted, and the object enhancement evaluation result is used for representing the enhancement processing effect of the to-be-processed object based on the enhancement processing mode; and if the object enhancement evaluation result meets a preset condition, performing enhancement processing on the to-be-processed object based on an enhancement processing mode. By means of the method, effective screening whether the objects are subjected to enhancement processing or not can be achieved, it is guaranteed that enhancement processing is only carried out on the objects generating the forward feedback effect, and therefore computing power resources are saved, the enhancement processing cost is reduced, and the forward feedback effect is improved.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD