Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

6 results about "Result Category" patented technology

A classification of a result.

Large language model reinforcement learning system, reinforcement learning method and related equipment

ActiveCN121072658ABiological modelsLinguistic modelResult Category
The invention provides a large language model reinforcement learning system and method and related equipment, and the system comprises a management module which is used for carrying out the unified management and calling of a multi-class value function, and specifically comprises an environment registration unit which is used for building a global registry, and storing a mapping relation between an environment function and corresponding meta-information and a mapping relation between the environment function and the value function; the environment running unit is used for positioning an environment function according to the unique identification information, instantiating a running environment, calling the environment function to obtain a running result, inputting an answer generated by a large language model into a value function for evaluation, and obtaining a reward result and a result category; the integration module is used for generating a reward signal adaptive to the reinforcement learning process based on the reward result of the value function and the reward result of the value model; and the training module is used for updating strategy parameters of the large language model based on the reward signal. According to the method, unified calling of various heterogeneous operating environments is realized, and the cross-task generalization ability of the model is improved.
Owner:BEIJING JIBU QIANLI TECHNOLOGY CO LTD

A large language model reinforcement learning system, a reinforcement learning method, and related devices

ActiveCN121072658BLinguistic modelResult Category
The application provides a large language model reinforcement learning system, a reinforcement learning method and related equipment, the system comprises: a management module for unified management and calling of multiple category value functions, specifically comprising: an environment registration unit for establishing a global registration table, storing environment functions and corresponding meta information and mapping relationship with value functions; an environment running unit for locating environment functions according to unique identification information and instantiating running environment, calling environment functions to obtain running results, and inputting the answers generated by the large language model into the value function for evaluation, obtaining reward results and result categories; an integration module for generating reward signals suitable for the reinforcement learning process based on the reward results of the value function and the reward results of the value model; a training module for updating the policy parameters of the large language model based on the reward signals. The application realizes unified calling of multiple heterogeneous running environments and improves the generalization ability of the model across tasks.
Owner:BEIJING JIBU QIANLI TECHNOLOGY CO LTD

Program running problem positioning method and device, equipment and storage medium

This invention relates to monitoring, and provides a method, apparatus, device, and storage medium for locating program execution problems. The method receives user requests, wherein the user request includes user identity information and execution information; generates an execution task based on the user identity information and the execution information; when the execution task is detected to be completed, obtains the response result of the execution task; performs multi-dimensional analysis on the response result to obtain the result category of the execution task in each dimension; if the result category includes a preset category, identifies the user who made the request, and obtains the input parameters of the user request from a preset message system based on the user; generates alarm information based on the input parameters and the preset category, which can accurately locate the operational problem encountered by the user. Furthermore, this invention also relates to blockchain technology, and the alarm information can be stored in the blockchain.
Owner:CHINA PING AN PROPERTY INSURANCE CO LTD

Method and device for inspecting glass containers according to at least two modalities in order to classify the containers according to glass defects

Method and device for inspecting glass containers according to at least two modalities in order to classify the containers according to glass defects. The invention relates to a method for inspecting glass containers (2) comprising the following steps: - inspecting each container using an inspection system (10) in order to obtain at least one analysis image according to a first modality corresponding to an absorption image (Ia) and at least one analysis image according to a second modality corresponding to a birefringence image (Ib) or a refraction image (Ir), - defining a list of classes (D1, D2,...Dk,…Dp) including at least glass defects, - ensure a matching of at least a portion of the analysis images according to the first modality and according to the second modality, - from at least one analysis image according to the first modality and at least one analysis image according to the second modality, matched, classify the analysis images using an image classifier (Cl) that determines membership in a result class from the list of classes, the image classifier having been trained by supervised learning, - classify the container according to the result class. Figure for the abstract: Fig. 1.
Owner:TIAMA SOCIETE ANONYME

Method, device, and medium for searching content and displaying searching results

ActiveUS12682003B2Result CategoryData mining
Provided are a content search method, apparatus, and device, and a storage medium. The present disclosure enables: receiving a search content; and displaying a plurality of answer viewpoints and first contents in a search result interface, wherein each answer viewpoint corresponds to one category of search results, the search results are results obtained by searching the search content, the first contents comprises keywords, the keywords are used for indicating reasons for displaying a target answer viewpoint among the plurality of answer viewpoints, and the keywords are extracted from a target category of search results corresponding to the target answer viewpoint.
Owner:DOUYIN VISION CO LTD