Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

3 results about "Service level requirement" patented technology

In Software Development / IT, a Service-Level Requirement (SLR) is a broad statement from a customer to a service provider describing their service expectations. A service provider prepares a service level agreement (SLA) based on the requirements from the customer. For example: A customer may require a server be operational (uptime) for 99.95% of the year excluding maintenance.

Request processing method and device based on large model, electronic equipment and storage medium

PendingCN121328705AInference methodsAlgorithmService level requirement
The invention provides a request processing method and device based on a large model, electronic equipment and a storage medium, and relates to the technical field of artificial intelligence, in particular to the technical fields of deep learning, large models and the like. According to the scheme, in response to a received large model reasoning request, a target service level requirement of the large model reasoning request is determined according to request parameters in the large model reasoning request; determining a target service instance pair from at least one candidate service instance pair matched with the target service level requirement; wherein the target service instance pair comprises a target pre-filling service instance determined based on a pre-filling delay and a target decoding service instance determined based on a load; adopting the target pre-filling service instance and the target decoding service instance to respond to the large model reasoning request based on the target service level requirement to obtain a response result; and sending a response result to the client associated with the large model reasoning request.
Owner:BEIJING BAIDU NETCOM SCI & TECH CO LTD

An end-to-end cloud multi-level multi-domain access control method

ActiveCN121217418BSecuring communicationPathPingService level requirement
The present application relates to the field of Internet of Things security and computing scheduling, and particularly relates to a kind of end edge cloud multistage multi-domain access control method, which method includes: first according to service level requirement and compliance clause generates intention constraint description and establishes session root and one-time session token;Each layer carries remote proof and execution commitment and submits bid in parallel, and control plane verifies to form a bid package set;Under the constraint, joint arrangement solution is determined, stage is settled and cross-domain data path is determined, and minimum data disclosure rule and minimum feasible plan are generated, and stage token is issued;Each stage executes according to token and generates execution proof, and control plane verifies and enters account after verification, and switches between backup plan and current plan according to verification and operation monitoring, and completes verifiable and traceable operation closed loop.
Owner:CHINA COMM INVESTMENT DIGITAL TECH (BEIJING) CO LTD

Micro-service-based medical AI model standardized pipe connection and dynamic scheduling method

PendingCN122455267AService level requirementDistributed computing
The application relates to the technical field of medical management and discloses a medical AI model standardization and dynamic scheduling method based on micro services, which extracts the environment dependence of a medical AI model file through a micro service, encapsulates the medical AI model file into a standardized micro service container image, collects bottom node load data by using a probe, introduces an exponential moving average algorithm for smooth filtering, receives a medical service request and completes authentication analysis through an API gateway, calculates a comprehensive matching score by combining the smooth node resource state, service priority and service level requirement through a scheduling micro service, screens an optimal target execution node, and performs elastic scaling based on the concurrent aggregation load rate by pulling the micro service container image through the target node. The application eliminates scheduling jitter caused by transient load, reduces deployment time consumption, and guarantees low-delay response of high-priority diagnosis tasks.
Owner:THE FIRST HOSPITAL OF LANZHOU UNIV