Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

3189 results about "Animation" patented technology

AI-based animation sub-mirror script automatic generation and visual preview method and system

The invention discloses an AI-based animation split script automatic generation and visual preview method and system, and the method comprises the following steps: 1, receiving a natural language script text inputted by a user, the natural language script text comprising scene description, role action, dialogue and shot indication information; step 2, performing semantic analysis and structured analysis on the script text based on a natural language processing technology, and identifying and extracting key narrative elements; by introducing an artificial intelligence technology, end-to-end automatic generation and interactive optimization from a character script to a dynamic split rehearsal video are realized, the system can deeply understand scenes, actions, role emotions and shot languages in the script, corresponding visual elements are automatically matched and generated, and the dynamic split rehearsal effect is improved. And the timeline and the rhythm conforming to the film and television grammar are constructed, so that the efficiency and the consistency of the split creation are greatly improved, and the professional threshold and the manufacturing cost are reduced.
Owner:NEW AXIS ANIMATION TECHNOLOGY DEVELOPMENT (BEIJING) CO LTD

Generative ai models for image rendering and inverse rendering

Embodiments of the present disclosure relate to rendering and inverse rendering using one or more generative models. “Rendering” refers to the process of generating a final visual image, video frame, or animation from a 2D or 3D model. “Inverse rendering” is a process that involves deducing or estimating the properties (e.g., material maps or other properties such as geometry, lighting, and textures) of a scene from observed images or visual data. Essentially, it aims to reverse the traditional rendering process. Various aspects of the present disclosure introduce editable light and material controls into generative models to allow for artistic creation. Various embodiments integrate generative models as a renderer for classic rendering pipelines to upcycle and enhance the style of rendered content.
Owner:NVIDIA CORP

Commodity display interaction visualization method and device

The invention relates to the field of commodity visualization, in particular to a commodity display interaction visualization method and device. The method comprises the following steps: collecting a multi-azimuth image of a commodity, carrying out three-dimensional texture modeling, and constructing a three-dimensional texture mapping model; performing material light rendering on the three-dimensional texture mapping model to generate a material rendering result; collecting an environment detection image of a commodity display environment, and performing environment illumination adaptation compensation on a material rendering result to obtain an illumination compensation rendering commodity; carrying out attribute information visual layout on the illumination compensation rendering commodity to obtain a commodity visual space; and carrying out interaction response animation analysis according to the commodity visualization space, carrying out multi-target parallel rendering, and executing commodity interaction visualization operation. The form and surface details of the commodity in the real world are accurately restored, the visual reality sense is improved, and the interactive experience feeling of browsing the commodity by a user is enhanced.
Owner:SHENZHEN XIAOYI SHUZHI TECH CO LTD

Method and system for generating 3D (three-dimensional) human motion under text driving by using 2D (two-dimensional) video

The invention discloses a method and a system for generating 3D (three-dimensional) human motion under text driving by utilizing a 2D (two-dimensional) video. The method comprises the following steps of: acquiring the video and preprocessing to obtain a two-dimensional key point sequence and text description; the two-dimensional key point sequence passes through a spatiotemporal feature adapter to obtain a potential spatiotemporal feature sequence, a residual vector quantizer quantizes and outputs a three-dimensional SMP L parameter sequence, and meanwhile, potential spatiotemporal features and a discrete Token sequence are mapped; preprocessing a text to extract a semantic vector, partially covering a Token sequence of a basic quantization layer, reconstructing a prediction sequence through a predictor in combination with the semantic vector, and obtaining a complete sequence through a refiner; constructing a total loss function and a text-to-action loss function to train the module; and inputting the text description and the basic quantization layer Token to a trained module, outputting a three-dimensional SMPL parameter sequence, and rendering to generate a three-dimensional human body grid and animation. According to the method, the end-to-end generation from the text to the three-dimensional SMPL action is realized only by two-dimensional key points and text description.
Owner:ZHEJIANG UNIV

Automatic lecturer video generation method based on AI speech synthesis and animation driving

The invention discloses a lecturer video automatic generation method based on AI speech synthesis and animation driving. The method comprises the following steps: performing structured analysis on a PPT or a text script through an improved interior point method and an incremental shortest path algorithm; performing semantic grouping by applying a full-dynamic parallel single-link clustering algorithm and generating an enhanced script with an expressive mark; a CosyVoice technology is combined with a low-rank approximation method to generate a high-quality voice data stream; establishing a mapping relation between contents and action expressions through semantic analysis, and generating a complete action expression instruction set; and driving the digital human model by using the msueTalk technology, and generating a final lecturer teaching video through a parallel rendering algorithm. According to the invention, the method achieves the efficient and automatic generation of the education video, remarkably improves the content production efficiency, reduces the production cost, and guarantees the specialty and expressive force of the teaching video.
Owner:SHENZHEN XUEYOU TECHNOLOGY CO LTD

Animation video generation method and system based on video script

The invention discloses a cartoon video generation method and system based on a video script, and relates to the technical field of video synthesis, and the method comprises the steps: based on structured script data, combining a predefined lens rule base and a reinforcement learning model, determining the type, duration and lens operation effect of a lens, and generating a lens splitting sequence through a dynamic lens splitting automatic generation mechanism, based on the split mirror sequence, generating an animation style key frame image through a diffusion model, selecting action data matched with the emotion label through a predefined action library, generating a voice waveform matched with the emotion label through a text-to-voice model, selecting a background music audio matched with the emotion label through a music library, and generating a multi-modal content stream; according to the method, the lens type, the duration and the lens operation effect are dynamically optimized through the reinforcement learning model, the defects of a traditional lens scheduling method based on static rule mapping in time continuity and narrative continuity are overcome, and the narrative fluency and the dynamic adaptability of the lens division sequence are obviously improved.
Owner:CHANGCHUN VOCATIONAL INST OF TECH

Digital human interaction system based on web terminal

The invention discloses a digital human interaction system based on a web end, relates to the technical field of digital human interaction, and aims to solve the problem of accumulated dislocation of browser end digital population animation and actual audio playback caused by multiple clocks and buffer scheduling. The system comprises a visual position cooperative control module, a session initialization module, an audio track and rendering canvas binding module, a multi-domain alignment time base cluster establishment module, construction of a time base cluster containing a system reference sub-time base and a content logic sub-time base, potential candidate anchor point generation module, an optimal anchor point selection module, synchronization error calculation, and judgment of a synchronization steady state, a fine adjustment state or a lost state. An optimal anchor point is screened to adjust the animation, and a visual effect studio dynamic maintenance module and an enhancement generation module assist in out-of-step processing and parameter optimization; through cooperation of multiple modules, accurate synchronization of audio and digital human animation is realized.
Owner:NANJING SUPERMIND INFORMATION TECHNOLOGY CO LTD

Mass data particle system rendering method and device based on WebGL

The invention discloses a WebGL-based mass data particle system rendering method and device, and belongs to the technical field of computer graphic processing. The method comprises the following steps: firstly, selecting the type of a particle emitter according to scene requirements, then constructing a primitive object based on points or quadrangles, and configuring three-dimensional coordinates and texture parameters; aggregation rendering is carried out on a single-frame primitive array through a WebGL vertex shader and a fragment shader, and the drawing calling frequency is reduced; the method comprises the steps of dynamically adjusting particle attributes, including position compensation, color gradient and explicit-implicit control, determining a rendering form and playing time of an animation according to requirements of an actual application scene, and supporting fine configuration of a single particle, including texture animation, dynamic scaling and rotation parameters. When a large amount of data is processed, on the premise of ensuring the rendering quality, the rendering efficiency can be remarkably improved, the consumption of a memory and GPU resources can be reduced, the response time of the system can be optimized, and meanwhile, the interaction experience of a user can be improved.
Owner:RES INST OF CHEM DEFENSE PLA ACAD OF MILITARY SCI

Devices, methods, and graphical user interfaces for attention based scrolling and object interactions

Some examples are directed to systems and methods for scrolling scrollable content in response to gaze-based inputs. Some examples are directed to systems and methods for scrolling scrollable content in response to gaze-based inputs and / or input provided by a respective input element. Some examples are directed to systems and methods for displaying virtual objects with an appearance that is based upon a duration of attention directed to the virtual objects. Some examples are directed to systems and methods for displaying animations of virtual objects. Some examples are directed to changing values of audio parameters based on attention of a user. Some examples are directed to changing gaze scrolling regions of content based on characteristics of the content.
Owner:APPLE INC

Techniques for automated generation and rigging of objects for animation

One embodiment of a method for generating animations includes generating one or more images of an object based on user input, generating textured geometry based on the one or more images, generating a weight map based on at least one image included in the one or more images, and generating an animation of the object based on the textured geometry, the weight map, and a skeleton.
Owner:DISNEY ENTERPRISES INC

VEM-Token emotion synchronization function hierarchical fusion method

A VEM-Token emotion synchronization function hierarchical fusion method is different from a traditional NLP-Token method, an emotion synchronization function VEM-sync is innovated for the first time, the emotion synchronization function VEM-sync is synchronized with a VEM-Token sequence of music beats, multiple high-dimensional emotions are directly described by adopting mathematical languages, and therefore deviation of discretized natural language characters on description of a high-dimensional emotion analog quantity function is avoided, and the emotion synchronization effect is improved. The method comprises the steps of defining VEM-sync and synchronous content, defining rhythm attributes, emotion attributes and emotion functions, adopting one or combination of multi-layer weighted scanning, a recurrent neural network, a long and short-term memory network, a self-attention mechanism and an RAG network generated by retrieval enhancement, and performing hierarchical fusion calculation to output a time emotion function of a vocal music file. According to the phonetic function or the emotional function, the effects of emotional texts, emotional expressions, emotional languages and emotional multi-dimensional animations are directly driven by crossing discrete text tokens, a model context protocol (MCP) and a function calling function are supported, and copyright management and encryption and decryption or interfaces are provided.
Owner:GREATER BAY AREA STAR BIOTECH (SHENZHEN) CO LTD

Method and device for generating an animation graph

ActiveUS12499599B1AnimationAlgorithmAnimation
In some implementations, the method includes: obtaining a plurality of animations; determining initial and end motion states for each of the plurality of animations; generating an animation graph including nodes for each of the plurality of animations by connecting, with a directional edge, a first node with an end motion state to a second node with an initial motion state that matches the end motion state of the first node; generating a transitional animation that is not included among the plurality of animations from an initial reference motion state to a target motion state that corresponds to a path that traverses the animation graph from a third node associated with the initial reference motion state to a fourth node associated with the target motion state; and updating the animation graph by removing one or more nodes from the animation graph based at least in part on the transitional animation.
Owner:APPLE INC

Dynamic generation method of digital animation character based on generative model

The invention discloses a digital animation role dynamic generation method based on a generative model. The method comprises the following steps: inputting a role original image and scene background information, extracting skeleton key points and scene features, analyzing an action sequence, predicting a motion track, calculating an optimal position of a role in a picture, and carrying out dynamic adjustment. And according to the role position, obtaining morphological feature data, evaluating the quality grade, and optimizing the coordination of the role image. And finally, comprehensively scoring by adopting a multi-dimensional quality evaluation system, and determining the visual presentation quality of the role. According to the method, deep fusion of role actions, positions and scenes is realized, the dynamic expressive force and visual coordination of animation pictures are improved, and an efficient solution is provided for generating high-quality animation contents.
Owner:HEBEI XIONGAN PEPSI HENGXING NETWORK TECHNOLOGY CO LTD

Artificial intelligence generated content gifting

A data processing system implements obtaining digital content as an output from the generative model; receiving a natural language prompt describing digital wrapping for the digital content, the digital wrapping to be presented to a recipient of the digital content; constructing a prompt based on the natural language prompt using the prompt construction unit; providing the second prompt to the generative model to cause the generative model to generate the digital wrapping; obtaining the digital wrapping as an output from the generative model; sending the digital content and the digital wrapping to a client device of a recipient; and causing the client device to present the digital wrapping on a second user interface of a client device and controls, which when activated, cause the client device to present an animation of the digital wrapping being removed and the digital content to be presented.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC