Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

5403results about "Details involving graphical user interface" patented technology

System and method for ai-driven multi-modal content generation and immersive interaction experiences

A system and method for creating complex, immersive, and interactive digital content is disclosed. The system integrates advanced artificial intelligence, multi-modal input processing, cloud-based shared environments, and immersive hardware to generate, optimize, and deliver rich interactive experiences. The platform supports content mashups, custom scenario generation, and adaptive AI behaviors, enabling the creation of unique and engaging digital environments across various media formats.
Owner:QOMPLX INC

Echoing a display of content based on familiarity of content type

An example operation includes one or more of logging user actions with respect to placement of objects of content on a user interface including respective content types of the objects of content, training an artificial intelligence (AI) model to learn location preferences for a content type on the user interface based on the logged user actions including the respective content types, receiving a request to open an object on the user interface with the content type, determining a display location on the user interface for the object based on execution of the AI model on the content type, and displaying the object at the determined display location on the user interface.
Owner:THE TORONTO DOMINION BANK

Virtual stylist

An example operation may include at least one of receiving, via a user interface of a device, an activation input from a user to initiate a session, capturing, by a camera of the device, a scan of a body of the user, wherein the capturing comprises recording at least one image and / or at least one video of the user, processing the at least one image and / or video to generate a three- dimensional model of the user comprising measurements and contours of the body, retrieving, from a database, at least one clothing item associated with the user, the at least one clothing item comprising dimensional attributes and texture attributes, rendering, by a graphics processing unit, the at least one clothing item onto the three-dimensional model to generate a visual representation, wherein the rendering simulates draping behavior, movement, and light interaction of the at least one clothing item relative to the three-dimensional model, and displaying, on the user interface, an interactive visualization comprising the visual representation of the three-dimensional model with the at least one clothing item from multiple viewing angles.
Owner:ELGORT PENELOPE

Three-dimensional reconstructions based on gaussian primitives

In implementation of techniques for three-dimensional reconstructions based on Gaussian primitives, a computing device implements a reconstruction system to receive a first digital image depicting an object from a first angle and a second digital image depicting the object from a second angle. The reconstruction system segments the first digital image and the second digital image into patches. The reconstruction system then generates, using a machine learning model, three-dimensional Gaussian primitives that predict parameters of points of the object in a three-dimensional space that correspond on a per-pixel basis to pixels of the patches. The reconstruction system then forms a three-dimensional reconstruction of the object for display in a user interface by merging the three-dimensional Gaussian primitives.
Owner:ADOBE INC

Electrical audio signal processing systems and devices

According to an aspect of the present invention, there is provided an electrical audio signal processing system and device, comprising: a computer graphics processing and selective visual display system with a screen; an eye tracking device; a processor; one or more computer memory devices; wherein the processor is arranged for operations comprising: measuring the user's eye movements to ascertain the specific word on which the user is fixated, by the eye tracking device; modifying the display at the user's current fixation point; applying a delay between the presentation of successive graphic elements based on the user's calculated rate to accommodate the user's required time; and presenting elements to the user at a rate based upon the user's required time.
Owner:DECHARMS RICHARD CHRISTOPHER

Systems and methods for providing semantics-based recommendations for three-dimensional content creation

Systems and methods are described for providing for display a user interface to facilitate creation of a three-dimensional (3D) environment. The disclosed techniques may generate the 3D environment based on one or more inputs received via the user interface, and perform visual processing of the 3D environment to obtain a natural language description of the 3D environment. The disclosed techniques may determine a context of the 3D environment based at least in part on the natural language description and may store the context in a data structure. A 3D content library may be queried, based on the natural language description of the 3D environment, to identify at least one recommended 3D object that is relevant to the stored context of the 3D environment. The disclosed techniques may providing for display, at the user interface, selectable option(s) to add the at least one recommended 3D object to the 3D environment.
Owner:ADEIA GUIDES INC

Ai-based visual content collage generation

A data processing system implements receiving, via a user interface of a client device, images for generating a collage image; generating captions for the images; constructing a first prompt by appending the captions to a first instruction string including instructions to a generative language model to extract a theme from the captions; providing the first prompt to the generative language model and receiving the theme therefrom; constructing a second prompt by appending the theme to a second instruction string including instructions to a text-to-image model to use the theme to create a background image with placeholders; providing the second prompt to the text-to-image model and receiving the background image therefrom; identifying the placeholders in the background image; creating the collage image by fitting the images into the identified placeholders; providing the collage image to the client device; and causing the user interface to display the collage image.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Assisting user(s) in interactions with large language model(s)

Implementations relate to a method implemented by one or more processors, the method including: receiving an input prompt for a large language model (LLM); generating first LLM output that is usable to generate a set of user interface (UI) each associated with a corresponding sub-prompt of the input prompt; causing, based on the first LLM output, the set of UI elements to be rendered at a user device; receiving further user input based on user interactions with one or more of the UI elements of the set of UI elements; and in response to determining that one or more termination conditions are satisfied: generating a final response to the input prompt based on generating second LLM output that is usable to generate the final response; and causing the final response to be rendered at the user device.
Owner:GOOGLE LLC

Systems and methods for providing artistic assistance on image capturing

Systems and methods for enhancing a live scene to be captured by a camera of a user device based on one or more best matched reference images that include similar subject matter and attributes of the live scene are disclosed. A preview of a live scene is displayed on the user device. Attributes of the live scene are identified. The live scene is segmented separately into a foreground and a background. Vector representations of the attributes of the segmented live scene are calculated and user device parameters are obtained, and both are used to identify reference images that include the attributes of the live scene. The identified reference images are displayed on the user device and a selection is made. The user device parameters are configured based on the selection made to capture the live scene to obtain similar effects of the selected reference image.
Owner:ADEIA IMAGING LLC

Avatar creation user interface

The present disclosure generally relates to creating and editing avatars, and navigating avatar selection interfaces. In some examples, an avatar feature user interface includes a plurality of feature options that can be customized to create an avatar. In some examples, different types of avatars can be managed for use in different applications. In some examples, an interface is provided for navigating types of avatars for an application.
Owner:APPLE INC

Presenting avatars in three-dimensional environments

In some embodiments, a computer system receives data representing a pose of at least a first portion of a user and causes presentation of an avatar that includes a respective avatar feature corresponding to the first portion of the user and presented having a variable display characteristic that is indicative of a certainty of the pose of the first portion of the user. In some embodiments, a computer system receives data indicating current activity of one or more users is activity of a first type and, in response, updates a representation of a user having a first appearance based on a first appearance template. The system receives second data indicating current activity of the one or more users and, in response, updates the appearance of the representation of the first user based on the current activity of the one or more users using the first or a second appearance template.
Owner:APPLE INC

Method and system for generating a 3D parametric mesh of an anatomical structure

There is provided a method and a system for generating a 3D parametric mesh of an anatomical structure of a patient for storing multi-domain data therein. A plurality of anatomical segments having been obtained from segmentation of a set of images of a patient are received. A 3D mesh comprising a plurality of concentric 3D mesh layers is received, where each concentric 3D mesh layer includes a same predetermined number of nodes. A set of nodes in the 3D mesh corresponding to a respective anatomical segment is determined to obtain a respective correspondence rule therebetween. The set of nodes is encoded with a set of features from the respective anatomical segment by using the correspondence rule to obtain a 3D parametric mesh, each node of the set of nodes in the 3D parametric mesh being associated with a respective plurality of feature channels comprises the set of features.
Owner:VITAA MEDICAL SOLUTIONS INC

Controllable image-to-video generation

Techniques are generally described for controllable image-to-video generation. In various examples, a first image representing at least a first object may be received. First input data including a selection of the first object in the first image for animation may be received. Second input data including at least a first bounding box indicating a target location of the first object may be received. A latent diffusion text-to-image model and the first image may be used to generate a first plurality of visual tokens. One or more first grounding tokens may be generated representing a location of the first bounding box. The latent diffusion text-to-image model may be used to generate a video animating the first object based on the first plurality of visual tokens and the one or more first grounding tokens.
Owner:AMAZON TECH INC

Patient registration for total hip arthroplasty procedure using pre-operative computed tomography (CT), intra-operative fluoroscopy, and / or point cloud data

ActiveUS12507972B2Image enhancementImage analysisPelvic regionPatient registration
A system for computer assisted navigation during surgery includes a computer platform that operates to register a target surgical area of a patient. In certain cases, a process includes: obtaining a pre-op CT image of a pelvic region of a patient and intra-operatively obtaining a point cloud data about the pelvic region with a navigated instrument, generating a 3D bone model which excludes non-targeted area such as a femur, and then merging the 3D bone model to the point cloud to register the target surgical area.
Owner:GLOBUS MEDICAL INC

Methods for participating in an artificial-reality application that coordinates artificial-reality activities between a user and at least one suggested user

Systems and methods are provided for facilitating an interactive artificial-reality activity. A method includes, after a user of a head-wearable device has opted-in to using an artificial-reality application to facilitate connecting with other participating users, determining, based on user-specific suggestion criteria, that a suggested user is located in an approved common space with the user and has opted-in to use the artificial-reality application. The method includes causing the head-wearable device to present an user interface (UI) element for linking the suggested user and the user. Upon the user selecting the UI element for linking the suggested user with the user, the method includes automatically causing the head-wearable device to provide visual-guidance UI elements to navigate the user to an interactive-activity location where the user and the suggested user will perform an artificial-reality activity while also displaying information about the artificial-reality activity to be performed.
Owner:META PLATFORMS TECHNOLOGIES LLC

Ai-based shape-adaptive consistent visual effect generation

A data processing system implements constructing a first prompt including a font mask of a reference character (RC) and a style prompt, sending the first prompt to a text2image model to iteratively generate salient content and concentrate the salient content within the font mask of RC as a first image of RC; concatenating two of the first images as a second image; generating a combined font mask of the font mask of RC and a font mask of a target character (TC); constructing a second prompt including the combined font mask and the second image, sending the second prompt to the model to iteratively generate salient content and in-paint the salient content within a half of the combined font mask as a third image of RC and TC; cropping a styled TC image from the third image using the font mask of TC; providing the styled TC image to a client device.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Systems for software-defined telescopes

A system can use a plurality of co-located telescopes to generate enhanced telescopic imagery of space objects. The system may be configured to receive telescopic imagery data of a plurality of space objects obtained from the plurality of co-located telescopes. The system can receive at least two imaging criteria. The available imaging criteria can include a target signal sensitivity comprising a minimum signal to noise ratio (SNR), a target number of spectral bands comprising a minimum number of spectral bands, a spectral range associated with one or more of the spectral bands, a location of the plurality of co-located telescopes, a target data cadence comprising a minimum number of frames per minute, a target number of space objects to be tracked, a target minimum spatial resolution, among others. The system can transmit instructions to the plurality of telescopes and generate enhanced telescopic imagery.
Owner:EXOANALYTIC SOLUTIONS INC

System and method for recommending resale alternatives for retail goods

A recommendation system identifies resale goods corresponding to retail goods being viewed by a user on a retail goods website. Retail goods metadata and images are extracted from the retail goods website using heuristics and LLM-based methods. The extracted retail goods data undergoes ML model-based product category classification, intelligent image cropping, and color detection. Vector embeddings are generated for images and text using ML models and compared to resale goods vector embeddings stored in a vector database. Multiple result sets are retrieved, fused, and re-ranked using LLM-based and preference-aware re-ranking. A data pipeline continuously loads, cleans, and processes resale goods inventory from resale goods websites.
Owner:PHIA HOLDINGS INC

Patient Registration For Total Hip Arthroplasty Procedure Using Pre-Operative Computed Tomography (CT), Intra-Operative Fluoroscopy, and / Or Point Cloud Data

PendingUS20250384569A1Image enhancementImage analysisPelvic regionPatient registration
A system for computer assisted navigation during surgery includes a computer platform that operates to register a target surgical area of a patient. In certain cases, a process includes: obtaining a pre-op CT image of a pelvic region of a patient and intra-operatively obtaining a point cloud data about the pelvic region with a navigated instrument, generating a 3D bone model which excludes non-targeted area such as a femur, and then merging the 3D bone model to the point cloud to register the target surgical area.
Owner:GLOBUS MEDICAL INC

Designing and optimizing adaptive shortcuts for extended reality

Disclosed herein is an extended reality system, and associated techniques, whereby personalized usage data of a user in one or more extended reality environments over a period of time can be collected and provided to a predictive model to determine optimal shortcut assignments for presentation to and subsequent use by the user. Determining the optimal shortcut assignments can involve estimating a plurality of interaction times for the user in the one or more extended reality environments, generating an optimized graphical user interface within the one or more extended reality environments with the optimal shortcut assignments reflected in the optimized graphical user interface, and rendering the optimized graphical user interface in the one or more extended reality environments to the user. The system may collect additional personalized usage data and periodically update the optimal shortcut assignments based on ongoing use of the system by the user.
Owner:META PLATFORMS TECHNOLOGIES LLC

Video generation for short-form content

Methods, systems, and apparatus, including computer programs encoded on computer storage media, for content generation are provided. One of the methods includes receiving one or more user input identifying information associated with one or more media elements and one or more characteristics of the video to be generated; and generating video content based on the received one or more user inputs, the generating comprising: identifying assets to include in the video, the assets including an avatar, generating a script for the video, and assembling a video layout.
Owner:LEMON INC(GB)

Three-dimensional digital human generation method and system capable of voice interaction

The invention belongs to the technical field of three-dimensional reconstruction, and discloses a three-dimensional digital human generation method and system capable of voice interaction. According to the invention, brand new speaking audios in different languages are automatically generated according to different languages of the input target text and the sampled human voice audios; the sequential stability and detail reduction capability of three-dimensional human motion are guaranteed by using multi-model joint estimation and a sequential loss function, and facial expression details and hand postures in the image can be accurately estimated. After the high-precision three-dimensional human body model is obtained through estimation, human body action and expression generation is carried out based on voice driving, accurate synchronization of actions and expressions generated through voice is achieved, and facial expression movement and body posture movement, namely a whole-body three-dimensional human body model, conforming to brand-new speaking audio are accurately generated; and finally, rendering the whole-body three-dimensional human body model into a real digital human capable of voice interaction by using a three-dimensional neural rendering model. According to the invention, the realization of single person picture input, high-precision three-dimensional digital person generation and voice interaction is facilitated.
Owner:NANJING UNIV OF SCI & TECH

Utilizing machine learning models to synthesize perturbation data to generate perturbation heatmap graphical user interfaces

The present disclosure relates to systems, non-transitory computer-readable media, and methods for embedding perturbation data via a machine learning model and filtering, aligning, and aggregating the embeddings to generate a genome-wide perturbation database for real-time generation of perturbation heatmaps. In particular, in one or more embodiments, the disclosed systems can receive a plurality of perturbation images portraying cells from a plurality of wells corresponding to a plurality of cell perturbations. Further, the systems can generate, utilizing a machine learning model, a plurality of well-level image embeddings from the plurality of perturbation images. Moreover, the systems can align, utilizing an alignment model, the plurality of well-level image embeddings to generate aligned well-level image embeddings. Additionally, the systems can aggregate, according to perturbations of one or more perturbation experiments, the well-level image embeddings to generate perturbation-level image embeddings. Furthermore, the systems can generate perturbation comparisons utilizing the perturbation-level image embeddings.
Owner:RECURSION PHARMACEUTICALS INC

Physiological monitoring soundbar

A soundbar for medical monitoring which may comprise a speaker, a sensor, and a hardware processor. The speaker can be configured to emit audio signals. The sensor can be configured to obtain sensor data relating to a physiology of a subject. The sensor can include a camera and the sensor data can include image data. The hardware processor can be configured to access the sensor data and determine a health status of the subject based on at least the sensor data.
Owner:MASIMO CORP

Patient Registration For Total Hip Arthroplasty Procedure Using Pre-Operative Computed Tomography (CT), Intra-Operative Fluoroscopy, and / Or Point Cloud Data

PendingUS20250384568A1Image enhancementImage analysisPelvic regionPatient registration
A system for computer assisted navigation during surgery includes a computer platform that operates to register a target surgical area of a patient. In certain cases, a process includes: obtaining a pre-op CT image of a pelvic region of a patient and intra-operatively obtaining a point cloud data about the pelvic region with a navigated instrument, generating a 3D bone model which excludes non-targeted area such as a femur, and then merging the 3D bone model to the point cloud to register the target surgical area.
Owner:GLOBUS MEDICAL INC

User interfaces for generating automatically-generated content

In some embodiments, an electronic device generates an automatically-generated visual media using one or more recognized concepts extracted from a prompt inputted by a user. The recognized concepts include personalized template subjects and / or prompt suggestions. While displaying the user interface including the recognized concepts, the electronic device receives one or more inputs to modify the recognized concepts. The electronic device generates multiple variants of the automatically-generated visual content using the one or more recognized concepts. The electronic device adds an automatically-generated visual content to a content entry field of an application, different than the automatically-generated visual media application, without opening the automatically-generated visual media application. The electronic device applies a visual effect to content that is generated using an artificial intelligence model. The electronic device displays visual information corresponding to an artificial intelligence model. The electronic device displays an animation including displaying a user interface with high dynamic range luminance.
Owner:APPLE INC

Patient Registration For Total Hip Arthroplasty Procedure Using Pre-Operative Computed Tomography (CT), Intra-Operative Fluoroscopy, and / Or Point Cloud Data

ActiveUS20250384570A1Image enhancementImage analysisPelvic regionPatient registration
A system for computer assisted navigation during surgery includes a computer platform that operates to register a target surgical area of a patient. In certain cases, a process includes: obtaining a pre-op CT image of a pelvic region of a patient and intra-operatively obtaining a point cloud data about the pelvic region with a navigated instrument, generating a 3D bone model which excludes non-targeted area such as a femur, and then merging the 3D bone model to the point cloud to register the target surgical area.
Owner:GLOBUS MEDICAL INC