Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

224 results about "Image content" patented technology

Object grasping method, computer-readable storage medium, electronic device

PendingCN122343449AFeature extractionRgb image
Embodiments of the present application disclose an object grabbing method, a computer readable storage medium and an electronic device. The method comprises: obtaining an RGB image collected by an image collection device for a target container in a grabbing work area, the RGB image comprising image content related to a plurality of objects stacked in the target container; calling a feature extraction model with the RGB image as input, and determining a target object to be grabbed and extracting feature information of the target object from the RGB image by the model; if the feature information indicates that a physical label is attached to the outer packaging of the target object, and the hardness information of the outer packaging material is not greater than the hardness information of the physical label, determining a region where the physical label is located as a grabbing region; determining a target end effector capable of performing a grabbing operation on the physical label; and controlling an intelligent grabbing device to perform a grabbing operation on the grabbing region of the target object by the target end effector, thereby realizing intelligent grabbing of the target object. This helps to improve the success rate of grabbing, and the grabbing strategy is more universal and compatible.
Owner:SHANGHAI HEMA ZHIYAN TECHNOLOGY CO LTD

A method and system for adaptive compression and breakpoint resume scheduling of wildlife images

This invention relates to the field of image processing and data transmission technology, specifically disclosing a method and system for adaptive compression and breakpoint resume scheduling of wildlife images. This invention acquires multi-dimensional data of wildlife images in real time, constructs a quantitative evaluation system, and dynamically optimizes data transmission and processing. First, raw data is collected from the wildlife image acquisition terminal, covering core information such as terminal identification, network status, and image content. Then, value assessment coefficients and breakpoint resume scheduling coefficients are calculated sequentially, and an image compression strategy is obtained accordingly. Adaptive compression and breakpoint resume transmission are then executed. Finally, verification, retransmission, and image enhancement optimization are performed on the server side, achieving closed-loop optimization across the entire link from terminal to server. This ensures high-definition transmission and reliable delivery of rare species images even in weak network environments, improving the efficiency and reliability of the monitoring system.
Owner:ZHEJIANG UNIHOME TECHNOLOGY CO LTD

Image processing method and device

An image processing method includes: in response to a target trigger operation, performing recognition processing on a target image, where the target image is an image input to a target application or an image currently displayed by the target application, and the target application is an application capable of performing the recognition processing or calling a target program file to perform the recognition processing; and outputting a recognition result for the target image, where the recognition result is capable of indicating source information of at least a portion of an image content of the target image.
Owner:LENOVO (BEIJING) LTD

A method and apparatus for image processing, an electronic device, and a storage medium

The application is suitable for the technical field of data processing, and provides a method and device for image processing, electronic equipment and storage medium, the method comprising: in response to a first operation initiated by a user, acquiring multiple images; for any image in the multiple images, determining an image order of the image in the multiple images according to a feature parameter of the image in at least one visual feature dimension; and generating an album based on the image order of the multiple images, wherein each image in the album is sorted based on the image order. In the embodiment of the application, the order of the images in the album is determined according to the difference degree between the visual feature dimensions, the similarity of the image content between adjacent images is improved, the continuity of switching between different images in the album by the user is improved, the fragmentation of the picture is reduced, and the use experience of the user is improved.
Owner:HUAWEI TECH CO LTD

Image content review method and system based on multi-layer scene graph structure

This invention discloses an image content review method and system based on a multi-layer scene graph structure, belonging to the field of image content review technology. The method includes: segmenting an image to obtain entities and their attribute descriptions; clustering entities into several entity semantic clusters based on the entity's visual image features, the textual semantic features of the attribute descriptions, and the normalized spatial coordinates of the entity's center point, and analyzing the relationships between entity pairs within each cluster; identifying higher-order relationships between cluster pairs based on the overall semantic summary of each entity semantic cluster, and generating the overall narrative intent of the image; performing role recognition and referential analysis on the text in the image to obtain text-image relationships; integrating the generated data into a hierarchical graph structure, establishing thought chain prompts, and inputting it into a large language model for comprehensive reasoning. This invention enables a layer-by-layer, in-depth understanding of image content from local entities to global intent, accurately identifying complex and non-compliant semantics in images to be reviewed.
Owner:ZHEJIANG UNIV

A screenshot method, device, apparatus and storage medium

This application provides a screenshot method, apparatus, device, and storage medium. The screenshot method can reconstruct a three-dimensional game scene matching the game screenshot captured by the player in the screenshot editing interface. It supports the player to adjust the reconstructed three-dimensional game scene, thereby generating a target screenshot matching the adjusted three-dimensional game scene. This allows the player to directly modify the actual image content of the game screenshot, effectively improving the fault tolerance of the game screenshot function and the editing freedom of the game screenshot.
Owner:NETEASE (HANGZHOU) NETWORK CO LTD

Content publishing method, apparatus, device, storage medium, and program product

The application relates to a content publishing method and device, computer equipment, a computer readable storage medium and a computer program product. The method comprises the following steps: displaying an image shooting interface; the image shooting interface comprises image content collected by a shooting component and shooting guide information; wherein the shooting guide information is determined based on first publishing content, and the first publishing content is determined based on the image content; in response to a shooting operation performed on the image shooting interface, determining a target image to be published; in response to a publishing operation performed on the target image to be published, publishing second publishing content; wherein the first publishing content and the second publishing content are content published on a content platform or distributed to a browsing account of the content platform based on a content distribution algorithm. Therefore, the obtained target image and the second publishing content are more in line with the aesthetics of the content platform and have more potential for dissemination, thereby improving the creation willingness of the user and the output efficiency of high-quality content.
Owner:XINGIN INFORMATION TECH (SHANGHAI) CO LTD

A visual multi-modal image detection method for AI training

The application discloses a visual multi-modal image detection method for AI training, and relates to the technical field of image detection.The method effectively fuses the data features of heterogeneous sensors through a cross-modal attention mechanism, significantly improves the detection robustness in complex environments, and realizes automatic conversion from visual features to text descriptions with the help of a multi-agent collaboration framework dynamically generated by a large language model.Combining a dynamic prompt mechanism and iterative self-consistent verification, the method avoids the deviation and omission of artificially designed prompt words, significantly improves the accuracy of anomaly detection and the generalization ability for unknown defects, automatically generates customized query prompts according to image content, and iteratively optimizes through gradient update.When the detection scene or product line changes, the method can quickly adapt itself, significantly reducing maintenance costs.
Owner:SAIER DIGITAL (BEIJING) TECH CO LTD

Projection device and operating method thereof

A projection device that is movable includes a projection portion that projects image contents. The projection device identifies, based on a call command for the projection device, a target location at which a projection is performed, identifies first image content corresponding to the target location, obtains setting information for projecting the first image content, moves to the target location while performing a presetting operation based on the setting information, and controls the projection portion to project the first image content based on the presetting operation being completed.
Owner:SAMSUNG ELECTRONICS CO LTD

A method, system, and storage medium for auditing instrument data records

This invention discloses a method, system, and storage medium for verifying instrument data records. Based on a received data verification command, each record in the instrument data record table is sequentially retrieved. The corresponding screen layout information is retrieved from the instrument information database according to the instrument model. Then, the display features in the screen image are identified and analyzed, and compared with the screen area shape and measurement data layout information stored in the instrument information database. If they are not similar, an instrument model entry error message is triggered. If they are similar, the identified instrument measurement parameters are compared with the recorded parameters in the data record. If they are not the same, a measurement parameter entry error message is triggered. By using screen image content to automatically verify and verify the instrument model and measurement parameters in electronic experimental records, the inefficiency and error-prone nature of existing manual verification methods are overcome.
Owner:MINGDU ZHIYUN (ZHEJIANG) TECH CO LTD

Distraction-free copilot entertainment display for a vehicle

This invention relates to a passenger-side entertainment display for a vehicle. The display includes a display surface for simultaneously generating personalized image content and a beam of light transmitting the image content for the driver and passenger, respectively, in their respective non-intersecting, nested display surface segments. The display also includes a cylindrical lens layer disposed on the display surface and configured such that the two beams exit the display in different directions, thereby reaching only the respective eye movement range of the driver or passenger. Here, the display is constructed and / or configured to display a distraction-free image to the driver, at least in its distraction-free operating mode, having a significantly lower pixel count, resolution, and / or sharpness than a high-quality entertainment image displayed to the passenger simultaneously.
Owner:BAYERISCHE MOTOREN WERKE AG

Display material automatic recognition method and system

ActiveCN117544758BSteroscopic systemsVirtual targetComputer graphics (images)
This invention provides an automatic display material identification method and system. Based on the real environment and historical display records of a head-mounted display device, it generates a set of real target display elements and a set of virtual target display elements, providing display material sources at both the real and virtual environment levels for subsequent adjustments to the displayed image content. It identifies permissible display change areas within the current display screen of the head-mounted display device, accurately locating subsequent content adjustments to ensure that the adjusted content does not affect the original display state and prevents interference. Furthermore, based on display control commands, it determines the set of display elements to be filtered, and identifies display elements that meet preset shape and contour conditions, ensuring that the content of the filtered display elements matches actual needs. Finally, after adjusting the visual characteristics of the filtered display elements, it integrates them into the screen area to improve the visual effect.
Owner:HUIZHIAN INFORMATION TECH CO LTD

Generative model reasoning using internal image and video generation

Implementations disclosed herein are directed to generative model (GM) reasoning that generates image(s) / video(s) as part of a chain-of-thought (CoT) in response to receiving certain user inputs that do not request any generative image content and / or generative video content. Processor(s) of a system can: receive user input, generate responsive content that is responsive to the user input, and cause the responsive content to be rendered. In generating the responsive content, the processor(s) can process, using the GM input, initial GM input to generate initial GM output, the initial GM input including at least the user input, and the initial GM output including at least a generative / video. In generating the responsive content, the processor(s) can further determine, based on processing at least the generative image / video, the responsive content. Thus, the processor(s) can generate the image(s) / video(s) to reason about the user input and / or the responsive content in these modalities.
Owner:GOOGLE LLC

System and method for manufacturing and maintenance

ActiveUS12669331B2Multi-imageSystem maintenance
A system and method for inspection maintenance and / or diagnosis of a variety of workpieces is provided. The system serves workers working on a workpiece, inspectors who are distal from the workers and / or can be used for remote training or for advanced diagnosis and / or repair. The system preferably includes a template of a set of one or more predefined required images of a workpiece required by an inspector to perform their inspection or diagnosis. The set of predefined required images is provided to the worker. The worker captures the images with an appropriate workpiece data capture device and provides them to the inspector for review. The inspector examines the provided images and either approves the workpiece based on their content, requests additional images for further examination and / or provides annotations and other information to the worker to address identified issues. The system maintains a database of all images and information.
Owner:INTERAPTIX INC

A small sample semantic segmentation method and system fusing class label semantics

ActiveCN118587440BCharacter and pattern recognitionSpeech segmentationImaging processing
The application relates to the field of image processing, and proposes a small sample semantic segmentation method and system fusing class label semantics, wherein a prior information generation module is designed, multi-modal data fusion of image data information and text data information serving as class labels is realized, a semantic segmentation model can more accurately understand image content, a multi-scale fusion module is further designed, original detail information of an image is further protected, the calculation performance and channel fusion capability of the semantic segmentation model are greatly improved, the parameter quantity of a decoder is greatly reduced while ensuring the speech segmentation precision, and the accuracy of target recognition and target positioning of semantic segmentation is greatly improved.
Owner:JIANGXI NORMAL UNIV

Test system and method for unmanned device controller

ActiveCN116643552Breduce dependenceGuaranteed attribute informationComputer graphics (images)Engineering
The application discloses a test system and method for an unmanned device controller, which comprises an image acquisition device configured to acquire first image data, and send the first image data to a video injection core board, wherein the first image data comprises at least first image attributes and first image content; an image storage device configured to send second image data to the video injection core board, wherein the second image data comprises second image content; and the video injection core board configured to analyze the first image attributes to obtain analyzed first image attributes, replace the first image content with the second image content to obtain target image data, and send the target image data to the unmanned device controller, wherein the target image data comprises at least the analyzed first image attributes and the second image content; and the unmanned device controller configured to output a control instruction based on the received target image data, thereby improving the test accuracy.
Owner:BEIJING JINGWEI HIRAIN TECH CO INC

Graphical user interface for generating learning plans for electronic devices

1. Name of the product in this design: Graphical User Interface for Generating Learning Plans for Electronic Devices. 2. Purpose of this design: An electronic device. 3. The key design feature of this product is its graphical user interface. 4. The image or photograph that best illustrates the design's key features: the front view. 5. The display screen panel adopts the conventional design, omitting the rear view, left view, right view, top view, and bottom view. 6. Purpose of the graphical user interface: for generating learning plans. 7. Human-computer interaction method of graphical user interface: Click "1V1 Customized Plan" in the main view to jump to interface change state diagram 1; select the subject and difficulty at the top of interface change state diagram 1, and select the option for encountering problems, then click "Generate Customized Plan" at the bottom of the interface to jump to interface change state diagram 2; after the content of interface change state diagram 2 is loaded, jump to interface change state diagram 3; select the course in interface change state diagram 3 to jump to interface change state diagram 4; select the learning time in interface change state diagram 4, click "Join Plan" at the bottom of the interface to jump to interface change state diagram 5, and generate the learning plan. 8. Other situations requiring explanation: In the graphical user interface, "X" can be a number, Chinese character, letter, or symbol, etc.; the gray area in the "box" at the lower left of the main view can be image content; the gray area in the lower middle part of the interface change state diagram 3 within the "box" can be image content; the light gray and dark gray areas on the right side of the interface change state diagram 4 can be image content; the gray area in the "box" at the lower right of the interface change state diagram 5 can be image content; the interface usage reference diagrams 1-6 are the main view and the interface change state diagrams 1-5, respectively.
Owner:GUANGDONG XIAOTIANCAI TECH CO LTD

Method for displaying image content on display surfaces arranged at different distances and display device

A method for operating a display device (102) for a vehicle comprises perspective-adjusting first image content to be displayed on a first display surface (106) using a first distance between an occupant's eye position (104) and the first display surface (106) to obtain perspective-adjusted first image content with a smooth transition to second image content to be displayed on a second display surface (108), and of second image content to be displayed on the second display surface (108) using a second distance between the eye position and the second display surface (108) to obtain perspective-adjusted second image content with a smooth transition to the first image content to be displayed on the first display surface.The perspective-adjusted first image content is displayed on the first display area (106) and the perspective-adjusted second image content is displayed on the second display area (108).
Owner:ROBERT BOSCH GMBH

Scene library based vehicle visual-only localization method

This application provides a vehicle pure vision localization method based on a scene library, relating to the field of vision localization. The method includes: performing early vision localization of the vehicle using a high-precision map to obtain an optimized image pose; performing quality detection on the image pose based on a deep learning model, and storing qualified image poses into a scene library; recalling matching historical images from the scene library based on the image content and rough position information of the current onboard camera image; and using the recalled historical images to assist or replace the high-precision map for pose estimation to obtain the vision localization result. The technical solution of this application forms a data closed loop in which image data accumulation and localization accuracy mutually promote each other, enabling the sustainable operation of the vision localization process.
Owner:JISHU TECHNOLOGY (WUHAN) CO LTD

A method for reducing power consumption of an led television

PendingCN122293811AImaging qualityMotion vector
This invention relates to the field of display control technology, specifically a method for reducing the power consumption of LED televisions. The method acquires illuminance using an ambient light sensor with integrated infrared suppression function, and generates effective ambient light feature values ​​through time-delay nonlinear hysteresis filtering; calculates the average image level and brightness histogram of video frames to generate image content features; calculates first and second backlight reference values ​​respectively, and uses a cross-state weight allocation strategy to dynamically map the target backlight adjustment value. Simultaneously, it extracts video motion vector information to bypass redundant calculations of static images; employs a nonlinear cosine gradient mathematical function to achieve a smooth transition of the physical backlight, and simultaneously uses a weighted piecewise linear enhancement curve for pixel-level visual lossless color compensation. This invention achieves a multi-dimensional dynamic balance between image quality protection, extreme energy saving, and computational optimization in complex scenarios.
Owner:GUANGZHOU CHAODING ELECTRONICS MANUFACTURING CO LTD

Image generation method and device based on spatiotemporal data interaction and electronic equipment

The application provides an image generation method and device based on space-time data interaction and electronic equipment, relates to the technical field of computer vision, and aims to realize efficient image generation. The method comprises the following steps: acquiring an image representation sequence of the last time; the image representation sequence comprises a visible representation and a mask representation, the mask representation represents unknown image content, and the visible representation represents known image content, which is used for providing image information for the mask representation to infer unknown image content; performing self-attention interaction-based coding on the visible representation to obtain a visible representation feature; performing cross-self-attention interaction-based decoding on the mask representation according to the visible representation feature to obtain a new image representation sequence; performing image representation sequence iteration generation for multiple times according to the above steps; and generating an image according to the new image representation sequence in the case that the new image representation sequence does not contain the mask representation.
Owner:TSINGHUA UNIVERSITY

Image content processing method and electronic device

PCT designated stageWO2026138892A1Computer graphics (images)Imaging quality
An image content processing method and an electronic device, used for improving the processing effect of image content processing of electronic devices, thereby improving the use experience of users. In the method, when a scenario involving image content processing is detected, the quality of an image used for image content processing is analyzed, and when it is detected that the quality of the image is low, prompt information can be displayed. In this way, guidance is performed on the basis of image quality, so that a user can be prompted to improve the photographing mode of a camera, thereby improving the quality of the image used for image content processing, and further improving the processing effect of image content processing.
Owner:HUAWEI TECH CO LTD

Image display control device, image display system, and image display control method

PendingJP2026110697ARadiologyNuclear medicine
This device detects abnormalities in the brightness control unit, which controls the brightness of multiple light sources, and is installed in an image display control device. [Solution] The image display control device includes: a brightness control unit that generates backlight control information used to control multiple light sources included in the backlight based on first image information indicating an input image; a pixel compensation unit that corrects the pixel values ​​included in the first image information based on the brightness of the multiple light sources to generate second image information indicating an output image; a first statistical acquisition unit that acquires first statistical data of the pixel values ​​included in the first image information; a second statistical acquisition unit that acquires second statistical data of the brightness values ​​of each light source included in the backlight control information; and an abnormality detection unit that detects an abnormality in the backlight control information generation process in the brightness control unit when the second statistical data does not fall within the range between the upper and lower limits determined by the first statistical data and no switching of image content has occurred.
Owner:SOCIONEXT INC

Adjusting image content to improve user experience

Various implementations disclosed herein include devices, systems, and methods that adjust image content to reduce motion sickness during an experience. For example, an example process may include determining a first zone and a second zone of a display, generating images of a three-dimensional (3D) environment, identifying image content of each of the images corresponding to the second zone of the display, and reducing at least one of contrast or spatial frequency of the image content of each of the images corresponding to the second zone of the display.
Owner:APPLE INC

Multistage search and results utilizing prestored image assets and adaptive caching to minimize machine learning and artificial intelligence data and energy costs

A data processing system implements an image generation system configured to operate in a first generation mode providing requested image contents based on prestored image assets without using an AI model to generate the requested image contents and a second generation mode generating the requested image contents using the AI model; receiving a first textual prompt first image content; analyzing the first textual prompt to determine whether the image generation system includes prestored image content that satisfy the prompt; operating the image generation system in the first generation mode to provide the first image content based on the first textual prompt based on the prestored image assets responsive to the image generation system including prestored image content that satisfies the first textual prompt; otherwise operating the image generation system in the second generation mode to generate the first image content; and providing the first image content to a client device.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Methods and systems for measuring brain reflexes, as well as their involvement and lifestyle modulatory effects.

PendingJP2026520105AInput/output for user-computer interactionPsychotechnic devicesAuditory stimuliAuditory startle
This disclosure relates to a method and system for measuring emotional engagement in response to auditory and / or visual stimuli, for example, using auditory startle response, prepulse suppression, spontaneous blinking, and / or eye movements. The method includes displaying a first video or image content. A second video is captured, including one or more of the individual's eyes, while the individual is viewing the first video or image content, and while the second video is being captured, the individual is exposed to one or more visual and / or auditory stimuli. The method may include determining a value representing eye closure based on the second video. The method may include calculating a metric based on these values.
Owner:BLINKLAB LTD

Image content analysis method and image analysis apparatus

An image content analysis method is applied to an image analysis apparatus and includes acquiring an image, utilizing an edge detection technology to compute an edge density of the image, utilizing a texture detection technology to compute a texture richness of the image, and analyzing the edge density and the texture richness to generate a richness score of the image.
Owner:VIVOTEK INC

Processing method and model training method, device and system of a picture-text reader

The application relates to the technical field of data processing, and discloses a processing method and a model training method of a picture-text reader, and a device and a system thereof. The method comprises the following steps: performing standardized analysis on a target file based on a preset multi-format analysis model in an AI NAS device to obtain standardized content; performing text analysis on the standardized text content to obtain target text content; performing image processing on the standardized image content to obtain target image content; displaying the target text content and / or the target image content based on a pre-constructed cross-device responsive interface; and updating the reading data of a target user for the target text content and / or the target image content based on the AI NAS device. It can be seen that, by implementing the application, the analysis and rendering tasks can be migrated to the AI NAS device, multi-format analysis is supported, the analysis efficiency can be improved, intelligent text analysis and image processing can be realized, the reading content can be optimized, multi-device interface display and cross-device synchronization are supported, the reading fluency of a user is improved, and the user experience is improved.
Owner:SHENZHEN GREEN CONNECTION TECH CO LTD