Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

103results about "Carrier editing" patented technology

Speaker thumbnail selection and speaker visualization in diarized transcripts for text-based video

Embodiments of the present invention provide systems, methods, and computer storage media for selection of the best image of a particular speaker's face in a video, and visualization in a diarized transcript. In an example embodiment, candidate images of a face of a detected speaker are extracted from frames of a video identified by a detected face track for the face, and a representative image of the detected speaker's face is selected from the candidate images based on image quality, facial emotion (e.g., using an emotion classifier that generates a happiness score), a size factor (e.g., favoring larger images), and / or penalizing images that appear towards the beginning or end of a face track. As such, each segment of the transcript is presented with the representative image of the speaker who spoke that segment and / or input is accepted changing the representative image associated with each speaker.
Owner:ADOBE INC

Sign language translation

A method may include obtaining, at a system, audio during a communication session that includes a first device and a second device. In these and other embodiments, the audio may originate at the second device. The method may further include selecting, automatically and independently by the system based on one or more features of the communication session, a first translation process instead of a second translation process to translate between sign language and language data during the communication session. The method may also include using the selected one of the first translation process and the second translation process to generate a video that includes sign language content based on the audio.
Owner:SORENSON IP HOLDINGS LLC

Sign language translation

A method may include obtaining a data stream with language data for translation by a translation system configured to translate between sign language and other language forms. The method may further include storing one or more portions of the data stream and directing a current portion of the data stream to the translation system. The method may also include providing the stored one or more of the portions of the data stream to the translation system and translating, by the translation system, the current portion of the data stream using the current portion of the data stream and the stored one or more of the portions of the data stream provided to the translation system.
Owner:SORENSON IP HOLDINGS LLC

Redacting videos and images to reduce PTSD and other negative effects on reviewers

An aspect of the current invention relates to a method to reduce the emotional impact on a viewer while allowing evaluation of graphic pornographic, pedophilic and / or violent content. In some embodiments the system has various functions that reduce the emotional impact and / or stress inducing factors of a media (e.g., an image and / or a video). For example, the functions may include changing and / or blurring certain parts of the image and / or changing the color and / or contrast of the image and / or presenting part of the media while withholding other parts and / or changing the media in different ways when it is repeatedly presented e.g., to reduce the effect of repeated seeing same distressing content.
Owner:NETABPARK

Optimal mapping of immersive video presentations between devices with different form factors

Mapping a source video stream to a screen of a target device includes extracting separate objects from the source video stream, identifying a subset of the objects that correspond to a presentation area, and creating a target video stream that is displayed on the target device by arranging the separate objects to enhance the presentation area on the screen of the target device based on a form factor, screen resolution, and / or aspect ratio of the target device. The presentation area may include a presenter and presentation materials. Creating the target video stream may include scaling different ones of the objects. Different scaling factors may be applied to different objects to enhance the presentation area on the screen of the target device. A separate video image corresponding to a linear mapping of the source video stream onto a target device may be provided in addition to the target video stream.
Owner:MMHMM INC

Reproduction information generating device, video editing device, and video editing program

To automatically edit a moving image attractively.SOLUTION: A reproduction information generation device for generating a moving image based on a plurality of image materials or information necessary for reproduction of the moving image comprises a retrieval unit which retrieves a plurality of cuts used for the moving image from the plurality of image materials using information on a theme of the moving image to be reproduced.SELECTED DRAWING: Figure 2
Owner:NIKON CORP

Video game engine assisted virtual studio production process

A production process involves predetermined number of cameras simultaneously filming a background at predetermined angles, and filming actors in a studio with the same number of cameras and the same angles, used in conjunction with a virtual studio system. In studio, the actors perform before a green screen and the virtual studio system composites the actors onto the background in real-time. Camera tracking allows the in-studio cameras to pan, tilt, focus, zoom, and make limited other movements as the virtual studio system adjusts display of the background in a corresponding manner, resulting in a realistic scene without transporting actors and crew to the background location.
Owner:PIPHER TIM

Audio generation method and system

To provide an audio generation method and a system.SOLUTION: A method for generating audio assets includes a step of receiving a plurality of input audio assets, a step of converting each of the input audio assets into an input graphic representation, a step of generating an input image from each of the input graphic representations, a step of feeding the input image to a generative model and extracting an output graphic representation from each of the output images in order to learn the generative model and generate one or more output images including the output graphic representations, and a step of converting the output graphic representation into an output audio asset.SELECTED DRAWING: Figure 2
Owner:SONY COMP ENTERTAINMENT EURO LTD

Electronic apparatus capable of performing synchronization between document and voice through matching between voice and editing object, and operation method thereof

Disclosed are an electronic apparatus capable of synchronizing between documents and voices through matching between a voice and an editing object, and operating method thereof. The present invention provides an electronic apparatus capable of synchronizing between documents and voices through matching between a voice and an editing object, and an operating method thereof, when a user edits a document while recording the voice to support the user to view editing objects inserted into the document in chronological order at the time when recording the voice.
Owner:FLEXCIL INC

Information recording apparatus, information recording method, and information recording program

To provide an information recording device that can record desired video data and voice data, and that can leave necessary voice data while paying attention to privacy protection.SOLUTION: Video data and voice data are acquired from a camera C and a microphone MC provided in a vehicle, event information indicating that an event such as rapid steering occurs in the vehicle is acquired, and after the acquired video data and voice data are recorded, the voice data recorded being associated with the event information for the recorded voice data is left in a recording unit R.SELECTED DRAWING: Figure 4
Owner:PIONEER IP

Work video editing system, video editing program and recording medium

To easily divide the entire video data in which a series of work is recorded into a plurality of parts for each work process, and to store the divided video data in chronological order. [Solution] The system comprises a memory unit 5 that stores a series of work processes including multiple element work processes that have been filmed in advance for one person, one machine, or one item as one work video file; an identification unit (control unit 10) that analyzes the video data that makes up one work video file stored in the memory unit 5 and automatically identifies the start timing of the multiple element work processes; a division unit (control unit 10) that divides the entire work video file into multiple element work period videos while positioning the video period from the previous start timing to the current start timing identified by the identification unit as one element work period video; and an editing unit (control unit 10) that associates each of the divided element work period videos with a work procedure and stores them as new work video files in the memory unit.
Owner:EPISOTECH CO LTD

Adaptive Audio Mixing

A system, apparatus and method for performing adaptive audio mixing are disclosed. A trained neural network dynamically selects and mixes pre-recorded human-composed music stems arranged in a mutually compatible set. Stem and track selection, volume mixing, filtering, dynamic compression, acoustic / reverberation characteristics, segue, tempo, beat matching and cross-fade parameters generated by the neural network are inferred from game scene characteristics and other dynamically changing factors. The trained neural network selects pre-recorded stems of artists and mixes the stems in a unique way in real time to dynamically adjust and change background music based on factors such as the game scenario, the player's unique storyline, scene elements, player's profile, interests, performance, adjustments made to game controls (e.g., music volume), number of viewers, comments received, player popularity, player's native language, player's influence and / or other factors. The trained neural network generates unique music that dynamically changes according to real-time conditions.
Owner:ADVANCED MICRO DEVICES INC +1

Interact with semantic video clips using interactive tiles.

Embodiments of the present invention relate to interacting with semantic video segments via interactive tiles. The embodiments relate to interactive tiles representing video segments that are segments of a video. In some embodiments, each interactive tile represents a different video segment from a specific video segment (e.g., a default video segment). Each interactive tile includes a thumbnail (e.g., the first frame of the video segment represented by the tile), a script from the beginning of the video segment, a visualization of detected faces in the video segment, and one or more faceted timelines that visualize the categories of detected features (e.g., visualizations of detected visual scenes, audio classifications, and visual artifacts). In some embodiments, interacting with a specific interactive tile can navigate to the corresponding portion of the video, add the corresponding video segment to the selection, and / or scan through the tile thumbnail.
Owner:ADOBE INC

Social media web dialog proxy

Examples are provided that relate to implementing actions on social media network content based on natural language input. One aspect includes a computing system configured to implement a social media network, the computing system including one or more processors and a storage device, the storage device including instructions executable to receive user input from a dialog agent, the user input including a natural language description of a request for an action of a content item, the dialogue agent is configured to participate in a dialogue using at least a language model and generate cues to the language model based on at least a user input. The instructions are further executable to input the cue word to the language model to generate an output describing an operation for implementing the action, invoke a backend service of the social media network to execute the command to implement the operation, and output a result of executing the command.
Owner:FACE CUTE CO LTD

Remote capture and viewing of workstation activity

Devices, systems, and methods for remote capture and viewing of workstation activity are provided. The system may comprise a worker computer, a data-streamer, a segmenter, a server, a requester, a stitcher, and a watcher computer. The segmenter creates segment files associated with segments of data-streams of captures of workstation activity. The stitcher creates data-streams of captures of workstation activity from such segment files. Gains in speed and efficiency are achieved by segmenting data-streams into segments for transmission, storage, and processing.
Owner:SMART TECH INC (CA)

Method and apparatus for processing video, and display device

Provided is a method for processing a video. The method for processing the video can segment an initial video into a plurality of video clips, and calculate color temperatures of the plurality of video clips. For each target video clip in the plurality of video clips, the method can determine a color temperature adjustment coefficient of the target video clip based on a color temperature of the target video clip and a color balance correction coefficient thereof, and adjust a color temperature of at least one video frame in the target video clip by using the color temperature adjustment coefficient.
Owner:BEIJING BOE TECH DEV CO LTD

Information processing device and information processing method

The present invention enables emotion data, which represents user emotion for each scene of moving image content, to be effectively used. A representative emotion scene is extracted by an extraction unit on the basis of emotion metadata having user emotion information for each scene of the moving image content. On the basis of the extracted representative emotion scene, playing back a portion of the moving image content or editing for taking out a portion of the moving image content can be effectively performed. For example, the extraction unit extracts the representative emotion scene on the basis of the type or degree of user emotion.
Owner:SONY GROUP CORP

Searching editing components based on text using a machine learning model

The present disclosure describes techniques for searching editing components based on text using a machine learning model. A plurality of visual embeddings indicative of a plurality of visual editing components is acquired by the machine learning model. The plurality of visual embeddings indicative of the plurality of visual editing components is projected into a common space by a first sub-model of the machine learning model. A text query input is received by a user. A text embedding indicative of the text query is generated. The text embedding is projected into the common space by a second sub-model of the machine learning model. At least one visual editing component among the plurality of visual editing components is determined based on the projected text embedding and the plurality of projected visual embeddings in the common space. Information indicative of the at least one visual editing component is displayed via a user interface.
Owner:LEMON INC(GB)

Network system, server apparatus, and control method thereof

A network system providing a service of uploading moving image data includes one or more memories, and one or more processors in communication with the one or more memories, wherein the one or more processors and the one or more memories are configured to acquire predetermined information when accepting upload of moving image data from a moving image data transfer apparatus, and generate moving image data for display of the moving image data at a time of uploading the moving image data from the moving image data transfer apparatus in accordance with the predetermined information.
Owner:CANON KK

Information processing device and information processing method

The present invention enables user's emotion for each scene of moving image content to be effectively used. The present invention generates, on the basis of user's emotion and video quality for each scene of moving image content A, correlation data obtained by associating the user's emotion and the video quality with each other. The present invention predicts, on the basis of video quality for each scene of moving image content B and the correlation data obtained by associating the user's emotion and the video quality related to moving image content A, the user's emotion for each scene of moving image content B. For example, the predicted user's emotion for each scene of moving image content B is displayed and used.
Owner:SONY GROUP CORP

Video generating method, electronic device and non-transitory computer-readable storage medium

A video generating method, an electronic device and a non-transitory computer-readable storage medium are provided. The video generating method includes: determining a reference video identifier specified by a current video editing task; determining reference video editing template information corresponding to the current video editing task based on the reference video identifier; and conducting video editing on a target multimedia material specified by the current video editing task based on the reference video editing template information, to obtain a target video.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Method and device for editing image in electronic device

An electronic device comprises a memory and a processor connected to the memory. The memory can store instructions such that, when executed by the processor, causes the electronic device to: use information related to original data and a hash function to generate a value (unique value) based on edited data obtained by modifying the original data being stored in the memory; use the generated value to generate an original storage route for the original data; and generate original meta information including information about the original storage route and the hash function.
Owner:SAMSUNG ELECTRONICS CO LTD

Audio output system and audio content output device

InactiveEP4521332A4Musical toysCarrier editing
An audio output system comprising: a key device including a communication tag; a main device configured to store a first playlist and to output audio content related to the communication tag based on the first playlist; a server configured to store a second playlist and to manage the first playlist based on the second playlist; and a user device configured to update the second playlist by accessing the second playlist, wherein if the second playlist is updated, the server provides a playlist update message to the main device, and wherein the main device updates the first playlist by comparing the first playlist and the updated second playlist with each other.
Owner:KOKOZI CO LTD

Information processing device, method, and program

PendingUS20260172630A1Image analysisCarrier editing
There is provided an information processing device, method, and a program enabling a part of a music piece that each performer is in charge of to be easily recognized. The information processing device includes: a video generation unit that generates, on the basis of text information including text of sound from a plurality of sound sources and one or a plurality of videos including the at least one sound source as an object, a presentation video in which with a color representing the sound source of sound that is being emitted, a part of the text corresponding to the sound that is being emitted, a figure, or a character sequence representing the sound source is displayed. The present technology can be applied to a video processing system.
Owner:SONY GROUP CORP

Dubbing interaction method and apparatus, computer device, and storage medium

A dubbing interaction method and an apparatus, a computer device, and a storage medium are provided. The method includes: showing character information associated with a text to be dubbed; in response to a selection operation for first character information, obtaining a first dubbing audio of a first user, and associating the first dubbing audio with the first character information, wherein the first dubbing audio is a dub for a text fragment of the text to be dubbed that is associated with the first character information; and obtaining an aggregated dubbing audio corresponding to the text to be dubbed based on the first dubbing audio associated with the first character information and an obtained second dubbing audio respectively corresponding to second character information associated with the text to be dubbed, and showing a first audio identifier corresponding to the aggregated dubbing audio.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

Content protection processing method

To provide a broadcast receiver capable of executing a function with a higher added value.SOLUTION: A broadcast receiver 20100 includes a digital broadcast reception section, a network communication section, a data string conversion section capable of mixing a data string of contents received by the digital broadcast reception section with a data string acquired by the network communication section to generate a new data string, a digital interface capable of outputting the contents to an external device, and a control section. The control state in which the content is output to the external device includes a first control state in which a new data row generated by mixing is output to the external device having the network communication function, a second control state in which a data row in a state where the mixing processing is not performed is output to the external device having the network communication function, and a third control state in which a new data row generated by mixing is output to the external device not having the network communication function.SELECTED DRAWING: Figure 29A
Owner:MAXELL LTD