Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

113 results about "Vocal music" patented technology

Vocal music is a type of music performed by one or more singers, either with instrumental accompaniment, or without instrumental accompaniment (a cappella), in which singing provides the main focus of the piece. Music which employs singing but does not feature it prominently is generally considered instrumental music (e.g. the wordless women's choir in the final movement of Holst's The Planets) as is music without singing. Music without any non-vocal instrumental accompaniment is referred to as a cappella.

VEM-Token beat capture and alignment model construction method

The invention discloses a VEM-Token beat capture and alignment model construction method, which is a deepening innovation that a vocal music file is segmented into VEM-Token lexical elements by adopting music beats based on a VEM-Token vocal music emotion multi-modal model method. The core of the method is to establish a rhythm model, a rhythm capture model and a rhythm alignment model of a vocal music file, the rhythm model separates singing sound, accompaniment sound and emotional fluctuation from a sample vocal music file through multiple filters and captures a start point and an end point of a rhythm in a frequency spectrum format file, and the rhythm alignment model performs rhythm alignment on the vocal music file through a start point fine tuning model and an end point fine tuning model. And the user imitation file and the sample file are enabled to complete beat alignment. Models including a rhythm basic model, harmonic impact, joint learning, harmonic frequency layering, dynamic time warping and the like are adopted to capture rhythms, and models including the basic model, starting point fine tuning, terminal point fine tuning, whole-course alignment verification, a rhythm editor, rhythm free playing, repeated alignment, a communication interface protocol and the like are adopted to construct. Therefore, the method is suitable for accessing an Agent music agent and an AI music application.
Owner:GREATER BAY AREA STAR BIOTECH (SHENZHEN) CO LTD

Vocal music training vowel pronunciation quality evaluation method based on auditory and visual spatio-temporal feature fusion

The invention provides a vocal music training vowel pronunciation quality evaluation method based on auditory and visual spatial-temporal feature fusion, and the method comprises the steps: collecting vowel pronunciation audio signals and corresponding videos of a singer, and constructing a multi-modal data set; generating a fractional order Mel spectrogram for the audio signal through short-time fractional order Fourier transform of an adaptive order; extracting time sequence features and spatial features of the fractional order Mel spectrogram, and fusing the time sequence features and the spatial features through a gating mechanism to generate audio spatio-temporal features; face visual features in the video are extracted and fused with the audio spatio-temporal features through a cross attention mechanism, and the cross attention mechanism is integrated with a periodic modeling network; the fused features are input into a classifier, a dynamic weight multi-mode cosine loss function training model is adopted, the dynamic weight multi-mode cosine loss function dynamically adjusts the sample weight through a confusion matrix, and the weight is increased for the samples with classification errors based on the historical frequency mistaken division times of the samples; and outputting a pronunciation quality evaluation result.
Owner:FUZHOU UNIV

VEM-Token emotion synchronization function hierarchical fusion method

A VEM-Token emotion synchronization function hierarchical fusion method is different from a traditional NLP-Token method, an emotion synchronization function VEM-sync is innovated for the first time, the emotion synchronization function VEM-sync is synchronized with a VEM-Token sequence of music beats, multiple high-dimensional emotions are directly described by adopting mathematical languages, and therefore deviation of discretized natural language characters on description of a high-dimensional emotion analog quantity function is avoided, and the emotion synchronization effect is improved. The method comprises the steps of defining VEM-sync and synchronous content, defining rhythm attributes, emotion attributes and emotion functions, adopting one or combination of multi-layer weighted scanning, a recurrent neural network, a long and short-term memory network, a self-attention mechanism and an RAG network generated by retrieval enhancement, and performing hierarchical fusion calculation to output a time emotion function of a vocal music file. According to the phonetic function or the emotional function, the effects of emotional texts, emotional expressions, emotional languages and emotional multi-dimensional animations are directly driven by crossing discrete text tokens, a model context protocol (MCP) and a function calling function are supported, and copyright management and encryption and decryption or interfaces are provided.
Owner:GREATER BAY AREA STAR BIOTECH (SHENZHEN) CO LTD

Sound model generation method and system for vocal music pronunciation training

The invention relates to the technical field of speech recognition, in particular to a sound model generation method and system for vocal music pronunciation training, and the method comprises the steps: collecting a preset number of sound signals of a trainer for different contents, and collecting a preset number of sound signals for each content; for any content and any modal component of any sound signal, obtaining each contrast component of the any modal component, and obtaining a feature difference array of the any modal component; obtaining a difference characteristic value between the any modal component and each contrast component thereof, obtaining an influence characteristic value and a reconstruction weight of the any modal component, and further obtaining each reconstructed sound signal; and obtaining a sound characteristic spectrum of each sound signal of the trainer, training a sound model, and evaluating the quality of new pronunciation of the trainer. The invention aims to improve the precision of the constructed vocal music pronunciation training sound model by improving the processing effect of the collected sound signals.
Owner:LUOYANG VOCATIONAL&TECHNICAL COLLEGE

Music teaching method and system based on vocal music live classroom

The invention discloses a music teaching method and system based on a vocal music live classroom, and relates to the technical field of music teaching, and the method comprises the steps: respectively collecting audio signals of a teacher end and a student end in the vocal music live classroom, carrying out the preprocessing, and generating frame-level acoustic feature vectors of teachers and students; based on frame-level acoustic feature vectors of teachers and students, generating a mapping feature vector, category probability and an adversarial sample set through mapping, domain discrimination and a CTC model, constructing a total loss function in combination with a loss function value of the model, and performing iterative optimization, gradient fixed-point rounding and feedback of model parameters; according to the invention, through constructing the total loss function, introducing the gradient inversion layer and the gradient fixed-point rounding mechanism, combining the three-state HMM state path and calculating the pronunciation quality logarithm score, the robustness and precision of student pronunciation evaluation in a vocal music live broadcast teaching scene are significantly improved, and thus a high-resolution teaching feedback basis is provided for teachers.
Owner:MUDANJIANG NORMAL UNIV

Construction method of vem-token vocal emotion multimodal magic modification model

The construction method of the VEM-Token vocal emotion multi-modal magic modification model is different from the natural language processing model NLP-Token which interprets music through text. Instead, the VEM-Token sound-to-text innovation model comes with vocal emotion multi-modal information. The model captures and aligns the music tempo of sample songs and user learning songs, identifies various modalities of vocal emotion, divides the file into word units based on tempo, obtains VEM parameters through supervised learning and reinforcement learning, and decomposes songs into vocals, accompaniment, and emotion. The magic modification model provides a multi-modal magic modification method for imitating sample songs, including vocals, accompaniment, emotion overtones, emotion fluctuations, learning to sing, voice cloning, lyrics, pitch calibration, ornaments, tempo length, rhythm speed, tempo strength, freedom, and multiple sample magic modification. It also provides member management, mobile and PC application systems, dedicated support hardware, and communication protocols including MIDI, making it easy to access popular AI large models, reducing model hallucinations, and forming AI vocal agents and AI karaoke.
Owner:GREATER BAY AREA STAR BIOTECH (SHENZHEN) CO LTD

Systems, devices, and methods for dynamic synchronization of a prerecorded vocal backing track to a live vocal performance

Disclosed are systems, methods, and devices, that overcome timing and self-expression limitations experienced by vocalists when using prerecorded vocal backing tracks to enhance live performances. The disclosed system, devices, and methods, dynamically synchronizes prerecorded vocal backing tracks with a live vocal stream by extracting vocal elements, such as phonemes, vector embeddings, or vocal audio spectra, from the live vocal performance in real-time. These extracted vocal elements are matched against corresponding timestamped vocal elements previously derived from the prerecorded vocal backing track, enabling precise real-time adjustment and alignment of the backing track timing to the live performance. Additionally, the system enhances expressive performance by identifying prosody factors, such as pitch, vibrato, accent, stress, dynamics, and level, in the live vocal performance, and dynamically adjusting corresponding prerecorded prosody factors within predefined ranges. This maintains naturalness and spontaneity in the vocalist's live performance, overcoming traditional limitations associated with prerecorded vocal backing tracks.
Owner:EIDOL CORP

Vocal music breath practice device for music teaching and learning

The invention relates to the technical field of music teaching and learning vocal music breath practicing, in particular to a music teaching and learning vocal music breath practicing device which comprises a base and further comprises an air channel a, the air channel a is formed in the base, a connecting sleeve a is fixedly connected to an upper side plate of the base, three threaded sleeves a are fixedly connected to the interior of the upper side plate of the base and internally engaged with a lower side plate of a suspension pipe, and the upper side plate of the suspension pipe is fixedly connected with the air channel a. A suspension ball is slidably connected into the suspension pipe. Suspension pipes are meshed in the threaded sleeves b, the threaded sleeves b are transversely, evenly and fixedly connected to the lower side of the connecting plate, the air filtering cylinder is connected through the connecting sleeves a and the connecting sleeves b, the suspension pipes are connected through the threaded sleeves b and the threaded sleeves a, and by arranging a connecting column and a handle, the device can be conveniently held by hand, and the device can be conveniently placed; a learner can conveniently use fragmented time to improve the breath control ability, the device is assembled through a plurality of modules, the device is convenient to carry, and then the exercise time and place are not limited.
Owner:LANZHOU PETROCHEMICAL VOCATIONAL & TECH UNIV

A breath training bottle for vocal instruction with variable diameter

This invention discloses a breath training bottle for vocal music teaching with variable pipe diameter, belonging to the field of vocal training bottle technology. Its key technical features include a water-storage bottle body, with a support leg fixed to the lower side of one side of the water-storage bottle body. A bottle body tube is connected to the upper side of one side of the bottle body tube, with a cap attached to the upper end of the bottle body tube. A fixed tube seat is installed on one side of the upper end of the cap, and an air blowing tube is installed inside the fixed tube seat. A pipe diameter adjustment component is installed at the lower end of the air blowing tube. The pipe diameter adjustment component includes a pipe diameter adjustment end threaded onto the lower end of the air blowing tube, with connecting arms symmetrically fixed at the bottom end of the pipe diameter adjustment end. Through the cooperation of the threaded pipe diameter adjustment end and the frustum-shaped adjustment plug, the depth of the plug entering the air blowing tube can be flexibly adjusted, achieving continuous adjustment of the air outlet pipe diameter. This solves the core defect of poor adaptability in traditional fixed-diameter vocal training bottles, covering the needs of multiple scenarios such as vocal music teaching, amateur practice, and vocal cord rehabilitation.
Owner:ANHUI NORMAL UNIV

Vocal music loudspeaker with light

The invention relates to the technical field of loudspeakers, in particular to a vocal music loudspeaker with light, which comprises a mounting panel, a loudspeaker opening is formed in the mounting panel, a loudspeaker body is mounted on the right side of the loudspeaker opening and comprises a frame, and a drum membrane, a first elastic metal ring and a second elastic metal ring are arranged on the left portion of the frame. The drum membrane fixing clamp is arranged between the first elastic metal ring and the second elastic metal ring, a baffle is fixedly arranged at the inner ring of the left end of the basin stand, heat dissipation holes are formed in the baffle, a metal substrate is arranged on the right side of the baffle, and the lamp strip is attached to the metal substrate; the right end of the metal substrate and the first elastic metal ring are arranged at an interval, and the interval is filled with sound absorption sponge. The outer edge of the tympanic membrane can drive the first elastic metal ring and the second elastic metal ring to follow up, the tympanic membrane is prevented from being damaged due to the fact that the tympanic membrane exceeds the vibration stroke, meanwhile, airflow exchange between the heat dissipation cavity and the outside is accelerated when the first elastic metal ring and the second elastic metal ring vibrate, and the heat dissipation effect of the heat dissipation cavity is improved.
Owner:XINYANG AGRI & FORESTRY UNIV

Natural emotional singing method and emotional singing audio processing method

PendingCN122347963AAcoustic areaMusic class
The application discloses a natural emotional singing method and an emotional singing audio processing method, and belongs to the singing information field, and aims to solve the problems of fragmented training elements and single feedback system of a traditional vocal music teaching method. The technical scheme comprises six training modules combined by action guidance, emotion awakening and sound area training, and a training system for improving singing emotional expression capability based on a deep learning model for real-time extraction of user singing audio emotional indexes and generation of dynamic visual feedback.
Owner:BEIJING FABIAN EDUCATION TECH CO LTD

Innovative karaoke sound system with combined neural network feedback suppression and coded vocal restoration

A method and apparatus comprising computer code configured to cause a processor or processors to receive an audio signal obtained from a microphone, input the audio signal into frequency-domain Kalman filter (FDKF), input the audio signal and an output from the FDKF into a neural network, estimate, based on the audio signal and the output from the FDKF, and removing feedback signals from the audio signal by the neural network, recover, by a codec receiving an output from the neural network, vocal quality of a target vocal signal, and output a version of the audio signal in which the target vocal signal is enhanced by removal of the feedback signals from the audio signal by the neural network and by recovery of the vocal quality by the codec.
Owner:TENCENT AMERICA LLC

Artificial intelligence-based vocal music learning intelligent assistance method and system

The application discloses an intelligent vocal music learning and assisting method and system based on artificial intelligence, and relates to the technical field of vocal music teaching.The method comprises the following steps: constructing a teacher pronunciation expression database comprising original pronunciation audio of multiple teachers, frame-level aligned real face videos and international phonetic transcription sequences, and generating teacher 3D face animation data for associated storage; acquiring learner audio data and extracting audio feature vectors; selecting the most similar target teacher based on audio feature similarity; inputting the learner audio into an audio-driven 3D face animation generation model to generate learner 3D face animation data; extracting the key frame sequences of the learner and the teacher for pronunciation actions; performing three-dimensional model difference calculation on the corresponding key frames to obtain frame-level difference quantization data, and generating visualized difference images.The application visualizes abstract vocal music pronunciation skills into intuitive 3D facial movements, realizes precise personalized guidance highly consistent with the individual characteristics of learners, and improves teaching efficiency.
Owner:SICHUAN NORMAL UNIV

Vocal music training method and system based on deep learning

The invention provides a vocal music training method based on deep learning. The vocal music training method comprises the steps of collecting vocal music training data through a multi-source sensor; preprocessing the vocal music training data, extracting acoustic features, expression features and physiological features, and constructing a multi-modal feature vector; the multi-modal feature vectors are input into a deep learning model, the deep learning model analyzes acoustic features through a convolutional neural network, analyzes physiological feature time sequence changes through a time recursive network, and fuses the multi-modal feature vectors through a cross-modal attention mechanism; based on the output of the deep learning model, real-time feedback information is generated, and the real-time feedback information comprises intonation correction suggestions, breathing rhythm guidance and emotion expression intensity scores; and dynamically adjusting the training difficulty according to the real-time feedback information, generating a personalized training plan, and updating parameters of the deep learning model. According to the embodiment of the invention, the problems of geographical limitation, low learning efficiency, difficulty in realizing personalized teaching and the like can be effectively solved.
Owner:LUOYANG INST OF SCI & TECH

Automatic debugging method and system, electronic equipment and storage medium

The invention provides an automatic debugging method and system, electronic equipment and a storage medium. The method is used for a sound sensing device and comprises the steps that a test audio is played to a user, and the test audio is any one of words, syllables, music, voice, vocal music and pure tone; receiving response information of the user; determining whether the response information is matched with the test audio; when the response information is not matched with the test audio, storing the test audio as a reference optimized audio; determining a target distinctive speech feature corresponding to the reference optimized audio; determining a target operation parameter corresponding to the target distinctive speech feature according to a corresponding relationship between the distinctive speech feature and the operation parameter of the sound sensing device; and adjusting the target operation parameters to optimize the sound sensing equipment. The method is favorable for reducing the debugging cost and simplifying the debugging process.
Owner:SHANGHAI LISTENT MEDICAL TECH CO LTD

Vocal practice microphone

ActiveCN309900089STesting MethodsVocal exercises
1. The name of the design product: vocal practice microphone. 2. The use of the design product: microphone for practicing vocal music. 3. The design points of the design product: in shape. 4. The picture or photo that best indicates the design points: perspective view.
Owner:王卉

Tone tuning method capable of displaying numbered musical notation singing names

The invention discloses an intonation device for displaying singing names. The intonation device comprises the following steps: A, selecting practicing tones, namely 12 tones; b, inputting a signal through an input unit; c, the single-chip microcomputer processing unit processes the input signal and quickly finds the code of the corresponding sound; and C, the single-chip microcomputer processing unit simulates and displays numbered musical notation singing name display of the standard tone according to the corresponding code. According to the method, the position of the numbered musical notation singing name can be simulated and displayed. The folk music enthusiasts only know the singing names of the numbered musical notation, so that the vocal music enthusiasts, the string music enthusiasts and the wind music enthusiasts can quickly link the made sounds with the pitch marks of the corresponding tones. The method is greatly helpful for a lover to quickly identify intonation deviation, and the intonation is remarkably improved.
Owner:赵冬

Recording device for vocal music teaching

The utility model relates to the technical field of vocal music teaching equipment, and discloses a voice recording device for vocal music teaching, which comprises a bottom plate, a voice recorder box is fixedly arranged on one side of the top end of the bottom plate and is connected with a sound pickup through a wire, and an adjusting box is fixedly arranged on the other side of the top end of the bottom plate and is connected with the sound pickup through a wire. Stand columns are slidably connected to the two sides of the top end of the adjusting box, a driving assembly for driving the stand columns to ascend and descend is arranged in the adjusting box, a cross beam is fixedly arranged at the top ends of the stand columns, and a cross rod is fixedly arranged in the middle of the cross beam; the stand column can support the cross beam, the driving assembly can control the stand column to stretch out and draw back automatically, the height of the sound pick-up is adjusted automatically through the cross beam in the stand column stretching-out and drawing-back process, operation is easy, good self-locking capacity is achieved, the stand column can be prevented from sliding off, stability is good, and safety and reliability are achieved. And supporting can be provided between the cross rod and the clamping assembly through the damping universal joint.
Owner:HUNAN INT ECONOMICS UNIV

Vocal music performance expression auxiliary correction mirror

ActiveCN223349910UPicture framesDomestic mirrorsCorrective contact lensCorrective lens
The utility model discloses a vocal music performance expression auxiliary correction mirror which comprises a base, and an adjusting assembly is arranged on the top of the base. The adjusting assembly comprises a first auxiliary block, one side of the rotating rod is fixedly connected with a second auxiliary block, the adjusting assembly further comprises an electric telescopic rod, one end of the electric telescopic rod is fixedly connected with a protruding block, and the top of the base is fixedly connected with a third auxiliary block. And one end of the convex block is rotationally connected with one end of the third auxiliary block. The utility model relates to the technical field of correction glasses. According to the vocal music performance expression auxiliary correction mirror, by starting an electric telescopic rod, a butt joint block can push a rotating rod to rotate, and finally the correction mirror body can be adjusted to adjust the angle, a triangular structure is formed among the rotating rod, the electric telescopic rod and a base, and all components at the three corners are rotationally connected; and when the length of the electric telescopic rod is changed, the number of angles in the triangle can be automatically and adaptively changed.
Owner:HUNAN INT ECONOMICS UNIV

Computer audio-visual method and system based on artificial intelligence, and computer readable medium

The invention relates to the cross technical field of computer technology and education informatization, in particular to a computer audio-visual method and system based on artificial intelligence and a computer readable medium, which are used for remote vocal music, dance or fitness action teaching. Image and audio data are converted into computable skeleton sequences and feature vectors through artificial intelligence, and are compared with standard data, so that the defects that traditional teaching depends on subjective experience of teachers, and real-time quantitative feedback cannot be realized are overcome. By constructing the unified virtual room and mapping all trainees into virtual characters therein, physical space isolation is broken, remote trainees can obtain the immediacy sense and team sense of'same-field training ', and learning motivation and interactivity are greatly improved. The evaluation result is instantly converted into visual elements (such as scores and special effects) in a virtual scene, visual and instant positive feedback or error correction prompts are provided for students, the skill acquisition process is accelerated, and a set of computer audio-visual method based on artificial intelligence is formed.
Owner:ZHENGZHOU UNIV

A vocal music practice assisting system based on vocal cord fatigue analysis

This application provides a vocal practice assistance system based on vocal cord fatigue analysis, comprising: a surface-mounted electrode for attaching to a detection target to acquire electromyographic (EMG) signals; an EMG signal acquisition module for acquiring the EMG signals transmitted by the surface-mounted electrode and performing analog-to-digital conversion; a data transmission and reception module for sending the EMG signals acquired by the EMG signal acquisition module to a host computer and receiving relevant information; a preprocessing module for preprocessing the EMG signals transmitted by the data transmission and reception module to reduce noise; a signal separation module for separating the preprocessed EMG signals; a feature extraction module for acquiring the time-domain and frequency-domain features of the EMG signals; and a fatigue analysis module for linearly fitting the time-domain and frequency-domain feature curves and analyzing the vocal cord fatigue trend through the slope of the linear fitting equation. This system can monitor and analyze the degree of vocal cord fatigue in real time and provide a convenient and reliable vocal practice assistance device for vocal professionals.
Owner:CHINA UNIV OF GEOSCIENCES (BEIJING)

Breath training device for vocal music teaching

ActiveCN223296452UMusicAcousticsMusic class
The utility model belongs to the technical field of vocal music teaching, and particularly relates to a breath training device for vocal music teaching. The breath training device comprises an air blowing cylinder, an air blowing pipe, a cover plate, a movable support, an indicating rod, a push rod, a round push plate, a piston and a movable column. Moving columns are arranged on the left side and the right side of the outer wall of the air blowing cylinder, moving supports are slidably installed on the moving columns, indicating rods are fixedly connected to the moving supports, a cover plate is installed at the top end of the air blowing cylinder, a push rod is slidably connected to the middle of the cover plate, a round push plate is installed at the top end of the push rod, and a piston is arranged at the bottom end of the push rod and slides in the air blowing cylinder. The vocal music learner keeps the blowing strength, the round push plate is kept parallel to the upper portion of the indicating rod, and the blowing stability and durability of the vocal music learner can be trained. The position of the indicating rod can be changed by adjusting the position of the movable support on the air blowing cylinder, and the air blowing stability capability that a vocal music learner keeps the round push plate at different positions can be trained.
Owner:EAST CHINA JIAOTONG UNIVERSITY

A vocal performance scoring method and system based on neural networks and audio-visual fusion

The present application relates to a kind of vocal performance scoring method and system based on neural network and audio-visual fusion, belong to vocal evaluation field.The method utilizes different neural networks to obtain expert score data in three dimensions, including audio score, emotion score and dressing score, then the scores in three dimensions are input into expert score fitting neural network, and finally the comprehensive score is obtained.The present application makes the evaluation result more real and effective, close to expert score, and the scoring process is more efficient and convenient.
Owner:CHONGQING UNIV OF POSTS & TELECOMM

Visual breath amount testing device

The utility model relates to a vocal music device, in particular to a breath amount testing device. The utility model provides a visual breath amount testing device which comprises a shell, a mainboard and a battery, the mainboard and the battery are arranged in the shell, a light-permeable display mask is arranged at one end of the shell, and continuous numbers are sequentially marked on the display mask; the mainboard is provided with a display module matched with the display mask in shape and a main control module for controlling the display module, the display module is provided with RGB lamp beads corresponding to the number positions on the display mask, and the RGB lamp beads are electrically connected with the main control module arranged on the mainboard; a pickup hole is formed in the shell, a pickup is arranged on the inner side of the pickup hole, and the pickup is electrically connected with the main control module; the main control module can turn on or present RGB lamp beads in the brightness and color change display module in sequence according to sound information and duration acquired by the sound pickup; by using the device, the breath amount can be objectively tested, and the rising sound vocal music training effect can be achieved.
Owner:FOSHAN YICHU ELECTRONIC LIGHTING CO LTD

A national vocal dialect prosody intelligent correction method and system

This invention proposes an intelligent error correction method and system for ethnic vocal music dialect prosody. The method includes: collecting multimodal data, extracting features to form a training dataset, constructing and updating a dynamic dialect prosody map, and performing cross-modal comparative analysis to obtain a joint representation of cross-modal error features. A neural vocoder direct-connect correction model is constructed, inputting the joint representation to generate a corrected Mel spectrum. A cross-modal generative adversarial network is constructed, including a generator and three discriminators. The generator generates a joint latent representation based on high-fidelity corrected audio and three features. The three discriminators respectively judge the naturalness of the audio, the compliance of the text tone, and the rhythmic coordination of the musical score. This invention solves the problem of manual error correction by constructing a prosody map, overcomes the bottleneck of manual detection by utilizing cross-modal comparative analysis, improves the generalization ability of the model by combining meta-learning algorithms, and optimizes the correction results with the help of adversarial networks, thereby reducing manual costs and improving the error correction efficiency of ethnic vocal music works.
Owner:GUIZHOU RADIO & TV UNIV

Isomorphic musical instrument and matching notation method

PCT designated stageWO2025214084A1Stringed musical instrumentsNotation recordingPianoOctave
The present application relates to a twelve-semitone isomorphic keyboard or soundboard group, which adapts to the natural arch of five fingers by using color allocation layouts, tone key shapes, concave-convex touch recognition points, etc., so as to achieve the technical effects of quickly identifying note positions, and simplifying and unifying fingering. Correspondingly, the structures of transmission apparatuses of a piano and an electronic organ are also changed. The present application further relates to a matching twelve-semitone independent-note isomorphic notation method, and software and apparatuses for music score creation, music score singing and note singing. A system endows a music score with various pattern identification detail features to match the music scores to note positions of a musical instrument on a one-to-one basis, such as a twelve-semitone pitch scale, a series of octave clefs, opposite notehead-stem directions between lines, zigzag bar lines and four sets of note systems, making music score information intuitive and concise, and reducing the workload of brain conversion during music score recognition, and thus increasing the reaction speed. The software and apparatuses for music score creation, music score singing and note singing are powerful tools for recording vocal music in a music score form and teaching music score singing.
Owner:LI WENGUO +1

Remote vocal music teaching multi-angle image synchronous analysis system

The invention relates to the technical field of remote education, in particular to a remote vocal music teaching multi-angle image synchronous analysis system. Comprising a student image data acquisition module, a course database, a curved surface manifold construction module, a geodesic deviation analysis module, a multi-modal interaction feedback module, an audio acquisition module, a pitch-attitude correlation analysis module, a teacher remote guidance module and the like. The difference between the student attitude and the standard attitude is accurately quantified through a geodesic deviation vector field, an optimal attitude correction path is calculated based on curvature optimization, real-time guidance is provided through multi-mode feedback of vision, hearing, touch and the like, the system further establishes association mapping, auxiliary analysis and guidance of acoustic features such as pitch and tone and attitude manifold features, and the system is applied to the field of teaching. Through the edge calculation and adaptive bandwidth management technology, the delay problem in remote teaching is solved.
Owner:YANGZHOU UNIV

Vocal music feature decoupling identification method and system based on mutual information minimization

The invention discloses a nasal sound recognition method based on acoustic feature decoupling, which is characterized in that a double-flow decoupling deep neural network is constructed, a mutual information minimization adversarial training strategy is introduced, and sound source features (vocal cord vibration) and sound channel features (oral cavity and nasal cavity adjustment) are forcibly separated from a single audio signal. On the basis, the high-resolution characteristic of CQT is used for accurately extracting acoustic characteristics related to the nasal sound, physiological offset correction is carried out in combination with personal acoustic fingerprints, and finally accurate quantification and recognition of different nasal sound states (normal, defect and skill) are achieved, so that an objective and visual feedback basis with artistic style adaptability is provided for vocal music teaching.
Owner:HUAZHONG UNIV OF SCI & TECH

Metaverse Personalized Digital Singer Generation System and Method Thereof

A metaverse personalized digital singer generation system and a method thereof. In the system, the server-end device receives a user voice, store the user voice as a personalized voice, capture an image of a user face to generate a facial image, generate a personalized digital singer displayed in a virtual scene through a 3D imaging technology, and convert the personalized voice into voice feature vectors, and use the voice feature vectors and the personalized voice as training data, input the training data to a generative AI model to train a generative pre-training model having the personal characteristics. When the user selects an original song for singing, the original song and the user singing voice and the prompt are inputted to the generative pre-training model, the remixed song matching a style of the original song is outputted, a vocal coaching is generated and displayed based on prompt.
Owner:SQ TECH (SHANGHAI) CORP +1