Systems and methods for conversational avatar systems are disclosed herein. The systems and methods may include receiving, via a computing
system, an input comprising speech data; generating, via the computing
system, a transcript of the speech data in real-time; analyzing, via the computing
system, the transcript in real-time; generating, via the computing system, a response to the transcript in real-
time based on the tagging; selecting, via the computing system, one or more avatar
animation gestures based on tone and
speech patterns of the generated response; synthesizing, via the computing system, an audible response based on the generated response;
synchronizing, via the computing system, the one or more avatar
animation gestures to the synthesized audible response to form a synchronized avatar
animation; and rendering, via the computing system, the synchronized avatar animation.