datasets: CREMA-D / RAVDESS / TESS / JL Corpus models: NVIDIA A2F - 3D, LAM audio2 expression, check the HF page for links
anywhere we can see it in action?
This dataset powers our v2 version so not published yet, the v1 with no strong facial expression is in many places ..check myned.ai avatar, or nyxclaw.ai