This dataset features 700 native Russian speakers captured on video, showcasing natural, unscripted speech along with gestures and emotions. With over 10 minutes of footage per participant in high resolution, it serves as a valuable resource for training multimodal AI models and enhancing emotion recognition technology.
The Storytelling Video Dataset offers a rich collection of high-quality video recordings featuring 700 native Russian speakers. Each participant shares unscripted narratives for over 10 minutes, capturing both full-body movements and an array of non-verbal cues such as gestures and facial expressions. This unique dataset is specifically designed to facilitate the training and development of multimodal AI models.
The dataset is ideal for applications in several advanced fields:
This dataset stands out in its ability to combine visual and auditory information, making it a robust resource for anyone engaged in AI training focusing on linguistic behavior and non-verbal communication.
GitHub: github.com/MaratDV/storytelling-video-dataset
Hugging Face: huggingface.co/datasets/MaratDV/video-dataset
License: Commercial use only — contact required
Contact: chinzad@gmail.com | Telegram: @Marat_DV
No comments yet.
Sign in to be the first to comment.