This story was originally published on HackerNoon at:
https://hackernoon.com/gigachat-the-ai-assistant-detects-emotions-and-finds-content-in-long-audio-files.
GigaChat Audio adds emotion recognition, long-audio understanding, and multilingual speech AI while open-sourcing new audio and speech recognition models.
Check more stories related to undefined at:
https://hackernoon.com/c/undefined.
You can also check exclusive content about
#gigachat-audio-model,
#gigachat3.1-audio-10b,
#emotion-recognition-ai-voice,
#long-audio-ai-summarization,
#arena-hard-audio-benchmark,
#open-source-audio-llm,
#multilingual-speech-model,
#good-company, and more.
This story was written by:
@jonstojanjournalist. Learn more about this writer by checking
@jonstojanjournalist's about page,
and for more stories, please visit
hackernoon.com.
GigaChat has upgraded its Audio AI to recognize emotions, analyze recordings up to three hours long, identify speakers, and generate timestamped summaries without converting speech to text first. It also introduces memory for voice interactions and open-sources GigaChat3.1-Audio-10B and GigaAM Multilingual, enabling developers to build speech recognition, transcription, and voice AI applications.