GigaChat the AI assistant detects emotions and finds content in long audio files
The updated GigaChat Audio model is available through the GigaChat AI assistant and as free open-source code for software developers across the globe
Users of the GigaChat AI assistant can enjoy the updated GigaChat Audio AI, a large language model capable of processing audio files and voice messages without converting speech to text first. The LLM has been trained to understand intonation and had its sound data processing potential enhanced to deliver higher performance.
The AI assistant with more empathy
The updated GigaChat can detect users’ positive or negative emotions based on intonation, voice characteristics, and pronunciation nuances to be more on-point in responses. For example, it will talk more gently to an irritated user or mirror the mood of someone sharing good news.
The model can process recordings up to three hours in length and navigate within them. You can ask it about the exact time when some specific...
Copyright of this story solely belongs to hackernoon.com. To see the full text click HERE