Episode

GigaChat the AI assistant detects emotions and finds content in long audio files

Podcast
The Good Tech Companies
Published
Jul 22, 2026
Duration seconds
229
Processing state
not_requested
Canonical source
https://share.transistor.fm/s/93425e1f
Audio
https://media.transistor.fm/93425e1f/fbba8f12.mp3
JSON
/v1/public/podcasts/the-good-tech-companies-6882802/episodes/gigachat-the-ai-assistant-detects-emotions-and-finds-content-in-long-audio-files
Markdown
/podcast/the-good-tech-companies-6882802/gigachat-the-ai-assistant-detects-emotions-and-finds-content-in-long-audio-files.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/the-good-tech-companies-6882802/episodes/gigachat-the-ai-assistant-detects-emotions-and-finds-content-in-long-audio-files/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/the-good-tech-companies-6882802/gigachat-the-ai-assistant-detects-emotions-and-finds-content-in-long-audio-files.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

This story was originally published on HackerNoon at: https://hackernoon.com/gigachat-the-ai-assistant-detects-emotions-and-finds-content-in-long-audio-files . GigaChat Audio adds emotion recognition, long-audio understanding, and multilingual speech AI while open-sourcing new audio and speech recognition models. Check more stories related to undefined at: https://hackernoon.com/c/undefined . You can also check exclusive content about #gigachat-audio-model , #gigachat3.1-audio-10b , #emotion-recognition-ai-voice , #long-audio-ai-summarization , #arena-hard-audio-benchmark , #open-source-audio-llm , #multilingual-speech-model , #good-company , and more. This story was written by: @jonstojanjournalist . Learn more about this writer by checking @jonstojanjournalist's about page, and for more stories, please visit hackernoon.com . GigaChat has upgraded its Audio AI to recognize emotions, analyze recordings up to three hours long, identify speakers, and generate timestamped summaries without converting speech to text first. It also introduces memory for voice interactions and open-sources GigaChat3.1-Audio-10B and GigaAM Multilingual, enabling developers to build speech recognition, transcription, and voice AI applications.