Atlas · GenAI 2026
Audio AI
Audio-Native AI
conceptPeak: 2025Audio & SpeechAI consensus: 1/3
Prerequisites
No prerequisites.
Recommended reference
Radford et al. (2023) 'Robust Speech Recognition via Large-Scale Weak Supervision' (Whisper paper) — foundational; plus OpenAI GPT-4o audio modality docs
Notes from AI deep research
Anthropic Opus
Whisper for transcription, GPT-4o natywne audio. Analiza tonow i emocji bez konwersji
Google Deep Think
Analiza tonów, szeptu, emocji w real-time [G#30]
Related skills
- → is subcategory of: Multimodal AI(1/3)
- ← is an instance of: Librosa(0/3)
- ← is subcategory of: Text-to-Speech(0/3)
- ← is subcategory of: Speech Recognition(0/3)
- ← is subcategory of: Audio Processing(0/3)