Native audio understanding, transcription, speech tone analysis, sound classification, and multi-speaker transcription.