Перейти к основному содержанию
Multimodal
stepfun
Released: Apr 2026 Updated: Aug 2026

Stepaudio 2.5 Asr

Stepaudio 2.5 Asr is a Multimodal model. Speech transcription model for accurate audio-to-text and captioning workflows
Vision
Audio
Tool Use
Reasoning
Citations

Key Specs

Context
-
Max Output
-
Best Price (Input / Output)
Free / Contact per 1M tokens
Chat with Model

Похожие модели