Speech Synthesis / Quote

Step-1o-Audio API Price Speech Synthesis

StepFun logo StepFun · step-1o-audio

Add to comparison

Input · Cache Miss
¥25 /1M tokens
Output
¥60 /1M tokens
Cache Hit Input
¥5 /1M tokens
BillingCNY · 1M tokens
ModalitiesAudio / Text
Deep Thinking—
Verified Date2026-10-01
Official pricing pageplatform.stepfun.com/docs/zh/guides/pricing/details ↗

Source text: the note below is the model's official text, kept in its original Chinese and not translated.

官方定价页「语音大模型·按 Token 计费」表:输入(缓存未命中)25 元、缓存命中 5 元、输出 60 元每百万 tokens,当前状态为按量计费。

Report a pricing issue

Found inaccurate pricing or model information? We will review it.

Monthly Cost Estimate

30 days · no cache hits · list price

Light Load
1M Input + 0.3M Output/ day
¥1,290/month
Medium Business
5M Input + 1.5M Output/ day
¥6,450/month
Heavy Agent
10M Input + 3M Output/ day
¥12,900/month

Free quotas, tiered pricing, batch discounts, and time-based pricing are not included; refer to official bills.

StepFun Related Models

FAQ

How much does the Step-1o-Audio API cost?

Input ¥25/1M tokens, output ¥60/1M tokens

Which input modalities does Step-1o-Audio support?

Audio / Text.

What is the approximate monthly cost of Step-1o-Audio?

Depends on volume: light use (1M input + 300K output per day) ≈ ¥1,290/month; medium (5M + 1.5M) ≈ ¥6,450/month; heavy agents (10M + 3M) ≈ ¥12,900/month. Cache hits, batch discounts and tiered pricing can lower this further.

How does Step-1o-Audio compare with peers on value?

StepFun Step-1o-Audio is competitively priced among comparable models; compare against similar models. See the Pricemon market table for a full comparison.

What use cases suit Step-1o-Audio?

Good for: voice-overs, audiobooks, social media, navigation systems, accessibility.

Does Step-1o-Audio have a free tier?

Check StepFun's official pricing page for the latest free quota, trial policy and tiered pricing. Policies differ by vendor and change over time.

← Back to Market