Seed-2.0-Mini
seed-2.0-miniSeed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal und Built by ByteDance Seed.
Prices updated
Input price
$0.10
per 1M tokens · Standard
Output price
$0.40
per 1M tokens · Standard
Input limit
262K
tokens
Output limit
131K
tokens
Input formats
Output formats
Seed-2.0-Mini price history
2 price records since 9/25/2026
Seed-2.0-Mini cost calculator
$4.00/month
Standard pricing
Overview
What is Seed-2.0-Mini?
Seed-2.0-Mini is a text & reasoning and vision model from ByteDance Seed. Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal und Its 262K context window and $0.10 input price make it a candidate for cost-sensitive, high-throughput applications.
Rates
Seed-2.0-Mini pricing
Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.
Input · Standard
$0.10
Per 1M tokens
Cached input · Standard
Not available
No listed cached-input rate
Capabilities
Seed-2.0-Mini Tools
Tools available when using Seed-2.0-Mini through supported provider APIs.
Function calling
Supported
Structured outputs
Supported
JSON mode
Supported
Reasoning
Supported
Built-in web search
Not supported
Log probabilities
Not supported
Deterministic seed
Not supported
Parallel tool calls
Not supported
Prompt caching
Not supported
Strengths and limitations
Strengths
- Low input cost at $0.10 per million tokens suits high-volume workloads.
- 262K context supports large documents, repositories and extended conversations.
- Supports text & reasoning and vision workloads in one model.
Limitations
- Generated tokens cost 4× more than input tokens, which matters for verbose responses.
- Large context capacity does not guarantee consistent retrieval across the entire prompt.
- Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.
Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.
Same provider
