Claude Fable 5.1
Anthropic
Context
1M
Benchmarks
Independent BenchLM leaderboard scores for frontier and open-weight models — so you can pick the right model for the job, not just the biggest name.
Frontier models
The same official APIs, on the same terms — one API key, unified JPY billing, and a single invoice through Tara Cloud.
Claude Fable 5.1
Anthropic
Context
1M
GPT-6 Astra
OpenAI
Context
1.05M
Claude Opus 5
Anthropic
Context
—
GPT-5.6 Sol
OpenAI
Context
1.05M
GPT-5.4
OpenAI
Context
1.05M
Claude Sonnet 5
Anthropic
Context
1M
Official APIs resold by Tara Cloud, with unified JPY billing.
Performance per need
85% of the leader's score
Qwen3.8 Max delivers 85% of the leader's benchmark performance with a 1M context window — a strong fit for high-volume production workloads.
78% of the leader's score
GLM-5.3-Flash reaches 78% of the leader's score with a compact footprint — built for throughput-sensitive pipelines.
Open-weight models
Open-weight models hosted on Tara Cloud's own GPU infrastructure in Japan — every prompt and completion stays in-country.
Qwen3.8 Max
Alibaba
Context
1M
GLM-5.3
Z.AI
Context
1M
GLM-5.3-Flash
Z.AI
Context
1M
Kimi K2.6
Moonshot AI
Context
256K
MiniMax M3
MiniMax
Context
1M
DeepSeek V3.2
DeepSeek
Context
128K
GPT-OSS 120B
OpenAI
Context
128K
All data stays in Japan.
Benchmark data: BenchLM (benchlm.ai), snapshot of September 8, 2026. Scores are indicative and updated periodically. benchlm.ai
Tell us about your workloads and we'll size the right mix of open-weight and frontier models — usually with a proposal within two business days.