Rankings by use case

LLM APIs with the largest context windows (August 2026)

The models FastMetal serves, ordered by how many tokens one request can hold.

Ordered by context window: the number of tokens one request can hold.

Links
1GPT-5.6 Lunaopenai logoOpenAIUnranked¥35.74¥214.441,050,000Details
2GPT-5.6 Solopenai logoOpenAIUnranked¥893.5¥5,3611,050,000DetailsCompare
3GPT-5.6 Terraopenai logoOpenAIUnranked¥357.4¥2,144.41,050,000DetailsCompare
4MiMo-V2.5XiaomiUnranked¥25.02¥50.041,050,000Details
5MiMo-V2.5-ProXiaomiUnranked¥77.73¥155.471,050,000DetailsCompare
6DeepSeek V4 Flashdeepseek logoDeepSeekUnranked¥16.08¥32.171,048,576DetailsCompare
7DeepSeek V4 Prodeepseek logoDeepSeekUnranked¥341.32¥684.421,048,576DetailsCompare
8Gemini 3.5 Flashgoogle logoGoogleUnranked¥268.05¥1,608.31,048,576DetailsCompare
9Gemini 3.7 Flashgoogle logoGoogleUnranked¥134.03¥670.131,048,576DetailsCompare
10GLM 5.2z-ai logoZ.aiUnranked¥250.18¥786.281,048,576DetailsCompare

Click a column heading to re-sort; the table opens in this category's ranking order. Prices are yen per 1M tokens before tax; context is in tokens.

Frequently asked questions

What is this ranking based on?
Ordered by context window: the number of tokens one request can hold.
How is pricing decided?
Each model has a yen price per 1M tokens (before tax), drawn from a prepaid balance as you use it. There is no subscription or monthly fee.
Can I use the top models with one API key?
Yes. Every model on this page is served from FastMetal's OpenAI-compatible endpoint under one API key; switching is a change to the model string.

Every model above runs on one API key

Create an account and add credit to call these models from the browser chat and the API. No monthly fee.

Other rankings