Models
FastMetal supports various LLM models for different use cases. View the models available on your dashboard, or query the /models endpoint.
Available Models
Models are configured by your administrator. Use the /models endpoint to list available models:
| Model | Description |
|---|---|
mistral-voxtral-mini-3b-2507 | Japan: Mistral's voxtral-mini-3b-2507 |
anthropic-claude-opus-4-8 | Global: Claude Opus 4.8 - Most intelligent, best for agents and coding |
anthropic-claude-opus-4-7 | Global: Claude Opus 4.7 - Most intelligent, best for agents and coding |
anthropic-claude-opus-4-6 | Global: Claude Opus 4.6 - Most intelligent, best for agents and coding |
anthropic-claude-sonnet-4-6 | Japan: Claude Sonnet 4.6 - Best balance of speed and intelligence |
anthropic-claude-haiku-4-5 | Japan: Claude Haiku 4.5 - Fastest with near-frontier intelligence |
minimax-m2.7 | Global: Minimax's M2.7 |
glm-5 | Global: Z.ai's GLM-5 |
llm-jp-3.1-8x13b-instruct4 | Japan: Japanese language model optimized for instruction following |
japan-gemma-4-31b | Japan: Google's Gemma 4 31B - hosted and run in Japan |
japan-qwen3.6-35b | Japan: Qwen 3.6 35B-A3B - hosted and run in Japan |
japan-kimi-k2.6 | Japan: Moonshot AI's Kimi K2.6 - hosted and run in Japan |
japan-kimi-k2.7-code | Japan: Moonshot AI's Kimi K2.7 Code - hosted and run in Japan |
gpt-oss-120b | Japan: gpt-oss-120b |
minimax-m3 | Global: MiniMax's M3 |
qwen3.6-27b | Global: Qwen 3.6 27B |
qwen3.8-27b | Global: Qwen 3.8 27B |
kimi-k2.6 | Global: Moonshot AI's Kimi K2.6 |
glm-4.7 | Global: Z.ai's GLM 4.7 |
glm-4.7-flash | Global: Z.ai's GLM 4.7 Flash |
glm-5.1 | Global: Z.ai's GLM 5.1 |
mimo-v2.5 | Global: Xiaomi's MiMo-V2.5 |
mimo-v2.5-pro | Global: Xiaomi's MiMo-V2.5-Pro |
deepseek-v4-flash | Global: DeepSeek's V4 Flash |
deepseek-v4-flash-0731 | Global: DeepSeek's V4 Flash 0731 |
deepseek-v4-pro | Global: DeepSeek's V4 Pro |
glm-5.2 | Global: Z.ai's GLM 5.2 |
glm-5.3 | Global: Z.ai's GLM 5.3 |
kimi-k3 | Global: Moonshot AI's Kimi K3 |
qwen3.7-max | Global: Qwen 3.7 Max |
qwen3.8-max | Global: Qwen 3.8 Max |
grok-4.5 | Global: xAI's Grok 4.5 |
grok-4.6 | Global: xAI's Grok 4.6 |
gpt-5.6-sol | Global: OpenAI's GPT-5.6 Sol |
gpt-5.6-terra | Global: OpenAI's GPT-5.6 Terra |
gpt-5.6-luna | Global: OpenAI's GPT-5.6 Luna |
gemini-3.5-flash | Global: Google's Gemini 3.5 Flash |
gemini-3.7-flash | Global: Google's Gemini 3.7 Flash |
inkling | Global: Thinking Machines' Inkling |
anthropic-claude-opus-5 | Global: Claude Opus 5 - Most intelligent, best for agents and coding |
anthropic-claude-sonnet-5 | Global: Claude Sonnet 5 - Best balance of speed and intelligence |
anthropic-claude-fable-5 | Global: Claude Fable 5 - Anthropic's flagship Mythos-class model |
muse-glimmer-30b | Global: Meta's Muse Glimmer 30B |
gemini-flash-lite-free | Global: Google's Gemini Flash Lite - free to try |
random-free | Global: Random - free to try |
google-nano-banana-2 Image | Global: Google Nano Banana 2 - fast, high-quality image generation |
bytedance-seedream-4.5 Image | Global: Generate images |
z-image-turbo Image | Global: Generate realistic images |
lustify-sdxl Image | Global: Generate animated images |
Model Pricing
Pricing is calculated per token. Each model has separate input and output token rates. Check the pricing page or your dashboard for current rates.
| Model | Input / 1M tokens | Output / 1M tokens |
|---|---|---|
| mistral-voxtral-mini-3b-2507 | ¥8.4 | ¥8.4 |
| anthropic-claude-opus-4-8 | ¥893.5 | ¥4,467.5 |
| anthropic-claude-opus-4-7 | ¥840 | ¥4,200 |
| anthropic-claude-opus-4-6 | ¥850 | ¥4,200 |
| anthropic-claude-sonnet-4-6 | ¥554.4 | ¥2,772 |
| anthropic-claude-haiku-4-5 | ¥184.8 | ¥924 |
| minimax-m2.7 | ¥53.61 | ¥214.44 |
| glm-5 | ¥178.7 | ¥571.84 |
| llm-jp-3.1-8x13b-instruct4 | ¥16 | ¥79 |
| japan-gemma-4-31b | ¥26 | ¥101 |
| japan-qwen3.6-35b | ¥32 | ¥158 |
| japan-kimi-k2.6 | ¥64 | ¥316 |
| japan-kimi-k2.7-code | ¥55 | ¥530 |
| gpt-oss-120b | ¥16 | ¥79 |
| minimax-m3 | ¥53.61 | ¥214.44 |
| qwen3.6-27b | ¥80.415 | ¥482.49 |
| qwen3.8-27b | ¥98.285 | ¥589.71 |
| kimi-k2.6 | ¥169.765 | ¥714.8 |
| glm-4.7 | ¥107.22 | ¥393.14 |
| glm-4.7-flash | ¥10.722 | ¥71.48 |
| glm-5.1 | ¥250.18 | ¥786.28 |
| mimo-v2.5 | ¥25.018 | ¥50.036 |
| mimo-v2.5-pro | ¥77.7345 | ¥155.469 |
| deepseek-v4-flash | ¥16.083 | ¥32.166 |
| deepseek-v4-flash-0731 | ¥25.018 | ¥50.036 |
| deepseek-v4-pro | ¥341.317 | ¥684.421 |
| glm-5.2 | ¥250.18 | ¥786.28 |
| glm-5.3 | ¥250.18 | ¥786.28 |
| kimi-k3 | ¥536.1 | ¥2,680.5 |
| qwen3.7-max | ¥263.5825 | ¥790.7475 |
| qwen3.8-max | ¥357.4 | ¥1,072.2 |
| grok-4.5 | ¥357.4 | ¥1,072.2 |
| grok-4.6 | ¥357.4 | ¥1,072.2 |
| gpt-5.6-sol | ¥893.5 | ¥5,361 |
| gpt-5.6-terra | ¥357.4 | ¥2,144.4 |
| gpt-5.6-luna | ¥35.74 | ¥214.44 |
| gemini-3.5-flash | ¥268.05 | ¥1,608.3 |
| gemini-3.7-flash | ¥134.025 | ¥670.125 |
| inkling | ¥178.7 | ¥723.735 |
| anthropic-claude-opus-5 | ¥893.5 | ¥4,467.5 |
| anthropic-claude-sonnet-5 | ¥357.4 | ¥1,787 |
| anthropic-claude-fable-5 | ¥1,787 | ¥8,935 |
| muse-glimmer-30b | ¥62.545 | ¥268.05 |
| gemini-flash-lite-free | ¥0 | ¥0 |
| random-free | ¥0 | ¥0 |
For full pricing details, see the pricing page.
Image Model Pricing
Image-generation models are billed per generated image rather than per token.
| Model | Per image |
|---|---|
| google-nano-banana-2 | ¥12.23 |
| bytedance-seedream-4.5 | ¥7.15 |
| z-image-turbo | ¥1.68 |
| lustify-sdxl | ¥1.68 |
Model Capabilities
Chat/Conversation — Multi-turn dialogue with context
All models support chat/conversationText Completion — Single-turn text generation
Text completion support