AionLabs · live price
Aion-RP 1.0 (8B) API pricing
Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...
- Input
- $0.8
- per 1M tokens
- Output
- $1.60
- per 1M tokens
- Cached input
- –
- per 1M tokens
- Context
- 33K
- 29K max output
List price from OpenRouter, which passes AionLabs's price through without markup. Refreshed hourly. Released February 4, 2025. Model id aion-labs/aion-rp-llama-3.1-8b.
What Aion-RP 1.0 (8B) costs in practice
1,000 chat messages
$1.60
1,000 input + 500 output tokens each
1,000 RAG questions
$7.20
8,000 input + 500 output tokens each
100 coding-agent steps
$2.72
30,000 input + 2,000 output tokens each, no caching
Aion-RP 1.0 (8B) price history
Aion-RP 1.0 (8B) price by host
- AionLabs$0.8 in · $1.60 out · 33K ctx · 100.0% uptime 24h
Where to buy AionLabs credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Aion-RP 1.0 (8B) cost calculator
Cheaper alternatives to Aion-RP 1.0 (8B)
| Model | Input / 1M | Output / 1M |
|---|---|---|
| OpenAI: GPT-6 Luna Pro | $0.1 | $0.5 |
| OpenAI: GPT-6 Luna | $0.1 | $0.5 |
| Qwen: Qwen3.8 Omni Flash | $0.15 | $0.47 |
| Z.ai: GLM 5.3 FlashX | $0.37 | $1.25 |
| DeepSeek: DeepSeek V4.1 Flash | $0.03 | $0.5 |
| Qwen: Qwen3.8 Flash | $0.15 | $0.47 |
Other AionLabs models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Aion-2.0 | $0.8 | $1.60 |
| Aion-3.0 | $3 | $6 |
| Aion-3.0-Mini | $0.7 | $1.40 |
| Aion 3.5 | $3 | $6 |
| Aion 3.5 Mini | $0.7 | $1.40 |