Models and Pricing
Latest Models
Section titled “Latest Models”The models below are currently available through the relaxAI API. All pricing is based on usage per 1M tokens.
Use the model names exactly as shown when specifying the model parameter in your API requests.
Mistral-7b-embedding
Embedding Model
Input Price: £0.20
Output Price: £0.00
Context Length: 32k tokens
Muse-Glimmer-30B
Multimodal model designed for high-throughput agentic workloads at low cost per token, specifically coding, tool use, and autonomous tasks.
Input Price: £0.18
Output Price: £0.66
Context Length: 131k tokens
DeepSeek-V41-Flash
Efficient MoE model optimised for high throughput agentic coding, tool-use, and long-context workloads (up to 1M tokens) with native multimodal input.
Input Price: £0.18
Output Price: £0.72
Context Length: 1049k tokens
Nemotron-3-Super
A reasoning model from NVIDIA, best for building specialized AI agents and long-context workflows.
Input Price: £0.22
Output Price: £0.67
Context Length: 262k tokens
DeepSeek-V4-Pro
Frontier MoE model designed for reasoning, coding, and long-context agentic tasks (up to 1M tokens)
Input Price: £1.17
Output Price: £2.33
Context Length: 1049k tokens