Skip to content

Models and Pricing

The models below are currently available through the relaxAI API. All pricing is based on usage per 1M tokens. Use the model names exactly as shown when specifying the model parameter in your API requests.

Mistral-7b-embedding

Embedding Model


Input Price: £0.20
Output Price: £0.00
Context Length: 32k tokens
Muse-Glimmer-30B

Multimodal model designed for high-throughput agentic workloads at low cost per token, specifically coding, tool use, and autonomous tasks.


Input Price: £0.18
Output Price: £0.66
Context Length: 131k tokens
DeepSeek-V41-Flash

Efficient MoE model optimised for high throughput agentic coding, tool-use, and long-context workloads (up to 1M tokens) with native multimodal input.


Input Price: £0.18
Output Price: £0.72
Context Length: 1049k tokens
Nemotron-3-Super

A reasoning model from NVIDIA, best for building specialized AI agents and long-context workflows.


Input Price: £0.22
Output Price: £0.67
Context Length: 262k tokens
DeepSeek-V4-Pro

Frontier MoE model designed for reasoning, coding, and long-context agentic tasks (up to 1M tokens)


Input Price: £1.17
Output Price: £2.33
Context Length: 1049k tokens