Skip to main content
June 30, 2026

Axon Eido 3 model family

Three new flagship models, all with a 1M-token context window, native multimodal input, and tool calling / structured outputs / web search.

Axon Eido 3 Pro — axon-eido-3-pro

The most intelligent model in the family. Built for long-running agents, tool calling, coding, and deep research. Frontier reasoning with native multimodal understanding.
  • Context: 1M tokens
  • Max output: 128k tokens
  • Input: $3.00 / 1M tokens
  • Output: $9.00 / 1M tokens
  • Cached input: $1.50 / 1M tokens

Axon Eido 3 Mini — axon-eido-3-mini

The everyday general-purpose workhorse for day-to-day tasks — chat, extraction, RAG, and agents — with the same 1M context, fast and cost-efficient.
  • Context: 1M tokens
  • Max output: 128k tokens
  • Input: $1.50 / 1M tokens
  • Output: $4.50 / 1M tokens
  • Cached input: $0.75 / 1M tokens

Axon Eido 3 Flash — axon-eido-3-flash

Fast, low-latency model suitable for low-complexity, day-to-day coding tasks and general use — with a 1M-token context and native multimodal input.
  • Context: 1M tokens
  • Max output: 128k tokens
  • Input: $0.30 / 1M tokens
  • Output: $0.90 / 1M tokens
  • Cached input: $0.15 / 1M tokens

Caching

Cached input tokens are automatically discounted 50% off the standard input price. Caching kicks in automatically for repeated prompt prefixes across all five Axon models.
December 16th, 2025

Axon Model 2.0

Major performance upgrade with reduced pricing and state-of-the-art benchmarks.
  • 2x Faster Speed: Optimized inference latency for rapid responses.
  • 89% LiveCodeBench: Achieving top-tier accuracy in code generation benchmarks.
  • New Pricing: Reduced to 1/1Minput(was1/1M input (was 2) and 4/1Moutput(was4/1M output (was 6).

Core Improvements

  • Parallel Processing: Enabled parallel client functionality for improved concurrent request handling.
  • Rate Limiting: Refined strategies and fixed context propagation in API responses.
November 20th, 2025

Cortex Intelligence Upgrade

Updated Cortex modules with advanced planning capabilities.
  • Enhanced Reasoning: New planning prompts and task memory helper.
  • Better Execution: Significantly improved agent reliability in complex tasks.
October 10th, 2025

Axon Playground Axon Playground is now available to test and explore the

capabilities of Axon models in the dashboard.
October 5th, 2025

New Axon-Mini model released Axon Mini 1 is now available for everyday

low-effort tasks with deep reasoning. Read more here.
September 30th, 2025

Deep Reasoner v2.0 Deep Reasoner v2.0 is now available with massive boost

in speed, reasoning and accuracy.
September 23rd, 2025

Global Rate Limits Global API Rate Limits are not applicable based on your

Tier. More details here.
September 11th, 2025

Tool call usage response format fixed Tool call usage response format is

now fixed for AI editors and tools to consume.
September 7th, 2025

Fix reasoning summary in response Breaking reasoning summary in response

format is not fixed.
September 3rd, 2025

Support for v1/messages API for Claude Code Deep Reasoner now outputs

less tokens to reduce the cost of the model.
September 1st, 2025

Reduced tokens for Deep Reasoner Deep Reasoner now outputs less tokens to

reduce the cost of the model.
August 25th, 2025

Improved Response Speed MatterAI Axon now supports improved response speed

to provide a high level overview of the reasoning process. This allows you to balance between accuracy and performance.
August 20th, 2025

Web Search & Web Fetch Tool Calling

MatterAI Axon now supports web search and web fetch tool calling to provide a high level overview of the reasoning process. This allows you to balance between accuracy and performance.
  • web_search - LLM will use web search to find relevant information
  • web_fetch - LLM will scrape the provided URL to find relevant information
Tool Calling is always available and enabled by default.
August 10th, 2025

Reasoning Summary in Response

MatterAI Axon now supports reasoning summary in response to provide a high level overview of the reasoning process. This allows you to balance between accuracy and performance.
  • auto - AI will decide whether to provide a summary or not
  • none - AI will never provide a summary
Reasoning summary is always enclosed in a <reasoning_start> REASONING <reasoning_end> tag.
August 5, 2025

Reasoning Effort Options

MatterAI Axon now supports reasoning effort options to control the level of reasoning and analysis performed by the AI model. This allows you to balance between accuracy and performance.
  • None - No reasoning, fastest
  • Low - Fastest, least accurate
  • Medium - Good balance
  • High - Most accurate, slowest
August 1, 2025

Private Beta Release MatterAI Axon private beta release