CodeRescue: Budget-Calibrated Recovery Routing for Coding Agents
By Qijia He, Jiayi Cheng, Chenqian Le, Rui Wang, Xunmei Liu, Yixian Chen, Jie Mei, Zhihao Wang, Xupeng Chen, Yuhuan Chen, Tao Wang
"Proposes a recovery routing system for coding agents using supervised learning and Conformal Risk Control to decide post-failure whether to retry cheaply or escalate, optimizing under budget constraints."
Abstract
Coding agents increasingly operate in executable environments where a failed attempt produces actionable feedback rather than merely an incorrect answer. Existing cost-aware systems typically treat such failures as cascade decisions: try a cheap model first, then escalate hard cases to a stronger and more expensive model. In coding, however, execution feedback can also make further cheap-model recovery worthwhile, raising a budgeted deployment question: when should an agent spend more cheap compute, and when should it escalate? We formulate this post-failure decision as recovery routing over heterogeneous actions and train a supervised router from execution rollouts. To make the same router usable under changing budgets, we add a Conformal Risk Control (CRC) layer that selects a deployment-time cost penalty without retraining and provides marginal expected-cost control under exchangeability. Across held-out failures from five coding benchmarks, cheap recovery and escalation exhibit complementary success patterns. The calibrated frontier improves over fixed actions, prompt-only routers, and a binary cascade baseline; in the main GPT-5.4-nano/GPT-5.4 setting, one CRC-calibrated frontier point exceeds always-escalate solve rate while using 35% of its mean recovery cost. Code is available at https://github.com/Qijia-He/agent-budget-control.
Technical Analysis & Implementation
Technical Breakdown§
Problem Formulation§
Coding agents execute code and receive execution feedback. After a failure, the agent can either retry with a cheap model (recovery) or escalate to a stronger, more expensive model. The goal is to minimize cost while maintaining high solve rate. The paper formalizes this as a routing problem: given a failure context $x$ (including code, error, execution trace), choose an action $a \in \mathcal{A}$ (e.g., cheap recovery, escalation) with associated cost $c(a)$ and success probability $p(a|x)$. A router $\pi: \mathcal{X} \to \Delta(\mathcal{A})$ maps contexts to action probabilities. The expected cost for a given solve rate is optimized via a supervised learning approach.
Model Architecture§
The router is a neural network $f_\theta(x)$ that outputs action scores. During training, it learns from execution rollouts: for each failure, multiple actions are tried, yielding success/failure and cost. The training loss is a weighted cross-entropy where positive weight is given to successful actions and negative weight to failures, with a cost penalty term controlled by hyperparameter $\lambda$. The exact loss is:
$$ \mathcal{L}(\theta) = \mathbb{E}_{(x, a, s) \sim \mathcal{D}} \left[ - s \cdot \log \pi_\theta(a|x) + \lambda \cdot c(a) \cdot \mathbb{1}[s=1] \right] $$
where $s \in \{0,1\}$ is success indicator, $c(a)$ is cost, and $\pi_\theta$ is softmax of $f_\theta$.
Conformal Risk Control (CRC) Layer§
To make the router robust to changing budgets, a post-hoc calibration step is added. Given a held-out calibration set, CRC selects a threshold $\tau$ on the cost penalty $\lambda$ such that the expected cost of the router is controlled at level $\alpha$ with high probability. Under exchangeability of calibration data, CRC guarantees:
$$ \mathbb{P}\left( \mathbb{E}[c(\pi_\tau)] \leq \alpha \right) \geq 1 - \delta $$
where $\pi_\tau$ is the router with penalty $\tau$. Practically, they compute empirical cost for a grid of $\tau$ values and pick the smallest $\tau$ such that the empirical cost is below $\alpha$.
Implementation Details§
- Router: 2-layer MLP with hidden size 256, ReLU activations, dropout 0.1.
- Training: Adam optimizer, learning rate 1e-4, batch size 64, early stopping.
- CRC grid: 100 equally spaced $\tau$ values between 0 and 10.
- Calibration set: 200 failure instances held out from training data.
Code Snippet (PyTorch)§
import torch
import torch.nn as nn
class RecoveryRouter(nn.Module):
def __init__(self, input_dim, num_actions):
super().__init__()
self.net = nn.Sequential(
nn.Linear(input_dim, 256),
nn.ReLU(),
nn.Dropout(0.1),
nn.Linear(256, num_actions)
)
def forward(self, x):
return torch.softmax(self.net(x), dim=-1)
# Training loop (simplified)
def train(model, dataloader, lambda_penalty):
optimizer = torch.optim.Adam(model.parameters(), lr=1e-4)
for x, action, success, cost in dataloader:
probs = model(x)
log_probs = torch.log(probs.gather(1, action.unsqueeze(1)))
success = success.float()
loss = - (success * log_probs).mean() + lambda_penalty * (cost * success).mean()
optimizer.zero_grad()
loss.backward()
optimizer.step()Results§
Across five coding benchmarks (e.g., HumanEval, MBPP), the method achieves a frontier of cost vs. solve rate. One operating point under calibration uses 35% of the cost of always-escalate while achieving comparable solve rate.
Interactive LLM Token & Cost Calculator
Estimate token usage and model pricing. Enter your prompt below to see how it is parsed into tokens and calculate the exact API cost for different providers.
Cost Breakdown (USD)
API Pricing Comparison (per Million Tokens)
| Model | Input | Output |
|---|---|---|
| DeepSeek V3.1 | $0.25 | $0.95 |
| Hermes 3 405B Instruct | $1.00 | $1.00 |
| GPT-4o-mini | $0.15 | $0.60 |
| Kimi K3 | $3.00 | $15.00 |
| GLM 4.7 Flash | $0.06 | $0.40 |
| Mistral Medium 3.1 | $0.40 | $2.00 |
| Muse Spark 1.1 | $1.25 | $4.25 |
| Nano Banana 2 (Gemini 3.1 Flash Image) | $0.50 | $3.00 |
| MiniMax M1 | $0.55 | $2.20 |
| GLM 5.2 | $0.79 | $2.47 |
| Nemotron 3 Ultra | $0.60 | $3.60 |
| Gemini 2.5 Flash | $0.30 | $2.50 |
| Claude Opus 4.8 (Fast) | $10.00 | $50.00 |
| GPT-5.2-Codex | $1.75 | $14.00 |
| Claude Sonnet 4.5 | $3.00 | $15.00 |
| Laguna S 2.1 | $0.10 | $0.20 |
| Sonar Reasoning Pro | $2.00 | $8.00 |
| Gemini 3.6 Flash | $1.50 | $7.50 |
| Gemini 3.5 Flash-Lite | $0.30 | $2.50 |
| Claude Sonnet 5 | $2.00 | $10.00 |
| Qwen Plus 0728 (thinking) | $0.26 | $0.78 |
| Claude Opus 4 | $15.00 | $75.00 |
| o4 Mini | $1.10 | $4.40 |
| GPT-4.1 Mini | $0.40 | $1.60 |
| Reka Flash 3 | $0.10 | $0.20 |
| Claude Opus 4.7 (Fast) | $30.00 | $150.00 |
| Gemini 3.1 Flash Lite | $0.25 | $1.50 |
| GPT Chat Latest | $5.00 | $30.00 |
| o1 | $15.00 | $60.00 |
| GPT-4o (2024-11-20) | $2.50 | $10.00 |
| Claude Sonnet 4.6 | $3.00 | $15.00 |
| Claude Opus 4.5 | $5.00 | $25.00 |
| GLM 4.5V | $0.60 | $1.80 |
| GPT-5 Chat | $1.25 | $10.00 |
| Mistral Large 2407 | $2.00 | $6.00 |
| GPT-5 Nano | $0.05 | $0.40 |
| gpt-oss-120b | $0.04 | $0.17 |
| MoonshotAI Kimi Latest | $3.00 | $15.00 |
| Google Gemini Flash Latest | $1.50 | $7.50 |
| GPT-5.5 Pro | $30.00 | $180.00 |
| Grok 4.20 Multi-Agent | $1.25 | $2.50 |
| Claude Haiku 4.5 | $1.00 | $5.00 |
| Qwen2.5 7B Instruct | $0.04 | $0.10 |
| Llama 3.2 3B Instruct | $0.05 | $0.34 |
| Qwen3.5-27B | $0.26 | $2.60 |
| Qwen3.5-122B-A10B | $0.26 | $2.08 |
| Gemini 3.1 Pro Preview Custom Tools | $2.00 | $12.00 |
| GPT-5.3-Codex | $1.75 | $14.00 |
| GPT-4 Turbo Preview | $10.00 | $30.00 |
| GPT-3.5 Turbo Instruct | $1.50 | $2.00 |
| Gemini 3.1 Flash | $0.25 | $1.50 |
| GPT-5.6 Luna Pro | $1.00 | $6.00 |
| GPT-5.6 Luna | $1.00 | $6.00 |
| Gemini 3.1 Pro Preview | $2.00 | $12.00 |
| Qwen3.5 Plus 2026-02-15 | $0.26 | $1.56 |
| Qwen2.5 Coder 32B Instruct | $0.66 | $1.00 |
| Qwen2.5 72B Instruct | $0.36 | $0.40 |
| Command R (08-2024) | $0.15 | $0.60 |
| GPT-4o (2024-08-06) | $2.50 | $10.00 |
| Mistral Nemo | $0.02 | $0.03 |
| GPT-4o-mini (2024-07-18) | $0.15 | $0.60 |
| KAT-Coder-Air V2.5 | $0.15 | $0.60 |
| KAT-Coder-Pro V2.5 | $0.74 | $2.96 |
| Llama 4 Maverick | $0.20 | $0.80 |
| GPT-4o (2024-05-13) | $5.00 | $15.00 |
| Kimi K2.6 | $0.68 | $3.42 |
| GLM 4.6V | $0.30 | $0.90 |
| Nova 2 Lite | $0.30 | $2.50 |
| Qwen3 VL 8B Thinking | $0.12 | $1.36 |
| Llama 3.2 11B Vision | $0.34 | $0.34 |
| Llama 4 Scout | $0.10 | $0.30 |
| Llama 3 8B Instruct | $0.14 | $0.14 |
| Mixtral 8x22B Instruct | $2.00 | $6.00 |
| Mistral Large | $2.00 | $6.00 |
| GPT-3.5 Turbo (older v0613) | $1.00 | $2.00 |
| MiniMax M2.7 | $0.25 | $1.00 |
| GPT-5.4 Nano | $0.20 | $1.25 |
| Sonar Pro | $3.00 | $15.00 |
| Sonar Deep Research | $2.00 | $8.00 |
| GLM 5 | $0.95 | $2.55 |
| Qwen3 30B A3B Instruct 2507 | $0.10 | $0.30 |
| Qwen3 Next 80B A3B Instruct | $0.10 | $1.10 |
| GLM 4.5 Air | $0.13 | $0.85 |
| Qwen3 Coder 480B A35B | $0.30 | $1.00 |
| Sonar | $1.00 | $1.00 |
| Claude 3 Haiku | $0.25 | $1.25 |
| Claude 3.5 Sonnet v2 | $3.00 | $15.00 |
| MiniMax M2-her | $0.30 | $1.20 |
| GPT-5.5 | $5.00 | $30.00 |
| GPT-3.5 Turbo 16k | $3.00 | $4.00 |
| Mistral Small 4 | $0.15 | $0.60 |
| UI-TARS 7B | $0.10 | $0.20 |
| GLM 5 Turbo | $1.20 | $4.00 |
| Mistral Small 3 | $0.10 | $0.30 |
| Qwen3 Max Thinking | $0.78 | $3.90 |
| Qwen3 Coder Next | $0.11 | $0.80 |
| Morph V3 Fast | $0.80 | $1.20 |
| Gemini 2.5 Pro Preview 06-05 | $1.25 | $10.00 |
| GPT-4o | $2.50 | $10.00 |
| Claude Fable 5 | $10.00 | $50.00 |
| Gemma 3n 4B | $0.06 | $0.12 |
| Qwen3.7 Plus | $0.32 | $1.28 |
| Gemini 2.5 Pro Preview 05-06 | $1.25 | $10.00 |
| Claude Opus 4.8 | $5.00 | $25.00 |
| DeepSeek V3.1 Terminus | $0.27 | $1.00 |
| o4 Mini High | $1.10 | $4.40 |
| Qwen3 30B A3B Thinking 2507 | $0.13 | $1.56 |
| Mistral Small 3.2 24B | $0.10 | $0.30 |
| GPT-3.5 Turbo | $0.50 | $1.50 |
| o3 Pro | $20.00 | $80.00 |
| MiniMax M3 | $0.30 | $1.20 |
| Step 3.7 Flash | $0.20 | $1.15 |
| Qwen3.7 Max | $1.48 | $4.42 |
| Step 3.5 Flash | $0.10 | $0.30 |
| Kimi K2.5 | $0.57 | $2.85 |
| gpt-oss-20b | $0.03 | $0.13 |
| Claude Opus 4.1 | $15.00 | $75.00 |
| o3 | $2.00 | $8.00 |
| Llama 3.1 8B Instruct | $0.05 | $0.08 |
| WizardLM-2 8x22B | $0.62 | $0.62 |
| Gemini 3.5 Flash | $1.50 | $9.00 |
| GLM 5V Turbo | $1.20 | $4.00 |
| DeepSeek V3.2 | $0.27 | $0.40 |
| Nano Banana Pro (Gemini 3 Pro Image Preview) | $2.00 | $12.00 |
| GPT-5.1 | $1.25 | $10.00 |
| GPT-5 Image Mini | $2.50 | $2.00 |
| Qwen3 8B | $0.12 | $0.46 |
| GPT-4 | $30.00 | $60.00 |
| Qwen3.6 Flash | $0.19 | $1.13 |
| DeepSeek V4 Pro | $0.43 | $0.87 |
| Mistral Large 3 | $0.50 | $1.50 |
| Grok 4.20 | $1.25 | $2.50 |
| GPT-5 Mini | $0.25 | $2.00 |
| DeepSeek V3 0324 | $0.27 | $1.12 |
| o1-pro | $150.00 | $600.00 |
| Llama 3.3 70B Instruct | $0.13 | $0.40 |
| Qwen Plus 0728 | $0.26 | $0.78 |
| Qwen3 235B A22B Thinking 2507 | $0.30 | $3.00 |
| Claude Opus 4.7 | $5.00 | $25.00 |
| GPT-5.4 Mini | $0.75 | $4.50 |
| Seed-2.0-Mini | $0.10 | $0.40 |
| Qwen3.5-Flash | $0.07 | $0.26 |
| GPT Audio | $2.50 | $10.00 |
| Yi-Lightning | $0.15 | $0.30 |
| GPT Audio Mini | $0.60 | $2.40 |
| Grok 4.3 | $1.25 | $2.50 |
| GPT-5.1 Chat | $1.25 | $10.00 |
| Seed-2.0-Lite | $0.25 | $2.00 |
| Qwen3.5 397B A17B | $0.39 | $2.34 |
| MiniMax M2.5 | $0.15 | $0.90 |
| GPT-5.1-Codex | $1.25 | $10.00 |
| Kimi K2 0711 | $0.57 | $2.30 |
| Command R | $0.15 | $0.60 |
| Solar Pro 3 | $0.15 | $0.60 |
| Mistral Small 3.1 24B | $0.35 | $0.56 |
| Mistral Medium 3 | $0.40 | $2.00 |
| GPT-5.6 Sol Pro | $5.00 | $30.00 |
| Claude Opus 4.6 | $5.00 | $25.00 |
| GPT-5.1-Codex-Max | $1.25 | $10.00 |
| GPT-5.6 Sol | $5.00 | $30.00 |
| Laguna XS 2.1 | $0.06 | $0.12 |
| Nex-N2-Mini | $0.03 | $0.10 |
| Ministral 3 14B 2512 | $0.20 | $0.20 |
| Fugu Ultra | $5.00 | $30.00 |
| Nex-N2-Pro | $0.25 | $1.00 |
| GPT-5 | $1.25 | $10.00 |
| Gemini 2.5 Pro | $1.25 | $10.00 |
| Grok 4.5 | $2.00 | $6.00 |
| Gemini 3.1 Pro | $2.00 | $12.00 |
| GPT-4.1 Nano | $0.10 | $0.40 |
| Granite 4.1 8B | $0.05 | $0.10 |
| Laguna M.1 | $0.20 | $0.40 |
| Hy3 preview | $0.06 | $0.21 |
| Google Gemini Pro Latest | $2.00 | $12.00 |
| Seed 1.6 Flash | $0.07 | $0.30 |
| MiniMax M2 | $0.30 | $1.20 |
| Llama 4 Maverick | $0.20 | $0.80 |
| Qwen3.6 35B A3B | $0.14 | $1.00 |
| Qwen3 VL 32B Instruct | $0.10 | $0.42 |
| Qwen3.6 Max Preview | $1.04 | $6.24 |
| GPT-5.4 Image 2 | $8.00 | $15.00 |
| Claude Opus Latest | $5.00 | $25.00 |
| GLM 5.1 | $0.97 | $3.04 |
| o3 Deep Research | $10.00 | $40.00 |
| o4 Mini Deep Research | $2.00 | $8.00 |
| Nova Lite 1.0 | $0.06 | $0.24 |
| Gemma 4 26B A4B | $0.07 | $0.34 |
| Nano Banana 2 (Gemini 3.1 Flash Image Preview) | $0.50 | $3.00 |
| Qwen3.5-35B-A3B | $0.14 | $1.00 |
| Ministral 3 8B 2512 | $0.15 | $0.15 |
| R1 0528 | $0.50 | $2.15 |
| Qwen 2.5-Coder 32B | $0.35 | $0.70 |
| MiMo-V2.5-Pro | $0.43 | $0.87 |
| Llama Guard 4 12B | $0.18 | $0.18 |
| MiMo-V2.5 | $0.14 | $0.28 |
| Qwen3 30B A3B | $0.13 | $0.52 |
| Gemma 4 31B | $0.12 | $0.37 |
| GLM 4.7 | $0.40 | $1.75 |
| Gemini 3 Flash Preview | $0.50 | $3.00 |
| Ministral 3 3B 2512 | $0.10 | $0.10 |
| Gemma 3 4B | $0.05 | $0.10 |
| GLM 4.6 | $0.50 | $2.00 |
| Qwen3 Max | $0.78 | $3.90 |
| Qwen3.6 Plus | $0.33 | $1.95 |
| Reka Edge | $0.10 | $0.10 |
| Nemotron 3 Super | $0.08 | $0.45 |
| GPT-5.4 Pro | $30.00 | $180.00 |
| GPT-5.4 | $2.50 | $15.00 |
| Nano Banana (Gemini 2.5 Flash Image) | $0.30 | $2.50 |
| Doubao Pro | $0.80 | $1.60 |
| Qwen3 VL 30B A3B Thinking | $0.13 | $1.56 |
| Mixtral 8x22B | $0.50 | $1.00 |
| GPT-5.6 Terra Pro | $2.50 | $15.00 |
| Qwen3.5-9B | $0.10 | $0.15 |
| GPT-5.6 Terra | $2.50 | $15.00 |
| Mercury 2 | $0.25 | $0.75 |
| GPT-5.3 Chat | $1.75 | $14.00 |
| Gemini 3.1 Flash Lite Preview | $0.25 | $1.50 |
| Hunyuan A13B Instruct | $0.14 | $0.57 |
| o3 Mini | $1.10 | $4.40 |
| GPT-5.2 Pro | $21.00 | $168.00 |
| Qwen3 VL 30B A3B Instruct | $0.13 | $0.52 |
| Codestral 2508 | $0.30 | $0.90 |
| Hy3 | $0.14 | $0.58 |
| Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) | $0.25 | $1.50 |
| Nano Banana Pro (Gemini 3 Pro Image) | $2.00 | $12.00 |
| KAT-Coder-Pro V2 | $0.30 | $1.20 |
| Palmyra X5 | $0.60 | $6.00 |
| Claude Fable Latest | $10.00 | $50.00 |
| Qwen3 Coder 30B A3B Instruct | $0.07 | $0.27 |
| Seed 1.6 | $0.25 | $2.00 |
| Nemotron 3 Nano 30B A3B | $0.05 | $0.20 |
| GPT-4.1 | $2.00 | $8.00 |
| Granite 4.0 Micro | $0.02 | $0.11 |
| Grok Build 0.1 | $1.00 | $2.00 |
| Qwen3 VL 8B Instruct | $0.12 | $0.46 |
| Kimi K2 0905 | $0.60 | $2.50 |
| Mistral Medium 3.5 | $1.50 | $7.50 |
| Anthropic Claude Haiku Latest | $1.00 | $5.00 |
| GPT-5 Pro | $15.00 | $120.00 |
| Anthropic Claude Sonnet Latest | $2.00 | $10.00 |
| DeepSeek V3.2 Exp | $0.27 | $0.41 |
| Qwen3.5 Plus 2026-04-20 | $0.30 | $1.80 |
| Qwen3.6 27B | $0.45 | $2.70 |
| Nova Premier 1.0 | $2.50 | $12.50 |
| Sonar Pro Search | $3.00 | $15.00 |
| DeepSeek R1 | $0.70 | $2.50 |
| Qwen 2.5 72B | $0.40 | $0.80 |
| GLM 4.5 | $0.60 | $2.20 |
| Kimi K2.7 Code | $0.82 | $3.75 |
| Lyria 3 Pro Preview | $0.00 | $0.00 |
| Gemma 3 12B | $0.05 | $0.15 |
| Command A | $2.50 | $10.00 |
| Gemini 2.5 Flash Lite | $0.10 | $0.40 |
| GPT-4o-mini Search Preview | $0.15 | $0.60 |
| Qwen3 235B A22B Instruct 2507 | $0.09 | $0.55 |
| Qwen3 32B | $0.08 | $0.28 |
| GPT-5.2 | $1.75 | $14.00 |
| Devstral 2 2512 | $0.40 | $2.00 |
| MiniMax M2.1 | $0.30 | $1.20 |
| GPT-5.2 Chat | $1.75 | $14.00 |
| Command R+ | $2.50 | $10.00 |
| GPT-5 Image | $10.00 | $10.00 |
| Qwen-Plus | $0.26 | $0.78 |
| Grok 4.20 | $1.25 | $2.50 |
| GPT-5.1-Codex-Mini | $0.25 | $2.00 |
| DeepSeek V3 | $0.20 | $0.80 |
| Command R7B (12-2024) | $0.04 | $0.15 |
| Llama 3.3 70B Instruct | $0.13 | $0.40 |
| Hermes 4 70B | $0.13 | $0.40 |
| DeepSeek V4 Flash | $0.10 | $0.20 |
| Llama 3.1 70B Instruct | $0.40 | $0.40 |
| GPT-4 Turbo | $10.00 | $30.00 |
| Kimi K2 Thinking | $0.60 | $2.50 |
| Hermes 4 405B | $1.00 | $3.00 |
| Jamba Large 1.7 | $2.00 | $8.00 |
| Morph V3 Large | $0.90 | $1.90 |
| ERNIE 4.0 | $1.20 | $2.40 |
| Voxtral Small 24B 2507 | $0.10 | $0.30 |
| gpt-oss-safeguard-20b | $0.07 | $0.30 |
| Gemini 2.5 Flash Lite Preview 09-2025 | $0.10 | $0.40 |
| Lyria 3 Clip Preview | $0.00 | $0.00 |
| Qwen3 VL 235B A22B Thinking | $0.26 | $2.60 |
| Qwen3 VL 235B A22B Instruct | $0.21 | $1.90 |
| GPT-4o Search Preview | $2.50 | $10.00 |
| Qwen2.5 VL 72B Instruct | $0.80 | $1.00 |
| R1 Distill Llama 70B | $0.80 | $0.80 |
| R1 | $0.70 | $2.50 |
| Qwen3 Coder Plus | $0.65 | $3.25 |
| MiniMax-01 | $0.20 | $1.10 |
| Qwen3 Coder Flash | $0.20 | $0.97 |
| Qwen3 Next 80B A3B Thinking | $0.10 | $0.78 |
| Phi 4 | $0.07 | $0.14 |
| Mistral Large 3 2512 | $0.50 | $1.50 |
| GPT-5 Codex | $1.25 | $10.00 |
| Claude Sonnet 4 | $3.00 | $15.00 |
| Qwen3 14B | $0.12 | $0.24 |
| ERNIE 4.5 VL 424B A47B | $0.42 | $1.25 |
| Qwen3 235B A22B | $0.46 | $1.82 |
| Gemma 2 27B | $0.65 | $0.65 |
| Mistral Large 2 | $0.60 | $1.80 |
| Llama 3.1 405B | $0.80 | $0.80 |
| Llama 3.1 8B | $0.04 | $0.04 |
| Gemma 3 27B | $0.10 | $0.30 |
| Saba | $0.20 | $0.60 |
| o3 Mini High | $1.10 | $4.40 |
| Nova Micro 1.0 | $0.04 | $0.14 |
| Hermes 3 70B Instruct | $0.70 | $0.70 |
| Nova Pro 1.0 | $0.80 | $3.20 |
| Inflection 3 Pi | $2.50 | $10.00 |
| Llama 3.2 11B Vision Instruct | $0.34 | $0.34 |
| Inflection 3 Productivity | $2.50 | $10.00 |
| Llama 3.2 1B Instruct | $0.03 | $0.20 |
| Gemini 2.0 Flash | $0.10 | $0.40 |
| Hunyuan Pro | $0.60 | $1.20 |
Accelerate your workflow with Araho
Need help choosing the right model for your product? We build AI-native MVPs.
Get your MVP built in weeks with top-tier AI developers.