Loading...
Loading...
Maestro Reasoning is Arcee's flagship analysis model: a 32 B‑parameter derivative of Qwen 2.5‑32 B tuned with DPO and chain‑of‑thought RL for step‑by‑step logic. Compared to the earlier 7 B preview, the production 32 B release widens the context window to 128 k tokens and doubles pass‑rate on MATH and GSM‑8K, while also lifting code completion accuracy. Its instruction style encourages structured "thought → answer" traces that can be parsed or hidden according to user preference. That transparency pairs well with audit‑focused industries like finance or healthcare where seeing the reasoning path matters. In Arcee Conductor, Maestro is automatically selected for complex, multi‑constraint queries that smaller SLMs bounce.
Price/1M
$1.50
521st cheapest
384% above median
Top 77%
Context Window
131K
145th largest
Top 63%
Input
$0.90
per 1M tokens
Output
$3.30
per 1M tokens
Blended
$1.50
per 1M tokens
Cheaper than 23% of models. Median price is $0.31/1M tokens.
Daily
$1.50
Monthly
$45.00
Context Window
131K
tokens
Larger than 37% of models
Max Output
32K
tokens
24% of context
Context Window Comparison