Context
1M
Max output
131K
Serving providers
1
Cheapest input
$1.50 /1M
Prices are on-demand list rates. Enterprise commitments (committed-use discounts, provisioned throughput, negotiated rates) can be materially lower.

Benchmarks

independent evaluations · source: Hugging Face leaderboards
Individual benchmarks · click a row for effort variants
BenchmarkBest scoreEffortCost / task
SWE-bench Verified86%default

Pricing by serving provider

standard tier · per 1M tokens · click a provider to expand
Serving providerInput /1MOutput /1MEndpoints
NovitaCheapest$1.50$4.501
EndpointStd inStd outCached inBatch in / outModel ID
StandardCheapest$1.50$4.50$0.30novita/mindai/macaron-v1-venti

Price history

input + output $/1M since we started tracking
Input Output

Price history is accruing. We record a point each day a price changes; the full backfill lands shortly.

Cost calculator

estimate your monthly spend on this model
$660
estimated / month

Model IDs

copy the exact identifier for your platform
novita/mindai/macaron-v1-venti

Frequently asked questions

Macaron V1 Venti pricing, context and availability

How much does Macaron V1 Venti cost?

Macaron V1 Venti costs $1.50 per 1M input tokens and $4.50 per 1M output tokens at its cheapest provider via Novita.

What is the context window of Macaron V1 Venti?

Macaron V1 Venti accepts up to 1M tokens of context and can return up to 131K output tokens.

What can Macaron V1 Venti do?

Macaron V1 Venti supports tool use, reasoning and prompt caching.

Other Other models

compare pricing across the Other lineup
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily

Every weekday

AI moves fast. Here's your debrief.

News, analysis, tools, and more.

For people who build with AI