Products

The Novo model family

Two models, one API. Pick the right balance of intelligence and speed for your use case.

Novo-4-Pro Flagship

Our most capable model. Advanced multi-step reasoning, precise instruction following, and stable long-context recall. Recommended for analysis, code generation, and agentic workflows.

128K context JSON mode Function calling Vision
Get API access →
Novo-4-Flash Efficient

Built for throughput. Streams at up to 60 tokens per second while keeping strong quality on everyday tasks — classification, extraction, chat, and drafting.

128K context 60 tok/s Function calling
Get API access →
Benchmarks

Measured on public evals

Full methodology is published with every technical report. Higher is better.

BenchmarkNovo-4-ProNovo-4-FlashCategory
MMLU-Pro82.475.1Knowledge
GPQA-Diamond61.749.3Reasoning
SWE-bench Verified48.933.2Code
LongBench-v254.241.8Long context
LiveCodeBench39.628.4Code

Scores from internal runs, Sep 2026. See the Novo-4 technical report for confidence intervals and eval details.

Try the Novo models now

Start on the free tier — upgrade when you ship to production.