Solar Pro 4: a 512K-context agent model at $0.30/1M input tokens

Upstage's drop-in, OpenAI-compatible model scores 57 on Terminal-Bench, flags gaps instead of bluffing, and runs 90% off to Sep 10 — plus Meta's 30B open agent.

Nowline AUG 17 2:00 AM banner

Top AI stories from the last hour

Top AI stories from the last hour

Copy markdown

  • It finishes multi-step jobs, not just chats

    Solar Pro 4 is tuned for agent loops rather than single answers: it scores 57 on Terminal-Bench v2.1 (live, multi-step shell tasks) and handles multi-turn tool use across knowledge bases. Upstage's whole pitch is end-to-end completion — the model runs the job to the finish instead of handing back a plan.

  • $0.30 in, $1.20 out — and 90% off until Sep 10

    Pricing is $0.30 per 1M input tokens and $1.20 per 1M output, with cached input at $0.06 — already cheap for a long-context agent model. A launch promo cuts all of it by 90% (roughly $0.03/$0.12) through Sep 10, 2026 on the Upstage Console and OpenRouter.

  • 512K context, and it says “I can’t verify” instead of bluffing

    The window is 512K tokens with up to 128K output, and it scores 71 on AA-LCR long-document reasoning. It's built to flag missing evidence rather than fill the gap — a hallucination guardrail that matters when you point an agent at contracts, codebases, or research.

  • Migration is a one-line endpoint swap

    The API is OpenAI-compatible: point your client at Upstage's endpoint and set the model to solar-pro4 — that's the whole change. It's also live on OpenRouter, SolarChat, Hermes Agent, and Upstage Studio, so you can A/B it against your current backend today.

  • Need the weights? Solar Open 2 self-hosts

    If you want on-prem control or data residency, Upstage's open-weights Solar Open 2 is the companion release. It matches Pro 4 on knowledge, math, and coding but trails on agent work (about 14 points lower on terminal tasks), so reach for Pro 4 when you need autonomy.

  • Elsewhere: Meta open-sources Muse Glimmer, a 30B agent for one GPU

    Meta returned to open weights with Muse Glimmer, an Apache-2.0 30B model tuned for agents that runs on a single consumer GPU. It's a commercial-friendly local alternative if you'd rather host than pay per token — no caps, no metered bill.