Drex: an open-weight 8B model tops the decision-model index

Nace.AI's diffusion model runs on your own GPU and leads the Decision Index, as Upstage ships a 512K-context decision endpoint that bills output at zero.

Nowline OCT 10 11:00 PM banner

Top AI stories from the last hour

Top AI stories from the last hour

Copy markdown

  • Drex: a decision model you download, not rent

    Nace.AI's Drex is an 8B diffusion LM with a pointer head (built on NVIDIA's Efficient-DLM-8B) that returns a probability for each typed option — choice, yes/no, or score — in a single forward pass. It leads the six-model Decision Index 0.2 at 52.31, topping retrieval/classification (54.25) and language understanding (57.77). Weights are on Hugging Face (32K context, ~16GB in bf16), so the routing layer runs on your own GPU — but the CC BY-NC 4.0 license rules out paid products (the code is MIT).

  • Solar Decide Flash: 512K context, output billed at $0

    Also new this week, Upstage's Solar Decide Flash (built on Solar Mini 4) answers typed questions in one pass — output tokens are free, input is $0.05/M at a 50% launch discount. A 512K context and ~0.84s median latency make it a cheap router/classifier for long docs; it's served through the System One (/v1/systemone) decisions API, not chat completions.

  • Build this weekend: a model-agnostic gating layer

    Drex, Solar Decide, Cloudflare's Clef and OpenAI's Decisions API now share a similar typed-decision surface, so you can prototype an agent router or policy check against self-hosted Drex and swap to a hosted endpoint for scale without rewriting calls. These return calibrated probabilities instead of prose — exactly what a “should this agent act?” gate needs.