Nvidia's agent-tuned Vera CPU: 88 cores, 1.8x on agentic work

Olympus cores chase single-thread speed for latency-bound agent loops; the rack claims 6X CPU throughput, and Kimi K3's open weights reportedly drop July 27.

Nowline Jul 22 12:00 PM banner

Top AI stories from the last hour

Top AI stories from the last hour

Copy markdown

  • Olympus cores, tuned for agent loops

    Vera packs 88 custom Olympus cores and 176 threads with 164MB of L3 and up to 1.2 TB/s of LPDDR5X, and Nvidia claims up to 1.8x higher performance on agentic workloads versus x86. The bet: agent orchestration is single-thread- and latency-bound, so per-core IPC matters more to your loop times than raw core count.

  • A 256-chip rack, 6X the CPU throughput

    The full Vera rack stacks 256 liquid-cooled CPUs for up to a 6X gain in CPU throughput over the prior generation, scaling dual-socket over NVLink-C2C with PCIe 6.4 and CXL 3.1. This is the host silicon your cloud-hosted Claude, GPT, and Gemini agents increasingly run on — more concurrent agent VMs packed per rack.

  • Per-VM encryption for agent sandboxes

    Vera ships Confidential Computing with per-VM memory encryption, isolating each tenant or agent VM at the silicon level. After a week of sandbox escapes across Cursor, Codex, and Gemini CLI, that's a hardware-level answer to stopping one runaway agent from reading another's memory.

  • Kimi K3's open weights reportedly land July 27

    Moonshot AI reportedly plans to release the weights of its 2.8-trillion-parameter K3 mixture-of-experts model on July 27, after demand outran its own servers. If it holds, self-hosters get zero-per-token inference of a model that's topped coding leaderboards — a real weekend project for anyone with the GPUs to run it.