#open-source

14 posts

Close-up of a black Gigabyte graphics card — the kind of GPU that runs Qwen3.8-27B at 4-bit in 24 GB of VRAM

Qwen 3.8-27B: 73 on Terminal Bench 2.1 and Still a 24 GB Model

Qwen 3.8-27B is the new dense flagship you can run on a single 24-32 GB GPU. It jumps to 73.0 on Terminal Bench 2.1 (up from 63.4), adds native image and hour-scale video understanding, and ships under Apache 2.0. What changed versus Qwen3.6-27B, how to run it at 4-bit, and the catches I would watch.

Rows of server racks with cabling in a modern data center aisle.

Your Agents Can Idle Free on a Managed Runtime

Microsoft's Agent Framework Harness and Foundry Hosted Agents hit GA August 3 — a consumption-billed runtime with per-session isolation and scale-to-zero. Here's what it costs and where CodeAct fits.

Close-up macro photography of a modern blue server circuit board with visible processor chips and copper traces.

Agents Now Call Physics Solvers Like Any Tool

NVIDIA embeds sparse linear algebra solvers into the Agent Toolkit—cuISS, cuDSS, cuEST—so autonomous design agents call physics simulation without leaving the flow. Free libraries, GPU-only execution.