Show HN: HarnessRouter: Unified interface for agent harnesses
A canonical API lets you run Codex, Claude Code, Hermes, and other managed harnesses as one backend instead of rebuilding integration code for each.
16 articles · 5 categories
The finishable daily brief
Monday, Aug 17, 2026
16 articles · 5 categories
read top to bottom · then stop
In 30 seconds
Today's agent-tooling news pointed toward consolidation: HarnessRouter shipped a canonical API for swapping coding-agent backends, DeepSeek open-sourced a permissively licensed coding harness, and SpaceXAI launched persistent cloud agents that can operate a browser and inbox on their own.
On the infra side, vLLM-Omni showed a 124GB model served on 64GB of memory via layerwise offload, and Grab published a rare hard adoption number — AI agents cut its mechanical analytics workload from 44% to 30% of analyst time in four months.
Builders shipped unifying and coordinating infrastructure for agent harnesses today: a common API across coding agents, a payments middleware for autonomous spend, a canvas UI for steering long runs, and a desktop-automation SDK.
A canonical API lets you run Codex, Claude Code, Hermes, and other managed harnesses as one backend instead of rebuilding integration code for each.
AgentCore Payments middleware signs x402 payments with deterministic per-session budgets and traces every transaction in LangSmith.
GitHub's canvas UI keeps long agentic coding sessions visible and steerable instead of losing intent in a scrolling chat log.
An open-source SDK gives agents desktop GUI automation — clicking, typing, and screen-reading across Windows, macOS, and Linux.
A frontier-adjacent player and a YC startup shipped new agent products today: persistent cloud agents from SpaceXAI, an open coding harness from DeepSeek, and a benchmarked router for voice-AI model stacks.
SpaceXAI's Grok Bot runs persistent agents on dedicated cloud computers that can operate websites, apps, and inboxes autonomously.
DeepSeek open-sourced an MIT-licensed harness for AI coding agents, adding a permissively licensed alternative to closed agent tooling.
Speko benchmarks speech-to-text, LLM, and text-to-speech combinations and picks the optimal stack for a given latency, cost, and quality constraint.
Infra news skewed toward serving bigger models on less hardware and hardening the compute stack behind them, plus a reminder that GPU vendors now profit from labs training their own models instead of buying inference.
vLLM-Omni's Distributed Layerwise Offload shards and streams DiT weights across devices, serving a 124GB Cosmos3 model on 64GB of HBM and charting a path past 200B parameters.
NVIDIA's Nemotron 3.5 Lightning, a 30B mixture-of-experts model with 3B active parameters built for high-volume agentic workloads, is now deployable via SageMaker JumpStart.
Nvidia's push toward custom silicon and open training stacks gives labs an incentive to build their own models instead of buying inference from Anthropic or OpenAI.
NVIDIA frames AI factories as revenue-generating infrastructure and details how it's hardening the compute stack that trains and serves models at that scale.
Practitioners published concrete patterns for judging non-deterministic agent output in production, plus a real adoption number for handing analytics work to agents.
Agentic fitness functions pair AI agents with versioned rubrics to evaluate judgment-heavy architectural qualities that deterministic rule checks can't capture.
A production pattern for LLM-powered selection systems: separate semantic extraction from deterministic decision logic and restrict output schemas to tame non-determinism.
Grab's AI agents cut mechanical analytics work from 44% of analyst time in February to 30% by June, combining agent autonomy with certified data and human oversight.
OpenAI published on two fronts today: sharpening its own defenses against AI-enabled attackers, and funding independent research into AI economic policy.
OpenAI details how it's hardening its own defenses as AI reshapes both attacker and defender capabilities in cybersecurity.
OpenAI is funding 14 independent projects exploring AI policy proposals aimed at expanding economic opportunity in the "Intelligence Age."
You are caught up for this edition