AIToday
Latent SpacePublished: Apr 30, 2026, 13:00 JST1 min read

Inference compute emerges as strategic bottleneck as AI agents reshape coding platforms and harness engineering becomes a key optimization layer

Inference compute emerges as strategic bottleneck as AI agents reshape coding platforms and harness engineering becomes a key optimization layer

3 Key Points

  1. OpenAI is expanding Codex from a coding tool into a general work surface with persistent context, integrations (Supabase, Figma plugin), and team rollout; launched Codex-only seats with $0 seat fee for eligible Business/Enterprise customers through end of June.

  2. WebSocket mode on OpenAI's Responses API keeps state warm across tool calls and yields up to 40% faster agentic workflows; Cursor released an SDK exposing its runtime, harness, and models for use in CI/CD and embedded agents; VS Code shipped semantic indexing, cross-repo search, and prompt evaluation extensions—shifting focus from raw model latency to agent-loop systems engineering and memory retrieval.

  3. Harness engineering research (Agentic Harness Engineering, HALO) shows gains in agent performance through revertible components and trace analysis: Terminal-Bench 2 pass@1 improved from 69.7% to 77.0% in ten iterations, and AppWorld improved from 73.7 to 89.5 on Sonnet 4.6; LangChain's Deep Agents product line introduced Harness Profiles for per-model prompt and tool tuning with built-in profiles for OpenAI, Anthropic, and Google models.

Ask the AI about this article →

Get AI news like this every morning

For example, today's edition would include:

  • askpolly raises $3M to turn social media chatter into market researchSiliconANGLE AI · 1h ago
  • AI agents compress cyberattacks to minutesSiliconANGLE AI · 1h ago
  • SanDisk pushes HBF ecosystem as AI inference growsDIGITIMES Asia · 1h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Next articleResearchers introduce atomic-quality probe for governing skill updates in compositional robot policies, addressing how library changes affect task success.