This Week in Agentic AI: July 27–August 3, 2026

A stateless MCP update brought renewed attention to agent integrations and enterprise scalability. Model releases emphasized efficiency, while small evaluation tools and scientific-computing examples gave builders practical ways to assess and apply coding agents.

Stateless MCP and integration tools

The MCP update introduced a stateless architecture, with coverage highlighting scalability and a formal deprecation policy. Simon Willison explored the specification and published tooling around MCP, including llm-mcp-client and work on Datasette integrations.

Efficiency and evaluation

OpenAI introduced GPT-5.6 with an emphasis on efficiency and agentic workflows. DeepSeek released V4 Flash 0731, while the open-source smevals tool offered a way to run small evaluation suites across models, prompts, and harnesses.

Scientific computing and security operations

OpenAI published a field report on scientists using coding agents to modernize scientific computing. Microsoft introduced an AI cybersecurity model and an agentic security system, and acquisitions involving Permiso and Oasis Security reflected continued investment in securing agent access.

Top stories this week

MCP & StandardsSimon Willison's Weblog · Jul 31, 2026

Stateless MCP 2.0 Specification Released, Inspires mcp-explorer and datasette-mcp

The Model Context Protocol (MCP) 2.0 specification, dubbed Stateless MCP, was released on July 28, 2026, marking the most significant update since its 2024 launch by Anthropic. It has reignited interest and led to the creation of new developer tools including mcp-explorer and datasette-mcp.

Why it matters for builders

MCP 2.0's stateless design removes the need for persistent server connections, making agent-tool integration simpler and more scalable. This directly benefits builders who can now create lightweight, composable agent tools without managing stateful backends.

MCPspecificationAnthropicagents
AI Agent LaunchesOpenAI News · Jul 29, 2026

GPT-5.6 improves AI efficiency and agentic workflows

OpenAI announces GPT-5.6, which enhances model efficiency and optimizes agentic workflows to reduce costs per intelligence unit.

Why it matters for builders

GPT-5.6's efficiency gains lower the cost of running agentic workflows, enabling developers to build more complex and responsive AI agents without breaking budget.

OpenAIGPT-5.6EfficiencyAgentic WorkflowsModel Update
AI Agent LaunchesSimon Willison's Weblog · Jul 31, 2026

DeepSeek V4 Flash 0731 Released with Enhanced Agentic Capabilities

DeepSeek released DeepSeek-V4-Flash-0731, a 304B-parameter model with substantially enhanced agentic capabilities. Priced at $0.14 per million input tokens and $0.27 per million output tokens, it ranks highly on cost-efficiency charts, outperforming larger models like MiniMax M3.

Why it matters for builders

Builders can leverage this model's strong agentic performance at low cost to create capable, scalable AI agents without excessive spending. Its cost-efficiency makes it a practical choice for SaaS and indie developers integrating agents into production workflows.

DeepSeekAgentic AIModel ReleasePricingCost-Efficiency
Open Source AgentsSimon Willison's Weblog · Jul 31, 2026

smevals: small eval suite for models, prompts, and harnesses

Simon Willison and Prime Radiant have released smevals, an open-source tool that lets you create and run small evaluation suites across different LLM configurations and grade the results.

Why it matters for builders

Provides a simple, reproducible way to evaluate LLMs with custom evals, potentially integrated with coding agents via uv, to streamline model selection and prompt engineering.

evalsLLMsopen-sourceuvmodel-comparison
Coding AgentsOpenAI News · Jul 28, 2026

Field report: Scientists modernize scientific computing with AI coding agents

OpenAI published a field report examining how scientists leverage AI coding agents to update legacy scientific computing, speeding up software development and enabling discoveries in fields like genomics.

Why it matters for builders

The report highlights practical patterns for integrating coding agents into scientific workflows, relevant for builders creating domain-specific AI tools for legacy system modernization.

Coding AgentsScientific ComputingGenomicsOpenAIReport

Explore the Best AI Agent Tools

Discover and compare launched AI agent tools on LaunchVault — or list your own and get discovered by builders and founders.