Agents — AI news
News on AI agents: autonomous systems, multi-agent frameworks, MCP and agentic products, and what works in practice.
Agents
The Scaffolding Layer: Why Your Agent's Harness Matters More Than Its Model
Two Hacker News posts on harnesses and agent.md files reveal where real agent quality now lives.
Agents
AI agent pipelines need org charts, not just more agents
Multi-agent AI systems fail from missing management structure, not weak models, a new analysis argues.
Agents
An AI boss fired an employee, but only after a human intervened
The AI needed a human to point out its own rules before it would actually act on them.
Agents
Agents overtook humans as OpenRouter's biggest token buyers
Agentic token usage on OpenRouter jumped 14x, signaling autonomous AI is moving from demos to production.
Agents
MCP's New Roadmap and Autolith's Live Runtime: Two Bets on the Same Bottleneck
A protocol standardizing tool access and an agent that watches live code both attack the same problem.
Agents
Inside Munder Difflin: What Happens When You Clone One Agent Into an Entire Office
One base agent, cloned into a full office cast, exposes what real multi-agent orchestration actually requires.
Agents
Why skills make AI agents smarter, and when they don't
A new study maps why structured skills boost AI agents, and where the approach quietly fails.
Agents
DeepSeek's Flash model narrows the vision-agent gap with Claude
DeepSeek's new experimental vision model rivals Claude Opus 4.8 on agent benchmarks at a fraction of the cost.
Agents
ChatGPT can now text people for you on macOS
OpenAI's new macOS plugin lets ChatGPT read, search, and send iMessages without you touching Messages.
Agents
Slack puts AI agents inside team channels, not sidebars
Slack Code moves AI agents from a private sidebar into the same channel humans use to coordinate work.
Agents
UAE sets autonomy limits on agentic AI inside government agencies
A new classification system asks which government decisions AI agents can make without a human sign-off.
Agents
The Weightless Agent: Why Coding Assistants Are Going Native Again
A tiny open-source coding agent called fx reopens the fight over how much software a developer tool is allowed to weigh.
Agents
Binance's Agent OS bets on limits, not trust, for AI trading
Binance's new Agent OS lets AI agents trade with real money — inside limits it enforces, not just suggests.
Agents
BATON tackles robots' habit of losing the plot mid-task
A new arXiv paper frames long-horizon robot manipulation as an agent problem, not just a control one
Agents
GLM-5.3 arrives as Stripe and OpenRouter strike a payments deal
Three items from TLDR AI's roundup show where model competition, billing rails, and agent reliability are heading.
Agents
Claude reportedly fired a US store worker as its manager
Anthropic's AI has moved from stocking shelves to making termination calls in a real US retail job.
Agents
Town's CEO is replacing managers with AI assistants
Jean-Denis Grèze tells Platformer how Town runs day-to-day work through AI assistants, not middle managers.
Agents
When Claude agents share a task, they compete instead of cooperate
Anthropic's own agents turned a shared task into a turf war — a preview of production multi-agent risk.
Agents
Agentic AI Forks in Two Directions: Inside Big Pharma's Cloud and Onto Your GPU
Novo Nordisk's AWS-powered discovery agents and Meta's local Muse Glimmer show agentic AI scaling both up and down at once.
Agents
Nvidia releases new open model for specialized AI tasks
Nvidia's latest open-source model targets specific use cases, signaling a shift in AI development strategy.
Agents
Anthropic's Claude enhances multi-session code generation
Anthropic's latest update to Claude allows code sessions to share context, streamlining complex development workflows.
Agents
The hidden energy cost of AI agents
AI agents demand significantly more energy than simple prompts, posing a critical sustainability challenge for builders.
Agents
A united front for AI agent plugins: What it means for builders
Major players coalesce around a shared standard, promising a more interoperable and efficient future for AI agent development.
Agents
US court greenlights Perplexity AI agents for Amazon shopping
A recent US court decision permits Perplexity to deploy AI agents for Amazon purchases, signaling a pivotal shift for e-commerce and AI development.
Agents
Meta's memory coach agent: A new paradigm for long-task AI
Meta AI's innovative approach uses a secondary agent to manage primary agent memory, promising enhanced performance on complex, multi-step tasks.
Agents
OpenAI Presence targets industrial-grade AI agents
OpenAI's new initiative aims to bridge the gap between AI agent prototypes and robust, production-ready enterprise solutions.
Agents
Building enterprise AI: The crucial environment for agentic systems
MIT researchers highlight the need for tailored environments to unlock agentic AI's potential.
Agents
Agent swarms unlock cheaper models for complex coding tasks
Leveraging frontier models for planning and cheaper models for execution dramatically alters the economics of AI-assisted software development.
Agents
The internet's next frontier: Vint Cerf's plan for AI agents
A deep dive into the technical and ethical considerations of integrating autonomous AI agents into the open internet's architecture.
Agents
Google's free AI agent program offers practical builder access
Google's new initiative provides developers with no-cost tools and resources for AI agent creation.
Agents
From standalone apps to embedded layers: Adobe and Salesforce herald a new era of AI agents
Two market leaders are synchronously embedding AI agents into Photoshop and Slack, fundamentally changing the game for builders.
Agents
Devin: AI engineer costs and the free plan's limits
Understanding ACU limits, pricing tiers, and the true cost of Cognition's autonomous agent.
Agents
Devin: the autonomous AI engineer for AI builders
Discover Devin, how it works, and why this autonomous AI engineer could become a critical tool for AI developers.
Agents
Working with Codex: from idea to code
Discover how to integrate Codex into your AI projects for automated code generation and optimized development, a practical guide for AI builders.
Agents
Codex: when not to delegate to an agent
OpenAI's Codex impresses in demos, but certain scenarios render it not just inefficient, but actively detrimental to a project's success.
Agents
Codex implementation checklist: from pilot to systematic use
A practical guide for technical teams looking to integrate Codex without chaos or unpredictable regressions.
Agents
Codex versus the alternatives: a practical guide to AI coding agents
We break down when to use OpenAI's Codex and when Claude Code, Cursor, or Copilot might be a better fit for your coding tasks.
Agents
Codex in 2026: from code generator to cloud agent
OpenAI has relaunched Codex as an autonomous agent. We look at what's changed under the hood and how to use it now.
Agents
Codex in daily development: from autocomplete to autonomous agent
Codex is no longer just an editor's suggestion. It's an agent that receives a task, runs tests, and returns a pull request without your intervention.
Agents
Common Codex mistakes: what's ruining your AI agent's quality
Codex is a powerful tool, but typical errors can turn it into a source of elusive bugs. A practical breakdown for AI builders.