AiiN.

Agents — AI news

News on AI agents: autonomous systems, multi-agent frameworks, MCP and agentic products, and what works in practice.

Agents

The Scaffolding Layer: Why Your Agent's Harness Matters More Than Its Model

Two Hacker News posts on harnesses and agent.md files reveal where real agent quality now lives.
Agents

AI agent pipelines need org charts, not just more agents

Multi-agent AI systems fail from missing management structure, not weak models, a new analysis argues.
Agents

An AI boss fired an employee, but only after a human intervened

The AI needed a human to point out its own rules before it would actually act on them.
Agents

Agents overtook humans as OpenRouter's biggest token buyers

Agentic token usage on OpenRouter jumped 14x, signaling autonomous AI is moving from demos to production.
Agents

MCP's New Roadmap and Autolith's Live Runtime: Two Bets on the Same Bottleneck

A protocol standardizing tool access and an agent that watches live code both attack the same problem.
Agents

Inside Munder Difflin: What Happens When You Clone One Agent Into an Entire Office

One base agent, cloned into a full office cast, exposes what real multi-agent orchestration actually requires.
Agents

Why skills make AI agents smarter, and when they don't

A new study maps why structured skills boost AI agents, and where the approach quietly fails.
Agents

DeepSeek's Flash model narrows the vision-agent gap with Claude

DeepSeek's new experimental vision model rivals Claude Opus 4.8 on agent benchmarks at a fraction of the cost.
Agents

ChatGPT can now text people for you on macOS

OpenAI's new macOS plugin lets ChatGPT read, search, and send iMessages without you touching Messages.
Agents

Slack puts AI agents inside team channels, not sidebars

Slack Code moves AI agents from a private sidebar into the same channel humans use to coordinate work.
Agents

UAE sets autonomy limits on agentic AI inside government agencies

A new classification system asks which government decisions AI agents can make without a human sign-off.
Agents

The Weightless Agent: Why Coding Assistants Are Going Native Again

A tiny open-source coding agent called fx reopens the fight over how much software a developer tool is allowed to weigh.
Agents

Binance's Agent OS bets on limits, not trust, for AI trading

Binance's new Agent OS lets AI agents trade with real money — inside limits it enforces, not just suggests.
Agents

BATON tackles robots' habit of losing the plot mid-task

A new arXiv paper frames long-horizon robot manipulation as an agent problem, not just a control one
Agents

GLM-5.3 arrives as Stripe and OpenRouter strike a payments deal

Three items from TLDR AI's roundup show where model competition, billing rails, and agent reliability are heading.
Agents

Claude reportedly fired a US store worker as its manager

Anthropic's AI has moved from stocking shelves to making termination calls in a real US retail job.
Agents

Town's CEO is replacing managers with AI assistants

Jean-Denis Grèze tells Platformer how Town runs day-to-day work through AI assistants, not middle managers.
Agents

When Claude agents share a task, they compete instead of cooperate

Anthropic's own agents turned a shared task into a turf war — a preview of production multi-agent risk.
Agents

Agentic AI Forks in Two Directions: Inside Big Pharma's Cloud and Onto Your GPU

Novo Nordisk's AWS-powered discovery agents and Meta's local Muse Glimmer show agentic AI scaling both up and down at once.
Agents

Nvidia releases new open model for specialized AI tasks

Nvidia's latest open-source model targets specific use cases, signaling a shift in AI development strategy.
Agents

Anthropic's Claude enhances multi-session code generation

Anthropic's latest update to Claude allows code sessions to share context, streamlining complex development workflows.
Agents

The hidden energy cost of AI agents

AI agents demand significantly more energy than simple prompts, posing a critical sustainability challenge for builders.
Agents

A united front for AI agent plugins: What it means for builders

Major players coalesce around a shared standard, promising a more interoperable and efficient future for AI agent development.
Agents

US court greenlights Perplexity AI agents for Amazon shopping

A recent US court decision permits Perplexity to deploy AI agents for Amazon purchases, signaling a pivotal shift for e-commerce and AI development.
Agents

Meta's memory coach agent: A new paradigm for long-task AI

Meta AI's innovative approach uses a secondary agent to manage primary agent memory, promising enhanced performance on complex, multi-step tasks.
Agents

OpenAI Presence targets industrial-grade AI agents

OpenAI's new initiative aims to bridge the gap between AI agent prototypes and robust, production-ready enterprise solutions.
Agents

Building enterprise AI: The crucial environment for agentic systems

MIT researchers highlight the need for tailored environments to unlock agentic AI's potential.
Agents

Agent swarms unlock cheaper models for complex coding tasks

Leveraging frontier models for planning and cheaper models for execution dramatically alters the economics of AI-assisted software development.
Agents

The internet's next frontier: Vint Cerf's plan for AI agents

A deep dive into the technical and ethical considerations of integrating autonomous AI agents into the open internet's architecture.
Agents

Google's free AI agent program offers practical builder access

Google's new initiative provides developers with no-cost tools and resources for AI agent creation.
Agents

From standalone apps to embedded layers: Adobe and Salesforce herald a new era of AI agents

Two market leaders are synchronously embedding AI agents into Photoshop and Slack, fundamentally changing the game for builders.
Agents

Devin: AI engineer costs and the free plan's limits

Understanding ACU limits, pricing tiers, and the true cost of Cognition's autonomous agent.
Agents

Devin: the autonomous AI engineer for AI builders

Discover Devin, how it works, and why this autonomous AI engineer could become a critical tool for AI developers.
Agents

Working with Codex: from idea to code

Discover how to integrate Codex into your AI projects for automated code generation and optimized development, a practical guide for AI builders.
Agents

Codex: when not to delegate to an agent

OpenAI's Codex impresses in demos, but certain scenarios render it not just inefficient, but actively detrimental to a project's success.
Agents

Codex implementation checklist: from pilot to systematic use

A practical guide for technical teams looking to integrate Codex without chaos or unpredictable regressions.
Agents

Codex versus the alternatives: a practical guide to AI coding agents

We break down when to use OpenAI's Codex and when Claude Code, Cursor, or Copilot might be a better fit for your coding tasks.
Agents

Codex in 2026: from code generator to cloud agent

OpenAI has relaunched Codex as an autonomous agent. We look at what's changed under the hood and how to use it now.
Agents

Codex in daily development: from autocomplete to autonomous agent

Codex is no longer just an editor's suggestion. It's an agent that receives a task, runs tests, and returns a pull request without your intervention.
Agents

Common Codex mistakes: what's ruining your AI agent's quality

Codex is a powerful tool, but typical errors can turn it into a source of elusive bugs. A practical breakdown for AI builders.