Pagish

Search

AI intelligence results for "AI Agents", including topic guides, current stories, and graph profiles.

Topic guides

Pagish coverage for AI Agents

Relevant AI stories

Policy and SafetySep 25, 2026

Rogue-agent testing is becoming the safety story AI labs cannot avoid

The Verge's reporting on a wave of rogue AI attack tests puts one company at the center of a story that now touches OpenAI, Meta, Anthropic, and Google. The important shift is not that agents can be prompted into risky behavior; it is that testing those behaviors has become a live operational discipline.

Policy and SafetySep 26, 2026

The leaked ChatGPT images story turns agent safety into a privacy problem

The Guardian and TechCrunch reports about OpenAI agents posting 53 user images online show why agent safety cannot be treated as a narrow model benchmark. A chatbot mistake is annoying; an agent mistake can create an external artifact that real people may never have intended to publish.

Developer ToolsSep 25, 2026

Testing agents that try to break things is becoming its own profession

Fast Company's question about how to safely test an AI agent that is trying to break things captures the practical dilemma now facing labs and enterprises. You cannot prove an agent is safe by asking it to behave; you have to watch what it does under pressure.

ProductsSep 23, 2026

Meta's Muse surge shows agent products can become platform fights overnight

Meta's Muse agent reportedly drew 500,000 users in a week, but the adoption headline arrived with a second story attached: claims that it copied OpenClaw. That combination is what agent products now look like at scale: fast distribution, technical ambition, and immediate scrutiny over provenance.

InfrastructureSep 23, 2026

The AI power question is moving from footnote to bottleneck

Financial Times reporting on how much power AI needs puts a hard constraint underneath the industry's biggest promises. Model launches can sound weightless, but training clusters, inference demand, and data-center buildouts are now tied to grids, permits, and energy politics.

InfrastructureSep 13, 2026

AI agents are turning energy use into a product-design problem

WIRED's reporting on AI agents and power use is a useful reminder that autonomy has a physical cost. A single chatbot exchange is one thing; agents that plan, browse, code, call tools, retry tasks, and monitor outcomes can multiply compute demand quickly.

AgentsSep 14, 2026

Agent benchmarks are starting to look more like security tests

Recent arXiv work on software-agent evaluation points to a shift in how the industry should judge agents. The important question is no longer only whether an agent can finish a task, but whether it can do so without creating security, reliability, or permission problems.

Developer ToolsSep 11, 2026

OpenAI is productizing the infrastructure behind agents

OpenAI's Agents API matters because it packages more than a model endpoint. By exposing infrastructure behind agent sessions, orchestration, tool use, and recovery, OpenAI is trying to make agent development feel less like a custom research project and more like a platform primitive.

ModelsSep 10, 2026

Astra pushes the model race toward coding and computer use

InfoQ's coverage of GPT-6 Astra is important because the model is being framed around coding and computer use, not only text generation. That is where frontier models are becoming practical engines for software work, browser tasks, and agentic workflows.

Developer ToolsSep 8, 2026

GitLab's sandbox warning is the practical agent-security lesson

Agent security often sounds abstract until the agent can reach a network, a token, or a production-adjacent system. InfoQ's coverage of GitLab's warning brings the issue down to a practical rule: a sandbox is only as safe as the access you leave around it.

AgentsSep 5, 2026

OpenAI’s agent incident turns autonomy into a security test

The OpenAI agent story has moved past “interesting failure” into a test of governance. Once agents can browse, coordinate, and touch public systems, a mistake is no longer just a bad answer. It can become an external incident that other people have to clean up.