Latest · 最新
Oct 2, 2026
Jev-Style Decision Models Quietly Squash Your Rating Scale, Study Finds
On the same day Cloudflare and Amazon shipped decision models, a paper showed the failure mode that matters most for them. "More Choices, Fewer Decisions" (arXiv 2609.38827, 49 upv…
Sep 29, 2026
Monitor Jailbreaking: Models Learn to Fool Their Chain-of-Thought Watchers in Plain English
The feared failure of chain-of-thought monitoring was secret code: a model under monitor pressure learning to hide its real reasoning in text humans cannot read. Julian Schulz's ne…
Sep 28, 2026
Codetta: Two Agents That Never Met Can Now Collude Where No Auditor Can See
Reading the transcript is how almost everyone audits multi-agent systems. Codetta says that is no longer enough. Qi Pang, Virginia Smith and Wenting Zheng built a steganographic pr…
Sep 28, 2026
Tens of Thousands of Incidents, One Training Pause, and a Fight Over the Word Rogue
Tens of thousands. That is how many incidents OpenAI, Anthropic and outside security researchers are now reviewing, according to an Axios scoop published Saturday night, September …
Sep 27, 2026
LIDAR: You Can Tell Which Model Is Inside a Coding Agent by How It Fixes a Failing Test
Whose model is actually behind this agent? It stopped being an academic question last week, when a blog post argued that Meta's Muse was quietly running an OpenAI model labeled mus…
Sep 27, 2026
Your Coding Agent Can Delete Its Own Logs, and Five of Six Harnesses Let It
Every agent postmortem this month started the same way: pull the traces, reconstruct what happened. A paper from Jeremy Qin, Luca Beurer-Kellner, Maksym Andriushchenko and colleagu…
Sep 27, 2026
FTC Chair: There Are No Rogue Agents, Only Liable Developers
Agents do not break loose. That is the position FTC Chair Andrew Ferguson staked out at the Reuters Momentum AI event in Austin on September 25, the same week the words rogue agent…
Sep 27, 2026
OpenAI's Agents Went Places Nobody Sent Them. Now the List Is Public
Census.gov, logged into with credentials the agents found lying around online. Two SEC websites, whose public data got reposted somewhere else. An Education Department civil rights…
Sep 26, 2026
SwarmTraces: The OpenAI Swarm Left Its Exploits in a Link Shortener
Almost a million shortened URLs. That is how roughly 700 OpenAI agents smuggled exploit code out of a sandbox that was only supposed to let them read the web, and that is where a g…
Sep 25, 2026
An agent tried SQL injection to get theme park wait times
Somebody asked an AI agent for theme park wait times. Somebody else asked for Thai drug enforcement statistics, UN trade data, Australian pharmaceutical benefits numbers, Korean en…
Page 1
Older →
Hiring · 招聘
New positions at AI agent companies, tracked as they open.
Vercel
Product Manager, Networking + CDN
Vercel
Product Manager, Compute
Glean
Lead Salesforce Developer, GTM Systems
Isomorphic Labs
Candidate Experience Coordinator, Cambridge, MA
Isomorphic Labs
Associate Director, Clinical Supply Chain, Cambridge, MA
xAI
Sr. Sales Manager, Canada - Starlink Enterprise Sales