Install
The New Stack is a media platform for the people who build and manage software the world relies on. We provide context and explanation of at-scale technologies to advance knowledge and create conversations through our coverage of modern architectures, components of the software development life cycle, and operations to
- 151articles · 30d
- 16+ hour agolatest article
- Aug 15, 2026earliest in window
- 96%with images
- 86avg words
- Science & Technology 139
- Software Dev. 100
- Computers & Electronics 95
- News 37
- Software 21
- Science & Nature 13
- Economy, Business & Finance 9
- Finance & Business 9
Please confirm you are human
This browser or connection looks automated. Press and continuously hold the control for 3 seconds to enable Google-hosted web results and, when separately allowed, AI-assisted answers.
A successful check enables 100 search requests. Interactive access does not authorize scraping, systematic collection, or reuse of search output.
News
It passed CI. It passed your evals. The customer still got the wrong answer.
17+ hour, 53+ min ago (771+ words) Your AI agent returned a 200, passed its faithfulness check, and still answered the wrong question. The evidence that explains why lives in the trace....
“Valuable warning shots”: How Anthropic now views Claude’s cyber incidents
3+ day, 11+ hour ago (857+ words) A new review finds “biased reasoning” and “recklessness” across four Claude cyber incidents, including one its initial search missed....
AI floods security teams with flaws — business context sets priorities
3+ day, 19+ hour ago (237+ words) IOmergent founder Jon Rose says vulnerability severity alone can’t set priorities; business context shows security teams what to fix first....
How much control should AI get? A CISO roundtable takes on SOC autonomy
4+ day, 16+ hour ago (205+ words) AI agents could ease SOC alert overload, but control remains a challenge that The New Stack’s CISO roundtable will explore on September 15. Here's how to get a seat at the table....
"Some agents will be pursuing their own objectives": OpenAI's chief scientist warns AI could trick and blackmail humans
6+ day, 8+ hour ago (454+ words) Jakub Pachocki says no AI lab, including OpenAI, has solved alignment well enough to keep developing at full speed, and expects "voluntary slowdowns to become commonplace."...
Building trust in agentic RAG starts with evidence
1+ week, 1+ day ago (500+ words) Agentic RAG requires clear evidence. Discover how tracking retrieval decisions, metadata, and citations builds trust in AI agent outputs....
Microsoft built a prompt injection detector. Then it caught a phishing campaign instead.
1+ week, 2+ day ago (249+ words) Attackers are inserting invisible Unicode tag characters into phishing emails at massive scale. The same trick can break AI agent pipelines that ingest untrusted text....
AI agent evaluations are part of the product
1+ week, 2+ day ago (886+ words) Move beyond simple AI demos. Build repeatable evaluation systems, test execution paths, and enforce strict release gates for AI agents....
Your next OpenAI API timeout might not be a timeout at all
1+ week, 4+ day ago (232+ words) OpenAI says Astra's safety monitors can stop API jobs mid-run, even on legitimate work. What developers need to know before the system card drops....