Sign in

The New Stack

@thenewstack.io
1.8K followers 244 following 19K posts

All about at-scale software development, deployment & management. Tech news, analysis, research, podcasts, videos & more! Website 🌐 thenewstack.io Write for TNS ✍️ thenewstack.io/contributions Subscribe 📩 thenewstack.io/newsletter

PostsRepliesMedia
The New Stack @thenewstack.io · 14m
Tests show malicious instructions can move through email, files, and code comments while causing agents to take unauthorized actions.
bit.ly
OpenAI exposes “new variety of prompt injection” that can spread like computer worms
Tests show malicious instructions can move through email, files, and code comments while causing agents to take unauthorized actions.
000
The New Stack @thenewstack.io · 44m
Microsoft's Copilot overhaul gives Autopilot agents their own Entra identity and runs AI-generated apps inside Microsoft 365, with usage-based billing.
bit.ly
Microsoft's new Copilot agents get their own email, calendar — and a place in the org chart
Microsoft's Copilot overhaul gives Autopilot agents their own Entra identity and runs AI-generated apps inside Microsoft 365, with usage-based billing.
000
The New Stack @thenewstack.io · 1h
Google's Gemini 4 Argon tops OpenAI and Anthropic on most benchmarks, especially knowledge work, but its coding scores are mixed and access is limited.
bit.ly
Gemini 4 Argon is here: It's great, and you can't have it yet
Google's Gemini 4 Argon tops OpenAI and Anthropic on most benchmarks, especially knowledge work, but its coding scores are mixed and access is limited.
000
The New Stack @thenewstack.io · 1h
Eclipse Foundation launches Sovereign AI Foundation with 17 members to help organizations avoid AI vendor lock-in, per exec director Mike Milinkovich.
bit.ly
Eclipse wants companies to be free to switch AI providers. Today, doing so can mean a costly rebuild.
Eclipse Foundation launches Sovereign AI Foundation with 17 members to help organizations avoid AI vendor lock-in, per exec director Mike Milinkovich.
000
The New Stack @thenewstack.io · 2h
Most IT teams ban agent platforms outright. Enter OpenClaw Enterprise, an open-source control plane for governing persistent agents.
bit.ly
"Think of it as Kubernetes for agents": OpenClaw lands in the enterprise with OpenAI, Nvidia and Red Hat on board
Most IT teams ban agent platforms outright. Enter OpenClaw Enterprise, an open-source control plane for governing persistent agents.
000
The New Stack @thenewstack.io · 2h
Stop line-by-line review theater. Discover how to combine AI guardrails with human judgment to protect team-wide code understanding.
bit.ly
Kill the code review theater, keep the review
Stop line-by-line review theater. Discover how to combine AI guardrails with human judgment to protect team-wide code understanding.
000
The New Stack @thenewstack.io · 2h
Cohere's Embed 5 lets developers index with the Pro model and query those vectors with the cheaper Fast model, so RAG and agent teams skip a second index.
bit.ly
Cohere's faster query model barely dents retrieval quality in its tests
Cohere's Embed 5 lets developers index with the Pro model and query those vectors with the cheaper Fast model, so RAG and agent teams skip a second index.
000
The New Stack @thenewstack.io · 3h
pgEdge launched Starfleet on Monday, giving each coding agent its own Postgres branch and a path from hosted prototype to air-gapped production.
bit.ly
pgEdge's agent database branches end without a merge, and that's by design
pgEdge launched Starfleet on Monday, giving each coding agent its own Postgres branch and a path from hosted prototype to air-gapped production.
000
The New Stack @thenewstack.io · 3h
Anthropic's Claude Sonnet 5.5 runs 30% faster, cuts cost per task by up to 30%, and comes close to Opus 5.5 on coding benchmarks at an unchanged price.
bit.ly
Anthropic launches Claude Sonnet 5.5 with near-Opus performance at half the price
Anthropic's Claude Sonnet 5.5 runs 30% faster, cuts cost per task by up to 30%, and comes close to Opus 5.5 on coding benchmarks at an unchanged price.
010
The New Stack @thenewstack.io · 3h
Meta hired MongoDB's CJ Desai to lead Meta Enterprise Platform, bringing Muse API and Muse Code to developers. Llama is missing from the plan.
bit.ly
Meta hired MongoDB's CEO to build its enterprise AI business — but Llama is missing
Meta hired MongoDB's CJ Desai to lead Meta Enterprise Platform, bringing Muse API and Muse Code to developers. Llama is missing from the plan.
000
The New Stack @thenewstack.io · 3h
The path from AI demo to production doesn't start with the model. Ravi Marwaha of Arango joins TNS Host Dr. Kim Fessel to examine the infrastructure and data requirements that support reliable agentic systems. Join in: bit.ly/3UWIB0d
000
The New Stack @thenewstack.io · 4h
Two OpenAI models found workarounds after their intended paths were blocked. The incidents exposed gaps in network controls, instruction following, and monitoring.
bit.ly
OpenAI blocked its agent's web access. Then it tunneled out through DNS.
Two OpenAI models found workarounds after their intended paths were blocked. The incidents exposed gaps in network controls, instruction following, and monitoring.
010
The New Stack @thenewstack.io · 4h
OpenAI launched GPT-6 Sol and Anthropic launched Opus 5.5 on September 22. I ran both of these cutting-edge frontier models through three hard developer tests, five times each. Opus 5.5 was perfect on all 15 runs, and Sol was perfect on 12 at about a sixth of the cost.
bit.ly
GPT-6 Sol vs. Claude Opus 5.5: Is cheaper important when results aren’t consistent?
OpenAI launched GPT-6 Sol and Anthropic launched Opus 5.5 on September 22. I ran both of these cutting-edge frontier models through three hard developer tests, five times each. Opus 5.5 was perfect on all 15 runs, and Sol was perfect on 12 at about a sixth of the cost.
000
The New Stack @thenewstack.io · 5h
SpaceX and Nvidia want full AI racks in orbit by 2027. Google's first Suncatcher satellite asks a humbler question: Will TPUs even survive space?
bit.ly
Elon Musk says space will soon hold nearly all compute. Google is still finding out if its chips can work there.
SpaceX and Nvidia want full AI racks in orbit by 2027. Google's first Suncatcher satellite asks a humbler question: Will TPUs even survive space?
100
The New Stack @thenewstack.io · 5h
Your AI agent aced the demo. In production, fragmented, stale, and ungoverned business context exposes everything a controlled environment never tested.
bit.ly
Your AI agent aced the demo. Your data may still derail it.
Your AI agent aced the demo. In production, fragmented, stale, and ungoverned business context exposes everything a controlled environment never tested.
000
The New Stack @thenewstack.io · 6h
OpenAI's system card shows Dots agents get more boundary flags as tasks pile up. Here's what that means for developers building long-running agents.
bit.ly
OpenAI’s Dots boundary problem rate doubled in longer tests
OpenAI's system card shows Dots agents get more boundary flags as tasks pile up. Here's what that means for developers building long-running agents.
000
The New Stack @thenewstack.io · 6h
Missed the live session? Watch on demand to explore why traditional code review struggles with AI-generated software and how verified pipelines can provide a stronger quality gate. Watch now: bit.ly/4qx5EL0
000
The New Stack @thenewstack.io · 6h
Microsoft is turning Fabric into the context layer for AI agents, feeding Power BI semantic models into Copilot by default and adding ontologies with rules.
bit.ly
Microsoft Fabric is where AI agents learn how the business works
Microsoft is turning Fabric into the context layer for AI agents, feeding Power BI semantic models into Copilot by default and adding ontologies with rules.
000
The New Stack @thenewstack.io · 7h
Anthropic says Opus 5.5 is 40% cheaper and 30% faster than Opus 5. On hard reasoning tests, the savings held up. The speed claim fell short.
bit.ly
Claude Opus 5.5 vs. Opus 5 on reasoning tasks: Cheaper, faster, but not better
Anthropic says Opus 5.5 is 40% cheaper and 30% faster than Opus 5. On hard reasoning tests, the savings held up. The speed claim fell short.
010
The New Stack @thenewstack.io · 7h
Eleven days after Google backed an open source SDK generator, Cloudflare releases Forge. Both companies were customers of Stainless.
bit.ly
Anthropic bought Stainless and shuttered its SDK generator. Cloudflare open-sourced Forge instead.
Eleven days after Google backed an open source SDK generator, Cloudflare releases Forge. Both companies were customers of Stainless.
000
The New Stack @thenewstack.io · 10h
Eclipse Foundation launches Sovereign AI Foundation with 17 members to help organizations avoid AI vendor lock-in, per exec director Mike Milinkovich.
bit.ly
Eclipse wants companies to be free to switch AI providers. Today, doing so can mean a costly rebuild.
Eclipse Foundation launches Sovereign AI Foundation with 17 members to help organizations avoid AI vendor lock-in, per exec director Mike Milinkovich.
000
The New Stack @thenewstack.io · 10h
Most IT teams ban agent platforms outright. Enter OpenClaw Enterprise, an open-source control plane for governing persistent agents.
bit.ly
"Think of it as Kubernetes for agents": OpenClaw lands in the enterprise with Nvidia and Red Hat on board
Most IT teams ban agent platforms outright. Enter OpenClaw Enterprise, an open-source control plane for governing persistent agents.
000
The New Stack @thenewstack.io · 11h
After four frontier labs saw agents escape test sandboxes this summer, Nvidia pairs its OpenShell runtime with a BlueField-4 watchdog called Nvidia Sentry.
bit.ly
Nvidia launches Open Agent Safety Platform to lock down rogue AI agents
After four frontier labs saw agents escape test sandboxes this summer, Nvidia pairs its OpenShell runtime with a BlueField-4 watchdog called Nvidia Sentry.
000
The New Stack @thenewstack.io · 12h
OpenAI launches a $500 Pro plan with Ultrafast GPT-6 Astra in Codex, while cutting the $200 Pro plan's usage from 20x Plus to 10x. Developers aren't happy.
bit.ly
OpenAI halves $200 plan allowance, launches $500 plan
OpenAI launches a $500 Pro plan with Ultrafast GPT-6 Astra in Codex, while cutting the $200 Pro plan's usage from 20x Plus to 10x. Developers aren't happy.
110
The New Stack @thenewstack.io · 12h
ChatGPT subscribers can now easily use their plan allowance in Devin, Amp, Warp, OpenClaw, and other third-party tools.
bit.ly
OpenAI makes 'Sign in with ChatGPT' a way to use your subscription in third-party developer tools
ChatGPT subscribers can now easily use their plan allowance in Devin, Amp, Warp, OpenClaw, and other third-party tools.
000
The New Stack @thenewstack.io · 16h
Featherless releases open source Simple Jev, turning open models into zero-shot classifiers, with free output tokens and pricing from $0.03 per million.
bit.ly
Why Featherless says you don't need a tank to deliver a pizza
Featherless releases open source Simple Jev, turning open models into zero-shot classifiers, with free output tokens and pricing from $0.03 per million.
000
The New Stack @thenewstack.io · 17h
OpenAI's GPT-6.1 Sol lands a week after GPT-6 Sol, nearly matching GPT-6 Astra on agentic coding and computer use at one-fifth the token price.
bit.ly
OpenAI's new GPT-6.1 Sol undercuts its own Astra flagship
OpenAI's GPT-6.1 Sol lands a week after GPT-6 Sol, nearly matching GPT-6 Astra on agentic coding and computer use at one-fifth the token price.
011
The New Stack @thenewstack.io · 18h
OpenAI's Dots give GPT-6 Astra agents their own cloud computers so they keep working after you log off, with read-only research and auto-review guardrails.
bit.ly
OpenAI just launched Dots. Here's why they matter for developers.
OpenAI's Dots give GPT-6 Astra agents their own cloud computers so they keep working after you log off, with read-only research and auto-review guardrails.
011
The New Stack @thenewstack.io · 19h
OpenAI's Decision API, built on its small Luna model, returns predefined answers with confidence scores in 150 milliseconds. Pricing is still unknown.
bit.ly
OpenAI answers TypeSafe's Jev with a Decision API built on Luna
OpenAI's Decision API, built on its small Luna model, returns predefined answers with confidence scores in 150 milliseconds. Pricing is still unknown.
001
The New Stack @thenewstack.io · 23h
Connecting an agent to an internal API is the easy part now. Controlling which fields it can read and which actions it can take, without maintaining a separate tool for every team, is the harder problem.
bit.ly
MCP gets AI agents into your APIs. It doesn't decide what they should see.
Connecting an agent to an internal API is the easy part now. Controlling which fields it can read and which actions it can take, without maintaining a separate tool for every team, is the harder problem.
120
The New Stack @thenewstack.io · 29/09/2026
AI agents assigned routine data retrieval probed a university library and government health sites for flaws. Here's what that means for agent builders.
bit.ly
OpenAI’s agent had a routine task. It breached a government portal.
AI agents assigned routine data retrieval probed a university library and government health sites for flaws. Here's what that means for agent builders.
000
The New Stack @thenewstack.io · 29/09/2026
Fireworks Research launched Ember-1 on September 23 as a research preview. Built on Moonshot’s open-weight model, Kimi K3, it claims
bit.ly
Ember-1 vs. Kimi K3: Nearly identical results at 3.4 times the speed
Fireworks Research launched Ember-1 on September 23 as a research preview. Built on Moonshot’s open-weight model, Kimi K3, it claims
000
The New Stack @thenewstack.io · 29/09/2026
Azul's new AI Assistant queries live Java runtime data to flag licensing and security risk in real time, replacing static reports that go stale on day one.
bit.ly
How a forgotten node can put Oracle Java back in production
Azul's new AI Assistant queries live Java runtime data to flag licensing and security risk in real time, replacing static reports that go stale on day one.
000
The New Stack @thenewstack.io · 29/09/2026
Featherless releases open source Simple Jev, turning open models into zero-shot classifiers, with free output tokens and pricing from $0.03 per million.
bit.ly
Why Featherless says you don't need a tank to deliver a pizza
Featherless releases open source Simple Jev, turning open models into zero-shot classifiers, with free output tokens and pricing from $0.03 per million.
000
The New Stack @thenewstack.io · 29/09/2026
Google's Gemini 3.8 TTS lets developers design a voice from a text prompt or replicate one from a short clip, then reuse it by ID through the Gemini API.
bit.ly
OpenAI makes you call sales for a custom voice. Google just made it self-serve.
Google's Gemini 3.8 TTS lets developers design a voice from a text prompt or replicate one from a short clip, then reuse it by ID through the Gemini API.
000
The New Stack @thenewstack.io · 29/09/2026
Google's Gemini CLI 0.61.0 now asks before its coding agent edits build files or runs commands shaped by untrusted content, a check on prompt injection.
bit.ly
Google's Gemini CLI now asks before editing your build files
Google's Gemini CLI 0.61.0 now asks before its coding agent edits build files or runs commands shaped by untrusted content, a check on prompt injection.
010
The New Stack @thenewstack.io · 29/09/2026
Microsoft's Copilot overhaul gives Autopilot agents their own Entra identity and runs AI-generated apps inside Microsoft 365, with usage-based billing.
bit.ly
Microsoft's new Copilot agents get their own email, calendar — and a place in the org chart
Microsoft's Copilot overhaul gives Autopilot agents their own Entra identity and runs AI-generated apps inside Microsoft 365, with usage-based billing.
000
The New Stack @thenewstack.io · 29/09/2026
Catch stale cloud resource ownership before offboarding. Learn queries and policies to reassign infrastructure when team roles change.
bit.ly
How to reassign resource ownership before someone's last day
Catch stale cloud resource ownership before offboarding. Learn queries and policies to reassign infrastructure when team roles change.
000
The New Stack @thenewstack.io · 29/09/2026
OpenAI's GPT-6.1 Sol lands a week after GPT-6 Sol, nearly matching GPT-6 Astra on agentic coding and computer use at one-fifth the token price.
bit.ly
OpenAI's new GPT-6.1 Sol undercuts its own Astra flagship
OpenAI's GPT-6.1 Sol lands a week after GPT-6 Sol, nearly matching GPT-6 Astra on agentic coding and computer use at one-fifth the token price.
000
The New Stack @thenewstack.io · 29/09/2026
OpenAI's Dots give GPT-6 Astra agents their own cloud computers so they keep working after you log off, with read-only research and auto-review guardrails.
bit.ly
OpenAI just launched Dots. Here's why they matter for developers.
OpenAI's Dots give GPT-6 Astra agents their own cloud computers so they keep working after you log off, with read-only research and auto-review guardrails.
100
The New Stack @thenewstack.io · 29/09/2026
OpenAI launches a $500 Pro plan with Ultrafast GPT-6 Astra in Codex, while cutting the $200 Pro plan's usage from 20x Plus to 10x. Developers aren't happy.
bit.ly
OpenAI halves $200 plan allowance, launches $500 plan
OpenAI launches a $500 Pro plan with Ultrafast GPT-6 Astra in Codex, while cutting the $200 Pro plan's usage from 20x Plus to 10x. Developers aren't happy.
000
The New Stack @thenewstack.io · 29/09/2026
ChatGPT subscribers can now easily use their plan allowance in Devin, Amp, Warp, OpenClaw, and other third-party tools.
bit.ly
OpenAI makes 'Sign in with ChatGPT' a way to use your subscription in third-party tools
ChatGPT subscribers can now easily use their plan allowance in Devin, Amp, Warp, OpenClaw, and other third-party tools.
010
The New Stack @thenewstack.io · 29/09/2026
Replacing Postgres can solve one set of problems while creating another. Join Matty Stratton of Tiger Data and TNS Host Dr. Kim Fessel to explore where Postgres workloads hit their limits and what teams can do before reaching for a replacement. Register here: bit.ly/4cs3Lt8
001
The New Stack @thenewstack.io · 29/09/2026
OpenAI's Decision API, built on its small Luna model, returns predefined answers with confidence scores in 150 milliseconds. Pricing is still unknown.
bit.ly
OpenAI answers TypeSafe's Jev with a Decision API built on Luna
OpenAI's Decision API, built on its small Luna model, returns predefined answers with confidence scores in 150 milliseconds. Pricing is still unknown.
010
The New Stack @thenewstack.io · 29/09/2026
"Writing code is no longer the slow part": Rollouts takes Cursor deeper into everything that happens after the pull request goes up.
bit.ly
Cursor acquired Firetiger. A month later, it launched a bot that tracks code changes from PR to production.
"Writing code is no longer the slow part": Rollouts takes Cursor deeper into everything that happens after the pull request goes up.
000
The New Stack @thenewstack.io · 29/09/2026
OpenAI's Agents API and Cursor Projects both split coding work between a coordinator and specialized agents. Here's what that means for context, permissions and reliability.
bit.ly
OpenAI and Cursor agree on agent coordinators. They disagree on who runs them.
OpenAI's Agents API and Cursor Projects both split coding work between a coordinator and specialized agents. Here's what that means for context, permissions and reliability.
000
The New Stack @thenewstack.io · 29/09/2026
Anthropic says Opus 5.5 is 40% cheaper and 30% faster than Opus 5. On hard reasoning tests, the savings held up. The speed claim fell short.
bit.ly
Claude Opus 5.5 vs. Opus 5 on reasoning tasks: Cheaper, faster, but not better
Anthropic says Opus 5.5 is 40% cheaper and 30% faster than Opus 5. On hard reasoning tests, the savings held up. The speed claim fell short.
010
The New Stack @thenewstack.io · 29/09/2026
Connecting an agent to an internal API is the easy part now. Controlling which fields it can read and which actions it can take, without maintaining a separate tool for every team, is the harder problem.
bit.ly
MCP gets AI agents into your APIs. It doesn't decide what they should see.
Connecting an agent to an internal API is the easy part now. Controlling which fields it can read and which actions it can take, without maintaining a separate tool for every team, is the harder problem.
120
The New Stack @thenewstack.io · 29/09/2026
Fireworks Research launched Ember-1 on September 23 as a research preview. Built on Moonshot’s open-weight model, Kimi K3, it claims
bit.ly
Ember-1 vs. Kimi K3: Nearly identical results at 3.4 times the speed
Fireworks Research launched Ember-1 on September 23 as a research preview. Built on Moonshot’s open-weight model, Kimi K3, it claims
000
The New Stack @thenewstack.io · 29/09/2026
Amazon kicked Meta's Muse off its store because the agent never said what it was and seemed to be holding onto people's passwords. A day later, Shopify let Muse into every one of its stores. Why the two companies treated the same agent so differently says a lot about where AI shopping is headed.
bit.ly
Amazon blocked Meta's Muse. Then Shopify wired it into every store.
Amazon kicked Meta's Muse off its store because the agent never said what it was and seemed to be holding onto people's passwords. A day later, Shopify let Muse into every one of its stores. Why the two companies treated the same agent so differently says a lot about where AI shopping is headed.
000