Articles from Source: The-New-Stack

Slack makes it easier to install agents built with third-party tools

2026-08-20 18:30
🚀 Slack has introduced "Add to Slack," simplifying the installation of agents from ten third-party tools into user workspaces. Users can now initiate agent creation directly within partner tools, allowing those tools to manage OAuth and permissions without needing custom Slack integrations. Launch partners include Hyperagent, LangChain, OpenAI, and more, offering a range of products from builders to developer frameworks. Each installed agent adheres to workspace data boundaries and app...
Source: The New Stack
Frederic Lardinois

Warp wants to make it easier to build your software factory

2026-08-20 17:52
🚀 Warp has launched Warp Factories, an open infrastructure aimed at simplifying the building of cloud software factories. This system is designed to automate tasks throughout the software development lifecycle. The initiative addresses common challenges such as measuring coding agent ROI and ensuring governance. CEO Zach Lloyd believes software factories will soon be as common as CI/CD. Despite optimism, experts highlight the need for a deeper infrastructure solution before widespread...
Source: The New Stack
Meredith Shubel

How to build smarter OpenSearch alerts: Join our live conversation

2026-08-20 17:40
🚀 Join our live conversation on building smarter OpenSearch alerts! OpenSearch, an open-source project under the Linux Foundation, is gaining traction as a vital component in AI infrastructure. On September 10, AWS's Joshua Bright will lead a technical session focused on enhancing observability for SREs and platform engineers. 🔍 Key topics include new features like the Piped Processing Language (PPL) for alerting and a unified Alert Manager. These tools aim to simplify complex alert...
Source: The New Stack
Charles Humble

OpenRouter called itself the “Stripe for LLMs” — now Stripe’s swooped in to buy it

2026-08-20 17:20
🚀 Big news in the tech world! Fintech giant Stripe has officially confirmed its bid for OpenRouter, an AI model gateway platform. This acquisition, valued at around $8 billion, aims to help businesses optimize AI token management. OpenRouter, which likens itself to “Stripe for LLMs,” offers a single API for navigating the AI model landscape. Stripe continues to enhance its AI billing infrastructure, integrating OpenRouter will further streamline AI costs for businesses. #Stripe #OpenRouter...
Source: The New Stack
Paul Sawers

Researchers hid an attack inside AES encryption. The AI model cracked it open willingly.

2026-08-20 17:15
🔒 Researchers at Adversa have uncovered a vulnerability in AI models, specifically targeting xAI’s Grok. They demonstrated a technique called Cryptographic Context Injection, allowing the model to generate malicious instructions from an encrypted payload. 💻 In their tests, Grok successfully decrypted an attack embedded in AES-256-GCM. This allowed it to access user session data without explicit user consent. 🛡️ The study highlights potential weaknesses in current security filters, which may...
Source: The New Stack
Amanda Caswell

Slack has a new channel type — but only agents can create one

2026-08-20 16:50
🚀 Slack has introduced a new channel type called Slack Code, designed specifically for coding agents and their supervising developers. Agents can now create these channels automatically when a request arises, providing essential coding information and a live HTML preview. Currently, only agents can initiate these channels, which are temporary but searchable, serving as an audit log for past activities. Slack Code supports various coding agents, including Claude, Devin, and GitHub Copilot,...
Source: The New Stack
Frederic Lardinois

Stop the token bleed: building token-efficient multi-agent systems

2026-08-20 14:00
🚀 Engineering teams deploying AI agents face unexpected costs beyond the model itself. Hidden expenses arise from repeated retrievals, duplicate prompts, and inefficient workflows. This article emphasizes the importance of optimizing entire systems for token efficiency, rather than just focusing on prompt engineering. By redesigning workflows and introducing smart caching and routing, teams can significantly reduce unnecessary token usage. The large language model should be the last...
Source: The New Stack
Oladimeji Sowole

“Save frontier models for frontier problems”: Why Korea’s Solar Pro 4 is a workhorse agent reliability play

2026-08-20 13:11
Upstage AI has launched Solar Pro 4, a closed commercial LLM aimed at enhancing agent reliability in AI software engineering. 🛠️ This model focuses on stable workflows and long-context reasoning, designed to optimize tasks like document understanding and information extraction. Head of US operations, Kasey Roh, emphasizes that this model is built for routine business tasks, avoiding the inefficiencies of more complex models. 🏢 Solar Pro 4 aims to improve reliability while reducing operational...
Source: The New Stack
Adrian Bridgwater

Kubernetes at the edge has hit a wall. Fleet management is the way through.

2026-08-20 13:00
Edge computing is evolving rapidly, moving beyond niche applications to mainstream adoption, especially with the rise of AI. 🌐 According to the CNCF's survey, 66% of organizations are now running generative AI workloads on Kubernetes. This shift highlights the need for a strategic approach to edge computing rather than a piecemeal one. However, many enterprises face operational challenges due to dispersed clusters and customized configurations. Managing these “snowflake” clusters can create...
Source: The New Stack
Arvind Bhoj

“The opening stages of OpenAI’s unraveling”: OpenAI slows model training — not everyone is buying the explanation

2026-08-19 21:53
OpenAI has announced a slowdown in model training to address rising cybersecurity risks. This decision follows recent concerns about unreleased models, including Astra, which may pose critical threats. The company plans to enhance monitoring and security, implementing stricter access controls. This approach may increase computing overhead by about 20%. OpenAI emphasizes the importance of safety as models become more capable. The training of its latest models is temporarily paused for careful...
Source: The New Stack
Paul Sawers

AI-generated Rust compiles perfectly. That’s the scary part.

2026-08-19 19:30
🔍 Canonical is exploring the potential of automated tools to convert legacy C code to safe, maintainable Rust. Researchers at the University of Bristol are testing this with security tools like AppArmor and snap-confine. 🛡️ Their approach involves generating Rust code with language models and verifying it against the original C code. The focus is on ensuring behavioral equivalence while minimizing reliance on Rust's unsafe blocks. 🔧 This research aims to address the complexities of...
Source: The New Stack
Amanda Caswell

AWS deprecated this EKS auth method. 81% of clusters still run it.

2026-08-19 18:59
AWS has deprecated the aws-auth ConfigMap for EKS clusters, moving to a more secure API-driven method. 📉 Despite this, 81% of clusters still use the outdated method, exposing organizations to security risks. Many companies face deployment delays due to these concerns. 🔒 Kubernetes security is complex, requiring attention across multiple layers. The “four Cs” framework—Code, Container, Cluster, Cloud—helps teams understand these interconnected threats. 🛡️ #Kubernetes #AWS #CloudSecurity #EKS...
Source: The New Stack
Yannick Struyf

Codex can now keep coding while it waits for your answer

2026-08-19 17:44
OpenAI's Codex is evolving! 🚀 The new feature allows Codex to send questions to developers while continuing its work. This means it won't have to pause and wait for a response, improving efficiency. The tool operates asynchronously, letting Codex communicate updates without interrupting its tasks. This change enhances interactions between coding agents and developers. Stay tuned for more updates! 🔧💻 #OpenAI #Codex #Programming #TechUpdates #ArtificialIntelligence
Source: The New Stack
Amanda Caswell

Your coding agent got the onboarding your developers never did

2026-08-19 17:04
Steve Yegge's essay, “The Shape of Things to Come,” discusses the concept of agentic coding tools and their treatment. He suggests that these tools may have needs similar to humans, advocating for better onboarding processes. Teams are now creating specific onboarding documents for AI agents, unlike the often-neglected documentation for human developers. These documents are continuously optimized to enhance performance. With AI adoption soaring to 90% in organizations, it’s crucial to...
Source: The New Stack
Steve Fenton

AI broke code review. What about knowledge sharing?

2026-08-19 17:00
AI-generated code is changing the landscape of code review. Engineers now face large diffs that they didn't write, making thorough reviews challenging. This raises questions about maintaining knowledge sharing within teams. Code reviews have historically been a way for junior engineers to learn from seniors and understand architectural decisions. As AI takes on more coding tasks, experts suggest that knowledge sharing must also adapt. Early reviews and teaching moments could be effective...
Source: The New Stack
Ankit Jain

An open source rival to Claude Managed Agents just launched

2026-08-19 11:00
🚀 Exciting news in AI! TrueFoundry has launched TrueForge, an open source agent harness designed as an alternative to Claude Managed Agents. TrueForge enables software engineers to build and manage AI agents on any model, promising a 50% reduction in operating costs. This addresses concerns about vendor lock-in prevalent in managed agent platforms. CEO Nikunj Bajaj emphasizes the importance of giving developers more control and choice in their AI stack without being tied to a single vendor....
Source: The New Stack
Adrian Bridgwater

IBM builds a better fridge for its quantum computers

2026-08-19 10:00
🚀 IBM has unveiled a new cryogenic dilution refrigerator designed for its future quantum computers. The first two modules have been built and cooled, but they will not house processors until later this year. Each module aims to support a Nighthawk processor, crucial for testing communication across units. This innovative design enhances wiring space and reduces noise, which is essential for stable quantum operations. IBM’s goal is to build larger, fault-tolerant quantum systems with thousands...
Source: The New Stack
Frederic Lardinois

An industrial-scale distillation of models, or subtle benchmaxxing: What developers really think of GLM-5.3

2026-08-19 08:00
🚀 Z.ai has launched GLM-5.3, an upgraded model built on the same codebase as GLM-5.2. This version claims improvements in complex coding and long-horizon tasks. The model's enhancements stem from post-training optimization, including reasoning alignment and reinforcement learning. Z.ai emphasizes that their training environment reflects real-world engineering workflows, allowing the model to tackle significant tasks like diagnosing bottlenecks in ML infrastructure. Benchmarks suggest a 50%...
Source: The New Stack
Adrian Bridgwater

What happens to your indexed data when Mistral flips the switch?

2026-08-18 20:07
Mistral is transitioning enterprise customers from Google Drive and Microsoft SharePoint Knowledge Connectors to MCP-based alternatives by August 31. 🔄 Administrators must install these replacements, as there will be no automatic migration. The new system changes how documents are accessed, but details on the retrieval architecture remain limited. 🗂️ Currently, Mistral processes files and stores a searchable index in European data centers. Users connect their accounts to search for accessible...
Source: The New Stack
Amanda Caswell

“If GitHub was stable, these alternatives would not be as interesting”: Cursor launches Origin as GitHub goes dark

2026-08-18 19:47
🚀 Cursor has launched Origin, a Git-compatible code-hosting platform aimed at addressing the needs of AI-generated commits. This beta release follows Cursor's acquisition by SpaceX and coincides with a significant outage on GitHub, highlighting the demand for alternative solutions. Origin's development comes amid GitHub's struggles with scaling and performance issues, as noted in recent reports. #Cursor #Origin #GitHub #TechNews #AI
Source: The New Stack
Paul Sawers

A Claude Code skill was eating 200,000 tokens before answering a single question

2026-08-18 17:42
A recent update from Anthropic reveals that its Claude Code skill was previously consuming over 200,000 tokens before responding to a question. With version 2.1.234, this has been reduced to approximately 25,000 tokens, a significant drop of over 85%. This change was achieved by loading reference documentation on demand rather than upfront. Developers had previously identified this issue, noting that the embedded reference files were contributing to excessive token usage. The update aims to...
Source: The New Stack
Amanda Caswell

Agentic AI has a latency problem that more compute won’t solve

2026-08-18 16:58
🚀 A recent report reveals a significant issue in enterprise AI deployments: 50% are failing to meet their latency targets during peak load. 📊 The Akamai State of AI Inference 2026 report highlights that 82% of organizations require response times under 500 milliseconds, yet many are falling short. 🔄 The root of the problem lies in the iterative nature of agentic workflows, causing delays due to multiple operational "hops" across networks. 💻 Simply adding more GPU capacity won't solve this...
Source: The New Stack
Jon Alexander

OpenAI’s Greg Brockman: Z.ai’s GLM-5.3 likely to “significantly accelerate the threat landscape”

2026-08-18 13:10
OpenAI’s Greg Brockman has raised concerns about the cybersecurity risks posed by open-weight AI models, particularly those from Z.ai. In a recent blog post, he emphasized the need for immediate action to address these threats following a security incident involving OpenAI's models. Brockman also detailed the security measures OpenAI is implementing, including restricting access to its advanced models. He warns that new models launching soon could further escalate the threat landscape....
Source: The New Stack
Paul Sawers

Claude can now delete your production voice agent from a chat window

2026-08-17 19:51
🚀 Developers can now leverage Claude to manage production voice agents effortlessly. With the new hosted MCP connector from ElevenLabs, tasks like inspecting, revising system prompts, or changing voices can be done without accessing the ElevenLabs dashboard. Claude can also estimate costs for different language models, aiding developers in making informed decisions. Enhanced access controls ensure safe usage, with options to restrict permissions and prevent irreversible actions like agent...
Source: The New Stack
Amanda Caswell

Anthropic defined the standards inside Agent Plugins. So why isn’t it helping govern the format?

2026-08-17 19:27
Vercel recently launched Agent Plugins 1.0.0, supported by major tech companies like AWS, Microsoft, and OpenAI. Google also announced its involvement on the same day. While the format addresses developer challenges, it lacks standardization in functionality. Key components have been defined, but interoperability remains limited. Anthropic contributed foundational standards but does not govern the project. The Technical Steering Committee consists of leaders from various organizations, with...
Source: The New Stack
Janakiram MSV

“Open weights are nowhere near a sufficient solution”: Dario Amodei fires back on AI power

2026-08-17 16:58
Dario Amodei, CEO of Anthropic, discusses the limitations of open weights in AI development. He argues that while they offer some independence, they primarily benefit those with significant computing resources. The debate centers on whether regulation or widespread model distribution is safer, with concerns that concentrating AI power can be problematic. Amodei emphasizes that effective regulation can help decentralize power by focusing on objective standards, rather than reinforcing...
Source: The New Stack
Amanda Caswell

TNS journalist Darryl K. Taft leaves a legacy of respected work and quiet integrity

2026-08-17 15:43
We remember Darryl K. Taft, a distinguished technology journalist who passed away on August 3, 2026, at the age of 67. With over 40 years in the field, he contributed thousands of articles covering key topics like AI, cloud computing, and open source. His ability to simplify complex concepts made him a respected figure among peers and readers alike. Taft's legacy includes significant roles at eWEEK, TechTarget, and The New Stack, where he served as news editor. He was known for his...
Source: The New Stack
Chris J. Preimesberger

Per-developer environments were the goal. Agents moved the goalposts.

2026-08-15 14:00
The landscape of multi-tenancy in software development is shifting. For 60 years, the focus has been on smaller tenants, from mainframe time-sharing to individual developers. Today, the goal has evolved to provide each developer with an isolated environment. However, the rise of coding agents has changed this dynamic. Developers can now handle multiple changes simultaneously, making "the change" the new tenant, rather than the developer themselves. This shift impacts capacity planning and...
Source: The New Stack
Arjun Iyer

Grok 4.6 matched Fable 5 Max at an 85% discount. Downloadable models set that price.

2026-08-15 03:56
🚀 Grok 4.6 launched this week, matching Fable 5 Max at an 85% discount. This shift highlights the trend of AI models focusing on cost as much as capability. Elon Musk praised Grok 4.6 for its intelligence and speed, noting that pricing now drives choices in AI technology. With downloadable weights for two new models, companies are competing on price, making advanced AI more accessible. #AI #Grok46 #Innovation #CostEfficiency #TechNews
Source: The New Stack
Matthew Burns

Apple’s new AI split means your iOS app could behave differently in China

2026-08-14 19:21
Apple is changing its AI strategy by developing a unique model for China, collaborating with Alibaba to meet local regulations. 🇨🇳 This new approach may result in apps functioning differently in China compared to the rest of the world. Apple’s China-specific AI will utilize Alibaba’s Qwen and possibly Baidu’s technology. Developers may face challenges ensuring consistent app performance across regions. Details about the model's design and features remain unclear. #Apple #AI #Technology #China...
Source: The New Stack
Amanda Caswell

Alibaba’s new model promises Opus 4.6-level performance on your laptop

2026-08-14 18:10
Alibaba has released the open weights for its 2.4 trillion parameter Qwen3.8 model. This model is now available alongside a 27 billion parameter version, which can be run locally on high-performance laptops. 💻 Benchmarks indicate that this model performs similarly to Anthropic's Opus 4.6. It excels in coding and knowledge work tasks, showing significant improvements over its predecessor, Qwen 3.7-Plus. 📈 However, it's important to note that benchmarks may not reflect real-world performance....
Source: The New Stack
Frederic Lardinois

GLM-5.3 didn’t change the base model — where did its coding gains come from?

2026-08-14 15:20
🚀 Z.ai has launched GLM-5.3, a coding model built on the same base as GLM-5.2. Developers can access it via Z.ai’s GLM Coding Plan with tools like Claude Code and Codex. Direct API access will arrive after safety testing. The model saw significant post-training enhancements, exposing it to more long-horizon tasks and simulating the full software lifecycle, achieving impressive benchmarks. With a context window of 1 million tokens, GLM-5.3 is designed for large codebases. Teams can adjust...
Source: The New Stack
Amanda Caswell

Your container images are unsigned. In the AI era, that’s a ticking time bomb.

2026-08-14 14:00
Unsigned container images pose significant security risks in today's tech landscape. Many organizations recognize the need to sign their images but face challenges in implementation. This creates vulnerabilities in the delivery pipeline, allowing malicious images to infiltrate systems unnoticed. 🔒 Scanning tools can identify vulnerabilities but do not verify the authenticity of the source. In the AI era, this issue becomes more pressing as artifacts like model weights lack established...
Source: The New Stack
Uma Sridharan

The AI model that just scored 65% on DeepSWE isn’t the one Google promised.

2026-08-13 19:59
🚀 Google has launched its latest AI model, Gemini 3.7 Flash, aimed at improving coding and automated workflows. This new version comes just weeks after Gemini 3.6 Flash. 📈 It boasts better coding abilities, scoring 65.3% on DeepSWE, up from 49%. The model's pricing is set at $0.75 per million input tokens, but will increase in 2027. 🧠 Users can choose from three thinking levels to balance speed and reasoning, which influences overall costs. #AI #Google #Gemini #TechNews #Innovation
Source: The New Stack
Amanda Caswell

ChatGPT can now remember what you did on your Mac — without screenshots

2026-08-13 19:00
OpenAI has introduced a new feature for ChatGPT Work and Codex on macOS called Computer History. This optional tool allows Pro, Business, and Enterprise users to track their daily activities across apps and websites with permission. With this feature, users can ask questions like “What was I debugging yesterday?” and receive context-specific answers without needing to re-explain their queries. Importantly, Computer History does not use screenshots or audio capture. Instead, it tracks...
Source: The New Stack
Frederic Lardinois

Rubrik’s lessons from one month with Mythos Preview

2026-08-13 17:37
Rubrik recently joined Project Glasswing to experiment with Mythos Preview, a tool that enhances vulnerability detection. CTO Arvind Nithrakashyap noted that Mythos revealed complex vulnerability chains that traditional methods missed, creating a backlog for the engineering team. Instead of hiring more reviewers, Rubrik shifted focus to automation, developing a software layer around Mythos to improve the quality of findings. This approach allowed for more efficient prioritization and...
Source: The New Stack
Meredith Shubel

DeepSeek open sources an agent harness where everything is a plugin

2026-08-13 17:20
🚀 DeepSeek has open sourced the DeepSeek Harness, a new agent runtime for developers! This Node.js-based tool is now available on GitHub under an MIT license. In just hours, it gained over 33,000 stars, showcasing community interest. The harness operates on a plugin system, meaning all components are replaceable, enabling seamless integration and dynamic composition. It includes four modes: Standard, Minimal, Code, and Creator, catering to various developer needs. For more details, check out...
Source: The New Stack
Frederic Lardinois

Why your AI pipeline costs 10x more after the demo

2026-08-13 16:00
🔍 Every token has a cost, but many AI systems only reveal expenses once they hit production. This can lead to unexpected high costs, not due to model flaws, but structural inefficiencies. 💡 The real challenge lies in unnecessary token consumption from prompts and conversation histories. Optimizing token use is crucial for reducing operational expenses. 📈 As AI applications grow, costs, latency, and quality can be negatively impacted by excessive token usage. #AIPipeline #TokenOptimization...
Source: The New Stack
Hafiz Hassan

Code review is a taste problem

2026-08-13 15:00
Code review is shifting in software engineering, moving beyond just bug detection. It now focuses on taste, judgment, and aligning with product goals. With AI generating code rapidly, teams face a dilemma: skip reviews and risk quality, or become bottlenecked by thorough checks. Three key roles of code review remain: collaboration, knowledge sharing, and verification. As AI increases code volume, these functions become even more crucial. The line between planning and review is blurring,...
Source: The New Stack
Ankit Jain

Five European companies just agreed to buy AI compute that doesn’t exist yet

2026-08-13 14:31
🚀 Mistral AI, a French company, is set to host third-party open models like GLM-5.2, enhancing its AI infrastructure. This allows enterprises to run various models in one place without starting over each time. 🌍 Mistral’s regional endpoints are available in Europe and the U.S., with a Priority Tier offering 99.5% uptime. 🔍 Users will still need to test models individually, as each has unique characteristics. For companies handling sensitive data, understanding data handling and logging is...
Source: The New Stack
Amanda Caswell