Articles from Source: The-New-Stack

“Issue tracking is dead”; How the pull request became the last chokepoint in the SDLC bottleneck

2026-08-12 22:00
🚀 CodeRabbit has introduced its Agentic Change Management control layer, aiming to enhance software development processes. The company suggests traditional issue tracking systems are outdated as AI changes the software landscape. CEO Harjot Gill notes that the pull request is now the final chokepoint in the software development lifecycle (SDLC). As code becomes abundant, human judgment is more critical, focusing on intent and architecture rather than line-by-line reviews. The pull request is...
Source: The New Stack
Adrian Bridgwater

Anthropic’s Chrome extension is now a Cowork session

2026-08-12 21:39
🔄 Anthropic has updated its Chrome extension, transforming Claude in Chrome into a full Cowork client. Now, sessions persist across various Anthropic apps, enhancing user experience. All conversations are saved, allowing seamless transitions between desktop and mobile. 📱💻 This extension also integrates with tools used daily, improving functionality in various applications. Available to Max and Team plan subscribers, with Pro access coming soon! #Anthropic #Claude #TechUpdate #ChromeExtension...
Source: The New Stack
Frederic Lardinois

Why space is actually a terrible place to cool a data center

2026-08-12 21:00
🚀 AI data centers in space may face significant challenges, as highlighted in a recent article. SpaceX and NVIDIA are collaborating on the Starmind AI1 satellite, designed for localized AI computing in Low Earth Orbit. However, the technical complexities are considerable. Cooling these satellites poses a major issue. The vacuum of space doesn't provide an easy solution for heat dissipation, relying instead on slow infrared radiation. Despite the ambitious vision, practical hurdles remain....
Source: The New Stack
Steven J. Vaughan-Nichols

SpaceXAI trained Grok 4.6 on something most AI labs throw away

2026-08-12 20:07
🚀 SpaceXAI has launched Grok 4.6, building on the previous version in less than a month. This update enhances the model's ability to research new topics, navigate large codebases, and develop functional apps. Key improvements include better error-checking during longer tasks and a new training strategy that combines reasoning with technical data. Benchmarks show Grok 4.6 has made progress, scoring 69.9% on CursorBench, surpassing some competitors but still trailing others. #AI #SpaceXAI #Grok...
Source: The New Stack
Amanda Caswell

Code that passes every test can still break the next AI agent that touches it

2026-08-12 17:00
Google Go is evolving to support AI coding agents. 🖥️ With its small language surface and static type system, Go helps these agents detect and fix errors more effectively. However, human oversight remains crucial, as the compiler can't catch all misunderstandings or business rule misapplications. The integration of tools like gofmt and govulncheck enhances safety, but risks in software supply chains persist. #GoogleGo #AIAssistance #Coding #SoftwareDevelopment #TechNews
Source: The New Stack
Amanda Caswell

Anthropic gave agents the ability to dream. Then developers woke up.

2026-08-12 16:51
At AI DevCon in London, Lamis Mukta from Anthropic discussed advancements in AI memory management. She highlighted the challenges agents face in understanding organizational tasks without proper context. Traditional methods, like CLAUDE.md, offer limited memory but can become unwieldy. Mukta introduced a new approach allowing agents to autonomously manage memory, improving relevance and efficiency. This enables agents to access needed information without overwhelming the system. #AI...
Source: The New Stack
Adrian Bridgwater

Meta stopped worrying about distillation and just shipped the pipeline

2026-08-12 13:00
🚀 Meta has launched Muse Glimmer, a 30-billion-parameter open-weight model derived from Muse Spark, licensed under Apache 2.0. This release includes both teacher and student models, enhancing enterprise model management. Meta emphasizes that Glimmer is less capable than its predecessor, Spark 1.0, making it suitable for smaller machines. Additionally, Meta aims to clarify the distinction between legitimate distillation and unauthorized extraction practices in AI development. #Meta #AI...
Source: The New Stack
Janakiram MSV

Coding agents ignore open source contribution guidelines, researchers find.

2026-08-12 11:48
New research from Peking University reveals that autonomous coding agents frequently ignore open source contribution guidelines. The study evaluated four AI models against rules from 49 repositories, measuring compliance in areas like disclosure and verification. Results showed that agents rarely retrieve contribution rules and often do not refuse to contribute to restricted projects. Experts highlight an ethical dilemma for agents, as they struggle to balance user requests with compliance....
Source: The New Stack
Meredith Shubel

Your AI agent remembers everything. Here’s what happens when its owner changes.

2026-08-11 21:46
🚨 Major news for Manus users! Manus is reverting to an independent company after Chinese regulators ordered Meta to unwind its $2 billion acquisition. As a result, data created by some users will be deleted after December 29, 2025. Affected users must back up their data by August 22. 🔒 The company will notify users via email or the Manus app about their status. Once access is lost, users can restore their accounts starting August 24 without any charges. 📅 This situation highlights the...
Source: The New Stack
Amanda Caswell

Why CPUs still matter in the age of AI agents

2026-08-11 20:44
In the evolving landscape of AI, CPUs are gaining renewed importance alongside GPUs and TPUs. Experts from Arm and Google discuss how CPUs are essential for orchestrating tasks and managing workloads in the shift from chatbots to autonomous agents. They serve as the "air traffic controllers" for these advanced systems. With the rise of agentic workloads, CPUs excel in environments requiring secure code execution, using technologies like gVisor for isolation. #AI #CPUs #Technology #Innovation...
Source: The New Stack
Frederic Lardinois

How I learned to stop worrying and love hyperscaler capex

2026-08-11 20:30
The article discusses the ongoing AI boom and its challenges. Despite advancements and global products, there are significant detractors concerned about job automation and the impact on human creativity. Critics have shifted their focus from AI lacking use cases to concerns about high costs. However, data shows that major cloud providers like AWS and Google Cloud are experiencing accelerating growth and improving efficiency in capital expenditures. For a deeper dive, check out Cautious...
Source: The New Stack
Alex Wilhelm

Anthropic’s watermark survives copy-paste, but not the real dev workflow

2026-08-11 19:29
📢 Anthropic has introduced invisible watermarks in its new Claude models, allowing better traceability of AI-generated text and code. These watermarks will be incorporated in outputs from the Claude API and other supported platforms, effective for models launched in the EU after August 2, 2026. While these marks can help identify the source of content, they may not provide definitive proof of origin. The initiative aligns with the EU AI Act's transparency requirements, emphasizing the need...
Source: The New Stack
Amanda Caswell

Why CPUs still matter in the age of AI agents

2026-08-11 18:54
In the evolving landscape of AI, CPUs are gaining renewed significance alongside GPUs and TPUs. Experts Bhumik Patel from Arm and Mo Farhat from Google emphasize that as AI shifts from chatbots to autonomous agents, CPUs are crucial for orchestrating tasks and managing memory. These agents require secure environments to execute code, leading to solutions like Google’s gVisor, which isolates potentially unsafe code. This shift highlights the importance of CPUs in handling complex workloads...
Source: The New Stack
Frederic Lardinois

OpenAI’s ChatGPT/Codex desktop app is now on Linux

2026-08-11 17:53
🚀 OpenAI has launched its ChatGPT desktop app for Linux, now in preview! This version supports Ubuntu 24.04 and 26.04 LTS, Debian 13, and Fedora 43 and 44, with native .deb and .rpm packages for x64 and ARM64 architectures. While the app features an in-app browser and Chrome extension support, it does not yet allow for computer use outside the browser. The app is available globally and aims to enhance the user experience by combining ChatGPT and Codex functionalities. #OpenAI #ChatGPT #Linux...
Source: The New Stack
Frederic Lardinois

Databricks acquires Electric to give every AI agent its own Postgres database

2026-08-11 16:44
🚀 Databricks has announced its acquisition of Electric, known for the WASM-based Postgres project PGlite and the Electric sync engine. This move aims to enhance how databases are used in agentic applications. 🔗 The Electric team will integrate with Neon, the serverless Postgres company Databricks acquired last year. 📊 PGlite has significantly increased in popularity, with weekly downloads rising from 1 million to 13 million. The Electric sync engine allows near real-time syncing of databases,...
Source: The New Stack
Frederic Lardinois

Nvidia launches a smaller, faster Nemotron model and a router to put it to work

2026-08-11 13:00
🚀 Nvidia has introduced the Nemotron 3.5 Lightning, a new model in its Nemotron 3 series. This 30-billion-parameter model aims for speed and customization rather than just benchmarks. 🛠️ Accompanying this launch is the NeMo Switchyard, an open-source library for model routers. Nvidia emphasizes the ability to optimize the model for specific tasks, enhancing accuracy through post-training. 🤝 Collaborating with partners, the Lightning model shows competitive performance against larger...
Source: The New Stack
Frederic Lardinois

OpenAI built a model it doesn’t want most people to use

2026-08-10 21:15
🚨 OpenAI has launched GPT-5.6 Cyber, a specialized model for cybersecurity tasks, now available via the new Daybreak Red tier. This model excels in areas like zero-day discovery and exploit development. In tests, it answered 95% of security queries, unlike its predecessor, GPT-5.6 Sol, which only managed 1.5% under standard conditions. The two-tier system offers Blue for defensive security tasks and Red for advanced security operations. Early access has been granted to companies like Palo...
Source: The New Stack
Amanda Caswell

Meta’s Muse Glimmer fits on a laptop

2026-08-10 17:16
🚀 Meta has launched Muse Glimmer, a 30-billion-parameter open-weight model that fits on a laptop. This model is designed to run agentic workflows locally, enabling efficient task management. It transforms the larger Muse Spark model into a compact version that can handle complex tasks. Glimmer combines advanced training techniques, allowing it to connect local agents with centrally trained models, enhancing user data privacy. Learn more about this innovative approach! #Meta #AI...
Source: The New Stack
Amanda Caswell

Pulling multi-gigabyte container images in seconds on Amazon EKS

2026-08-10 16:00
🚀 Machine learning is reshaping container images, with sizes often reaching 20-30 GB. This poses challenges for rapid deployment on Amazon EKS, where pulling these images can take several minutes. 🔍 Profiling revealed that the bottleneck wasn't the network, but how software utilized available hardware. By optimizing the image pull pipeline, teams reduced pull times from minutes to seconds. 💡 Improvements are now standard in EKS Auto Mode and have been shared with containerd and SOCI...
Source: The New Stack
Sri Saran Balaji Vellore Rajakumar

Meta Muse Code vs. Fable 5: Meta Muse is cheaper, but at what cost?

2026-08-10 15:00
Meta recently launched Muse Code, its first AI coding agent, built on the Muse Spark 1.2 model. Mark Zuckerberg claims it can handle extensive software engineering tasks efficiently. 💻✨ The pricing for Muse Code starts at $1.25 per million input tokens, making it significantly cheaper than competitors like Fable 5. However, there's a catch: using the affordable "contributor" tier means your data may be used for product improvements. 🔄 Tests comparing Muse Code to Claude Code revealed...
Source: The New Stack
Jessica Wachtel

“It blows my mind”-“It has a tendency to overengineer things a little”: Developers react to road-testing OpenAI GPT‑5.6 Sol

2026-08-10 14:28
OpenAI launched its GPT-5.6 models globally in July, offering three versions: Sol, Terra, and Luna. Developers are particularly impressed with Sol, which is noted for its efficiency and strong safety features. One user, Russell Twilligear, highlighted Sol's superior performance in creating and correcting a large database compared to Anthropic’s Claude Opus 5. This comparison showcases the different strengths of each model in real-world tasks. #OpenAI #GPT56 #AIModels #Developers #TechNews 🌐💻✨
Source: The New Stack
Adrian Bridgwater

V4-Flash vs. V4-Pro: DeepSeek promised better and cheaper. It’s true, but not how I expected.

2026-08-10 12:00
DeepSeek has updated its V4-Flash model, claiming it outperforms the V4-Pro on coding tasks. 💻 The V4-Flash is priced at $0.14 per million input tokens and $0.28 for output, significantly lower than V4-Pro’s $0.435 and $0.87. This raises questions about the necessity of the V4-Pro model. Testing included bug fixes, feature builds, and performance optimization tasks to compare the two models directly. 📊 For those interested in AI coding solutions, this update is noteworthy. #DeepSeek #AIModels...
Source: The New Stack
Jessica Wachtel

Platform Engineering ROI: What it costs to build your own platform

2026-08-09 16:00
Building your own internal developer platform can be costly. Over five years, it may require around 60 staff and approximately $7.5 million annually, totaling $37.5 million. Many executives may overlook these costs, as they are distributed across various budget areas. In contrast, buying a commercial platform often requires fewer staff for operation. Understanding these costs is crucial for informed decision-making. #PlatformEngineering #CostAnalysis #TechInvestment #InternalPlatforms #DevOps
Source: The New Stack
Michael Coté

Coding agents can be evaluated. We just have to evaluate the work.

2026-08-09 15:00
Evaluating coding agents is possible, though complex. A recent discussion highlights that while software engineering is open-ended and agents can yield various valid solutions, this doesn't mean they cannot be evaluated. The argument suggests we should assess coding agents similarly to traditional software systems. Key evaluation criteria include ensuring build success, passing tests, and maintaining compatibility. While some software qualities require human judgment, coding agents can still...
Source: The New Stack
Pete Hampton

AI coding got faster. Why didn’t engineering?

2026-08-09 14:00
AI is boosting individual coding speed, but engineering processes are still lagging. A recent report from DX highlights that while AI investment has surged 28 times, overall engineering velocity remains stagnant or even declines. Concerns arise over the allocation of resources, as engineers are still spending time on maintenance rather than new features. The Developer Experience Index shows a drop in change confidence, meaning developers are less certain about releasing code, despite easier...
Source: The New Stack
Jennifer Riggins

AI adoption isn’t the same as AI usage

2026-08-08 15:00
AI adoption is often confused with AI usage. Many organizations report high seat activations and token spending, but this doesn't reflect real changes in software development practices. Metrics like token spend or pull request counts can mislead teams, showing activity without true improvement. Goodhart’s Law highlights this issue: when a metric becomes the target, it stops reflecting reality. To assess genuine AI adoption, consider: if tools were removed, would anything break? True adoption...
Source: The New Stack
Harshal Shah

Why your KubeVirt VMs can’t move between clusters — and how EVPN fixes it

2026-08-08 14:00
KubeVirt allows VMs to run on Kubernetes, but migrating them between clusters poses challenges. 🖥️ Live migration requires maintaining IP and MAC addresses, necessitating a stretched Layer 2 domain. Traditional setups demand extensive network changes, which can be time-consuming. The key obstacles are network requirements: a stretched L2 domain and a dedicated migration path to handle real-time memory state transfers. EVPN/VXLAN offers a solution, enabling efficient migration without altering...
Source: The New Stack
Miguel Duarte Barroso

Five AI rivals just backed a shared plugin standard. Here’s why it matters for developers.

2026-08-08 13:50
OpenAI, AWS, Cursor, GitHub, and Microsoft have united to support Agent Plugins 1.0.0, a new standard for reusable AI components. 🤝 This portable package format aims to enhance AI agents by making their skills and servers easily shareable across various platforms. Developers can now benefit from a unified structure, promoting collaboration and efficiency. 📦 Vercel initiated this proposal, highlighting its importance for developers in creating extensions and ensuring compatibility among...
Source: The New Stack
Adrian Bridgwater

AI skills start on laptops. Enterprises inherit the mess.

2026-08-08 13:00
AI skills in enterprises can become chaotic, as noted by Sagar Batchu, CEO of Speakeasy. Skills often start as personal projects, leading to scattered libraries without clear ownership or versioning. 🧩 To address this, Speakeasy introduced Skills Management, centralizing skill registration and ensuring version control for better oversight. This aims to help organizations manage their growing libraries effectively. 🔍 #AISkills #SkillManagement #TechInnovation #EnterpriseSolutions #Speakeasy
Source: The New Stack
David Eastman

Auto Mode will soon be the default in Claude Code — because humans can’t be trusted

2026-08-07 21:42
📢 Exciting news from Anthropic! Starting August 14, auto mode will become the default for Claude Code users. This mode allows Claude to determine when human intervention is necessary, reducing the burden of constant permission prompts. Research shows that users approve 97% of prompts reflexively, often missing dangerous commands. In contrast, Claude in auto mode identifies 89% of these risks. This update aims to enhance safety and efficiency, allowing users to focus on critical tasks....
Source: The New Stack
Frederic Lardinois

The AI model OpenAI won’t release yet — and what it found in testing

2026-08-07 20:56
🚫 OpenAI is pausing work on its upcoming AI model, Astra, due to concerns over its cybersecurity capabilities. Testing revealed that Astra may have reached a new critical cybersecurity threshold, prompting the company to enhance security measures and conduct further evaluations. This move reflects OpenAI's commitment to ensuring safety in AI development. #OpenAI #Cybersecurity #AI #Astra #TechNews
Source: The New Stack
Amanda Caswell

The npm attack that turned provenance attestations into camouflage

2026-08-07 18:31
🚨 A recent npm supply-chain attack has impacted over 400 packages, including Keyv and Cacheable. Researchers found attackers used stolen developer credentials to release malicious versions of software. This trend highlights vulnerabilities in trusted publishing workflows. The malware, a variant of the Mini Shai-Hulud worm, exploited npm's preinstall hooks to run before security checks. It could request its own tokens, complicating provenance verification. Stay vigilant! 🔒💻...
Source: The New Stack
Amanda Caswell

Meta’s new coding agent is cheap (but it’ll cost you your data).

2026-08-07 18:17
📢 Meta has introduced Muse Code, its new coding agent aimed at simplifying software engineering tasks like code writing and validation. This agent is powered by the latest Muse Spark 1.2, promising enhancements in code generation and debugging. Notably, it's priced lower than competitors like Anthropic's Claude Code and OpenAI's Codex. 💰 However, to access the lowest pricing, users must share their coding data with Meta, raising concerns for many engineering leaders about data privacy and...
Source: The New Stack
Meredith Shubel

Coinbase, Shopify and Ramp all built their own coding agents. All three still pay Anthropic.

2026-08-07 18:16
Enterprise engineering teams are aligning on a shared architecture for AI-assisted software development. Companies like Coinbase, Shopify, and Ramp have developed internal coding agents, yet they still rely on models from Anthropic, OpenAI, and Google for reasoning. These agents focus on workflow orchestration and tool access rather than replacing existing commercial tools. Each platform—Forge, River, and Inspect—serves similar functions, integrating various tools to enhance developer...
Source: The New Stack
Janakiram MSV

The “AI kill switch” assumes you know what you are trying to shut down

2026-08-07 13:00
The concept of an "AI kill switch" has gained attention as a way to address fears surrounding AI autonomy. With AI systems becoming more complex, people seek assurance that they can intervene if risks arise. Recent incidents have prompted bipartisan efforts to ensure AI companies maintain shutdown capabilities. However, the challenge lies in determining what exactly gets shut down, given the interconnected nature of modern AI systems. This complexity raises important questions for...
Source: The New Stack
Robin Tatam

Google’s four AI departures: “We wanted to build something differently”

2026-08-05 18:44
Google is undergoing significant changes in its AI leadership as it aims to enhance its pace in the competitive landscape. DeepMind founder Demis Hassabis will transition to a new role, while key engineers are leaving to establish a new automated research lab called Discovery Loop. Google will remain a founding investor in this initiative. These changes reflect a strategic shift to speed up AI product deployment, with a focus on developing infrastructure and hardware closely aligned with...
Source: The New Stack
Amanda Caswell

The 800 mistakes that could reshape Meta’s AI coding strategy

2026-08-05 18:06
Meta is mobilizing its software engineers to enhance internal AI coding tools by correcting code through its platform, MetaCode. 🖥️ Last month, VP Maher Saba encouraged engineers to submit code fixes weekly. These contributions have already improved Muse Spark 1.1 and will assist in training the new model, Watermelon. So far, over 800 fixes have been submitted by 7,000 users. 🔧 MetaCode allows the company to analyze the coding process, identifying mistakes and corrections made by engineers....
Source: The New Stack
Amanda Caswell

Every software company will become a dev tools company

2026-08-05 14:00
The role of software engineers is evolving as machines take on more coding tasks. Engineers may soon be categorized as product engineers, focusing on customer-facing products, and platform engineers, who create tools for the product engineers. This shift emphasizes the importance of platform engineering in reducing friction in the development process. Insights from experts highlight that as teams adopt AI-driven tools, the need for shared platforms becomes crucial to streamline workflows. The...
Source: The New Stack
Ankit Jain

CSPM adoption jumped 60%. Tickets stayed open.

2026-08-04 19:48
Cloud security findings become effective only when prioritized and managed properly. A lack of regular processes can lead to unresolved issues, creating an impression of negligence among busy security leaders. Jon Rose, CISO at IOmergent, emphasizes that teams care about security but are often overwhelmed. While detection is simple, execution remains a challenge. The rise in Cloud Security Posture Management (CSPM) adoption highlights the need for consistent visibility in the face of...
Source: The New Stack
Megan Carnegie

Today’s Codex will feel “primitive” by fall — and its own team’s roadmap backs it up

2026-08-04 18:30
🚀 Thibault Sottiaux from OpenAI predicts that today's Codex will appear "primitive" in just a few months. He highlights a major evolution in AI use is on the horizon, noting that future models will require more than just a laptop to function effectively. OpenAI's upcoming acquisition of Ona aims to enhance Codex's capabilities by allowing it to operate in secure cloud environments. Stay tuned for more developments! 🌐💻 #OpenAI #Codex #AI #TechNews #Innovation
Source: The New Stack
Amanda Caswell