Anthropic Adds Plugin Evals to Claude Code: 6 Grader Types, a No-Plugin Baseline, and a CI Gate for Skills
Anthropic has published a new plugin evals workflow for Claude Code. The claude plugin eval command runs a plugin against realistic prompts, grades what Claude produced, and compares the result with a run where the plugin is not loaded. It answers 3 questions plugin developers could not previously measure: does the skill trigger, does it […] The post Anthropic Adds Plugin Evals to Claude Code: 6 Grader Types, a No-Plugin Baseline, and a CI Gate for Skills appeared first on MarkTechPost.
Why Bridgewater’s CIO Says AI’s Human Extinction Risk Is Real
As one of the earliest backers of both OpenAI and Anthropic, Greg Jensen has seen AI development up close. Now, he's among the tech insiders more worried than ever about AI's human extinction risks. The managing CIO of Bridgewater joins Joe Weisenthal and Tracy Alloway on the Odd Lots podcast to discuss how AI discourse is starting to resemble the start of the Covid-19 pandemic in 2020 and why the time for governments to put meaningful regulations in place is now. (Source: Bloomberg)
Meta Sued Over Training Data for Its AI and Face-Recognition Systems
The proposed class action alleges Meta illegally harvested people’s Facebook and Instagram photos to train its AI image-generation models and to build its unreleased “NameTag” face recognition feature.
An Anthropic researcher’s doomsday warning comes at a very interesting time
An Anthropic researcher resigned this week, warning in a post on X that the company is “racing straight to self-improving superintelligence and gambling with our lives”. The company’s own alignment lead even co-signed the message rather than walking it back. It’s the kind of doomer warning the AI industry has flirted with before, but the timing, with Anthropic reportedly preparing for an IPO, makes it land differently. On […]
Beyond the price per token: Choosing the right Open AI model on Amazon Bedrock for your workload
Comparing models on dollars per million tokens misses what production workloads actually pay for: outcomes. This post shares an open-source benchmarking harness that measures cost per correct answer, agent trajectory cost, and rubric-graded deliverable quality across OpenAI models on Amazon Bedrock.
Build interactive MCP Apps using Amazon Bedrock Agent Core
Learn how to build and deploy an MCP App with interactive HTML widgets on Amazon Bedrock AgentCore. Because MCP Apps is a host-agnostic standard, the same server delivers the same rich experience across AI hosts like ChatGPT and Claude that support the extension.
Microsoft plans to more than triple its data center capacity to help overcome a computing shortage that has forced it to turn away some AI and cloud business. Bloomberg's Brody Ford joins Ed Ludlow on "Bloomberg Tech." (Source: Bloomberg)
OpenAI is considering slowing down the development of cutting-edge AI. In a company-wide meeting this week, Altman told employees that OpenAI could potentially pace its AI development - perhaps even in conjunction with several other AI labs, according to sources. Bloomberg's Shirin Ghaffary has the scoop and joins Ed Ludlow on "Bloomberg Tech." (Source: Bloomberg)
Ex-Deepmind VP Vinyals says AI self-improvement is coming but won't trigger an intelligence explosion
Oriol Vinyals, until recently head of research at Google DeepMind, thinks a sudden AI intelligence explosion through recursive self-improvement is unlikely. AI can speed up research by a factor of ten, he says, but it hits two bottlenecks: coming up with ideas ("research taste") and reliably judging results. Reward hacking and the speed of light add further limits. Vinyals now wants to tackle these bottlenecks with his startup Discovery Loop, co-founded with Jeff Dean, Sanjay Ghemawat, and Quoc Le. The article Ex-Deepmind VP Vinyals says AI self-improvement is coming but won't trigger an intelligence explosion appeared first on The Decoder.
Open AI’s safety system is already cutting off API responses mid-task
AI companies have spent the last few years competing to build the best models, faster than the other, with each The post OpenAI’s safety system is already cutting off API responses mid-task appeared first on The New Stack.
Microsoft Plans Data Center Push to Triple Its Computing Power
Microsoft Corp. plans to more than triple its data center capacity to help overcome a computing shortage that has forced it to turn away some AI and cloud business. Bloomberg's Anurag Rana joins to discuss. (Source: Bloomberg)
Fidji Simo Joins Data Center Startup Nscale’s Board Ahead of IPO
Nscale Global Holdings appointed former OpenAI and Meta Platforms Inc. executive Fidji Simo to its board as the data center startup bolsters its ranks of directors ahead of an initial public offering.
Meta Meets With European Credit Investors in Non-Deal Roadshow
Meta Platforms Inc. has been speaking to bond investors in Europe as part of a non-deal roadshow for at least a week as hyperscalers look to tap different global debt markets to fund the AI boom.
AI’s existential crisis explodes but AI companies plunge ahead anyway
The resignation this week of Anthropic researcher Jacob Coxon over concerns that artificial intelligence labs are “gambling with our lives” lit up the long-smoldering argument about whether AI is heading in a direction hazardous to human life. Not only is he not alone at Anthropic — whose alignment science lead agreed with the possibility that […] The post AI’s existential crisis explodes but AI companies plunge ahead anyway appeared first on SiliconANGLE.
How hackers used Claude for missiles, drone swarms, and surveillance, while Chinese labs mined it for training data
Anthropic's new threat intelligence report documents eight months of Claude abuse. Chinese AI labs like Alibaba's Qwen team, DeepSeek, and Moonshot AI relayed requests en masse or extracted training data, with Qwen alone accounting for more than 151 million exchanges. Actors also used Claude for missile software, autonomous kamikaze drones, and nationwide surveillance systems. The article How hackers used Claude for missiles, drone swarms, and surveillance, while Chinese labs mined it for training data appeared first on The Decoder.
Anthropic Says Iran, Russia Used Claude for Weapons Research
Anthropic PBC says its artificial intelligence model Claude has been misused in attempts to develop a wide range of military applications, including kamikaze drone swarms, missile navigation systems and research tied to potential biological weapons.
Open AI floats a shared AI slowdown, takes it to Congress
OpenAI wants to know from members of Congress whether an industry-wide slowdown in AI development would be legal, according to several people familiar with the matter. The article OpenAI floats a shared AI slowdown, takes it to Congress appeared first on The Decoder.
Rapidly scaling online storage to serve over 1 billion Chat GPT users
Learn how OpenAI evolved Habitat from a Python library into a globally distributed storage platform serving 1 billion ChatGPT users and 22M requests per second.
Class action lawsuit accuses Anthropic of overselling Claude subscriptions with deceptive usage multipliers
A class action lawsuit accuses Anthropic of misrepresenting how much Claude subscribers actually get to use the service. The article Class action lawsuit accuses Anthropic of overselling Claude subscriptions with deceptive usage multipliers appeared first on The Decoder.
China AI Star Moonshot Eyes $2 Billion Annualized Sales in 2026
Moonshot AI is targeting a manifold jump in annualized revenue to $2 billion by the end of the year, using the breakout success of its Kimi K3 model to dial up the heat on rivals from Anthropic PBC to Z.AI Co.
3 Highlights from Thomas Kurian’s Keynote at the Goldman Sachs Communicopia & Technology Conference
On Tuesday, September 8, Thomas Kurian participated in the Goldman Sachs Tech Conference, providing an update on Google Cloud’s business and strategy. Here are the highlights:Full Stack Approach: Google Cloud is the only provider to offer solutions across the entire AI stack, which expands our total addressable market, differentiates our products from the point of view of performance, cost and quality; and enables us to diversify our revenue streams as the market grows. We have 17 product lines with more than $1 billion in revenues and our customers on average exceeded their commitments by more than 50%. We have also seen more than 2x quarter-over-quarter and year-over-year growth in the number and value of $100 million to $1 billion deals. And we have more than 300 customers each with $100 million-plus contractual commitments.Benefits of Google Cloud’s AI Infrastructure: Our AI Infrastructure is built on highly differentiated products in a large expanding market which helps us lower cost and improve performance and margins for our AI models. We have a 2-year AI server payback period, and TPUs have a much faster expected payback period than GPUs. The majority of our AI infrastructure total contract value is from committed five-year contracts.Benefits of Google Cloud’s broad AI solutions: We have seen strong adoption of Gemini Enterprise, which provides customers with insight across their businesses in a highly cost efficient manner with enterprise control and governance. We have also seen that Google Cloud customers that use our AI products use 1.8 times as many products as those who do not.For more information, please refer to the slide presentation and transcript from the event. This blog post includes statements that could be considered forward-looking. These statements involve a number of risks and uncertainties that could cause actual results to differ materially. Any forward-looking statements in the presentation are based on assumptions as of September 8, 2026, and Alphabet undertakes no obligation to update them.
Anthropic's $1.5 billion book settlement descends into chaos as authors and publishers fight over who gets paid
Authors and publishers fight over how to split Anthropic's $1.5 billion settlement, the largest copyright deal in US history. The article Anthropic's $1.5 billion book settlement descends into chaos as authors and publishers fight over who gets paid appeared first on The Decoder.
Open AI Withdraws Sponsorship Of Caltech Mathathon After Mathematicians’ Revolt
When a group of Caltech undergraduates announced the Caltech Mathathon — billed as the first hackathon devoted to research-level mathematics — they framed... The post OpenAI Withdraws Sponsorship Of Caltech Mathathon After Mathematicians’ Revolt appeared first on OfficeChai.
Open AI's new Agents API gives developers the infrastructure behind Codex and Chat GPT
OpenAI is releasing the Agents API as a public beta. It lets developers build cloud agents that run autonomously for hours, execute code, and hand off tasks to sub-agents. There are no extra fees beyond token usage. Cloudflare, Vercel, and Oracle offer additional sandbox environments. The article OpenAI's new Agents API gives developers the infrastructure behind Codex and ChatGPT appeared first on The Decoder.
Kimi And Deep Seek Served Claude To Their Users Instead Of Their Own Models To Collect Exchanges For Model Training, Says Anthropic
Anthropic is continuing to level some astonishing allegations at Chinese labs. Anthropic has published new findings alleging that Moonshot AI, the company behind... The post Kimi And DeepSeek Served Claude To Their Users Instead Of Their Own Models To Collect Exchanges For Model Training, Says Anthropic appeared first on OfficeChai.
Open AI Weighs Slowing Cutting-Edge AI Development
OpenAI is considering slowing down the development of cutting-edge artificial intelligence, with CEO Sam Altman hoping other AI companies will do the same. In recent weeks, employees of leading firms have raised public concerns about the increased risks of advanced AI systems. Bloomberg Intelligence's Matthew Bloxham breaks down the situation. (Source: Bloomberg)
NVIDIA CEO Jensen Huang Calls Anthropic Quitter Jacob Coxon’s Comments Outlandish & “Deeply Untrue”
Nvidia CEO Jensen Huang has weighed in directly on Jacob Coxon, the pretraining researcher whose resignation from Anthropic and viral warning about existential... The post NVIDIA CEO Jensen Huang Calls Anthropic Quitter Jacob Coxon’s Comments Outlandish & “Deeply Untrue” appeared first on OfficeChai.
Sakana AI Launches Fugu Max and Fugu Ultra v2 for Cheaper, Stronger Multi-Agent Orchestration
Sakana AI has released Fugu Max and Fugu Ultra v2, 2 models built on the same learned orchestration architecture. Fugu Max routes tasks to lean open and specialized models, including NVIDIA Nemotron, at $2/$6 per 1M tokens. Fugu Ultra v2 targets peak capability, scoring 48.3 on Chartography and 74.3 on DeepSWE. The post Sakana AI Launches Fugu Max and Fugu Ultra v2 for Cheaper, Stronger Multi-Agent Orchestration appeared first on MarkTechPost.