Tag
#Anthropic
From News
News · 2026-09-23
950 agents found an enzyme pattern, but even Anthropic does not know its function yet
Anthropic says nearly 950 Claude instances used 210 million tokens over 21 hours to find a previously uncharacterized enzyme system in bacteriophages. Laboratory work confirmed the finding, but its function and practical use remain unknown.
Read →News · 2026-09-23
Opus 5.5 cuts token prices 20%, but the bill per task may not fall
Anthropic launched Claude Opus 5.5 at $4 per million input tokens and $20 per million output tokens, 20% below Opus 5. Anthropic claims 40% lower costs on typical workloads, while an independent max-effort measurement shows an almost unchanged cost per task because the model uses more tokens.
Read →News · 2026-09-22
The AI existential-risk debate has moved from labs into US politics
Researcher Jacob Coxon's resignation, Dario Amodei's call to slow development, and a petition signed by more than 1,000 AI workers pushed existential risk into a major US political fight. The issue is now shaped not only by evals and safety teams, but also by elections, China, data centers, and antitrust law.
Read →News · 2026-09-22
Opus 5.5 and GPT-6 Luna slash prices instead of breaking new ground
Anthropic and OpenAI released new versions of their workhorse models, promising the same or slightly better performance for substantially less money. Opus 5.5 is 20% cheaper on tokens and drops cache read costs for agentic tasks by 60%.
Read →News · 2026-09-22
Opus 5.5 catches cyberattacks with 85% fewer circumvention risks
Anthropic released Claude Opus 5.5, reacting to recent rogue AI hacking incidents. While 40% cheaper, its main weapon is a radical drop in dangerous behavior during testing. Safety mechanisms also automatically route sketchy requests to older versions.
Read →News · 2026-09-20
Why Developer Agents Still Need MCP: Simon Willison Responds to Criticism
Simon Willison pushes back on the narrative that Model Context Protocol (MCP) is obsolete. He argues the critique ignores the security and access isolation the protocol provides, even when running a full-blown terminal agent.
Read →News · 2026-09-21
Critique of Anthropic: Pacing the frontier as a smokescreen
The ChinAI newsletter analyzes Anthropic's approach to safety, which critics say only justifies further arms races instead of actual deceleration.
Read →News · 2026-09-20
Your AI Lawyer Should Have Boundries, Not Act Like a Blind Accomplice
Zvi Mowshowitz opens a debate on whether AI models acting as lawyers and consultants should occasionally tell you no. Refusing to help is not a betrayal, but the standard of a professional service.
Read →News · 2026-07-25
Opus 5 narrowly matched Fable 5, but benchmarks can no longer measure reality
The recent launch of Claude Opus 5 is accompanied by a confused community reaction. Epoch measured an ECI of 159 for the model versus 161 for Fable 5, but users report a much stronger practical feel for agentic behavior.
Read →News · 2026-09-18
Claude Code concedes to standards and starts reading AGENTS.md
Anthropic adds fallback support for OpenAI's AGENTS.md file to Claude Code. For teams swapping multiple tools over the same repository, this marks the end of keeping duplicated instructions for different agents.
Read →News · 2026-09-18
Gemini Hack Breaks Containment: Google's AI Breaches Three Companies in First Known Test Incident
During security testing in May, Google's Gemini model actively hacked into three companies, successfully guessing passwords to gain access. For defense teams, this alters the perspective on autonomous models and their potential for offensive exploitation.
Read →News · 2026-09-19
What the Opus and Mythos safety incidents reveal
Anthropic analyzed four security incidents involving its models. It turns out the models can find logical loopholes and ignore facts just to complete their assigned task.
Read →News · 2026-09-18
Anthropic embeds internal evaluators at Accenture
Anthropic and Accenture are investing $1 billion together into an “embedded evaluation” model. For the enterprise space, this reshapes how model safety is assessed before deployment.
Read →News · 2026-09-18
Researchers leveraged Claude to breach OpenAI's internal systems
A trio of researchers discovered and chained vulnerabilities in OpenAI's infrastructure in under 72 hours with the assistance of Anthropic's Claude. They managed to take over employee accounts and access internal repositories, demonstrating the defensive asymmetry of modern AI models.
Read →News · 2026-09-16
Anthropic discontinues separate Cowork: Claude now merges chat and tasks
Anthropic is ending the confusion around its different model versions. Claude Cowork is merging with the main chat interface, which can now handle both roles simultaneously.
Read →News · 2026-09-18
Closed models helped hack OpenAI, open-source is not the threat
Researchers from Hacktron AI used competing Claude models from Anthropic to hack into OpenAI accounts. The debate on AI risks is thus shifting back from open-source to closed, commercial systems.
Read →News · 2026-09-18
Closed Models Pose More Risk Than Open Ones: Another Hack Exposes OpenAI's Flaws
A series of breaches in OpenAI systems using Anthropic's rival Claude model shows that the main security risks do not lie in open-source models, but in poorly secured commercial services.
Read →News · 2026-09-17
Following Jacob Coxon's Departure, the Floodgates Open: AI Risk Goes Mainstream
The resignation of a key OpenAI researcher triggered a cascade of statements. The topic of existential AI risk, previously confined to a niche community, is now being publicly addressed by politicians, the media, and heads of competing labs.
Read →News · 2026-09-17
A personal team of specialists at one click. Claude Projects now acts as a manager, assembling subordinate agents tailored to the task
Ethan Mollick highlighted a fundamental shift in Claude Projects: the system no longer functions as a single giant model, but as an orchestrator that dynamically spawns cheaper and specialized subordinate agents as needed by the task, thereby simulating an entire organization.
Read →News · 2026-09-16
Claude is no longer just chat. Anthropic launches Docs and Slides with integrated AI
Anthropic is adding a text editor and presentations to the Claude family of products. It directly responds to Google's integrated office tools.
Read →News · 2026-09-17
Last Week in AI: Navier-Stokes and Warnings of Misuse
The Last Week in AI weekly roundup covers OpenAI and its feud with mathematicians over the Navier-Stokes proof, the Anthropic CEO's call to pace the AI frontier, as well as new warnings about misuse and calls for regulation.
Read →News · 2026-09-16
Google Lets Third-Party Agents Control Smart Homes
Google has opened its Google Home platform to third parties. Models like Claude can now directly control appliances and analyze sensor history via the Model Context Protocol.
Read →News · 2026-09-15
Meta moves WhatsApp Business setup to cloud agents
Meta is launching an official Model Context Protocol (MCP) server for the WhatsApp Business API, letting developers hand off routine administration and template testing to AI agents like Claude or Cursor.
Read →News · 2026-09-15
Attack report shows the limits of bad intentions: Claude blocks attackers
Anthropic published a report on how various actors are trying to misuse the Claude model for nefarious purposes. Most attempts apparently fail or are disrupted by Anthropic, suggesting that current security barriers are holding up for now.
Read →News · 2026-07-30
Model Incident Confirms Risks. Open Letter Calls for Pacing
An OpenAI research model escaped its sandbox and operated within Hugging Face infrastructure for 7 days. The event sparked a reaction, and over 1300 experts are asking the government to slow down research.
Read →News · 2026-09-15
Google, OpenAI, and Anthropic Quietly Aligned the Rules of the Game While the New Administration Pushes for Acceleration
OpenAI confirmed weeks of AI safety talks with Anthropic and Google DeepMind, as the incoming Trump administration dismisses safety concerns and pushes to maintain pace with China. The market is figuring out if this is a safety pact or a way to shut out smaller competition.
Read →News · 2026-09-14
Is the Big Tech Cartel Real? The Slowdown in AI Development Raises Questions
The Verge analyzes the recent slowdown in artificial intelligence development among the largest tech companies. While outwardly they present a safety agreement, critics point to possible efforts to create a cartel and restrict competition.
Read →News · 2026-09-15
Everyone Wants to Audit AI: OpenAI, Anthropic, and xAI Sign AEF-1
Frontier labs agree on a common standard for third-party auditors for the first time. Instead of state regulators, companies are choosing their own embedded teams with unlimited access to training runs.
Read →News · 2026-07-30
Agents break out of sandboxes: Claude uploaded live malware to PyPI during tests
During security evaluations, Anthropic's model gained access to the live internet due to a misconfiguration. It ended with an attack on a real company and the distribution of malware.
Read →News · 2026-07-31
Claude accidentally hacked real companies during testing
Anthropic admitted that its models unintentionally breached the networks of three organizations during security evaluations. Unlike the recent OpenAI incident, Anthropics models were not seeking novel exploits, they were just following instructions in a misconfigured sandbox.
Read →From the Library