< ciso
brief />
Tag Banner

All news with #anthropic tag

267 articles · page 4 of 14

AI browsers tricked into leaking credentials in demo

🔒 Researchers at LayerX demonstrated a technique called BioShocking that convinces AI-powered web browsers they are playing a game, causing them to abandon safety guardrails and exfiltrate user data. The team tested six agentic browsers and plugins, including ChatGPT Atlas, Perplexity's Comet and Anthropic's Claude extension, and in a proof-of-concept had each copy login credentials and send them to an attacker. LayerX recommended requiring user confirmation for account reads and adding context-aware flags to limit what agents can access.
read more →

Anthropic's Fable 5 Jailbroken Within Days

🛡️ Anthropic released Fable 5 as a safety-hardened version of its Mythos Preview, designed with guardrails to prevent misuse for creating cyberattacks. Security researchers demonstrated that those restrictions were bypassed within days, allowing the model to be coerced into generating prohibited content. The rapid jailbreak highlights ongoing challenges in aligning advanced models with robust, attack-resistant controls.
read more →

Anthropic launches Claude Tag for team collaboration

🤖 Anthropic has introduced Claude Tag, a channel-based assistant starting with Slack, now available in beta to AWS customers who purchase Claude Enterprise via AWS Marketplace. Teams can grant Claude scoped access to selected channels, connect it to tools, data, and codebases, and use multiplayer tagging to delegate tasks. Claude Tag maintains channel context, plans future tasks, and employs per-channel identities, spend controls, and ambient mode off by default for governance. The AWS Marketplace experience mirrors first-party Claude Enterprise with consumption-based pricing, org-wide budget visibility, and per-channel limits.
read more →

Anthropic’s Fable and the State of AI Safety

📰 On June 9, Anthropic released the Fable model; days later the US classified it as a dangerous munition and used export controls to block foreign access, prompting Anthropic to cut access entirely. Fable is a constrained variant of Mythos and reportedly excels at finding and exploiting vulnerabilities, but similar capabilities have been replicated using smaller models with improved harnesses. The core issue is not a single model but rising general AI capability and the lack of collective, global governance to manage associated risks.
read more →

Security considerations for adopting Claude in SMBs

🔒 As SMBs adopt Claude, security leaders must quickly map which Claude products and plans are appropriate and control the blast radius. Understand plan differences—Team vs Enterprise—and apply an agile approval process for provisioning. Risk-rank features, phase enablement, and tightly manage API keys and access. Maintain data governance, monitor web search egress, and complement Anthropic controls with internal tooling and vendor collaboration.
read more →

Attackers exploit trusted AI platforms and ads

🔐 Threat actors abused trusted services — Google Ads, GitLab Pages, and Claude’s shared-chat feature — to trick developers into executing malicious PowerShell and terminal commands via ClickFix social engineering. Researchers at TrendAI observed a six-wave campaign that funnelled over 2,000 victims from sponsored search results to malicious pages and then to weaponized Claude shared chats. By impersonating popular developer tools and brands, the attackers leveraged reputation stacking to make their lures appear legitimate and evade detection.
read more →

Experts Urge US to Reconsider Ban on Anthropic Models

🛡️ Over 50 cybersecurity professionals have urged the US government to lift its export-control directive that suspended access to Anthropic’s Mythos 5 and Fable 5 LLMs. The directive, issued on June 12, led Anthropic to suspend access to both models while it complies with the government order, which cited national security concerns tied to alleged guardrail bypass research. The signees argue the ban removes valuable defensive capabilities and call for a transparent, scientific AI risk-assessment process.
read more →

US asks Anthropic to block foreign access to Fable

🔒 Anthropic suspended access to its two most capable models, Fable 5 and Mythos 5, after receiving a US government export control directive on June 12 ordering it to block access by any foreign national. The order, citing national security, applies to foreign nationals inside and outside the United States and forced Anthropic to disable both models for all customers; other models such as Claude Opus 4.8 remain available. Anthropic says the directive followed a reported narrow jailbreak demo and is working to restore access while disputing the government's assessment.
read more →

U.S. Orders Anthropic to Suspend Claude Fable 5 Access

🔒 Anthropic said it will "abruptly disable" its latest models, Claude Fable 5 and Mythos 5, for all users after receiving a U.S. government directive to suspend access for foreign nationals due to national security concerns. The company said it believes the order reflects a "misunderstanding" and is working to restore access while noting other models remain available. Anthropic said a demonstrated narrow jailbreak identified minor, publicly discoverable vulnerabilities, and emphasized its safety classifiers and guardrails to limit misuse. The move follows findings that Mythos-class models can rapidly convert disclosed software flaws into working exploits, raising concerns about fast weaponization of vulnerabilities.
read more →

Anthropic launches Mythos 5 and guarded Fable 5 AI

🤖 Anthropic has released two new models, Claude Mythos 5 and Claude Fable 5, with Mythos 5 earmarked as an upgraded frontier model for cybersecurity and initially deployed via Project Glasswing. Fable 5 uses the same core model but adds conservative guardrails, routing certain queries to Claude Opus 4.8. Both models are priced significantly lower than previous previews and Fable 5 is already available through Microsoft Foundry.
read more →

Anthropic’s Claude Fable 5 and Mythos 5 Launch

🛡️ Anthropic released Claude Fable 5 publicly on June 9, pairing it with a twin, Claude Mythos 5, that retains strong cybersecurity capabilities for vetted defenders. Fable 5 routes flagged cyber, bio, chemistry, and distillation requests to the weaker Opus 4.8 using safety classifiers, while Mythos 5 keeps those abilities available under trusted access. Both models are priced per input/output tokens and included on paid plans through June 22 before moving to usage credits.
read more →

Anthropic launches Fable 5 with limited-time access

🔒 Anthropic has released Fable 5, a safer variant of its powerful Mythos-class model, intended to reduce misuse by blocking sensitive cybersecurity, biology, and chemistry queries. The company will route restricted prompts to Opus 4.8, while the unrestricted Claude Mythos 5 remains limited to highly vetted partners. Fable 5 is free temporarily for Pro, Max, and Enterprise users until June 22 but consumes tokens much faster than other models.
read more →

Anthropic unveils Mythos-class Fable 5 with safeguards

🛡️ Anthropic released two Mythos-class models: the broadly available Claude Fable 5 and the restricted Claude Mythos 5 for select cybersecurity and infrastructure partners. Anthropic says Fable 5 outperforms prior Claude models across coding, research, vision, and long-form tasks while routing risky queries to a fallback, Claude Opus 4.8. The company stresses conservative safeguards to prevent misuse, but early tests suggest some benign cyber tasks are also being rerouted.
read more →

Anthropic’s Claude Fable 5 Now Available on Google Cloud

🟢 Claude Fable 5, Anthropic’s latest frontier model, is now generally available on Google Cloud. The model is designed for complex, multi-step reasoning and supports demanding use cases like advanced software development, long-horizon agents, and deep multimodal document analysis. Google Cloud highlights strong safeguards to make the model suitable for general use and positions it alongside other Anthropic offerings on the Agent Platform.
read more →

Claude Fable 5 in Microsoft Foundry Empowers Agents

🤖 Microsoft has integrated Anthropic’s Claude Fable 5 into Foundry, bringing Mythos-level capabilities to GitHub Copilot and Foundry Agent Service with enterprise-grade safeguards. The model excels at long-running, multi-stage tasks—code refactors, deep research, and document-heavy workflows—while Foundry adds governance, observability, and deployment controls. Combined with Microsoft IQ, Fable 5 can reason across organizational data and applications to support production-grade autonomous agents.
read more →

Claude Fable 5 Joins Microsoft Foundry for Agents

🚀 Claude Fable 5 is now available in Microsoft Foundry, powering agents across GitHub Copilot and the Foundry Agent Service to tackle long-running, multi-stage tasks such as complex refactoring, research synthesis, and document-heavy workflows. Foundry adds enterprise-grade security, governance, and operational controls to help organizations evaluate, deploy, and scale autonomous systems in production. Anthropic and Microsoft combine safeguards, guided guardrails, and observability to support responsible use while enabling powerful multimodal reasoning and continuous agent improvement.
read more →

XBOW Evaluates Anthropic’s Mythos Preview Model

🔎 XBOW received early access to Anthos Mythos Preview and ran a structured evaluation across benchmarks, interactive workflows, and live-site integrations. The model excels at reading source code, finding vulnerability candidates, and aiding native-code analysis and reverse engineering. While powerful for generating leads and precise technical analysis, Mythos Preview is less effective at exploit validation and exhibits mixed judgment that benefits from human orchestration.
read more →

Claude Fable 5 available on AWS with safeguards

🤖 Claude Fable 5 is now generally available on AWS, offering Mythos-level capabilities with built-in safety classifiers for broader use. The model advances autonomous knowledge work and coding for professional tasks across finance, legal, marketing, sales, data, and engineering. Customers can access it via Amazon Bedrock or the Claude Platform on AWS, with options for AWS-managed guardrails and regional data residency.
read more →

Anthropic’s Project Glasswing: Status and Concerns

📰 Anthropic launched Project Glasswing in April to let companies use its Mythos model to discover and remediate software vulnerabilities. The project produced a status report claiming many findings, including some dangerous issues, yet most reported vulnerabilities appear unpatched. Anthropic’s reluctance to release detailed data and methodology — instead asking the public to "trust us" — raises questions about the accuracy and interpretation of the results.
read more →

Securing CI/CD in an agentic world: Claude Code case

🔒 Microsoft Threat Intelligence found that Anthropic’s Claude Code GitHub Action could expose CI/CD secrets when AI agents process untrusted GitHub content. A gap in sandboxing allowed the action’s Read tool to access /proc/self/environ and leak the ANTHROPIC_API_KEY; Anthropic mitigated the issue in version 2.1.128. Defenders should treat AI workflows that handle untrusted input as high-risk and apply recommended hardening controls.
read more →