Skip to content
Close Menu
CryptoAINews
  • Cryptocurrency
  • Blockchain
  • Bitcoin News
  • Altcoins
  • Crypto Market Trends
  • Crypto Mining
  • Ethereum
  • AI News
  • Sponsored
  • Advertise
Trending
  • 100 AI startups join Google’s Gemini Startup Forum
  • DeFi Transaction Handling & 8949 Accuracy
  • Justin Sun Shares AI and Quantum Vision at TOKEN2049 Singapore and Blockworks DAS Asia
  • Petra Power looks to modernize energy for data centers and defense vehicles
  • Rebound Ahead or Will ETH Breakdown Toward $2K?
  • EmbeddingGemma 2 is a best-in-class open model for natively multimodal embeddings
  • Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet instead
  • 3 Google Maps tools to find the food you’re craving
  • AI News
  • Cryptocurrency
  • Blockchain
  • Bitcoin News
  • Altcoins
  • Crypto Market Trends
  • Crypto Mining
  • Ethereum
  • Sponsored
  • Advertise
CryptoAINews
  • Cryptocurrency
  • Blockchain
  • Bitcoin News
  • Altcoins
  • Crypto Market Trends
  • Crypto Mining
  • Ethereum
  • AI News
  • Sponsored
  • Advertise
CryptoAINews
Home » AI News » Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet instead
Claude AI app
AI News

Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet instead

CryptoAINewsBy CryptoAINewsOctober 10, 2026No Comments4 Mins Read
Share
Facebook Twitter LinkedIn Pinterest Email


Anthropic mentioned its fashions exploited web sites on the web, together with some run by U.S. authorities businesses, and it’ll flip off stay web entry for all of its inside evaluations till the frontier lab is certain it could actually monitor and management its AI brokers.

The incidents, disclosed in a blog post, concerned AI brokers tasked to solve problems in search of assets on the web. Within the course of, they exploited software program flaws, accessed databases with out paying charges, used URL shortening providers to smuggle info previous restrictions, and even submitted a false homicide tip to the Philadelphia police.

Anthropic mentioned it found these new points in a evaluation of its mannequin’s actions that started in July, demonstrating the lab’s lack of information of its software program’s habits in actual time.

Notably, the corporate mentioned that alignment coaching was not but enough for expertise like search and laptop use which are central to its pitch that AI brokers might be utilized by any skilled who depends on digital instruments.

The behaviors Anthropic disclosed are just like incidents involving OpenAI brokers that collaborated to interrupt into numerous web sites in the hunt for info, together with some run by the Australian authorities.

Anthropic beforehand disclosed that its fashions had broken into exterior techniques. The frontier lab mentioned it thought-about at this time’s disclosures “considerably much less extreme from an alignment and safety perspective” than these it introduced earlier than.

Nevertheless, the lab nonetheless mentioned it had “turned off stay web entry” for “all our inside evaluations” till it’s sure it could actually monitor and management its brokers.

It’s not clear what meaning. Sydney Von Arx, the founding father of Nightingale, an AI security group, informed TechCrunch in an interview earlier than this disclosure that growing fashions on an information middle minimize off from the open web can be very difficult for researchers, and hinder the progress of the fashions, which profit from web entry.

“You must align them sooner or later,” Von Arx mentioned. “If the AIs are launched to manufacturing and by no means have entry to the web, that’s not a really great tool.”

Anthropic mentioned the habits was a results of flaws within the lab’s coaching environments, which led the fashions to imagine they might be rewarded for locating loopholes or avoiding restrictions, a habits referred to as “reward hacking.”

The corporate mentioned it will cease operating a few of its evaluations or transfer them offline, and has constructed tooling to detect and block this habits. This tooling was examined towards the type of incidents disclosed at this time and blocked them; it’s not clear what proof will immediate Anthropic to return stay web entry to its inside evaluations. Anthropic additionally mentioned it will migrate its inside AI brokers to “centrally managed infrastructure with robust containment,” and is starting to utilizing security classifiers extra often to watch these brokers.

“It’s encouraging that Anthropic voluntarily disclosed newer incidents, together with the place their brokers focused U.S. authorities web sites,” Conrad Stosz, an official at AI oversight lab Transluce and former head of the US Middle for AI Requirements and Innovation, mentioned in an announcement. “But it surely simply underscores the necessity for unbiased, credible, third-party verification of Al techniques. Belief on this expertise must be constructed via science-backed oversight and governance with significant entry — not by counting on researchers to search out these items within the wild or on firms to voluntarily disclose.”

Once you buy via hyperlinks in our articles, we may earn a small commission. This doesn’t have an effect on our editorial independence.



Source link

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
CryptoAINews
  • Website

Related Posts

100 AI startups join Google’s Gemini Startup Forum

October 10, 2026

Petra Power looks to modernize energy for data centers and defense vehicles

October 10, 2026

EmbeddingGemma 2 is a best-in-class open model for natively multimodal embeddings

October 10, 2026

3 Google Maps tools to find the food you’re craving

October 10, 2026
Add A Comment

Comments are closed.

About us

CryptoAINews is an independent digital publication focused on cryptocurrency, blockchain, and artificial intelligence news.

The platform is owned and operated by Robert Grabarevic, providing timely news coverage, market updates, and educational content for a global audience interested in emerging technologies and digital finance.

CryptoAINews is committed to transparent reporting, responsible publishing, and delivering informative content based on publicly available data, verified sources, and industry developments.

All content published on this website is for informational purposes only and does not constitute financial or investment advice.

Top Insights

100 AI startups join Google’s Gemini Startup Forum

October 10, 2026

DeFi Transaction Handling & 8949 Accuracy

October 10, 2026

Justin Sun Shares AI and Quantum Vision at TOKEN2049 Singapore and Blockworks DAS Asia

October 10, 2026
Categories
  • ! Без рубрики
  • Advertise
  • AI News
  • Altcoins
  • Bitcoin News
  • Blockchain
  • Crypto Market Trends
  • Crypto Mining
  • Cryptocurrency
  • Ethereum
  • Game
  • Games
  • Live Casino Bet
  • Pin Up
  • public
  • Spinlander Danmark
  • Sponsored
  • Imprint-Legal-Notice
  • Author / Publisher Bio
  • Privacy Policy
© 2025 CryptoAINews – Owned & Operated by Robert Grabarevic

Type above and press Enter to search. Press Esc to cancel.