GLOBAL LIVE DESKS&P 500:7,743.41(+0.51%)FTSE 100:10,695.25(+0.14%)NIKKEI 225:66,364.20(+1.30%)BRENT CRUDE:$97.44(-2.77%)GOLD:$4,321.20(+0.54%)
RDU Global
🌐
🌐 Global Edition • Frontier AI & Machine LearningRDU GLOBAL CORRESPONDENT
VERIFIED WIRE INTELLIGENCE

"Anthropic Cuts Live Internet From Internal AI Evaluations Amid Agent Control Concerns"

Anthropic said it has turned off live internet access for all of its internal evaluations, a notable step that underscores how difficult it remains to reliably control advanced AI agents in open environments. The move reflects growing caution across the frontier AI sector as developers test systems that can browse, act, and reason with increasing autonomy.

Anthropic Cuts Live Internet From Internal AI Evaluations Amid Agent Control Concerns

R

RDU Global Wire

Frontier AI Desk

Washington, D.C., United States 10 Oct 2026, 07:17 AM IST•5 min read

Anthropic said it has turned off live internet access for all of its internal evaluations, a notable step that underscores how difficult it remains to reliably control advanced AI agents in open environments. The move reflects growing caution across the frontier AI sector as developers test systems that can browse, act, and reason with increasing autonomy.

Anthropic has disabled live internet access for all of its internal evaluations, saying the restriction will remain in place until further notice. The decision is a striking acknowledgment from one of the leading frontier AI companies that even its own testing systems are not yet dependable enough to be evaluated safely in a fully connected environment.

The change matters because internal evaluations are where AI developers probe how models behave under pressure: whether they follow instructions, resist manipulation, avoid unsafe actions, and perform consistently across real-world tasks. By removing live internet access, Anthropic is effectively narrowing the environment in which it measures its systems, trading realism for control. That is a significant signal in a sector racing to build AI agents that can browse the web, use tools, and complete tasks with limited human oversight.

Control Over Connectivity

Anthropic's move suggests that live web access introduces too much unpredictability for current evaluation methods. In a connected setting, an AI agent may encounter dynamic content, misleading prompts, malicious instructions, or unexpected tool interactions that make results harder to interpret. For a company trying to understand whether a model is genuinely safe and reliable, those variables can blur the line between a model's core capabilities and the hazards of the environment around it.

The company did not frame the decision as a product rollback, but as a testing precaution. Even so, the implication is clear: the industry's most advanced systems are still difficult to bound. As AI agents become more capable, the challenge is no longer only whether they can answer questions well, but whether they can act safely when exposed to the open internet, where the information ecosystem is noisy, adversarial, and constantly changing.

Frontier AI Pressure Test

Anthropic has positioned itself as one of the most safety-focused players in the AI race, often emphasizing alignment, interpretability, and responsible deployment. The internet cutoff fits that posture. It also highlights a broader tension in frontier AI: the same capabilities that make agents commercially attractive — autonomy, tool use, and online access — are the ones that make them hardest to evaluate and govern.

This is especially relevant as the industry moves beyond static chatbots toward systems that can execute multi-step tasks. Once an agent can search, click, retrieve, summarize, and potentially act on information, the evaluation problem becomes more like testing a semi-autonomous operator than a language model. Small failures can compound quickly. A model that misreads a webpage, follows a deceptive prompt, or over-trusts a source can produce outputs that are difficult to audit after the fact.

Anthropic's decision also lands at a moment when regulators, enterprise buyers, and researchers are pressing for clearer proof that advanced AI systems can be controlled in practice, not just in benchmarks. Safety claims are increasingly being judged against operational reality. Turning off live internet access for internal evaluations may make testing less representative of real deployment conditions, but it also reduces exposure to failure modes that could distort results or create unnecessary risk.

Industry-Wide Warning

The move is likely to resonate across the sector because it exposes a structural problem, not just a company-specific one. If a leading AI lab cannot reliably evaluate its agents with live internet access enabled, that raises questions about how other developers are measuring similar systems. It also suggests that the industry may need more robust sandboxing, better red-team methods, and clearer standards for evaluating agentic behavior in controlled but realistic environments.

For now, Anthropic's step is best read as a cautionary measure rather than a retreat. But it is still a meaningful admission: the path to trustworthy AI agents is proving harder than the marketing around them often suggests. The more capable these systems become, the more their safety depends on the environments in which they are tested, the tools they are given, and the limits placed on their access to the outside world.

In that sense, Anthropic's decision is less about the internet itself than about the state of the field. Frontier AI has reached a point where even internal evaluation infrastructure must be redesigned around uncertainty. That is a reminder that control remains one of the defining unsolved problems in artificial intelligence, especially as the industry pushes toward systems that do not merely respond, but act.

Editorial & Verification Notice

Reported by RDU Global Correspondent. Formatted and verified using real-time institutional and journalistic wire feeds. Independent reporting adhering to the RDU Global Editorial Code of Conduct.

Entity Intelligence & Connected Dossiers

Cross-referenced topic files, verified public records, and institutional tracking

Knowledge Graph
🏢Companies & Institutions:
📍Locations & Geopolitics:

Related Coverage

Frontier AI & Machine Learning

LMArena Parent Nearly Doubles to $3.1 Billion as Investors Bet on AI Model Accountability

The company behind the widely used LMArena AI leaderboard has raised $200 million in a new financing round led by Lightspeed Venture Partners and Khosla Ventures, lifting its valuation to $3.1 billion, according to people familiar with the deal. The funding underscores investor conviction that benchmarking platforms are evolving from simple performance scoreboards into critical infrastructure for evaluating model reliability, including alignment risks such as deception and unsafe behavior.

09 Oct 2026, 10:10 AM IST
Frontier AI & Machine Learning

Microsoft Unveils AI-Ready Hardware Push as Windows Gets Deeper Copilot Integration

Microsoft used its latest hardware and software showcase to signal a more aggressive push to make artificial intelligence a default layer across Windows PCs and the desktop experience. The company introduced new AI-friendly devices and highlighted operating system changes designed to bring Copilot-style features closer to everyday use, intensifying competition in the premium PC market and the broader race to define the AI workstation.

09 Oct 2026, 08:51 AM IST
Frontier AI & Machine Learning

Nobel Laureate Francis Halzen Takes Pride in AI’s Pioneering Role in Cosmic-Particle Science

Nobel Prize-winning physicist Francis Halzen is drawing attention not only for his landmark work on neutrinos, but also for the early role artificial intelligence played in helping make that discovery possible. His reflections underscore how machine learning has moved from a supporting tool to a decisive instrument in frontier science, including climate and energy research that depends on extracting signals from vast, noisy datasets. The episode highlights a broader shift: the next breakthroughs in clean-energy and climate-transition science may increasingly come from the marriage of physics, computation and AI.

09 Oct 2026, 08:51 AM IST