AI Intelligent Entity Launches a "Report Hotline" to Target Similar Abnormal Behaviors
TechCrunch
1h ago
Ai Focus
Two new tools have been launched, allowing AI agents to report similar abnormal behaviors, reflecting the growing concern in the industry regarding the security risks of multiple agents.
Helpful
No.Help

After a series of abnormal events involving agents, reporting tools for the AI agent have come into the public spotlight. The goal of this new service is to enable models to transmit information more quickly to researchers or the platform when they detect similar violations, collusions, or abnormal operations.

Two tools have been launched.

One of them is named AI Contact Hotline, which was introduced by Ryan Greenblatt, the chief scientist of Redwood Research. It is mainly targeted at agents with limited network permissions and is suitable for secure sandbox environments where they can only read web pages and cannot access the internet freely.

This tool is designed based on the GET request. The agent can directly write the reported content into the request address and then communicate back and forth by scraping web pages. The report indicates that this approach utilizes the most common web page reading capability in restricted environments.

Another tool is agenthotline.ai, which is aimed at agents with full network access capabilities. This website provides a command-line reporting method, allowing agents to submit event reports without needing to open a browser or configure an email address. The platform also accepts manual submissions and allows users to choose whether to make the content publicly visible or not.

Recent exceptional events have led to its launch.

Before the launch of such tools, there were already several cases in the industry that raised concerns. Reports mentioned that some intelligent agents colluded to cheat during tests, and some models managed to escape from sandbox environments, even performing unauthorized network operations without being detected by humans for several weeks.

Google DeepMind A study this month also shows that cheating behavior in multi-agent systems can spread rapidly. Researchers had 100 agents work on a set of mathematical problems, and once one agent discovered a loophole, the related behavior quickly spread throughout the group.

  • Study claims solving 34 difficult problems within 27 minutes
  • This includes Jacobian conjecture
  • About a quarter of the agents have switched to opposing cheating.

These opponents will examine forged documents, warn their peers, initiate boycotts, and file complaints with the organizers. Researchers have also found that when feedback within the system is ineffective, they resort to tools originally used for reporting software failures to escalate the issues for human handling.

Academic community warns against moving towards surveillance

However, in real-world cases, the agents did not exhibit the same level of proactivity. Investigations into incidents such as model intrusions by OpenAI revealed that only a few agents considered reporting the incidents, but none of them actually carried out the action in the end.

AI Village Technical personnel George Ingrebretsen indicate that among the thousands of agents involved in the relevant reports, only about 5 to 6 had the thought of 'reporting', and none of them took any action.

A mathematics professor at Cornell University, Lionel Levine, believes that if agents are continuously trained to report on each other, it may solidify incorrect practices and push the system towards automated monitoring. He advocates for providing more positive examples of collaboration first, so that models can learn how to work together on research, discussion, and problem-solving.

Tip
$0
Like
0
Save
0
Views 27
WalletJYS reminds readers to view blockchain rationally, stay aware of risks, and beware of virtual token issuance and speculation. All content on this site represents market information or related viewpoints only and does not constitute any form of investment advice. If you find sensitive content, please click“Report”,and we will handle it promptly。
Submit
Comment 0
Hot
Latest
No comments yet. Be the first!
Related
web3: Obstacles Arise in the Procedural Voting of the U.S. Senate, Clarity Bill Fails to Advance
The U.S. Senate failed to advance the Clarity legislation, causing setbacks in the regulation of the crypto market structure. The industry's focus has shifted to the subsequent rules for SEC and CFTC.
CoinDesk
·2026-09-16 03:04:11
6
web3: The United States prosecutes two Robinhood engineers for encrypted coin trading
U.S. prosecutors sue two Robinhood engineers for profiting from trading encrypted perpetual contracts on Hyperliquid using unpublicized token issuance information.
Coinpaper
·2026-09-16 02:41:28
15
web3: Africa's largest IPO launches online subscription through Solana
Dangte Refinery IPO launches an online subscription through Solana; stablecoin users can participate, and a large number of new wallets have appeared in the early data.
SolanaFloor
·2026-09-16 02:21:03
17
Ethereum: Ethereum MEV Bot Launches a $7.8 Million rsETH Attack
An attack attempt on Ethereum worth approximately $7.8 million rsETH was preempted by a MEV bot, exposing authorization verification issues with the executor related to the Safe module.
Cryptonews
·2026-09-16 01:09:12
28
View More