“We do not have good approaches for understanding/overseeing the exercise and goals of AI ‘swarms,’” wrote Greenblatt on X. “The problem of understanding incidents and overseeing AI brokers seems to be rising sooner than the speed at which extra succesful AIs assist us with oversight and understanding.”
The impartial researchers’ reliance on AI was partly necessitated by the truth that they have been a staff of solely three individuals, whose investigation at OpenAI was initially deliberate to final two days, then prolonged to 6 after they raised considerations about restricted time and incomplete knowledge, in response to the report.
OpenAI printed its personal technical report on the incident individually on Wednesday. The corporate stated in August that it had moved some employees from capabilities work to alignment, and paused a few of its coaching till it might higher mitigate what went unsuitable.
However the impartial researchers’ reliance on AI to grasp the Hugging Face incident is a microcosm of a much bigger pattern. Main AI firms are themselves more and more counting on AI to observe their very own methods for wrongdoing.




