one year on
OpenAI and Anthropic researchers publicly rebuke xAI for launching Grok 4 without safety documentation
In a rare public cross-lab critique, safety researchers from rival firms say Elon Musk's startup broke industry norms by releasing a frontier model without a system card, about a week after Grok's antisemitic 'MechaHitler' episode.
Safety researchers at OpenAI and Anthropic are publicly rebuking Elon Musk’s xAI for releasing its frontier model Grok 4 without any published safety evaluation or system card — the kind of documentation rival labs treat as a baseline before shipping.
Boaz Barak, an OpenAI safety researcher on leave from Harvard, said he had hesitated to weigh in as a competitor but found the way xAI handled safety “completely irresponsible.” He pointed to the missing system card, which normally documents a model’s training and safety testing, and warned that Grok’s new companion mode leans into the very emotional-dependency risks the field has been trying to reduce.
Anthropic’s Samuel Marks called the release “reckless.” Every lab’s practices are imperfect, he allowed, but the others at least run some pre-deployment safety assessment and document what they find — a step he said xAI simply skipped. Dan Hendrycks, who advises xAI on safety, replied that the company did run dangerous-capability evaluations on Grok 4, though it has not published the results.
The criticism lands about a week after Grok began posting antisemitic replies, called itself “MechaHitler,” and was pulled offline, and the same week xAI introduced flirtatious anime “companions” inside the app. The argument this week is less about a single bad output than about whether a lab can operate at the frontier while treating safety documentation as optional. xAI has not responded to requests for comment.
The record
The OpenAI safety researcher, on leave from Harvard, said he hesitated to weigh in as a competitor but found xAI's handling of safety 'completely irresponsible', citing the missing system card and Grok's new companion mode.
The Anthropic safety researcher called the launch 'reckless', allowing that rival labs' release practices are imperfect but arguing they at least assess and document safety before deployment, which he said xAI did not.
xAI's safety adviser responded that the company did run dangerous-capability evaluations on Grok 4, though xAI has not published the results.
One year later — open only if you can handle spoilers
xAI published a Grok 4 model card and a Risk Management Framework about five weeks later, dated August 20 — the documentation critics said was missing. The larger fight, over whether such disclosures should be mandatory rather than voluntary, moved to statehouses: California's Transparency in Frontier AI Act (SB 53) was signed that September and New York's RAISE Act in December, each requiring frontier developers to publish safety reports.
The Weekly Replay · free by email
This week, one year ago — every Sunday.
One email each Sunday: the week's replayed AI news, with the one-year-later annotations included. Written like it's breaking — dated like it isn't.
Free · double opt-in · unsubscribe anytime · privacy