one year on
Meta unveils Galactica demo, drawing criticism over fabricated and biased scientific text
The company's large language model for science, trained on 48 million papers, is drawing criticism after researchers show it fabricates papers and produces offensive output.
Meta announced on November 15 a new large language model called Galactica, trained on 48 million scientific papers, textbooks, and encyclopedias. The company promoted it as a tool that “can summarize academic papers, solve math problems, generate Wiki articles, write scientific code, annotate molecules and proteins, and more.” Researchers and the public were invited to try the demo.
Within hours, scientists began sharing examples of Galactica fabricating papers—sometimes attributing them to real authors—and generating authoritative-sounding nonsense, such as a wiki article on “the benefits of eating crushed glass” or a history of bears in space. The model also produced racist and biased content when prompted with certain topics, and refused to generate text on “racism” or “AIDS,” returning a content filter message.
Chief AI scientist Yann LeCun defended the model, tweeting “Type a text and Galactica will generate a paper with relevant references, formulas, and everything.” The episode mirrors Microsoft’s 2016 Tay chatbot debacle, where a public demo was shut down after users turned it into a racist bot. It highlights the gap between the promise of large language models and their current inability to reliably separate fact from fiction.
The record
Said the model was 'wrong or biased but sounded right and authoritative' and called it dangerous.
Said he was 'astounded and unsurprised,' and that people don't grasp such systems can't work as hyped.
Called the model fun and impressive but said it was unfortunate it was touted as a practical research tool.
Described the ability of LLMs to mimic text as 'a superlative feat of statistics' in a Substack post.
Defended the model, later tweeting 'Galactica demo is off line for now. It’s no longer possible to have some fun by casually misusing it. Happy?'
One year later — open only if you can handle spoilers
Galactica's implosion became an early cautionary tale about releasing LLMs with confident hallucinations, often cited before ChatGPT's launch two weeks later. The episode reinforced concerns that the industry's "move fast" approach would repeat with newer models.
The Weekly Replay · free by email
This week, one year ago — every Sunday.
One email each Sunday: the week's replayed AI news, with the one-year-later annotations included. Written like it's breaking — dated like it isn't.
Free · double opt-in · unsubscribe anytime · privacy