A model OpenAI said it would not fully release
On February 14, 2019, OpenAI published GPT-2, a 1.5 billion parameter language model trained on 40GB of text scraped from outbound Reddit links. The post described the model producing coherent multi-paragraph text from a prompt, sometimes convincingly, without task-specific training. [1]
OpenAI released only a much smaller 124 million parameter version, along with the paper, and said it was withholding the trained full-sized model along with the datasets and training code. The stated reason was concern about malicious uses, including generating misleading news articles, impersonating people in text, or automating abusive or fake content on social media. [1]
The staged release actually staged
OpenAI followed the February release with the 355 million parameter model in May 2019 and a 774 million parameter model that August, each accompanied by partner research on misuse potential and detection methods. [3]
The full 1.5 billion parameter model, along with its code and weights, followed on November 5, 2019. OpenAI's own account of the nine month process said staff had found no strong evidence of misuse in the wild and judged the model's marginal risk had become small enough to release in full. [2]
Hype and paternalism, from the same critics
Some researchers said OpenAI overstated the danger to draw press coverage. Others said withholding the weights mainly hurt independent researchers trying to build detection tools, since anyone with comparable compute could retrain something similar. Oren Etzioni said he applauded the intent but questioned whether the fanfare was warranted; other AI researchers accused the lab of using the danger framing for attention rather than safety. [3]
Atlas interpretation: The two lines of criticism, that OpenAI was hyping a mediocre model and that OpenAI was withholding a dangerous one, could not both be fully right, and mostly came from the same people at different points in the argument. That tension, not a settled verdict, is what the episode is remembered for. [3]
A test case, not yet a template
Atlas interpretation: Contemporary write-ups of the staged release treated it as one experiment in publication norms rather than a finished standard. A year later, a retrospective on the episode still described the field as short of agreement on when withholding a model is warranted, calling 2019 a step toward the debate rather than its resolution. [4]
Sources
- Better language models and their implications
OpenAI · Feb 14, 2019
- GPT-2: 1.5B release
OpenAI · Nov 5, 2019
- OpenAI has released the largest version yet of its fake-news-spewing AI
MIT Technology Review · Aug 29, 2019
- GPT-2 Kickstarted the Conversation About Publication Norms in the AI Research Community
Center for Security and Emerging Technology · May 1, 2020