The “Dangerous” OpenAI Text Generator Recreate by Two Researchers
On Thursday, a pair of master graduates from Computer Science rolled out an AI text generator based upon GPT-2, an Elon Musk-backed OpenAI program that the company withheld from public release citing concerns over the societal impact of it.
Quick answer: In 2019 two researchers, Aaron and Vanya, recreated OpenAI's GPT-2 language model using about $50,000 of free Google cloud computing credits. They wanted to show that anyone could build such software, and said it carried no risk to society yet. OpenAI had held GPT-2 back, citing fear of misuse.
Key takeaways
- The 2019 story is about two researchers rebuilding GPT-2 with free cloud credit to argue that such models were not an immediate risk.
- It shows that a model once called too dangerous to release can be reproduced by small teams.
- Treat this as news history; it says little about today's AI models.
However, the 2 researchers, Aaron and Vanya believe that the software does not possess any risk to society — not yet. According to Wired, the duo wanted to prove that anyone can develop the software, regardless of their economical status.
In order to replicate GPT-2, the duo used $50,000 worth of free cloud computing from Google. The research graduates also fed millions of webpages to the ML software, gathered by digging up links shared on Reddit.
Just like OpneAI GPT-2, the newly created software analyzed the language patterns and could be used up for many tasks — Translation, chatbots, coming up with unprecedented answers and more. However, the foremost alarming concern among specialists has been the creation of synthetic text, consequently, Fake news.
David Luan, vice president of engineering at OpenAI once told Wired, “It could be that someone who has malicious intent would be able to generate high-quality fake news”. Owing to this and other dangers, the team decided to withhold the model. However, it did put out a research paper.
Previously, there have been iterations of GPT-2. In fact, few people have released language models online based upon the OpenAI software. Of course, they’re not using the first model that used “8 million web pages”, however, it still uses the previous versions. You can try it out yourself.
While they’re smart for playing around, they don’t appear to provide logical statements. Wired, who tested out the original GPT-2 and new model as well, writes, “Machine learning software picks up the statistical patterns of language, not a true understanding of the world.”
To take this further with guided labs and an instructor, see our practical LLM engineering labs.
Related reading
- Is ChaosGPT a Threat to Cybersecurity? Examining the Risks, Potential Dangers, and Ethical Concerns of Autonomous AI
- What are the new features of OpenAI GPT-5, and how will it impact the AI industry in 2026?
- OpenAI Introduces AI Image Generation in ChatGPT | A Game-Changer for Creativity and Content Creation with GPT-4o
Frequently Asked Questions
What's Your Reaction?
Like
0
Dislike
0
Love
0
Funny
0
Wow
0
Sad
0
Angry
0