Model
gpt-neo
Featured in threads
Tracks
- Benchmarks & progress 1
- Safety & alignment 1
- Open weights & ecosystem 1
- Models & capabilities 1
TruthfulQA measures whether models repeat human falsehoods
On 817 questions designed to elicit common misconceptions, the best model tested was truthful only 58% of the time against 94% for humans, and larger models scored worse.
Benchmarks & progress · Safety & alignment
EleutherAI releases GPT-Neo
The 1.3B and 2.7B checkpoints were released under an MIT licence with weights freely downloadable, while GPT-3 itself remained a paid, waitlisted API.
Open weights & ecosystem · Models & capabilities