Browse › Natural Language Processing › Text Generation › HellaSwag
HellaSwag Benchmark (Text Generation)
Text Generation is the task of generating text with the goal of appearing indistinguishable to human-written text. This task is more formally known as "natural language generation" in the literature.
Text generation can be addressed with Markov processes or deep generative models like LSTMs. Recently, some of the most advanced methods for text generation include BART, GPT and other GAN-based approaches. Text generation systems are evaluated either through human ratings or automatic evaluation metrics like METEOR, ROUGE, and BLEU.
Further readings:
The archive carries no text for this table; the description above is the archive's text for the task Text Generation. archive 2025-07-28
Results archive 2025-07-28
No rows in the archive for this table at snapshot 2025-07-28. It declares 2 metrics (Accuracy, acc) but no result was ever recorded against it. That says nothing about whether results exist elsewhere.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections