{"url":"/dataset/klej","name":"KLEJ","full_name":null,"description_markdown":"The KLEJ benchmark (Kompleksowa Lista Ewaluacji Językowych) is a set of nine evaluation tasks for the Polish language understanding task.\r\n\r\nKey benchmark features:\r\n\r\n- It contains a diverse set of tasks from different domains and with different objectives.\r\n- Most tasks are created from existing datasets but the authors also released the new sentiment analysis dataset from an e-commerce domain.\r\n- It includes tasks which have relatively small datasets and require extensive external knowledge to solve them. It promotes the usage of transfer learning instead of training separate models from scratch.\r\n\r\nThe name KLEJ (English: GLUE) is an abbreviation for Kompleksowa Lista Ewaluacji Językowych (English: Comprehensive List of Language Evaluations) and refers to the [GLUE benchmark](/dataset/glue).\r\n\r\nSource: [KLEJ](https://klejbenchmark.com/)","description_withheld":null,"homepage":"https://klejbenchmark.com/","introduced_date":"2020-05-01","introduced_date_note":null,"introduced_by":{"paper":"/paper/klej-comprehensive-benchmark-for-polish","title":"KLEJ: Comprehensive Benchmark for Polish Language Understanding","first_author":"Piotr Rybak","url":null},"license":null,"modalities":[{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[{"name":"Question Answering","url":"/task/question-answering","datasets_with_task":"/datasets/task/question-answering"},{"name":"Natural Language Understanding","url":"/task/natural-language-understanding","datasets_with_task":"/datasets/task/natural-language-understanding"}],"languages":[{"name":"Polish","url":"/datasets/language/polish"}],"variants":["KLEJ"],"data_loaders":[],"num_papers_in_archive":13,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}