{"url":"/dataset/l3cubemahasent","name":"L3CubeMahaSent","full_name":null,"description_markdown":"L3CubeMahaSent  is a large publicly available Marathi Sentiment Analysis dataset. It consists of marathi tweets which are manually labelled.\r\n\r\nThis dataset contains a total of 18,378 tweets which are classified into three classes - Positive (1), Negative (-1) and Neutral (0). All tweets are present in their original form, without any preprocessing.\r\n\r\nOut of these, 15,864 tweets are considered for splitting them into train, test and validation datasets. This has been done to avoid class imbalance in the dataset.\r\nThe remaining 2,514 tweets are also provided in a separate sheet.","description_withheld":null,"homepage":"https://github.com/l3cube-pune/MarathiNLP","introduced_date":"2021-03-21","introduced_date_note":null,"introduced_by":{"paper":"/paper/l3cubemahasent-a-marathi-tweet-based","title":"L3CubeMahaSent: A Marathi Tweet-based Sentiment Analysis Dataset","first_author":"Atharva Kulkarni","url":null},"license":{"name":"CC BY-NC-SA 4.0","url":"https://creativecommons.org/licenses/by-nc-sa/4.0/"},"modalities":[{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[{"name":"Sentiment Analysis","url":"/task/sentiment-analysis","datasets_with_task":"/datasets/task/sentiment-analysis"}],"languages":[{"name":"Marathi","url":"/datasets/language/marathi"}],"variants":["L3CubeMahaSent"],"data_loaders":[{"repo":"https://github.com/l3cube-pune/MarathiNLP","url":"https://github.com/l3cube-pune/MarathiNLP","frameworks":[]}],"num_papers_in_archive":8,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}