Datasets › Parakweet Lab's Email Intent Data Set
Parakweet Lab's Email Intent Data Set
This resource contains training and test data for detecting "intent" sentences in email messages. This data comes from the Enron email corpus. Each labeled example is one sentence from an email. We define "intent" here to correspond primarily to the categories "request" and "propose" in the paper:
Cohen, William W., Vitor R. Carvalho, and Tom M. Mitchell. "Learning to Classify Email into``Speech Acts''." EMNLP. 2004.
In some cases, we also apply the positive label to some sentences from the "commit" category if they contain datetime, which makes them useful. Detecting the presence of intent in email is useful in many applications, e.g., machine mediation between human and email.
Training data: 4213 cases with 1631 positives Test data: 991 cases with 277 positives
Also see https://github.com/vseledkin/enron_intent_dataset_verified
Benchmarks archive 2025-07-28
No leaderboard in the archive resolves to this dataset.
Papers archive 2025-07-28
No paper in the archive has a leaderboard row on this dataset.
Dataset loaders archive 2025-07-28
No loader listed in the archive.
Tasks archive 2025-07-28
No task tagged in the archive.
License archive 2025-07-28
No licence recorded in the archive. Absence here is not a statement about the dataset's terms.
Modalities archive 2025-07-28
No modality tagged.
Languages archive 2025-07-28
No language tagged.
Variants archive 2025-07-28
- Parakweet Lab's Email Intent Data Set
1 variant name, as the archive lists them.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections