{"url":"/task/coreference-resolution","name":"Coreference Resolution","slug":"coreference-resolution","description_markdown":"Coreference resolution is the task of clustering mentions in text that refer to the same underlying real world entities.\n\nExample:\n\n```\n               +-----------+\n               |           |\nI voted for Obama because he was most aligned with my values\", she said.\n |                                                 |            |\n +-------------------------------------------------+------------+\n```\n\n\"I\", \"my\", and \"she\" belong to the same cluster and \"Obama\" and \"he\" belong to the same cluster.","categories":[{"name":"Natural Language Processing","url":"/area/natural-language-processing"}],"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","slug_source":"archive_url"},"counts":{"papers_tagged":880,"papers_with_code":288,"benchmarks":16,"benchmark_tables_in_archive":16,"benchmark_tables_shown":16,"benchmark_tables_withheld_as_spam":0,"benchmark_definition":"a leaderboard table with at least one row; benchmark_tables_shown also counts the zero-row tables; benchmark_tables_in_archive adds the tables withheld as spam","datasets":45,"subtasks":2,"parent_tasks":0},"benchmarks":[{"leaderboard":"/sota/coreference-resolution-on-winograd-schema","slug":"coreference-resolution-on-winograd-schema","dataset":"Winograd Schema Challenge","dataset_url":"/dataset/wsc","rows_in_archive":82,"metrics":["Accuracy"],"first_row_in_archive_order":{"model":"PaLM 540B (fine-tuned)","paper_title":"PaLM: Scaling Language Modeling with Pathways","paper_url":"/paper/palm-scaling-language-modeling-with-pathways-1","paper_date":"2022-04-05","arxiv_id":"2204.02311","code_links":[{"title":"lucidrains/CoCa-pytorch","url":"https://github.com/lucidrains/CoCa-pytorch"},{"title":"lucidrains/PaLM-pytorch","url":"https://github.com/lucidrains/PaLM-pytorch"},{"title":"google/paxml","url":"https://github.com/google/paxml"},{"title":"foundation-model-stack/fms-fsdp","url":"https://github.com/foundation-model-stack/fms-fsdp"},{"title":"lucidrains/PaLM-jax","url":"https://github.com/lucidrains/PaLM-jax"},{"title":"chrisociepa/allamo","url":"https://github.com/chrisociepa/allamo"},{"title":"conceptofmind/PaLM-flax","url":"https://github.com/conceptofmind/PaLM-flax"}],"syntology":{"n":37,"n_ran":30,"n_unverified":7,"n_pointer_only":0}}},{"leaderboard":"/sota/coreference-resolution-on-ontonotes","slug":"coreference-resolution-on-ontonotes","dataset":"OntoNotes","dataset_url":"/dataset/ontonotes-5-0","rows_in_archive":26,"metrics":["F1"],"first_row_in_archive_order":{"model":"Maverick_mes","paper_title":"Maverick: Efficient and Accurate Coreference Resolution Defying Recent Trends","paper_url":"/paper/2407-21489","paper_date":"2024-07-31","arxiv_id":"2407.21489","code_links":[{"title":"sapienzanlp/maverick-coref","url":"https://github.com/sapienzanlp/maverick-coref"}],"syntology":{"n":11,"n_ran":5,"n_unverified":6,"n_pointer_only":11}}},{"leaderboard":"/sota/coreference-resolution-on-conll-2012","slug":"coreference-resolution-on-conll-2012","dataset":"CoNLL 2012","dataset_url":"/dataset/conll-1","rows_in_archive":18,"metrics":["Avg F1"],"first_row_in_archive_order":{"model":"Maverick_mes","paper_title":"Maverick: Efficient and Accurate Coreference Resolution Defying Recent Trends","paper_url":"/paper/2407-21489","paper_date":"2024-07-31","arxiv_id":"2407.21489","code_links":[{"title":"sapienzanlp/maverick-coref","url":"https://github.com/sapienzanlp/maverick-coref"}],"syntology":{"n":11,"n_ran":5,"n_unverified":6,"n_pointer_only":11}}},{"leaderboard":"/sota/coreference-resolution-on-gap-1","slug":"coreference-resolution-on-gap-1","dataset":"GAP","dataset_url":"/dataset/gap","rows_in_archive":5,"metrics":["Overall F1","Masculine F1 (M)","Feminine F1 (F)","Bias (F/M)","F1"],"first_row_in_archive_order":{"model":"Coref-MTL","paper_title":null,"paper_url":null,"paper_date":"","arxiv_id":null,"code_links":[],"syntology":null}},{"leaderboard":"/sota/coreference-resolution-on-dwie","slug":"coreference-resolution-on-dwie","dataset":"DWIE","dataset_url":"/dataset/dwie","rows_in_archive":3,"metrics":["Avg. F1"],"first_row_in_archive_order":{"model":"REXEL","paper_title":"REXEL: An End-to-end Model for Document-Level Relation Extraction and Entity Linking","paper_url":"/paper/rexel-an-end-to-end-model-for-document-level","paper_date":"2024-04-19","arxiv_id":"2404.12788","code_links":[{"title":"amazon-science/e2e-docie","url":"https://github.com/amazon-science/e2e-docie"}],"syntology":{"n":3,"n_ran":3,"n_unverified":0,"n_pointer_only":0}}},{"leaderboard":"/sota/coreference-resolution-on-wikicoref","slug":"coreference-resolution-on-wikicoref","dataset":"WikiCoref","dataset_url":"/dataset/wikicoref","rows_in_archive":3,"metrics":["F1"],"first_row_in_archive_order":{"model":"Maverick_mes","paper_title":"Maverick: Efficient and Accurate Coreference Resolution Defying Recent Trends","paper_url":"/paper/2407-21489","paper_date":"2024-07-31","arxiv_id":"2407.21489","code_links":[{"title":"sapienzanlp/maverick-coref","url":"https://github.com/sapienzanlp/maverick-coref"}],"syntology":{"n":11,"n_ran":5,"n_unverified":6,"n_pointer_only":11}}},{"leaderboard":"/sota/coreference-resolution-on-conll12","slug":"coreference-resolution-on-conll12","dataset":"CoNLL12","dataset_url":"/dataset/conll-1","rows_in_archive":2,"metrics":["Average F1","B3","CEAFϕ4","MUC"],"first_row_in_archive_order":{"model":"DeepStruct multi-task w/ finetune","paper_title":"DeepStruct: Pretraining of Language Models for Structure Prediction","paper_url":"/paper/deepstruct-pretraining-of-language-models-for-1","paper_date":"2022-05-21","arxiv_id":"2205.10475","code_links":[{"title":"cgraywang/deepstruct","url":"https://github.com/cgraywang/deepstruct"}],"syntology":{"n":13,"n_ran":7,"n_unverified":6,"n_pointer_only":0}}},{"leaderboard":"/sota/coreference-resolution-on-litbank","slug":"coreference-resolution-on-litbank","dataset":"LitBank","dataset_url":"/dataset/litbank","rows_in_archive":2,"metrics":["Avg F1","F1"],"first_row_in_archive_order":{"model":"Maverick_incr","paper_title":"Maverick: Efficient and Accurate Coreference Resolution Defying Recent Trends","paper_url":"/paper/2407-21489","paper_date":"2024-07-31","arxiv_id":"2407.21489","code_links":[{"title":"sapienzanlp/maverick-coref","url":"https://github.com/sapienzanlp/maverick-coref"}],"syntology":{"n":11,"n_ran":5,"n_unverified":6,"n_pointer_only":11}}},{"leaderboard":"/sota/coreference-resolution-on-ontogum","slug":"coreference-resolution-on-ontogum","dataset":"OntoGUM","dataset_url":"/dataset/ontogum","rows_in_archive":2,"metrics":["Avg F1"],"first_row_in_archive_order":{"model":"MTL-coref","paper_title":"Incorporating Singletons and Mention-based Features in Coreference Resolution via Multi-task Learning for Better Generalization","paper_url":"/paper/incorporating-singletons-and-mention-based","paper_date":"2023-09-20","arxiv_id":"2309.11582","code_links":[{"title":"yilunzhu/coref-mtl","url":"https://github.com/yilunzhu/coref-mtl"}],"syntology":null}},{"leaderboard":"/sota/coreference-resolution-on-preco","slug":"coreference-resolution-on-preco","dataset":"PreCo","dataset_url":"/dataset/preco","rows_in_archive":2,"metrics":["F1"],"first_row_in_archive_order":{"model":"Maverick_incr","paper_title":"Maverick: Efficient and Accurate Coreference Resolution Defying Recent Trends","paper_url":"/paper/2407-21489","paper_date":"2024-07-31","arxiv_id":"2407.21489","code_links":[{"title":"sapienzanlp/maverick-coref","url":"https://github.com/sapienzanlp/maverick-coref"}],"syntology":{"n":11,"n_ran":5,"n_unverified":6,"n_pointer_only":11}}},{"leaderboard":"/sota/coreference-resolution-on-stm-coref","slug":"coreference-resolution-on-stm-coref","dataset":"STM-coref","dataset_url":null,"rows_in_archive":2,"metrics":["CoNLL F1"],"first_row_in_archive_order":{"model":"BFCR + SpanBERT + Transfer Learning","paper_title":"Coreference Resolution in Research Papers from Multiple Domains","paper_url":"/paper/coreference-resolution-in-research-papers","paper_date":"2021-01-04","arxiv_id":"2101.00884","code_links":[{"title":"arthurbra/stm-coref","url":"https://github.com/arthurbra/stm-coref"}],"syntology":null}},{"leaderboard":"/sota/coreference-resolution-on-xwinograd-en","slug":"coreference-resolution-on-xwinograd-en","dataset":"XWinograd EN","dataset_url":null,"rows_in_archive":2,"metrics":["Accuracy"],"first_row_in_archive_order":{"model":"mT0-13B","paper_title":"Crosslingual Generalization through Multitask Finetuning","paper_url":"/paper/crosslingual-generalization-through-multitask","paper_date":"2022-11-03","arxiv_id":"2211.01786","code_links":[{"title":"bigscience-workshop/xmtf","url":"https://github.com/bigscience-workshop/xmtf"}],"syntology":{"n":4,"n_ran":1,"n_unverified":3,"n_pointer_only":0}}},{"leaderboard":"/sota/coreference-resolution-on-xwinograd-fr","slug":"coreference-resolution-on-xwinograd-fr","dataset":"XWinograd FR","dataset_url":null,"rows_in_archive":2,"metrics":["Accuracy"],"first_row_in_archive_order":{"model":"mT0-13B","paper_title":"Crosslingual Generalization through Multitask Finetuning","paper_url":"/paper/crosslingual-generalization-through-multitask","paper_date":"2022-11-03","arxiv_id":"2211.01786","code_links":[{"title":"bigscience-workshop/xmtf","url":"https://github.com/bigscience-workshop/xmtf"}],"syntology":{"n":4,"n_ran":1,"n_unverified":3,"n_pointer_only":0}}},{"leaderboard":"/sota/coreference-resolution-on-docred-ie","slug":"coreference-resolution-on-docred-ie","dataset":"DocRED-IE","dataset_url":"/dataset/docred-ie","rows_in_archive":1,"metrics":["Avg F1"],"first_row_in_archive_order":{"model":"REXEL","paper_title":"REXEL: An End-to-end Model for Document-Level Relation Extraction and Entity Linking","paper_url":"/paper/rexel-an-end-to-end-model-for-document-level","paper_date":"2024-04-19","arxiv_id":"2404.12788","code_links":[{"title":"amazon-science/e2e-docie","url":"https://github.com/amazon-science/e2e-docie"}],"syntology":{"n":3,"n_ran":3,"n_unverified":0,"n_pointer_only":0}}},{"leaderboard":"/sota/coreference-resolution-on-quizbowl","slug":"coreference-resolution-on-quizbowl","dataset":"Quizbowl","dataset_url":"/dataset/quizbowl","rows_in_archive":1,"metrics":["F1"],"first_row_in_archive_order":{"model":"longdoc S (OntoNotes + PreCo + LitBank)","paper_title":"On Generalization in Coreference Resolution","paper_url":"/paper/on-generalization-in-coreference-resolution","paper_date":"2021-09-20","arxiv_id":"2109.09667","code_links":[{"title":"shtoshni/fast-coref","url":"https://github.com/shtoshni/fast-coref"},{"title":"shtoshni92/fast-coref","url":"https://github.com/shtoshni92/fast-coref"}],"syntology":null}},{"leaderboard":"/sota/coreference-resolution-on-the-arrau-corpus","slug":"coreference-resolution-on-the-arrau-corpus","dataset":"The ARRAU Corpus","dataset_url":null,"rows_in_archive":1,"metrics":["Avg F1"],"first_row_in_archive_order":{"model":"dali-full-anaphora","paper_title":"A Cluster Ranking Model for Full Anaphora Resolution","paper_url":"/paper/a-cluster-ranking-model-for-full-anaphora","paper_date":"2019-11-21","arxiv_id":"1911.09532","code_links":[{"title":"juntaoy/dali-full-anaphora","url":"https://github.com/juntaoy/dali-full-anaphora"}],"syntology":null}}],"datasets":[{"url":"/dataset/wsc","name":"WSC","full_name":"Winograd Schema Challenge","num_papers_in_archive":361},{"url":"/dataset/ontonotes-5-0","name":"OntoNotes 5.0","full_name":"","num_papers_in_archive":254},{"url":"/dataset/conll-1","name":"CoNLL","full_name":"","num_papers_in_archive":187},{"url":"/dataset/winobias","name":"WinoBias","full_name":"WinoBias","num_papers_in_archive":134},{"url":"/dataset/gap-coreference-dataset","name":"GAP Coreference Dataset","full_name":"","num_papers_in_archive":104},{"url":"/dataset/conll-2012-1","name":"CoNLL-2012","full_name":"","num_papers_in_archive":89},{"url":"/dataset/ecb","name":"ECB+","full_name":"extension to the EventCorefBank","num_papers_in_archive":76},{"url":"/dataset/gap","name":"GAP","full_name":"GAP Benchmark Suite","num_papers_in_archive":60},{"url":"/dataset/quoref","name":"Quoref","full_name":"Quoref","num_papers_in_archive":50},{"url":"/dataset/english-web-treebank","name":"English Web Treebank","full_name":"English Web Treebank","num_papers_in_archive":42},{"url":"/dataset/xp3","name":"xP3","full_name":"","num_papers_in_archive":34},{"url":"/dataset/wikicoref","name":"WikiCoref","full_name":"","num_papers_in_archive":28},{"url":"/dataset/litbank","name":"LitBank","full_name":"LitBank","num_papers_in_archive":23},{"url":"/dataset/preco","name":"PreCo","full_name":"","num_papers_in_archive":20},{"url":"/dataset/dwie","name":"DWIE","full_name":"Deutsche Welle corpus for Information Extraction","num_papers_in_archive":18},{"url":"/dataset/map","name":"MAP","full_name":"Maybe Ambiguous Pronoun","num_papers_in_archive":14},{"url":"/dataset/gum","name":"GUM","full_name":"Georgetown University Multilayer corpus","num_papers_in_archive":13},{"url":"/dataset/parcorfull","name":"ParCorFull","full_name":"Parallel Corpus Annotated with Full Coreference","num_papers_in_archive":11},{"url":"/dataset/photobook","name":"PhotoBook","full_name":"","num_papers_in_archive":11},{"url":"/dataset/clevr-dialog","name":"CLEVR-Dialog","full_name":"CLEVR-Dialog","num_papers_in_archive":10},{"url":"/dataset/ontogum","name":"OntoGUM","full_name":"","num_papers_in_archive":9},{"url":"/dataset/quizbowl","name":"Quizbowl","full_name":"","num_papers_in_archive":8},{"url":"/dataset/winogender-schemas","name":"Winogender Schemas","full_name":"","num_papers_in_archive":8},{"url":"/dataset/definite-pronoun-resolution-dataset","name":"Definite Pronoun Resolution Dataset","full_name":"","num_papers_in_archive":7},{"url":"/dataset/bipar","name":"BiPaR","full_name":"BiPaR","num_papers_in_archive":6},{"url":"/dataset/gicoref","name":"GICoref","full_name":"Gender Inclusive Coreference","num_papers_in_archive":6},{"url":"/dataset/vispro","name":"VisPro","full_name":"VisPro","num_papers_in_archive":6},{"url":"/dataset/wikicrem","name":"WikiCREM","full_name":"WikiCREM","num_papers_in_archive":6},{"url":"/dataset/amalgum","name":"AMALGUM","full_name":"A Machine Annotated Lookalike of GUM","num_papers_in_archive":5},{"url":"/dataset/sp-10k","name":"SP-10K","full_name":"","num_papers_in_archive":5},{"url":"/dataset/a-game-of-sorts","name":"A Game Of Sorts","full_name":"","num_papers_in_archive":4},{"url":"/dataset/xwino","name":"XWINO","full_name":"","num_papers_in_archive":4},{"url":"/dataset/multireqa","name":"MultiReQA","full_name":null,"num_papers_in_archive":2},{"url":"/dataset/poc","name":"PoC","full_name":"Points of correspondence","num_papers_in_archive":2},{"url":"/dataset/comet","name":"Comet","full_name":"","num_papers_in_archive":1},{"url":"/dataset/coresearch","name":"CoreSearch","full_name":"","num_papers_in_archive":1},{"url":"/dataset/docred-ie","name":"DocRED-IE","full_name":"","num_papers_in_archive":1},{"url":"/dataset/marmara-turkish-coreference-resolution-corpus","name":"Marmara Turkish Coreference Resolution Corpus","full_name":"","num_papers_in_archive":1},{"url":"/dataset/masc","name":"MASC","full_name":"Manually Annotated Sub-Corpus","num_papers_in_archive":1},{"url":"/dataset/mudoco-queryrewrite","name":"MuDoCo_QueryRewrite","full_name":"The MuDoCo dataset with Query Rewrite Annotations","num_papers_in_archive":1},{"url":"/dataset/scico","name":"SciCo","full_name":"Scientific Concept Induction Corpus","num_papers_in_archive":1},{"url":"/dataset/winonb","name":"WinoNB","full_name":"","num_papers_in_archive":1},{"url":"/dataset/winopron","name":"WinoPron","full_name":"","num_papers_in_archive":1},{"url":"/dataset/contracat","name":"ContraCAT","full_name":"Contrastive Coreference Analytical Templates (for Machine Translation)","num_papers_in_archive":0},{"url":"/dataset/popcorn","name":"POPCORN","full_name":"POPCORN: Fictional and Synthetic Intelligence Reports for Named Entity Recognition and Relation Extraction Tasks","num_papers_in_archive":0}],"subtasks":[{"url":"/task/coreference-resolution-1","name":"coreference-resolution"},{"url":"/task/cross-document-coreference-resolution","name":"Cross Document Coreference Resolution"}],"parent_tasks":[],"papers":{"order":"repositories listed in the archive (desc), then date (desc); the archive holds no stars","population":"papers tagged with this task that list at least one repository in the archive","shown":30,"of":288,"tagged_in_all":880,"items":[{"url":"/paper/attention-is-all-you-need","title":"Attention Is All You Need","date":"2017-06-12","arxiv_id":"1706.03762","repositories_listed":595,"syntology":{"n":946,"n_ran":600,"n_unverified":346,"n_pointer_only":451}},{"url":"/paper/bert-pre-training-of-deep-bidirectional","title":"BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding","date":"2018-10-11","arxiv_id":"1810.04805","repositories_listed":534,"syntology":{"n":659,"n_ran":204,"n_unverified":455,"n_pointer_only":149}},{"url":"/paper/language-models-are-few-shot-learners","title":"Language Models are Few-Shot Learners","date":"2020-05-28","arxiv_id":"2005.14165","repositories_listed":67,"syntology":{"n":65,"n_ran":15,"n_unverified":50,"n_pointer_only":4}},{"url":"/paper/exploring-the-limits-of-transfer-learning","title":"Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer","date":"2019-10-23","arxiv_id":"1910.10683","repositories_listed":57,"syntology":{"n":31,"n_ran":2,"n_unverified":29,"n_pointer_only":0}},{"url":"/paper/deep-contextualized-word-representations","title":"Deep contextualized word representations","date":"2018-02-15","arxiv_id":"1802.05365","repositories_listed":46,"syntology":{"n":58,"n_ran":23,"n_unverified":35,"n_pointer_only":25}},{"url":"/paper/language-models-are-unsupervised-multitask","title":"Language Models are Unsupervised Multitask Learners","date":"2019-02-14","arxiv_id":null,"repositories_listed":21,"syntology":null},{"url":"/paper/deberta-decoding-enhanced-bert-with","title":"DeBERTa: Decoding-enhanced BERT with Disentangled Attention","date":"2020-06-05","arxiv_id":"2006.03654","repositories_listed":14,"syntology":{"n":13,"n_ran":4,"n_unverified":9,"n_pointer_only":3}},{"url":"/paper/winogrande-an-adversarial-winograd-schema","title":"WinoGrande: An Adversarial Winograd Schema Challenge at Scale","date":"2019-07-24","arxiv_id":"1907.10641","repositories_listed":10,"syntology":null},{"url":"/paper/scaling-instruction-finetuned-language-models","title":"Scaling Instruction-Finetuned Language Models","date":"2022-10-20","arxiv_id":"2210.11416","repositories_listed":9,"syntology":{"n":17,"n_ran":8,"n_unverified":9,"n_pointer_only":2}},{"url":"/paper/finetuned-language-models-are-zero-shot","title":"Finetuned Language Models Are Zero-Shot Learners","date":"2021-09-03","arxiv_id":"2109.01652","repositories_listed":8,"syntology":{"n":1,"n_ran":0,"n_unverified":1,"n_pointer_only":0}},{"url":"/paper/palm-scaling-language-modeling-with-pathways-1","title":"PaLM: Scaling Language Modeling with Pathways","date":"2022-04-05","arxiv_id":"2204.02311","repositories_listed":7,"syntology":{"n":37,"n_ran":30,"n_unverified":7,"n_pointer_only":0}},{"url":"/paper/spanbert-improving-pre-training-by","title":"SpanBERT: Improving Pre-training by Representing and Predicting Spans","date":"2019-07-24","arxiv_id":"1907.10529","repositories_listed":6,"syntology":{"n":15,"n_ran":3,"n_unverified":12,"n_pointer_only":4}},{"url":"/paper/sdnet-contextualized-attention-based-deep","title":"SDNet: Contextualized Attention-based Deep Network for Conversational Question Answering","date":"2018-12-10","arxiv_id":"1812.03593","repositories_listed":6,"syntology":null},{"url":"/paper/stanza-a-python-natural-language-processing","title":"Stanza: A Python Natural Language Processing Toolkit for Many Human Languages","date":"2020-03-16","arxiv_id":"2003.07082","repositories_listed":5,"syntology":{"n":30,"n_ran":16,"n_unverified":14,"n_pointer_only":28}},{"url":"/paper/multi-task-identification-of-entities","title":"Multi-Task Identification of Entities, Relations, and Coreference for Scientific Knowledge Graph Construction","date":"2018-08-29","arxiv_id":"1808.09602","repositories_listed":5,"syntology":null},{"url":"/paper/higher-order-coreference-resolution-with","title":"Higher-order Coreference Resolution with Coarse-to-fine Inference","date":"2018-04-15","arxiv_id":"1804.05392","repositories_listed":5,"syntology":{"n":23,"n_ran":2,"n_unverified":21,"n_pointer_only":0}},{"url":"/paper/pythia-a-suite-for-analyzing-large-language","title":"Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling","date":"2023-04-03","arxiv_id":"2304.01373","repositories_listed":4,"syntology":null},{"url":"/paper/mind-the-gap-a-balanced-corpus-of-gendered","title":"Mind the GAP: A Balanced Corpus of Gendered Ambiguous Pronouns","date":"2018-10-11","arxiv_id":"1810.05201","repositories_listed":4,"syntology":null},{"url":"/paper/gender-bias-in-coreference-resolution","title":"Gender Bias in Coreference Resolution","date":"2018-04-25","arxiv_id":"1804.09301","repositories_listed":4,"syntology":null},{"url":"/paper/gender-bias-in-coreference-resolution-1","title":"Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods","date":"2018-04-18","arxiv_id":"1804.06876","repositories_listed":4,"syntology":null},{"url":"/paper/end-to-end-neural-coreference-resolution","title":"End-to-end Neural Coreference Resolution","date":"2017-07-21","arxiv_id":"1707.07045","repositories_listed":4,"syntology":{"n":3,"n_ran":3,"n_unverified":0,"n_pointer_only":3}},{"url":"/paper/hungry-hungry-hippos-towards-language","title":"Hungry Hungry Hippos: Towards Language Modeling with State Space Models","date":"2022-12-28","arxiv_id":"2212.14052","repositories_listed":3,"syntology":{"n":15,"n_ran":7,"n_unverified":8,"n_pointer_only":0}},{"url":"/paper/ask-me-anything-a-simple-strategy-for","title":"Ask Me Anything: A simple strategy for prompting language models","date":"2022-10-05","arxiv_id":"2210.02441","repositories_listed":3,"syntology":{"n":2,"n_ran":2,"n_unverified":0,"n_pointer_only":0}},{"url":"/paper/designing-effective-sparse-expert-models","title":"ST-MoE: Designing Stable and Transferable Sparse Expert Models","date":"2022-02-17","arxiv_id":"2202.08906","repositories_listed":3,"syntology":{"n":5,"n_ran":5,"n_unverified":0,"n_pointer_only":5}},{"url":"/paper/the-multiberts-bert-reproductions-for","title":"The MultiBERTs: BERT Reproductions for Robustness Analysis","date":"2021-06-30","arxiv_id":"2106.16163","repositories_listed":3,"syntology":{"n":2,"n_ran":2,"n_unverified":0,"n_pointer_only":0}},{"url":"/paper/learning-to-ignore-long-document-coreference","title":"Learning to Ignore: Long Document Coreference with Bounded Memory Neural Networks","date":"2020-10-06","arxiv_id":"2010.02807","repositories_listed":3,"syntology":null},{"url":"/paper/an-annotated-dataset-of-coreference-in","title":"An Annotated Dataset of Coreference in English Literature","date":"2019-12-03","arxiv_id":"1912.01140","repositories_listed":3,"syntology":null},{"url":"/paper/multi-hop-question-answering-via-reasoning","title":"Multi-hop Question Answering via Reasoning Chains","date":"2019-10-07","arxiv_id":"1910.02610","repositories_listed":3,"syntology":null},{"url":"/paper/a-hybrid-neural-network-model-for-commonsense","title":"A Hybrid Neural Network Model for Commonsense Reasoning","date":"2019-07-27","arxiv_id":"1907.11983","repositories_listed":3,"syntology":{"n":1,"n_ran":0,"n_unverified":1,"n_pointer_only":0}},{"url":"/paper/huixiangdou-cr-coreference-resolution-in","title":"Labeling supervised fine-tuning data with the scaling law","date":"2024-05-05","arxiv_id":"2405.02817","repositories_listed":2,"syntology":null}],"syntology_records":18,"syntology_note":"a paper without a record is not a recorded non-run: it may lack an arXiv id or simply be absent from the graph layer"},"description_links":{"kept":0,"unwrapped_to_text":0,"bare_urls_linked":0,"relative_images_dropped":0,"rule":"internal links are kept only when the target slug exists in the catalog"},"syntology":{"read_at":"2026-09-24T18:15:14+00:00","claim":"Per-sample execution status on synthesized fixtures ('ran N of M samples'); not a correctness claim and not a ranking signal.","status_vocabulary":{"ran_honours":"ran, honoured the contract we drafted","ran_violates":"ran, violated the contract we drafted","ran_draft_wrong":"ran; our contract draft was wrong, not the code","ran_fixture":"ran; our fixture could not drive it","ran":"ran on a synthesized input","unverified":"unverified (harvested, no recorded run)"}},"not_shown":{"libraries":"the archive has no per-task library table","trend_sparklines":"the Trend column of the benchmarks table was a rendered image; it is not in the archive","social_and_latest_sorts":"stars and social signals are not in the archive"}}