Browse State-of-the-Art › Selection bias
Selection bias
143 papers with code · 0 benchmarks · 2 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
2 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 143 papers with code (365 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
24 Mar 2020 8 repositories listed Syntology ran 4 of 40 samples · 36 unverified · 1 pointer-only (licence)The emerging area of computational pathology (CPath) is ripe ground for the application of deep learning (DL) methods to healthcare due to the sheer volume of raw pixel data in whole-slide images (WSIs) of cancerous…
-
21 Apr 2018 6 repositories listedTo the best of our knowledge, this is the first public dataset which contains samples with sequential dependence of click and conversion labels for CVR modeling.
-
1 Nov 2020 4 repositories listedMost existing works focus on worst-case or average-case lower bounds for the number of interventions required to orient a DAG.
-
26 Jun 2019 3 repositories listedBased on the original definition of MDI by Breiman et al.
-
13 Dec 2024 2 repositories listed Syntology ran 2 of 7 samples · 5 unverifiedWe propose EVOlutionary Selector (EVOS), an efficient training paradigm for accelerating Implicit Neural Representation (INR).
-
16 Sep 2023 2 repositories listedLLMs are increasingly powerful and widely used to assist users in a variety of tasks.
-
3 Apr 2023 2 repositories listedTherefore, the degree of overfitting for clutter reflects the non-causality of deep learning in SAR ATR.
-
31 Jan 2023 2 repositories listedWe investigate the mathematical capabilities of two iterations of ChatGPT (released 9-January-2023 and 30-January-2023) and of GPT-4 by testing them on publicly available datasets, as well as hand-crafted ones, using a…
-
30 Oct 2022 2 repositories listed Syntology ran 6 of 12 samples · 6 unverified · 6 pointer-only (licence)In this paper, we find that this problem usually occurs when the positions of support samples are in the vicinity of task centroid -- the mean of all class centroids in the task.
-
30 Sep 2022 2 repositories listedModern language modeling tasks are often underspecified: for a given token prediction, many words may satisfy the user's intent of producing natural language at inference time, however only one word will minimize the…
-
26 Apr 2021 2 repositories listedAlgorithms make a growing portion of policy and business decisions.
-
15 Apr 2021 2 repositories listed Syntology ran 1 of 8 samples · 7 unverifiedTo overcome the above-mentioned disfluencies, we propose All Unmasked Likelihood (AUL), a bias evaluation measure that predicts all tokens in a test case given the MLM embedding of the unmasked input.
-
2 Dec 2019 2 repositories listedTo address these drawbacks, we formalize a method for automating the selection of interesting PDPs and extend PDPs beyond showing single features to show the model response along arbitrary directions, for example in raw…
-
15 Jul 2019 2 repositories listedAt the moment, two methodologies for dealing with bias prevail in the field of LTR: counterfactual methods that learn from historical data and model user behavior to deal with biases; and online methods that perform…
-
15 May 2019 2 repositories listedNatural Language Sentence Matching (NLSM) has gained substantial attention from both academics and the industry, and rich public datasets contribute a lot to this process.
-
30 Apr 2019 2 repositories listed Syntology ran 1 of 23 samples · 22 unverifiedThe challenges for this problem are two-fold: on the one hand, we have to derive a causal estimator to estimate the causal quantity from observational data, where there exists confounding bias; on the other hand, we…
-
20 Jun 2025 1 repository listedLarge language models (LLMs) are increasingly used to generate feedback, yet their impact on learning remains underexplored, especially compared to existing feedback methods.
-
9 Jun 2025 1 repository listedIn this paper, we release this assumption and propose a learning algorithm based on likelihood maximization to learn a prediction model.
-
4 Jun 2025 1 repository listedDisaggregated evaluation across subgroups is critical for assessing the fairness of machine learning models, but its uncritical use can mislead practitioners.
-
29 Apr 2025 1 repository listedTo address the former, conventional techniques that require detailed knowledge in the form of causal graphs have been proposed.
-
14 Apr 2025 1 repository listedMultimodal representation learning, exemplified by multimodal contrastive learning (MMCL) using image-text pairs, aims to learn powerful representations by aligning cues across modalities.
-
10 Mar 2025 1 repository listed Syntology ran 0 of 1 samples · 1 unverifiedEven when recognized, the existing paradigm for interventional causal discovery still fails to address it.
-
27 Nov 2024 1 repository listedSelection bias poses a critical challenge for fairness in machine learning, as models trained on data that is less representative of the population might exhibit undesirable behavior for underrepresented profiles.
-
29 Oct 2024 1 repository listedExperimental results reveal linguistic inequalities: 1) high-resource languages stand out in Monolingual Knowledge Extraction; 2) Indo-European languages lead RALMs to provide answers directly from documents,…
-
28 Oct 2024 1 repository listed Syntology ran 0 of 2 samples · 2 unverified · 2 pointer-only (licence)Industrial recommendation systems (RS) rely on the multi-stage pipeline to balance effectiveness and efficiency when delivering items from a vast corpus to users.
-
24 Oct 2024 1 repository listedTo assess the comparative performance of HRF against other widely adopted ensemble methods, we conducted tests on 52 datasets, comprising both real-world and synthetic data.
-
21 Oct 2024 1 repository listedDiverging from previous methods focused on each local pixel value, the CORAL-Correlation Consistency Module (CCM) in the CORN leverages second-order statistical information to capture global structural features by…
-
20 Oct 2024 1 repository listedThe use of large language models (LLMs) as automated evaluation tools to assess the quality of generated natural language, known as LLMs-as-Judges, has demonstrated promising capabilities and is rapidly gaining…
-
11 Oct 2024 1 repository listedIn this paper, we propose a novel causal diffusion model called DiffPO, which is carefully designed for reliable inferences in medicine by learning the distribution of potential outcomes.
-
30 Sep 2024 1 repository listedFairness in machine learning seeks to mitigate model bias against individuals based on sensitive features such as sex or age, often caused by an uneven representation of the population in the training data due to…
Syntology lines on 7 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections