Browse State-of-the-Art › Mixture-of-Experts › Papers, page 9
Mixture-of-Experts
Papers archive 2025-07-28
archive papers tagged: 1,312 · with a code link: 516 · where Syntology ran a sample: 216 (184 with a run with no instrument failure, 32 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (216 of 1,312 tagged: 184 with a run with no instrument failure, 32 where every run was a failure of Syntology's instrument)
Page 9 of 14: papers 801 to 900 of 1,312, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Yi-Lightning Technical Report2 Dec 2024 0 repositories listed
-
HiMoE: Heterogeneity-Informed Mixture-of-Experts for Fair Spatial-Temporal Forecasting30 Nov 2024 0 repositories listed
-
Mixture of Experts for Node Classification30 Nov 2024 0 repositories listed
-
MQFL-FHE: Multimodal Quantum Federated Learning Framework with Fully Homomorphic Encryption30 Nov 2024 0 repositories listed
-
LaVIDE: A Language-Vision Discriminator for Detecting Changes in Satellite Image with Map References29 Nov 2024 0 repositories listed
-
On the effectiveness of discrete representations in sparse mixture of experts28 Nov 2024 0 repositories listed
-
27 Nov 2024 0 repositories listed
-
Mixture of Cache-Conditional Experts for Efficient Mobile Device Inference27 Nov 2024 0 repositories listed
-
Mixture of Experts in Image Classification: What's the Sweet Spot?27 Nov 2024 0 repositories listed
-
UOE: Unlearning One Expert Is Enough For Mixture-of-experts LLMS27 Nov 2024 0 repositories listed
-
Enhancing Code-Switching ASR Leveraging Non-Peaky CTC Loss and Deep Language Posterior Injection26 Nov 2024 0 repositories listed
-
LDACP: Long-Delayed Ad Conversions Prediction Model for Bidding Strategy25 Nov 2024 0 repositories listed
-
MH-MoE: Multi-Head Mixture-of-Experts25 Nov 2024 0 repositories listed
-
Lifelong Knowledge Editing for Vision Language Models with Low-Rank Mixture-of-Experts23 Nov 2024 0 repositories listed
-
KAAE: Numerical Reasoning for Knowledge Graphs via Knowledge-aware Attributes Learning20 Nov 2024 0 repositories listed
-
MERLOT: A Distilled LLM-based Mixture-of-Experts Framework for Scalable Encrypted Traffic Classification20 Nov 2024 0 repositories listed
-
Ultra-Sparse Memory Network19 Nov 2024 0 repositories listed
-
MoE-Lightning: High-Throughput MoE Inference on Memory-constrained GPUs18 Nov 2024 0 repositories listed
-
Lynx: Enabling Efficient MoE Inference through Dynamic Batch-Aware Expert Selection13 Nov 2024 0 repositories listed
-
Sparse Upcycling: Inference Inefficient Finetuning13 Nov 2024 0 repositories listed
-
Imitation Learning from Observations: An Autoregressive Mixture of Experts Approach12 Nov 2024 0 repositories listed
-
PERFT: Parameter-Efficient Routed Fine-Tuning for Mixture-of-Expert Model12 Nov 2024 0 repositories listed
-
Towards Vision Mixture of Experts for Wildlife Monitoring on the Edge12 Nov 2024 0 repositories listed
-
Adaptive Conditional Expert Selection Network for Multi-domain Recommendation11 Nov 2024 0 repositories listed
-
WDMoE: Wireless Distributed Mixture of Experts for Large Language Models11 Nov 2024 0 repositories listed
-
NeKo: Toward Post Recognition Generative Correction Large Language Models with Task-Oriented Experts8 Nov 2024 0 repositories listed
-
Advancing Robust Underwater Acoustic Target Recognition through Multi-task Learning and Multi-Gate Mixture-of-Experts5 Nov 2024 0 repositories listed
-
FedMoE-DA: Federated Mixture of Experts via Domain Aware Fine-grained Aggregation4 Nov 2024 0 repositories listed
-
Facet-Aware Multi-Head Mixture-of-Experts Model for Sequential Recommendation3 Nov 2024 0 repositories listed
-
HOBBIT: A Mixed Precision Expert Offloading System for Fast MoE Inference3 Nov 2024 0 repositories listed
-
RS-MoE: Mixture of Experts for Remote Sensing Image Captioning and Visual Question Answering3 Nov 2024 0 repositories listed
-
PMoL: Parameter Efficient MoE for Preference Mixing of LLM Alignment2 Nov 2024 0 repositories listed
-
MoE-I²: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition1 Nov 2024 0 repositories listed
-
Stereo-Talker: Audio-driven 3D Human Synthesis with Prior-Guided Mixture-of-Experts31 Oct 2024 0 repositories listed
-
MALoRA: Mixture of Asymmetric Low-Rank Adaptation for Enhanced Multi-Task Learning30 Oct 2024 0 repositories listed
-
Stealing User Prompts from Mixture of Experts30 Oct 2024 0 repositories listed
-
Efficient and Effective Weight-Ensembling Mixture of Experts for Multi-Task Model Merging29 Oct 2024 0 repositories listed
-
Neural Experts: Mixture of Experts for Implicit Neural Representations29 Oct 2024 0 repositories listed
-
ProMoE: Fast MoE-based LLM Serving using Proactive Caching29 Oct 2024 0 repositories listed
-
Efficient Mixture-of-Expert for Video-based Driver State and Physiological Multi-task Estimation in Conditional Autonomous Driving28 Oct 2024 0 repositories listed
-
FinTeamExperts: Role Specialized MOEs For Financial Analysis28 Oct 2024 0 repositories listed
-
Mixture of Parrots: Experts improve memorization more than reasoning24 Oct 2024 0 repositories listed
-
MoMQ: Mixture-of-Experts Enhances Multi-Dialect Query Generation across Relational and Non-Relational Databases24 Oct 2024 0 repositories listed
-
ExpertFlow: Optimized Expert Activation and Token Allocation for Efficient Mixture-of-Experts Inference23 Oct 2024 0 repositories listed
-
Faster Language Models with Better Multi-Token Prediction Using Tensor Decomposition23 Oct 2024 0 repositories listed
-
MiLoRA: Efficient Mixture of Low-Rank Adaptation for Large Language Models Fine-tuning23 Oct 2024 0 repositories listed
-
Robust and Explainable Depression Identification from Speech Using Vowel-Based Ensemble Learning Approaches23 Oct 2024 0 repositories listed
-
Optimizing Mixture-of-Experts Inference Time Combining Model Deployment and Communication Scheduling22 Oct 2024 0 repositories listed
-
ViMoE: An Empirical Study of Designing Vision Mixture-of-Experts21 Oct 2024 0 repositories listed
-
MENTOR: Mixture-of-Experts Network with Task-Oriented Perturbation for Visual Reinforcement Learning19 Oct 2024 0 repositories listed
-
Enhancing Generalization in Sparse Mixture of Experts Models: The Case for Increased Expert Activation in Compositional Tasks17 Oct 2024 0 repositories listed
-
EPS-MoE: Expert Pipeline Scheduler for Cost-Efficient MoE Inference16 Oct 2024 0 repositories listed
-
On the Risk of Evidence Pollution for Malicious Social Text Detection in the Era of LLMs16 Oct 2024 0 repositories listed
-
Understanding Expert Structures on Minimax Parameter Estimation in Contaminated Mixture of Experts16 Oct 2024 0 repositories listed
-
MoE-Pruner: Pruning Mixture-of-Experts Large Language Model using the Hints from Its Router15 Oct 2024 0 repositories listed
-
Quadratic Gating Functions in Mixture of Experts: A Statistical Insight15 Oct 2024 0 repositories listed
-
Transformer Layer Injection: A Novel Approach for Efficient Upscaling of Large Language Models15 Oct 2024 0 repositories listed
-
Ada-K Routing: Boosting the Efficiency of MoE-based LLMs14 Oct 2024 0 repositories listed
-
Learning to Ground VLMs without Forgetting14 Oct 2024 0 repositories listed
-
Scalable Multi-Domain Adaptation of Language Models using Modular Experts14 Oct 2024 0 repositories listed
-
ContextWIN: Whittle Index Based Mixture-of-Experts Neural Model For Restless Bandits Via Deep RL13 Oct 2024 0 repositories listed
-
MoIN: Mixture of Introvert Experts to Upcycle an LLM13 Oct 2024 0 repositories listed
-
AT-MoE: Adaptive Task-planning Mixture of Experts via LoRA Approach12 Oct 2024 0 repositories listed
-
12 Oct 2024 0 repositories listed Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
10 Oct 2024 0 repositories listed
-
Upcycling Large Language Models into Mixture of Experts10 Oct 2024 0 repositories listed
-
Functional-level Uncertainty Quantification for Calibrated Fine-tuning on LLMs9 Oct 2024 0 repositories listed
-
Toward generalizable learning of all (linear) first-order methods via memory augmented Transformers8 Oct 2024 0 repositories listed
-
Probing the Robustness of Theory of Mind in Large Language Models8 Oct 2024 0 repositories listed
-
Scaling Laws Across Model Architectures: A Comparative Analysis of Dense and MoE Models in Large Language Models8 Oct 2024 0 repositories listed
-
Realizing Video Summarization from the Path of Language-based Semantic Understanding6 Oct 2024 0 repositories listed
-
A Dynamic Approach to Stock Price Prediction: Comparing RNN and Mixture of Experts Models Across Different Volatility Profiles4 Oct 2024 0 repositories listed
-
Structure-Enhanced Protein Instruction Tuning: Towards General-Purpose Protein Understanding with LLMs4 Oct 2024 0 repositories listed
-
Efficient Residual Learning with Mixture-of-Experts for Universal Dexterous Grasping3 Oct 2024 0 repositories listed
-
Neutral residues: revisiting adapters for model extension3 Oct 2024 0 repositories listed
-
On Expert Estimation in Hierarchical Mixture of Experts: Beyond Softmax Gating Functions3 Oct 2024 0 repositories listed
-
Revisiting Prefix-tuning: Statistical Benefits of Reparameterization among Prompts3 Oct 2024 0 repositories listed
-
EC-DIT: Scaling Diffusion Transformers with Adaptive Expert-Choice Routing2 Oct 2024 0 repositories listed
-
The Labyrinth of Links: Navigating the Associative Maze of Multi-modal LLMs2 Oct 2024 0 repositories listed
-
Upcycling Instruction Tuning from Dense to Mixture-of-Experts via Parameter Merging2 Oct 2024 0 repositories listed
-
MoS: Unleashing Parameter Efficiency of Low-Rank Adaptation with Mixture of Shards1 Oct 2024 0 repositories listed
-
UniAdapt: A Universal Adapter for Knowledge Calibration1 Oct 2024 0 repositories listed
-
30 Sep 2024 0 repositories listed
-
IDEA: An Inverse Domain Expert Adaptation Based Active DNN IP Protection Method29 Sep 2024 0 repositories listed
-
SciDFM: A Large Language Model with Mixture-of-Experts for Science27 Sep 2024 0 repositories listed
-
Boosting Code-Switching ASR with Mixture of Experts Enhanced Speech-Conditioned LLM24 Sep 2024 0 repositories listed
-
Leveraging Mixture of Experts for Improved Speech Deepfake Detection24 Sep 2024 0 repositories listed
-
Toward Mixture-of-Experts Enabled Trustworthy Semantic Communication for 6G Networks24 Sep 2024 0 repositories listed
-
Multi-Modal Generative AI: Multi-modal LLM, Diffusion and Beyond23 Sep 2024 0 repositories listed
-
Routing in Sparsely-gated Language Models responds to Context21 Sep 2024 0 repositories listed
-
Multi-omics data integration for early diagnosis of hepatocellular carcinoma (HCC) using machine learning20 Sep 2024 0 repositories listed
-
Robust Audiovisual Speech Recognition Models with Mixture-of-Experts19 Sep 2024 0 repositories listed
-
GRIN: GRadient-INformed MoE18 Sep 2024 0 repositories listed
-
Mixture of Diverse Size Experts18 Sep 2024 0 repositories listed
-
LPT++: Efficient Training on Mixture of Long-tailed Experts17 Sep 2024 0 repositories listed
-
Adaptive Segmentation-Based Initialization for Steered Mixture of Experts Image Regression16 Sep 2024 0 repositories listed
-
Integrating AI's Carbon Footprint into Risk Management Frameworks: Strategies and Tools for Sustainable Compliance in Banking Sector15 Sep 2024 0 repositories listed
-
DA-MoE: Towards Dynamic Expert Allocation for Mixture-of-Experts Models10 Sep 2024 0 repositories listed
-
STUN: Structured-Then-Unstructured Pruning for Scalable MoE Pruning10 Sep 2024 0 repositories listed
-
Adapted-MoE: Mixture of Experts with Test-Time Adaption for Anomaly Detection9 Sep 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.