Methods › General › Ensembling › MoE
Mixture of Experts
MoE
Introduced by Neel Kanwal et al. in Equipping Computational Pathology Systems with Artifact Processing Pipelines: A Showcase for Computation and Performance Trade-offs
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
The archive carries no description for this method.
Papers archive 2025-07-28
30 shown of 366, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
Mixture of Experts in Large Language Models 15 Jul 2025 · 0 repositories · arXiv:2507.11181
-
MoFE-Time: Mixture of Frequency Domain Experts for Time-Series Forecasting Models 9 Jul 2025 · 1 repository · arXiv:2507.06502
-
Growing Transformers: Modular Composition and Layer-wise Expansion on a Frozen Substrate 8 Jul 2025 · 1 repository · arXiv:2507.07129
-
Speech Quality Assessment Model Based on Mixture of Experts: System-Level Performance Enhancement and Utterance-Level Challenge Analysis 8 Jul 2025 · 0 repositories · arXiv:2507.06116
-
Learning Robust Stereo Matching in the Wild with Selective Mixture-of-Experts 7 Jul 2025 · 1 repository · arXiv:2507.04631
-
Sub-MoE: Efficient Mixture-of-Expert LLMs Compression via Subspace Expert Merging 29 Jun 2025 · 0 repositories · arXiv:2506.23266
-
Latent Prototype Routing: Achieving Near-Perfect Load Balancing in Mixture-of-Experts 26 Jun 2025 · 1 repository · arXiv:2506.21328
-
SAFEx: Analyzing Vulnerabilities of MoE-Based LLMs via Stable Safety-critical Expert Identification 20 Jun 2025 · 0 repositories · arXiv:2506.17368
-
Less is More: Undertraining Experts Improves Model Upcycling 17 Jun 2025 · 0 repositories · arXiv:2506.14126
-
LoRA-Mixer: Coordinate Modular LoRA Experts Through Serial Attention Routing 17 Jun 2025 · 0 repositories · arXiv:2507.00029
-
Ring-lite: Scalable Reasoning via C3PO-Stabilized Reinforcement Learning for LLMs 17 Jun 2025 · 0 repositories · arXiv:2506.14731
-
Utility-Driven Speculative Decoding for Mixture-of-Experts 17 Jun 2025 · 0 repositories · arXiv:2506.20675
-
EAQuant: Enhancing Post-Training Quantization for MoE Models via Expert-Aware Optimization 16 Jun 2025 · 0 repositories · arXiv:2506.13329
-
Serving Large Language Models on Huawei CloudMatrix384 15 Jun 2025 · 0 repositories · arXiv:2506.12708
-
Ming-Omni: A Unified Multimodal Model for Perception and Generation 11 Jun 2025 · 1 repository · arXiv:2506.09344Syntology ran 4 of 14 samples · 10 unverified
-
M2Restore: Mixture-of-Experts-based Mamba-CNN Fusion Framework for All-in-One Image Restoration 9 Jun 2025 · 0 repositories · arXiv:2506.07814
-
MoE-GPS: Guidlines for Prediction Strategy for Dynamic Expert Duplication in MoE Load Balancing 9 Jun 2025 · 0 repositories · arXiv:2506.07366
-
SMAR: Soft Modality-Aware Routing Strategy for MoE-based Multimodal Large Language Models Preserving Language Capabilities 6 Jun 2025 · 0 repositories · arXiv:2506.06406
-
FlashDMoE: Fast Distributed MoE in a Single Kernel 5 Jun 2025 · 2 repositories · arXiv:2506.04667Syntology ran 2 of 7 samples · 5 unverified
-
Out-of-Distribution Graph Models Merging 4 Jun 2025 · 0 repositories · arXiv:2506.03674
-
Decoding Knowledge Attribution in Mixture-of-Experts: A Framework of Basic-Refinement Collaboration and Efficiency Analysis 30 May 2025 · 0 repositories · arXiv:2505.24593
-
Mastering Massive Multi-Task Reinforcement Learning via Mixture-of-Expert Decision Transformer 30 May 2025 · 1 repository · arXiv:2505.24378
-
Mixture-of-Experts for Personalized and Semantic-Aware Next Location Prediction 30 May 2025 · 0 repositories · arXiv:2505.24597
-
On the Expressive Power of Mixture-of-Experts for Structured Complex Tasks 30 May 2025 · 0 repositories · arXiv:2505.24205
-
Advancing Expert Specialization for Better MoE 28 May 2025 · 0 repositories · arXiv:2505.22323
-
EvoMoE: Expert Evolution in Mixture of Experts for Multimodal Large Language Models 28 May 2025 · 0 repositories · arXiv:2505.23830
-
HiDream-I1: A High-Efficient Image Generative Foundation Model with Sparse Diffusion Transformer 28 May 2025 · 2 repositories · arXiv:2505.22705Syntology ran 2 of 3 samples · 1 unverified
-
FLAME-MoE: A Transparent End-to-End Research Platform for Mixture-of-Experts Language Models 26 May 2025 · 1 repository · arXiv:2505.20225
-
MoESD: Unveil Speculative Decoding's Potential for Accelerating Sparse MoE 26 May 2025 · 0 repositories · arXiv:2505.19645
-
Mosaic: Data-Free Knowledge Distillation via Mixture-of-Experts for Heterogeneous Distributed Environments 26 May 2025 · 0 repositories · arXiv:2505.19699
Tasks archive 2025-07-28
20 shown of 254 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
| Task | Papers |
|---|---|
| Mixture-of-Experts | 312 |
| GPU | 38 |
| Language Modelling | 34 |
| Computational Efficiency | 30 |
| Language Modeling | 29 |
| Large Language Model | 16 |
| Quantization | 14 |
| CPU | 13 |
| Multi-Task Learning | 13 |
| Image Classification | 12 |
| parameter-efficient fine-tuning | 12 |
| Scheduling | 11 |
| MMLU | 10 |
| Question Answering | 10 |
| Decoder | 9 |
| Diversity | 9 |
| image-classification | 9 |
| Continual Learning | 7 |
| Math | 7 |
| Retrieval | 7 |
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections