norgitov/ trends
Technology · people · ideas
Back to discovery/arXiv48 minutes ago

OMP-MoE: Efficient Expert Pruning for Mixture-of-Experts LLMs via Orthogonal Matching Pursuit

OMP-MoE is a training-free compression framework for Mixture-of-Experts (MoE) large language models that reduces expert redundancy through orthogonal matching pursuit. It uses a greedy approach to select experts that minimize reconstruction error and optimizes cross-layer allocation with a water-filling strategy, achieving significant speed improvements while maintaining high performance.

Open original
SIGNAL FROM THE SOURCE
29
September 2026
Tracking sinceSeptember 29, 2026
Momentum—More observations needed
Discussion—No comment count provided
PublishedSeptember 29, 2026Dezhi Li, Lujun Li, Qiyuan Zhu, Hao Gu, Bei Liu, Sirui Han, Yike Guo
BEHIND THE NUMBERS

How interest changes

The publication is the signal

This official source does not publish popularity metrics. The story is refreshed from its RSS feed.

Real observations only. History before source connection is not reconstructed.

WHY IT MAY MATTER

This method could be useful for optimizing the efficiency of large language models by reducing computational resources needed without significant performance loss.

A useful discovery?
KEEP EXPLORING

Connected ideas

Explore topic