norgitov/ trends
Technology · people · ideas
Back to discovery/Daily Papers23 hours ago

Block Sparse Attention with Log-Linear Complexity

The paper introduces PISA, a block-sparse attention mechanism with log-linear complexity, designed to scale language models for long contexts. It uses a pyramid Top-K selection strategy to efficiently identify relevant keys, reducing computational cost compared to traditional methods.

Open original
SIGNAL FROM THE SOURCE
13
source votes
Tracking sinceSeptember 28, 202613 source votes

This source has not updated recently. Its observation time is shown above.

Momentum+3.96/hover 3.03 h
Discussion—Read comments ↗
PublishedSeptember 25, 2026Bohao Tang, Zhen Qin, Yuqi Pan
BEHIND THE NUMBERS

How interest changes

History starts here

The chart will appear after repeat observations. The current metric comes from the source.

13 source votes

Real observations only. History before source connection is not reconstructed.

WHY IT MAY MATTER

This approach could be useful for improving the efficiency of language models when processing long sequences, reducing computational costs without significant loss in performance.

A useful discovery?
KEEP EXPLORING

Connected ideas

Explore topic