FlashPrefill V2: Block-Sparse Prefill Attention for Long-Context LLM Serving Paper • 2608.19758 • Published 1 day ago • 4
AVA-Encoder: Towards Agent-Native Video Representation Learning Paper • 2608.12313 • Published 9 days ago • 40
ASI-Bench: At the Dawn of Artificial Superintelligence Paper • 2608.17271 • Published 3 days ago • 56
Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization Paper • 2608.16072 • Published 4 days ago • 145
MegaParts: Scaling Part-Aware 3D Object Generation to 300 Parts via Token-Efficient Autoregressive Modeling Paper • 2608.14783 • Published 7 days ago • 19
Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus Paper • 2608.12149 • Published 9 days ago • 30
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 11 days ago • 336
Skaling: Chinchilla's Exponents Meet Kaplan's Coupling Paper • 2608.07222 • Published 14 days ago • 10
Sol-Attn: Accelerating Video Generation Inference via On-the-Fly Attention Sparsification Paper • 2607.24027 • Published 25 days ago • 37