Beyond Relevance-Centric Retrieval: Rubric-Oriented Document Set Selection and Ranking Paper • 2607.19747 • Published 3 days ago • 27
Text Template Tokens Are Implicit Semantic Registers in Diffusion Transformers Paper • 2607.19139 • Published 4 days ago • 72
Mean Mode Screaming: Mean--Variance Split Residuals for 1000-Layer Diffusion Transformers Paper • 2605.06169 • Published May 7 • 238
Inference-Scale Complexity in ANN-SNN Conversion for High-Performance and Low-Power Applications Paper • 2409.03368 • Published Mar 5, 2025 • 1
RT-Lynx: Putting the GEMM Sparsity In a Right Way for Diffusion Models Paper • 2605.26632 • Published May 26 • 16
Rethinking Cross-Layer Information Routing in Diffusion Transformers Paper • 2605.20708 • Published May 20 • 112
Full Attention Strikes Back: Transferring Full Attention into Sparse within Hundred Training Steps Paper • 2605.16928 • Published May 16 • 99
EdgeRazor: A Lightweight Framework for Large Language Models via Mixed-Precision Quantization-Aware Distillation Paper • 2605.04062 • Published Apr 10 • 31