Drawing2CAD: Sequence-to-Sequence Learning for CAD Generation from Vector Drawings Paper β’ 2508.18733 β’ Published Aug 26, 2025 β’ 11
Bottom-up Policy Optimization: Your Language Model Policy Secretly Contains Internal Policies Paper β’ 2512.19673 β’ Published Dec 22, 2025 β’ 66
Less is More: Recursive Reasoning with Tiny Networks Paper β’ 2510.04871 β’ Published Oct 6, 2025 β’ 518
The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain Paper β’ 2509.26507 β’ Published Sep 30, 2025 β’ 551
Drivel-ology: Challenging LLMs with Interpreting Nonsense with Depth Paper β’ 2509.03867 β’ Published Sep 4, 2025 β’ 213
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey Paper β’ 2509.02547 β’ Published Sep 2, 2025 β’ 239
VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model Paper β’ 2509.09372 β’ Published Sep 11, 2025 β’ 259
Sharing is Caring: Efficient LM Post-Training with Collective RL Experience Sharing Paper β’ 2509.08721 β’ Published Sep 10, 2025 β’ 665
AgentFly: Fine-tuning LLM Agents without Fine-tuning LLMs Paper β’ 2508.16153 β’ Published Aug 22, 2025 β’ 162
view article Article CPU Optimized Embeddings with π€ Optimum Intel and fastRAG +4 peterizsak, mber, danf, echarlaix, mfuntowicz, moshew β’ Mar 15, 2024 β’ 14
Learning to Skip the Middle Layers of Transformers Paper β’ 2506.21103 β’ Published Jun 26, 2025 β’ 18
view article Article BM25 for Python: Achieving high performance while simplifying dependencies with *BM25S*β‘ xhluca β’ Jul 9, 2024 β’ 85
Absolute Zero: Reinforced Self-play Reasoning with Zero Data Paper β’ 2505.03335 β’ Published May 6, 2025 β’ 191
Medical SAM 2: Segment medical images as video via Segment Anything Model 2 Paper β’ 2408.00874 β’ Published Aug 1, 2024 β’ 52