ActiveUltraFeedback: Efficient Preference Data Generation using Active Learning Paper • 2603.09692 • Published Mar 10 • 5
Vidu S1: A Real-Time Interactive Video Generation Model Paper • 2607.03118 • Published 24 days ago • 139
For-Value: Efficient Forward-Only Data Valuation for finetuning LLMs and VLMs Paper • 2508.10180 • Published Apr 25 • 19
Learning to Trigger: Reinforcement Learning at the Large Hadron Collider Paper • 2606.23993 • Published 30 days ago • 5
Learning to Trigger: Reinforcement Learning at the Large Hadron Collider Paper • 2606.23993 • Published 30 days ago • 5
Learning to Trigger: Reinforcement Learning at the Large Hadron Collider Paper • 2606.23993 • Published 30 days ago • 5
view article Article Jupyter Agents: training LLMs to reason with notebooks +1 baptistecolle, hannayukhymenko, lvwerra • Sep 10, 2025 • 67