view article Article Making Knowledge Distillation Cheap Enough to Run at Scale MultiverseComputingCAI β’ 15 days ago β’ 37
view article Article DAVIDAU DELETED MY POST AND BANNED ME FOR ASKING ABOUT BENCHMARKS β HERE IS THE FULL ANALYSIS (9B + 27B + 390 MODELS) DedeProGames β’ 13 days ago β’ 15
FastUMI-100K: Advancing Data-driven Robotic Manipulation with a Large-scale UMI-style Dataset Paper β’ 2510.08022 β’ Published Oct 9, 2025 β’ 1
SWE-RL Collection Software Engineering Tasks for RL - Images served by Prime's registry β’ 9 items β’ Updated Jul 22 β’ 2
Immersion in the GitHub Universe: Scaling Coding Agents to Mastery Paper β’ 2602.09892 β’ Published Feb 10 β’ 6
view article Article From GRPO to DAPO and GSPO: What, Why, and How NormalUhr β’ Aug 9, 2025 β’ 140
PivotRL: High Accuracy Agentic Post-Training at Low Compute Cost Paper β’ 2603.21383 β’ Published Mar 22 β’ 20
GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models Paper β’ 2508.06471 β’ Published Aug 8, 2025 β’ 213
π€ Smol-Data Collection Tried and tested mixes for strong pretraining. Inspired by https://huggingface.co/blog/codelion/optimal-dataset-mixing β’ 14 items β’ Updated Mar 2 β’ 18
SmolLM3 pretraining datasets Collection datasets used in SmolLM3 pretraining β’ 15 items β’ Updated Aug 12, 2025 β’ 56
Embarrassingly Simple Self-Distillation Improves Code Generation Paper β’ 2604.01193 β’ Published Apr 1 β’ 56
view article Article Model2Vec: Distill a Small Fast Model from any Sentence Transformer Pringled β’ Oct 14, 2024 β’ 105
view article Article LLM Inference on Edge: A Fun and Easy Guide to run LLMs via React Native on your Phone! medmekk, marcsun13 β’ Mar 7, 2025 β’ 99
view article Article Welcome Gemma 4: Frontier multimodal intelligence on device +5 merve, pcuenq, sergiopaniego, burtenshaw, Steveeeeeeen, alvarobartt, SaylorTwift β’ Apr 2 β’ 921
view article Article Cosmopedia: how to create large-scale synthetic data for pre-training Large Language Models +1 loubnabnl, anton-l, davanstrien β’ Mar 20, 2024 β’ 115