view article Article Up to 3.2x Faster Inference with LFM2.5-DSpark LiquidAI • about 12 hours ago • 14
view article Article LFM2.5 Q4\_0 Checkpoints from Quantization-Aware Distillation LiquidAI • 1 day ago • 33
📚 LLM pretraining datasets Collection A collection of datasets for LLM pretraining • 9 items • Updated May 5, 2025 • 27
Apodex Discovery: Reality Benchmarks and Environments for Evaluating and Building Discoverative Artificial Intelligence Paper • 2608.11341 • Published 10 days ago • 35
Principled Analysis of Deep Reinforcement Learning Evaluation and Design Paradigms Paper • 2607.07769 • Published Jul 8 • 11
Weaves, Wires, and Morphisms: Formalizing and Implementing the Algebra of Deep Learning Paper • 2604.07242 • Published 9 days ago • 4
LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks Paper • 2608.01964 • Published 18 days ago • 176
SequenceMatch: Imitation Learning for Autoregressive Sequence Modelling with Backtracking Paper • 2306.05426 • Published Jun 8, 2023 • 1
🤏 Smol-Data Collection Tried and tested mixes for strong pretraining. Inspired by https://huggingface.co/blog/codelion/optimal-dataset-mixing • 14 items • Updated Mar 2 • 18
Zero-Mem: Zero-Token Memory Operations for LLM Agents Paper • 2607.29377 • Published 21 days ago • 12
Progressive Agent Skill Generation via Reinforcement Learning Paper • 2608.01678 • Published 18 days ago • 59
view article Article LFM2.5-Encoders for Fast Long-Context Inference on CPU LiquidAI • 24 days ago • 68
RADLADS: Rapid Attention Distillation to Linear Attention Decoders at Scale Paper • 2505.03005 • Published May 5, 2025 • 36
view article Article Kimi K3 Model Overview: 2.8T Parameters, MXFP4 Quantization, and What the Open Weights Mean for the Community ResterChed • Jul 17 • 197