Community Blog & Articles
NEW Articles from Team or Enterprise organizations will get promoted to the main section. Kimi K3 Model Overview: 2.8T Parameters, MXFP4 Quantization, and What the Open Weights Mean for the Community
ResterChed
• • 87
Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers
Introducing Cosmos 3 Edge
NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval
Aether-7B-5Attn: A 100% Open-Source Sovereign Foundation Model — and a Controlled Experiment in Heterogeneous Attention
FINAL-Bench
• • 21
Be Ready Before the Attack: A Practical Guide to Self-Hosting an Open Model for Cyber Defense
jeffboudier
• • 14
Hugging Face on AMD Instinct MI455X: First Transformers Results
badaoui
• • 12
KV Caching Explained: Optimizing Transformer Inference Efficiency
not-lain
• • 379
POCKET: a 35-billion-parameter model that runs on your iPhone — and on your PC with no GPU
FINAL-Bench
• • 9
J-Space: Yet Another LLM Mind Reader?
dlouapre
• • 34
One Adapter, Both Modalities: Field Notes from Building and Serving a Multimodal Reranker
lightonai
• • 16
The influx of specialist models on the Open SLM Leaderboard
Banaxi-Tech
• • 7
Uncensor any LLM with abliteration
mlabonne
• • 880
Introduction to State Space Models (SSM)
lbourdois
• • 238
Code a simple RAG from scratch
ngxson
• • 366
From GRPO to DAPO and GSPO: What, Why, and How
NormalUhr
• • 132
Tokenization is Killing our Multilingual LLM Dream
Introducing North Mini Code: Cohere’s First Model For Developers
Data for Agents
nvidia
• • 25
We Just Surgically Changed What Your Model Believes
ApolloRaines
• • 3