Instella-MoE ✨ Collection Family of fully open 16B MoE LLM with 2.8B active params per token, trained on AMD Instinct™ MI300 & MI325 GPUs. • 6 items • Updated 4 days ago • 11
Laguna M.1 Collection Our first M-class coding agent model, designed for long-horizon work. Apache 2.0. • 4 items • Updated 7 days ago • 23
Gemma 4 Collection Gemma 4 is Google's new model family including including E2B, E4B, 26B-A4B, and 31B. • 43 items • Updated 10 days ago • 250
Zamba2-VL Collection A suite of vision-language models based on Zamba2. • 3 items • Updated Jun 9 • 5
Gemma 4 QAT Collection Gemma 4 QAT (Quantization-Aware Training) for 3x less memory use and near original accuracy. • 16 items • Updated 10 days ago • 113
Self-Improving Language Models with Bidirectional Evolutionary Search Paper • 2605.28814 • Published May 27 • 62
Talkie 1930 Collection Models based on Talkie 1930; conversions to hf transformers format and context extension to 32k. • 4 items • Updated Jun 1 • 3
jina-embeddings-v5-omni Collection Multimodal (text + image + video + audio) embedding models aligned with jina-embeddings-v5-text-*. Two sizes, four task variants each. • 27 items • Updated 1 day ago • 36