view article Article Meta is back with Muse Glimmer: local, agentic, multimodal, and open source +2 pcuenq, merve, burtenshaw, ariG23498 • 5 days ago • 93
view article Article SmolLM3: smol, multilingual, long-context reasoner +21 eliebak, cmpatino, anton-l, edbeeching, m-ric, nouamanetazi, akseljoonas, guipenedo, hynky, clefourrier, SaylorTwift, kashif, qgallouedec, hlarcher, glutamatt, Xenova, reach-vb, ngxson, craffel, lewtun, loubnabnl, lvwerra, thomwolf • Jul 8, 2025 • 789
Running 3.98k The Ultra-Scale Playbook 🌌 3.98k The ultimate guide to training LLM on large GPU Clusters
view article Article Welcome Gemma 4: Frontier multimodal intelligence on device +5 merve, pcuenq, sergiopaniego, burtenshaw, Steveeeeeeen, alvarobartt, SaylorTwift • Apr 2 • 919
view article Article How I contributed a new model to the Transformers library using Codex nielsr • Mar 30 • 53
view article Article SigLIP 2: A better multilingual vision language encoder +1 ariG23498, merve, qubvel-hf • Feb 21, 2025 • 224
view article Article LoRA training scripts of the world, unite! linoyts, multimodalart • Jan 2, 2024 • 80
view article Article Advanced Flux Dreambooth LoRA Training with 🧨 diffusers linoyts • Oct 21, 2024 • 42
Running Featured 561 Vision Arena (Testing VLMs side-by-side) 🖼 561 Explore Vision Arena visual AI demo online
Running Agents 358 VBench Leaderboard 📊 358 Submit video model evaluation results to a public benchmark
view article Article Using LoRA for Efficient Stable Diffusion Fine-Tuning pcuenq, sayakpaul • Jan 26, 2023 • 84
Running on Zero Agents Featured 169 IDEFICS2 Playground 🐨 169 Chat with a visual AI that answers questions about images
view article Article Introducing Idefics2: A Powerful 8B Vision-Language Model for the community +1 Leyo, HugoLaurencon, VictorSanh • Apr 15, 2024 • 191