Community Blog & Articles
NEW Articles from Team or Enterprise organizations will get promoted to the main section. Kimi K3 Model Overview: 2.8T Parameters, MXFP4 Quantization, and What the Open Weights Mean for the Community
ResterChed
• • 167
The State of Simulation for Physical AI: An Overview
Introducing Cosmos 3 Edge
Hugging Face on AMD Instinct MI455X: First Transformers Results
badaoui
• • 14
Be Ready Before the Attack: A Practical Guide to Self-Hosting an Open Model for Cyber Defense
jeffboudier
• • 16
POCKET: a 35-billion-parameter model that runs on your iPhone — and on your PC with no GPU
FINAL-Bench
• • 12
KV Caching Explained: Optimizing Transformer Inference Efficiency
not-lain
• • 383
NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval
nvidia
• • 57
The influx of specialist models on the Open SLM Leaderboard
Banaxi-Tech
• • 7
Uncensor any LLM with abliteration
mlabonne
• • 882
Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers
nvidia
• • 80
Kimi K3, previewed: inside the first open 3T-class model
Introduction to State Space Models (SSM)
lbourdois
• • 238
Code a simple RAG from scratch
ngxson
• • 367
From GRPO to DAPO and GSPO: What, Why, and How
NormalUhr
• • 133
J-Space: Yet Another LLM Mind Reader?
dlouapre
• • 35
Tokenization is Killing our Multilingual LLM Dream
Efficient LLM Pretraining: Packed Sequences and Masked Attention
sirluk
• • 73
What is test-time compute and how to scale it?
Kseniase
• • 125
Sensitivity Aware Mixed Precision Quantization V1
badaoui
• • 29