view article Article Welcome Inkling by Thinking Machines +2 burtenshaw, merve, pcuenq, ariG23498 • 6 days ago • 107
view article Article Native-speed vLLM transformers modeling backend hmellor, lysandre • 13 days ago • 56
Nemotron 3 Ultra: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning Paper • 2606.15007 • Published Jun 12 • 19
OpenThinker-Agent2 Collection OpenThinker-Agent2: agentic SFT/RL datasets and 8B/32B models (cold-start SFT, RL, and the OpenThinkerAgent-32B release). • 11 items • Updated Jun 11 • 9
view article Article Building Moon Bot: A Slack-Native Coding Agent Backed by HuggingFace Buckets huggingface • 27 days ago • 47
view article Article I fine-tuned a model for free from one prompt, with TRL and the Google Colab CLI sergiopaniego • Jun 15 • 4
view article Article Harness, Scaffold, and the AI Agent Terms Worth Getting Right sergiopaniego, ariG23498 • May 25 • 134
view article Article The Open Source Community is backing OpenEnv for Agentic RL +18 burtenshaw, spisakjo, lysandre, darktex, willcb, qjoy, pawalt, cwing-nv, danielhanchen, andrewzhou, thegovind, shimmyshimmer, Hamid-Nazeri, Sanyam, zkwentz, emre0, lewtun, sergiopaniego, banghua, unseenmars • Jun 8 • 106
view article Article Profiling in PyTorch (Part 1): A Beginner's Guide to torch.profiler +3 ariG23498, sayakpaul, sergiopaniego, ror, pcuenq • May 29 • 149
Repo2RLEnv — Verifiable RL Environments Collection Verifiable RL environments built from real GitHub repos. One dataset per pipeline. Source: https://github.com/huggingface/Repo2RLEnv • 5 items • Updated Jun 18 • 1
view article Article Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL +6 aminediroHF, qgallouedec, kashif, lewtun, edbeeching, albertvillanova, lvwerra, sergiopaniego • May 27 • 43
RFDetr Collection RF-DETR checkpoints converted to be used with 🤗 Transformers • 15 items • Updated May 27 • 17
🧬 Carbon Collection Carbon 500M, 3B, 8B genomic models and GGUF variants for llama.cpp • 7 items • Updated Jun 2 • 44
view article Article KV Caching Explained: Optimizing Transformer Inference Efficiency not-lain • Jan 30, 2025 • 376