view article Article Post-training NVIDIA Nemotron 3.5 Lightning for enterprise domains lujangusface • 4 days ago
view article Article EAGLE-3 speculative decoding for NVIDIA Nemotron 3.5 Lightning lujangusface • 4 days ago
thoughtworks/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Eagle3 Text Generation • 0.2B • Updated 4 days ago • 309 • 1
thoughtworks/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Eagle3 Text Generation • 0.2B • Updated 4 days ago • 309 • 1
view article Article We Pitted the Cheapest TPU Against an NVIDIA L4. Here's What 6 Experiments Revealed. lujangusface • Apr 17 • 1
view article Article 1.37x Faster on Alibaba's 80B Code Model: EAGLE3 for Qwen3-Coder-Next lujangusface • Apr 15
view article Article 1.7x Faster on a 218B Model: EAGLE3 Speculative Decoding for GLM-4.7 lujangusface • Apr 15 • 1
view article Article 2x Faster on a 229B MoE: EAGLE3 Speculative Decoding for MiniMax-M2.5 lujangusface • Apr 9 • 3
view article Article Google Released Gemma-4 Four Days Ago. We Already Made It 1.72× Faster. lujangusface • Apr 7 • 3
view article Article Google Released Gemma-4 Four Days Ago. We Already Made It 1.72× Faster. lujangusface • Apr 7 • 3