Flash-BoN: Instant Drafts for Inference-Time Scaling in Diffusion Models Paper • 2607.04461 • Published Jul 5 • 11
Capability Self-Assessment: Teaching LLMs to Know Their Limits Paper • 2606.00251 • Published May 29 • 10
ARGUS: Hallucination and Omission Evaluation in Video-LLMs Paper • 2506.07371 • Published Jun 9, 2025 • 8
Flash-BoN: Instant Drafts for Inference-Time Scaling in Diffusion Models Paper • 2607.04461 • Published Jul 5 • 11
Capability Self-Assessment: Teaching LLMs to Know Their Limits Paper • 2606.00251 • Published May 29 • 10
ToMoE: Converting Dense Large Language Models to Mixture-of-Experts through Dynamic Structural Pruning Paper • 2501.15316 • Published Jan 25, 2025 • 3
ToMoE: Converting Dense Large Language Models to Mixture-of-Experts through Dynamic Structural Pruning Paper • 2501.15316 • Published Jan 25, 2025 • 3
Not All Prompts Are Made Equal: Prompt-based Pruning of Text-to-Image Diffusion Models Paper • 2406.12042 • Published Jun 17, 2024 • 8