Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Yi Wang's picture

Yi Wang

CokeWang
3 32 3
dipankarsarkar's profile picture nisar333's profile picture
·

AI & ML interests

Agent, Time-series, LLM, Multimodal LLM

Recent Activity

posted an update about 8 hours ago
LoopArena: Which Models Make Good Runtime Controllers for Coding Agents? Long-running coding agents are often guided by another model that reviews progress, chooses the next assignment, requests verification, and decides when to stop. LoopArena evaluates that model as the Controller while keeping the coding Worker and execution setup fixed. The benchmark covers next-step decisions, repeated control over task slices, and complete software tasks. We have released the benchmark data, evaluation code, and v0.1.0 results. https://huggingface.co/papers/2608.28281 https://github.com/AMAP-ML/LoopArena
upvoted a paper 4 days ago
DreamX-Creator: Democratizing Native Audio-Video Generation at 2K Resolution
submitted a paper 4 days ago
LoopArena: Benchmarking Models as Runtime Controllers for Loop Engineering
View all activity

Organizations

AGI Lab's profile picture
CokeWang 's papers 4
arxiv:2608.28281
arxiv:2604.17295
arxiv:2602.09856
arxiv:2601.10477
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs