Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Yi Cui's picture

Yi Cui

onekq
24 20 1
amadanielcasmir's profile picture Fukrapehlwan's profile picture siyengfeng's profile picture
·
  • onekq_ai
  • onekq
  • yicui

AI & ML interests

Benchmark, Code Generation Model

Recent Activity

posted an update about 5 hours ago
There has been a leaked memo (now struck down) from the founder of DeepSeek. I'm not here to circulate it, but comment on the minimum-effort evolutionary path he proposed. LLM->CoT->Agent->Self-improvement->Singularity->Physical This makes sense to me: even at the agent stage I learn world models much faster than when I learned LLM at the LLM stage. But this means humans are still needed beyond the digital singularity, until robots can close their own loop: eval, manufacturing, self improvement, i.e. physical singularity.
posted an update 1 day ago
The Nvidia paper came down to this: remove synchronization barriers. DeepSeek has already done that with DeepEP (which this paper cited) one layer above. https://huggingface.co/papers/2607.16100 Nevertheless this is great. More users will benefit from this: Nvidia NCCL > SGLang/vLLM DeepEP
upvoted a paper 2 days ago
Every Microsecond Matters: Achieving Near Speed-of-Light Latency in GPU Collectives
View all activity

Organizations

MLX Community's profile picture ONEKQ AI's profile picture CSC Generation's profile picture

onekq 's datasets

None public yet
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs