Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Yi Cui's picture

Yi Cui

onekq
24 20 1
Dcas89's profile picture nhlayisekobvuma's profile picture astrologerbot's profile picture
·
  • onekq_ai
  • onekq
  • yicui

AI & ML interests

Benchmark, Code Generation Model

Recent Activity

posted an update 1 day ago
My biggest takeaway from Ben Thompson interview is the legacy of this round of "bubble". For railroad, it was the track, and dot com, fiber. Compared to these two, GPUs depreciate too fast. The answer is power plants.
posted an update 3 days ago
NVidia?!
posted an update 5 days ago
Back of the envelope calculation: Ox Alpha has been out for a week with 130K users and 7T tokens burned. Method 1: Assuming Ox Alpha is the GLM 5.3 class, it has roughly the same active param count as DeepSeek V4 Pro (~40B vs 49B), derive the cost by the floor price DS has ever published. Method 2: Assuming the users are concentrated within an 8-hour working window each day, derive number of H800 nodes needed (~700 nodes at $2/GPU-hour). In both methods, I assume 90/10 IO split and 60% cache hit. Both methods come to $2M. I think the ROI is awesome: (1) publicity and (2) data harvesting.
View all activity

Organizations

MLX Community's profile picture ONEKQ AI's profile picture CSC Generation's profile picture

onekq 's datasets

None public yet
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs