Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Yi Cui's picture

Yi Cui

onekq
24 20 1
duartevn's profile picture arpenxd's profile picture Mugestu49's profile picture
ยท
  • onekq_ai
  • onekq
  • yicui

AI & ML interests

Benchmark, Code Generation Model

Recent Activity

posted an update about 19 hours ago
Back of the envelope calculation: Ox Alpha has been out for a week with 130K users and 7T tokens burned. Method 1: Assuming Ox Alpha is the GLM 5.3 class, it has roughly the same active param count as DeepSeek V4 Pro (~40B vs 49B), derive the cost by the floor price DS has ever published. Method 2: Assuming the users are concentrated within an 8-hour working window each day, derive number of H800 nodes needed (~700 nodes at $2/GPU-hour). In both methods, I assume 90/10 IO split and 60% cache hit. Both methods come to $2M. I think the ROI is awesome: (1) publicity and (2) data harvesting.
repliedto their post 1 day ago
My guess on Ox Alpha -> GLM
repliedto their post 1 day ago
My guess on Ox Alpha -> GLM
View all activity

Organizations

MLX Community's profile picture ONEKQ AI's profile picture CSC Generation's profile picture

liked a Space almost 2 years ago
Runtime error
Agents
23

Quant Request

๐Ÿฆ€
23

Submit Hugging Face model links for quantization requests

Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs