Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Yixuan Wei's picture

Yixuan Wei

EasonWei
5 7 16
Srjnnnn's profile picture amar2q3o's profile picture qiangicy's profile picture
·
  • weiyx16

AI & ML interests

None yet

Organizations

OneModel's profile picture Xwin-LM's profile picture DeepSeek's profile picture

upvoted a collection 5 months ago

DeepSeek-V4

Collection
10 items • Updated 5 days ago • 906
upvoted 3 papers about 1 year ago

FP4 All the Way: Fully Quantized Training of LLMs

Paper • 2505.19115 • Published May 25, 2025 • 4

Group Sequence Policy Optimization

Paper • 2507.18071 • Published Jul 24, 2025 • 323

The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text

Paper • 2506.05209 • Published Jun 5, 2025 • 66
upvoted a paper over 1 year ago

The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models

Paper • 2505.22617 • Published May 28, 2025 • 132
upvoted a paper almost 2 years ago

On Memorization of Large Language Models in Logical Reasoning

Paper • 2410.23123 • Published Oct 30, 2024 • 18
upvoted a collection almost 2 years ago

Qwen2.5

Collection
Qwen2.5 language models, including pretrained and instruction-tuned models of 7 sizes, including 0.5B, 1.5B, 3B, 7B, 14B, 32B, and 72B. • 43 items • Updated Mar 2 • 736
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs