Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Kai Zheng's picture

Kai Zheng

tangmen
2 5 4
agentlans's profile picture kokemaya's profile picture AidanShi0331's profile picture
·

AI & ML interests

None yet

Recent Activity

authored a paper 18 days ago
TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training
upvoted a paper 19 days ago
TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training
submitted a paper 19 days ago
TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training
View all activity

Organizations

WizardLM Team's profile picture

upvoted a paper 19 days ago

TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training

Paper • 2607.05804 • Published 21 days ago • 18
upvoted a paper about 1 month ago

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct

Paper • 2606.23543 • Published Jun 22 • 6
upvoted a paper 4 months ago

OffSeeker: Online Reinforcement Learning Is Not All You Need for Deep Research Agents

Paper • 2601.18467 • Published Jan 26 • 1
upvoted a paper 5 months ago

RubricBench: Aligning Model-Generated Rubrics with Human Standards

Paper • 2603.01562 • Published Mar 2 • 64
upvoted a collection over 2 years ago

WizardLM

Collection
0 items • Updated Apr 17, 2025 • 108
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs