Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
WildEval
non-profit
wild_eval
WildEval
Activity Feed
Request to join this org
Follow
17
AI & ML interests
None defined yet.
Recent Activity
ChengsongHuang
Â
authored
a paper
about 22 hours ago
Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling
ChengsongHuang
Â
authored
a paper
about 22 hours ago
MoCo: A One-Stop Shop for Model Collaboration Research
ChengsongHuang
Â
authored
a paper
about 22 hours ago
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories
View all activity
Team members
9
spaces
1
pinned
Runtime error
Agents
6
Zebra Logic Bench
🦓
Explore and evaluate Zebra Logic models
models
0
None public yet
datasets
9
Sort:Â Recently updated
WildEval/ZebraLogic
Viewer
•
Updated
Feb 4, 2025
•
4.26k
•
2.01k
•
18
WildEval/G-PlanET
Viewer
•
Updated
Aug 1, 2024
•
1.42k
•
11
•
1
WildEval/ZeroEval
Viewer
•
Updated
Jul 23, 2024
•
4.61k
•
156
WildEval/WildBench-V2
Viewer
•
Updated
May 22, 2024
•
2.05k
•
86
WildEval/WildBench-Results-v2-internal
Viewer
•
Updated
May 21, 2024
•
30k
•
82
WildEval/WildBench-Results-V2
Viewer
•
Updated
May 20, 2024
•
10.2k
•
38
WildEval/WildBench-v2-dev
Viewer
•
Updated
Apr 19, 2024
•
5.99k
•
5
WildEval/WildBench-dev
Viewer
•
Updated
Apr 19, 2024
•
14.1k
•
5
•
1
WildEval/NaturalChats
Viewer
•
Updated
Apr 18, 2024
•
641k
•
4