Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🔄
In a Training Loop
Peng Wang
stillarrow
1
63
149
Follow
yomir's profile picture
weizhepei's profile picture
singhagrima1705-glitch's profile picture
5 followers
·
39 following
https://peter-peng-w.github.io/
AI & ML interests
None yet
Recent Activity
liked
a dataset
1 day ago
FrontisAI/NatureBench
upvoted
a
paper
7 days ago
Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States
upvoted
a
paper
27 days ago
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement
View all activity
Organizations
None yet
stillarrow
's datasets
1
Sort:Â Recently updated
stillarrow/MATH
Viewer
•
Updated
Sep 25, 2025
•
26.5k
•
30