Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Xinyu Zhu
TianHongZXY
3
21
14
Follow
wanng's profile picture
weic22's profile picture
Twwilght's profile picture
9 followers
·
8 following
https://zhuxinyu.top
tianhongzxy
TianHongZXY
AI & ML interests
Large Language Models; Reasoning; Reinforcement Learning
Recent Activity
upvoted
a
paper
1 day ago
EnvHarness: Awakening Static Worlds for Agent Learning
authored
a paper
about 1 month ago
Self-Guided Test-Time Training for Long-Context LLMs
upvoted
a
paper
about 1 month ago
Self-Guided Test-Time Training for Long-Context LLMs
View all activity
Organizations
TianHongZXY
's models
12
Sort: Recently updated
TianHongZXY/CHIMERA-4B-RL
Text Generation
•
4B
•
Updated
Jun 24
•
13
•
4
TianHongZXY/CHIMERA-4B-SFT
Text Generation
•
4B
•
Updated
Jun 24
•
17
•
2
TianHongZXY/Qwen3-4B-NSR
4B
•
Updated
Dec 6, 2025
•
15
TianHongZXY/Qwen2.5-Math-7B-GRPO
8B
•
Updated
Jul 28, 2025
•
5
TianHongZXY/OpenR1-Math-46k-8192-Qwen2.5-7B-Instruct-GRPO-clip_0.28
Updated
Jul 8, 2025
TianHongZXY/Qwen2.5-Math-7B-W-REINFORCE
8B
•
Updated
Jun 1, 2025
•
9
•
1
TianHongZXY/Qwen3-4B-GRPO
4B
•
Updated
May 31, 2025
•
6
TianHongZXY/Qwen3-4B-PPO
4B
•
Updated
May 31, 2025
•
9
TianHongZXY/Qwen3-4B-PSR
4B
•
Updated
May 31, 2025
•
8
•
1
TianHongZXY/Qwen2.5-Math-7B-PPO
8B
•
Updated
May 31, 2025
•
8
TianHongZXY/Qwen2.5-Math-7B-PSR
8B
•
Updated
May 31, 2025
•
9
TianHongZXY/Qwen2.5-Math-7B-NSR
8B
•
Updated
May 30, 2025
•
9
•
2