Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Xinyu Zhu
TianHongZXY
3
20
14
Follow
weic22's profile picture
weizhepei's profile picture
taki555's profile picture
9 followers
·
8 following
https://zhuxinyu.top
tianhongzxy
TianHongZXY
AI & ML interests
Large Language Models; Reasoning; Reinforcement Learning
Recent Activity
authored
a paper
6 days ago
Self-Guided Test-Time Training for Long-Context LLMs
upvoted
a
paper
13 days ago
Self-Guided Test-Time Training for Long-Context LLMs
new
activity
about 1 month ago
TianHongZXY/CHIMERA-4B-RL:
Add paper link and model metadata
View all activity
Organizations
TianHongZXY
's models
12
Sort: Recently updated
TianHongZXY/CHIMERA-4B-RL
Text Generation
•
4B
•
Updated
Jun 24
•
7
•
4
TianHongZXY/CHIMERA-4B-SFT
Text Generation
•
4B
•
Updated
Jun 24
•
13
•
•
2
TianHongZXY/Qwen3-4B-NSR
4B
•
Updated
Dec 6, 2025
•
4
TianHongZXY/Qwen2.5-Math-7B-GRPO
8B
•
Updated
Jul 28, 2025
•
6
TianHongZXY/OpenR1-Math-46k-8192-Qwen2.5-7B-Instruct-GRPO-clip_0.28
Updated
Jul 8, 2025
TianHongZXY/Qwen2.5-Math-7B-W-REINFORCE
8B
•
Updated
Jun 1, 2025
•
9
•
1
TianHongZXY/Qwen3-4B-GRPO
4B
•
Updated
May 31, 2025
•
6
TianHongZXY/Qwen3-4B-PPO
4B
•
Updated
May 31, 2025
•
5
TianHongZXY/Qwen3-4B-PSR
4B
•
Updated
May 31, 2025
•
8
•
1
TianHongZXY/Qwen2.5-Math-7B-PPO
8B
•
Updated
May 31, 2025
•
5
TianHongZXY/Qwen2.5-Math-7B-PSR
8B
•
Updated
May 31, 2025
•
4
TianHongZXY/Qwen2.5-Math-7B-NSR
8B
•
Updated
May 30, 2025
•
8
•
2