Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Jinyang Wu
Jinyang23
18
37
2
Follow
zzzzhw's profile picture
John-SYSU's profile picture
Piggy12's profile picture
14 followers
·
8 following
https://jinyangwu.github.io/
jinyangwu
AI & ML interests
large language models, reasoning, agentic rl
Recent Activity
authored
a paper
about 2 hours ago
From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search
upvoted
a
paper
about 6 hours ago
Kimi K3: Open Frontier Intelligence
upvoted
a
paper
about 8 hours ago
From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search
View all activity
Organizations
None yet
Jinyang23
's models
19
Sort: Recently updated
Jinyang23/Seed-AlfWorld-3B
Text Generation
•
3B
•
Updated
11 days ago
•
572
•
2
Jinyang23/EKGPO-ScienceWorld-7B
Updated
20 days ago
Jinyang23/EKGPO-WebShop-7B
Updated
24 days ago
Jinyang23/Journal-AlfWorld-7B
Updated
25 days ago
Jinyang23/STARK-WebShop-7B
Updated
25 days ago
Jinyang23/Journal-ScienceWorld-7B
Updated
25 days ago
Jinyang23/STARK-WebShop-1.5B
Updated
25 days ago
Jinyang23/EKGPO-WebShop-1.5B
Updated
25 days ago
Jinyang23/EKGPO-ScienceWorld-1.5B
Updated
26 days ago
Jinyang23/EKGPO-AlfWorld-1.5B
Updated
26 days ago
Jinyang23/Journal-WebShop-7B
Updated
28 days ago
Jinyang23/EKGPO-AlfWorld-7B
Updated
28 days ago
Jinyang23/OPID-ALFWorld-1.7B
Reinforcement Learning
•
2B
•
Updated
Jun 26
•
41
•
3
Jinyang23/Maestro-4B
5B
•
Updated
May 22
•
33
Jinyang23/mm-agentic-tool-use
Updated
Mar 24
Jinyang23/ATLAS-RL
Reinforcement Learning
•
3B
•
Updated
Mar 5
•
14
Jinyang23/Spark-1.5B-ScienceWorld
Reinforcement Learning
•
2B
•
Updated
Jan 30
•
11
Jinyang23/Spark-1.5B-WebShop
Reinforcement Learning
•
2B
•
Updated
Jan 30
•
10
•
1
Jinyang23/Spark-1.5B-ALFWorld
Reinforcement Learning
•
2B
•
Updated
Jan 30
•
8