Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Jinyang Wu
Jinyang23
21
41
2
Follow
Rohitbobli's profile picture
Randolphzeng's profile picture
Zethive's profile picture
19 followers
·
8 following
https://jinyangwu.github.io/
jinyangwu
AI & ML interests
large language models, reasoning, agentic rl
Recent Activity
authored
a paper
about 14 hours ago
VideoGAIA: A Benchmark for General AI Assistants on Agentic Video Understanding
upvoted
a
paper
about 21 hours ago
VideoGAIA: A Benchmark for General AI Assistants on Agentic Video Understanding
published
a model
7 days ago
Jinyang23/SEED-ALFWorld-3B-SFT
View all activity
Organizations
None yet
Jinyang23
's models
20
Sort: Recently updated
Jinyang23/SEED-ALFWorld-3B-SFT
3B
•
Updated
23 days ago
•
20
Jinyang23/Seed-AlfWorld-3B
Text Generation
•
3B
•
Updated
Jul 17
•
419
•
2
Jinyang23/EKGPO-ScienceWorld-7B
Updated
Jul 8
Jinyang23/EKGPO-WebShop-7B
Updated
Jul 4
Jinyang23/Journal-AlfWorld-7B
Updated
Jul 3
Jinyang23/STARK-WebShop-7B
Updated
Jul 3
Jinyang23/Journal-ScienceWorld-7B
Updated
Jul 3
Jinyang23/STARK-WebShop-1.5B
Updated
Jul 3
Jinyang23/EKGPO-WebShop-1.5B
Updated
Jul 3
Jinyang23/EKGPO-ScienceWorld-1.5B
Updated
Jul 2
Jinyang23/EKGPO-AlfWorld-1.5B
Updated
Jul 2
Jinyang23/Journal-WebShop-7B
Updated
Jun 30
Jinyang23/EKGPO-AlfWorld-7B
Updated
Jun 30
Jinyang23/OPID-ALFWorld-1.7B
Reinforcement Learning
•
2B
•
Updated
Jun 26
•
13
•
3
Jinyang23/Maestro-4B
5B
•
Updated
May 22
•
38
Jinyang23/mm-agentic-tool-use
Updated
Mar 24
Jinyang23/ATLAS-RL
Reinforcement Learning
•
3B
•
Updated
Mar 5
•
10
Jinyang23/Spark-1.5B-ScienceWorld
Reinforcement Learning
•
2B
•
Updated
Jan 30
•
18
Jinyang23/Spark-1.5B-WebShop
Reinforcement Learning
•
2B
•
Updated
Jan 30
•
13
•
1
Jinyang23/Spark-1.5B-ALFWorld
Reinforcement Learning
•
2B
•
Updated
Jan 30
•
23