Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Kishan Vavdara
kishan51
1
3
3
Follow
kishan511
AI & ML interests
LLM finetuning and Agents.
Recent Activity
updated
a model
about 1 month ago
kishan51/llm-zero-lite-experiments
published
a model
about 1 month ago
kishan51/llm-zero-lite-experiments
upvoted
an
article
about 1 month ago
From GRPO to DAPO and GSPO: What, Why, and How
View all activity
Organizations
None yet
kishan51
's models
11
Sort: Recently updated
kishan51/llm-zero-lite-experiments
Reinforcement Learning
•
Updated
Jun 21
kishan51/variable_grpo_final
Text Generation
•
Updated
Apr 12
•
3
kishan51/variable_grpo_checkpoint500
Text Generation
•
Updated
Apr 12
•
1
kishan51/variable_grpo_checkpoint400
Text Generation
•
Updated
Apr 12
•
1
kishan51/variable_grpo_checkpoint300
Text Generation
•
Updated
Apr 12
•
2
kishan51/variable_grpo_checkpoint200
Text Generation
•
Updated
Apr 12
•
1
kishan51/variable_grpo_checkpoint100
Text Generation
•
Updated
Apr 12
•
2
kishan51/binary_grpo_final
Text Generation
•
Updated
Apr 12
•
2
kishan51/binary_grpo_checkpoint500
Text Generation
•
Updated
Apr 12
•
5
kishan51/soft_grpo_final
Text Generation
•
Updated
Apr 12
•
1
kishan51/qwen05b-gptoss-phase2-resized
Text Generation
•
0.5B
•
Updated
Jan 30
•
3