Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Alexander Gurung
PRO
agurung
Follow
John6666's profile picture
MichaelR207's profile picture
21world's profile picture
5 followers
·
3 following
alex-gurung
AI & ML interests
None yet
Organizations
agurung
's models
121
Sort: Recently updated
agurung/lcb-med-hard-ft-qwen3-4b-sft-iid-24-highlr-v1
Updated
Apr 28
•
12
agurung/lcb-med-hard-ft-qwen3-4b-rft-mixed-24-highlr-v1
Updated
Apr 28
•
11
agurung/lcb-med-hard-ft-qwen3-4b-rft-iid-24-highlr-v1
Updated
Apr 28
•
9
agurung/lcb-med-hard-ft-qwen3-4b-dft-rft-mixed-24-highlr-v1
Updated
Apr 28
•
7
agurung/lcb-med-hard-ft-qwen3-4b-dft-rft-iid-24-highlr-v1
Updated
Apr 28
•
11
agurung/lcb-med-hard-ft-qwen3-4b-dft-mixed-24-highlr-v1
Updated
Apr 28
•
9
agurung/lcb-med-hard-ft-qwen3-4b-dft-iid-24-highlr-v1
Updated
Apr 28
•
7
agurung/lcb-med-hard-openrlhf-base-model-base-openrlhf-best7-4096-ep20-clean
Updated
Apr 28
agurung/lcb-med-hard-openrlhf-rft-mixed-24-step0050-openrlhf-sft-iid-cmp16
Updated
Apr 27
agurung/lcb-med-hard-openrlhf-sft-iid-24-step0038-openrlhf-sft-iid-cmp16
Updated
Apr 27
agurung/coconut-gemma-3-4b-ff-reward-filtered
Updated
Apr 15
agurung/coconut-qwen3-4b-ff-reward-filtered
Updated
Apr 15
agurung/coconut-gemma-3-1b-gsm-hard
Updated
Apr 14
agurung/coconut-gemma-3-4b-ff
4B
•
Updated
Apr 14
•
9
agurung/coconut-qwen3-4b-ff
4B
•
Updated
Apr 14
•
7
agurung/flawed-fictions-gemma-3-4b-litereason-sft-positive
5B
•
Updated
Apr 14
•
1
agurung/flawed-fictions-qwen3-4b-litereason-sft-positive
4B
•
Updated
Apr 14
•
8
agurung/colar-gemma-3-1b-gsm-hard-rl
Reinforcement Learning
•
1.0B
•
Updated
Apr 9
•
3
agurung/colar-gemma-3-1b-gsm-hard-sft
1.0B
•
Updated
Apr 9
•
3
agurung/colar-gemma-3-4b-ff-sft
4B
•
Updated
Apr 9
•
2
agurung/colar-qwen3-4b-ff-rl
Reinforcement Learning
•
4B
•
Updated
Apr 9
•
6
agurung/colar-qwen25-7b-ff-post-rl
Reinforcement Learning
•
8B
•
Updated
Apr 9
•
3
agurung/colar-qwen25-7b-ncp-post-rl
Reinforcement Learning
•
8B
•
Updated
Apr 9
•
2
agurung/colar-qwen25-7b-ncp-post-sft
8B
•
Updated
Apr 9
•
2
agurung/flawed-fictions-qwen3-4b-litereason
Reinforcement Learning
•
4B
•
Updated
Mar 21
agurung/flawed-fictions-qwen3-4b
Reinforcement Learning
•
4B
•
Updated
Mar 21
•
2
agurung/colar-qwen3-4b-ff-post-rl
Reinforcement Learning
•
4B
•
Updated
Mar 20
•
7
agurung/colar-qwen25-7b-ff-post-sft
8B
•
Updated
Mar 15
•
2
agurung/qwen-coconut-ff-v2
8B
•
Updated
Mar 15
•
5
agurung/ncp-qwen25-7b-lengthpenalty
Reinforcement Learning
•
8B
•
Updated
Mar 11
•
10
Previous
1
2
3
4
5
Next