Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🏗️
Building on HF
9.2
TFLOPS
Sergio Paniego
PRO
sergiopaniego
234
205
131
Follow
hshetty's profile picture
caraman's profile picture
BasitMustafa's profile picture
1,973 followers
·
133 following
https://sergiopaniego.github.io/
sergiopaniego
sergiopaniego
sergio-paniego-blanco
AI & ML interests
None yet
Recent Activity
updated
a bucket
about 3 hours ago
FineEnvs/watercolour-trackio-hps-only-bucket
upvoted
an
article
about 8 hours ago
tokenizers v1: encode, decode and scaling, measured
updated
a dataset
about 14 hours ago
sergiopaniego/opencode-rollout-trace
View all activity
Organizations
sergiopaniego
's models
141
Sort: Recently updated
sergiopaniego/qwen3-1.7b-mbpp-grpo
Text Generation
•
2B
•
Updated
11 days ago
•
494
sergiopaniego/rl-envs-youtube-livestream-4-scripts
Updated
11 days ago
sergiopaniego/qwen3-1.7b-wordle-grpo
Text Generation
•
2B
•
Updated
Aug 19
•
152
sergiopaniego/Qwen3-4B-claude-code-local-grpo
Text Generation
•
4B
•
Updated
Jul 31
•
12
sergiopaniego/Qwen3-4B-claude-code-deepcoder-grpo
Text Generation
•
4B
•
Updated
Jul 31
•
12
sergiopaniego/pelican-svg-grpo-Qwen3-1.7B-judged
Text Generation
•
2B
•
Updated
Jul 29
•
11
sergiopaniego/pelican-svg-grpo-Qwen3-1.7B
Text Generation
•
2B
•
Updated
Jul 29
•
15
sergiopaniego/Qwen3-8B-opencode-deepcoder-grpo
Text Generation
•
8B
•
Updated
Jul 28
•
24
•
1
sergiopaniego/grpo-youtube-livestream-3-scripts
Reinforcement Learning
•
Updated
Jul 27
•
2
sergiopaniego/Qwen3.5-4B-sdpo-math-hints
Updated
Jul 10
•
1
sergiopaniego/Qwen3.5-4B-sdpo-math-gold
Updated
Jul 10
sergiopaniego/Qwen3.5-4B-sdpo-math-baseline
Updated
Jul 10
sergiopaniego/sdpo-hints
Updated
Jul 10
sergiopaniego/pi-mono-youtube-livestream-2-scripts
Updated
Jul 6
•
2
sergiopaniego/gemma-4-E2B-offpolicy-kd-lr1e4
Updated
Jul 6
sergiopaniego/gemma-4-E2B-offpolicy-kd-lr2e4
Updated
Jul 6
sergiopaniego/gemma-4-E2B-offpolicy-kd-lr5e5
Updated
Jul 6
sergiopaniego/qwen3-0.6b-pimono-gkd-lr2e5
Text Generation
•
0.6B
•
Updated
Jul 2
•
13
sergiopaniego/qwen3-0.6b-pimono-gkd-lr1e5
Text Generation
•
0.6B
•
Updated
Jul 2
•
13
sergiopaniego/qwen3-0.6b-pimono-gkd-lr5e5
Text Generation
•
0.6B
•
Updated
Jul 2
•
13
sergiopaniego/qwen3-0.6b-pimono-logit-kd-lr1e5
Text Generation
•
0.6B
•
Updated
Jul 2
•
9
sergiopaniego/qwen3-0.6b-pimono-logit-kd-lr2e5
Text Generation
•
0.6B
•
Updated
Jul 2
•
11
sergiopaniego/qwen3-0.6b-pimono-logit-kd-lr5e5
Text Generation
•
0.6B
•
Updated
Jul 2
•
9
sergiopaniego/Qwen2.5-0.5B-Instruct-text-to-sql-qlora
Updated
Jun 15
sergiopaniego/browsergym-grpo-functiongemma-270m-it
Text Generation
•
0.3B
•
Updated
May 29
•
18
•
2
sergiopaniego/qwen3-grpo-requests
Updated
May 19
sergiopaniego/reasoning-gym-chain-sum-Qwen3-1.7B-sft
Text Generation
•
2B
•
Updated
May 4
•
10
sergiopaniego/reasoning-gym-chain-sum-Qwen3-1.7B
Text Generation
•
2B
•
Updated
Apr 28
•
17
sergiopaniego/carla-vlm-gemma-test
Updated
Apr 15
sergiopaniego/carla-vlm-qwen35-test
Updated
Apr 13
Previous
1
2
3
...
5
Next