Image-Text-to-Text
Transformers
Safetensors
gemma4
text-generation-inference
unsloth
reasoning
conversational

Dataset suggestion

#1
by DFveloper - opened

Hello, I'm the big fan of this model.

I think the model's multilingual ablity is not good.
so I generated the dataset for this model, and double checked with native speaker(16y lived in Korea).
in now, the dataset is just Korean, but it will be more languages.
I hope you to train with this two datasets.
sorry for bad english, but I mean it.

Korean dataset:
https://hf.co/datasets/DFveloper/claude-opus-4.6-4.7-korean-8.7k
English dataset(not mine):
https://hf.co/datasets/angrygiraffe/claude-opus-4.6-4.7-reasoning-8.7k

p.s: I finetuned Gemma 4 26B with this datasets, but it wasn't easy. You're incredible.

I verified the dataset, and found some errors.
so, I filtered and human-verified for last 3 days. now it's perfectly fine.
https://hf.co/datasets/DFveloper/claude-opus-4.6-4.7-Korean-4k

TeichAI org

I'll include it in the next tune. will most likely downsample it to not overpower the other data. Thanks for the tip :)

I trained Gemma 4 26B with this Datasets, and got 11th place on Korean national leaderboard.
9 times of trial and error.
Total Runpod cost: over $200

avg KMMLU-Pro CLIcK HLE(Ko) MuSR(Ko) Com2-main(Ko)
53.8% 59.1% 78.6% 6.4% 60.4% 64.6%

more info on(needs translator): https://leaderboard.aihub.or.kr/leaderboard

TeichAI org

I trained Gemma 4 26B with this Datasets, and got 11th place on Korean national leaderboard.
9 times of trial and error.
Total Runpod cost: over $200

avg KMMLU-Pro CLIcK HLE(Ko) MuSR(Ko) Com2-main(Ko)
53.8% 59.1% 78.6% 6.4% 60.4% 64.6%

more info on(needs translator): https://leaderboard.aihub.or.kr/leaderboard

Thats amazing! Nice work and I understand the trial and error. Currently doing GRPO and has been 40 hours of non stop trial and error and failure then little wins.

I trained Gemma 4 26B with this Datasets, and got 11th place on Korean national leaderboard.
9 times of trial and error.
Total Runpod cost: over $200

avg KMMLU-Pro CLIcK HLE(Ko) MuSR(Ko) Com2-main(Ko)
53.8% 59.1% 78.6% 6.4% 60.4% 64.6%

more info on(needs translator): https://leaderboard.aihub.or.kr/leaderboard

Thats amazing! Nice work and I understand the trial and error. Currently doing GRPO and has been 40 hours of non stop trial and error and failure then little wins.

Wow, thanks for the kind words! It really was a journey of trial and error (and a painful Runpod bill 😂). I totally feel you on the GRPO struggles—40 hours of non-stop trials is insane. Hope those 'little wins' turn into a huge breakthrough soon! Keep pushing!

TeichAI org

The painful Runpod bill is so relatable XD. Seriously great work! @DFveloper

Sign up or log in to comment