AI Radar
Support
LiveUpdated 2026-09-21 17:06 UTC

I am finally learning how to fine-tune models using @UnslothAI doing a "hello world" test.

I am finally learning how to fine-tune models using @UnslothAI doing a "hello world" test. Base model ->…I am finally learning how to fine-tune models using @UnslothAI doing a "hello world" test. Base model ->…

This is a AI post classified by Jev as AI infra & evals (a tutorial), kept by the AI Radar because it carries real work, not commentary.

I am finally learning how to fine-tune models using @UnslothAI doing a "hello world" test. Base model -> unsloth/Qwen3-4B-Instruct-2507 It's a 4B params, ~8GB on disk. It's a good model for learn as you can train it relatively fast and see results immediately. training data -> mlabonne/FineTome-100k It's 100k ChatML conversations, cleaned and curated (a well-known community SFT set). That's what I am "teaching" the model: high-quality Q&A / explanation-style conversations. What's LoRA? That's the core trick. Instead of updating all 4B weights (needs ~4×8GB+ for gradients/optimizer - won

Posted by Javier • priv/acc (4.4k followers) 1 h ago · 5 likes · 144 views · view the original post on X. Kept by the AI Radar as AI infra & evals.

More AI work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 39.3k posts from 5k X accounts over the last 14 days, 1.6k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-21 17:06 UTC. Full method.