Home / 🏋️ Training / Fine-tuning

🏋️ Training / Fine-tuning

29 View all · Daily curated AI / LLM open-source intelligence, with plain-language notes and license checks.

Saivineeth147/lora-speedrun

Speedrunning LoRA fine-tuning: frozen task, frozen hardware, public wall-clock leaderboard. modded-nanogpt for fine-tuning.

🧱 AI Foundation Stack ✓ Commercial OK ★ 119

openbmb/UltraData-SFT-2605

· task_categories:text-generation, task_categories:question-answering, language:en

🧱 AI Foundation Stack License unclear

Anthropic/hh-rlhf

· license:mit, size_categories:100K<n<1M, format:json

🧱 AI Foundation Stack License unclear

Enping-Hu/minimind-deep-dive

从 MiniMind 源码读起,再延伸到现代大模型技术体系的中文学习笔记。主线逐行精读预训练 / SFT / DPO / PPO / GRPO 与训练机制;附录 17 篇进阶卷覆盖量化、投机解码、RLHF 全景、模型代际史等 MiniMind 没涉及、但进阶绕不开的主题。

🧱 AI Foundation Stack ✓ Commercial OK ★ 95

sapientinc/HRM-Text

HRM-Text is a 1B text generation model based on the HRM architecture, strengthened by task completion and latent space reasoning.

🧱 AI Foundation Stack ✓ Commercial OK ★ 1,710

gvkhosla/pi-tinker

Fine-tune open-source models with Tinker from inside Pi — managed improve loops, data prep, evals, smoke tests, deploy snippets, and checkpoint chat.

🧱 AI Foundation Stack ✓ Commercial OK ★ 22

JuliusBrussee/cavegemma

LoRA fine-tune Gemma 4 31B to speak caveman-mode natively. Style: github.com/JuliusBrussee/caveman

🧱 AI Foundation Stack License unclear ★ 100

MatthewK78/Rose

🌹 Rose: Range-Of-Slice Equilibration PyTorch optimizer. Stateless optimization through range-normalized gradient updates.

🧱 AI Foundation Stack ✓ Commercial OK ★ 77

hlpun/Train-in-Silence

The first Task-Aware MCP server and automated VRAM calculator for LLM fine-tuning. Instantly snipe the cheapest, fastest GPUs across 10+ cloud providers.

🧱 AI Foundation Stack ✓ Commercial OK ★ 101

sunrainyg/RandOpt

Official Codebase for "Neural Thickets: Diverse Task Experts Are Dense Around Pretrained Weights" (ICML 2026 Spotlight)

🧱 AI Foundation Stack License unclear ★ 612

arkorlab/arkor

TypeScript framework for fine-tuning

🧱 AI Foundation Stack ✓ Commercial OK ★ 30

Jackohhhh/MedLLM-Finetuning

一个开箱即用、用于二分类任务的大语言微调模型框架。An out-of-the-box LLM fine-tuning framework for medical binary classification.

🧱 AI Foundation Stack ✓ Commercial OK ★ 21

shlokchorge/Fine-Tuning-QWEN2vl

Fine-tuned Qwen2-VL-7B for LaTeX OCR using LoRA and Unsloth on the LaTeX OCR dataset. Built augmentation pipeline (rotation, noise, contrast jitter), ran LoRA rank sweep (r=8/16/32), and evaluated across CER, Token F1, BLEU-4, and Exact Match. Deployed as a Gradio Space with live metric computation.

🧱 AI Foundation Stack License unclear ★ 25