conradlocke/krea2-identity-edit
· image-editing, lora, comfyui
Home / 🏋️ Training / Fine-tuning
29 View all · Daily curated AI / LLM open-source intelligence, with plain-language notes and license checks.
· image-editing, lora, comfyui
text-generation · transformers, safetensors, qwen3
· lora, krea, krea-2
text-generation · transformers, safetensors, nemotron_labs_audex
· video-generation, lora, ic-lora
video-to-video · ltx-video, ic-lora, ltx-2.3
· static, region:us
· gradio, region:us
· gradio, mcp-server, region:us
Speedrunning LoRA fine-tuning: frozen task, frozen hardware, public wall-clock leaderboard. modded-nanogpt for fine-tuning.
· task_categories:text-generation, language:code, license:cc-by-4.0
· task_categories:text-generation, task_categories:question-answering, language:en
· gradio, mcp-server, region:us
Fine-tuning LLM — lora, qlora, unsloth, fine tune tutorial.
· license:mit, size_categories:100K<n<1M, format:json
从 MiniMind 源码读起,再延伸到现代大模型技术体系的中文学习笔记。主线逐行精读预训练 / SFT / DPO / PPO / GRPO 与训练机制;附录 17 篇进阶卷覆盖量化、投机解码、RLHF 全景、模型代际史等 MiniMind 没涉及、但进阶绕不开的主题。
HRM-Text is a 1B text generation model based on the HRM architecture, strengthened by task completion and latent space reasoning.
Estimate whether a Hugging Face model fits and fine-tunes on your local GPU.
Fine-tune open-source models with Tinker from inside Pi — managed improve loops, data prep, evals, smoke tests, deploy snippets, and checkpoint chat.
· language:en, language:de, language:it
LoRA fine-tune Gemma 4 31B to speak caveman-mode natively. Style: github.com/JuliusBrussee/caveman
LLM微調指南,涵蓋資料集和模型選擇。
A curated collection of papers and resources on On-Policy Distillation for Large Language Models.
🌹 Rose: Range-Of-Slice Equilibration PyTorch optimizer. Stateless optimization through range-normalized gradient updates.
The first Task-Aware MCP server and automated VRAM calculator for LLM fine-tuning. Instantly snipe the cheapest, fastest GPUs across 10+ cloud providers.
Official Codebase for "Neural Thickets: Diverse Task Experts Are Dense Around Pretrained Weights" (ICML 2026 Spotlight)
TypeScript framework for fine-tuning
一个开箱即用、用于二分类任务的大语言微调模型框架。An out-of-the-box LLM fine-tuning framework for medical binary classification.
Fine-tuned Qwen2-VL-7B for LaTeX OCR using LoRA and Unsloth on the LaTeX OCR dataset. Built augmentation pipeline (rotation, noise, contrast jitter), ran LoRA rank sweep (r=8/16/32), and evaluated across CER, Token F1, BLEU-4, and Exact Match. Deployed as a Gradio Space with live metric computation.