llm-fine-tuning
Here are 125 public repositories matching this topic...
Collection of resources for finetuning Large Language Models (LLMs).
-
Updated
Jan 12, 2025
synthetic dataset generation workflow using local file resources for finetuning llms.
-
Updated
Oct 9, 2025 - Python
Fine tune LLM with HuggingFace
-
Updated
Jun 18, 2026 - Jupyter Notebook
Sustain-LC is a benchmarking environment for traditional and reinforcement learning based controls as well as LLM based control
-
Updated
Aug 7, 2025 - Jupyter Notebook
Distributed Reinforcement Learning for LLM Fine-Tuning with multi-GPU utilization
-
Updated
Mar 12, 2025 - Python
The Personal Knowledge Graph You Didn’t Know You Already Wrote
-
Updated
Apr 20, 2026 - Python
A curated collection of evolution strategies and zeroth-order optimization for large language models.
-
Updated
Oct 2, 2026 - TeX
This repository contains code associated with Neuro-LIFT: A Neuromorphic, LLM-based Interactive Framework for Autonomous Drone FlighT at the Edge
-
Updated
Apr 25, 2025 - Python
A sacred space for heartfelt conversations, where wisdom flows freely and memories gently fade like whispers at sunset.
-
Updated
Dec 11, 2025 - HTML
CLI tool for generating high-quality synthetic datasets for LLM fine-tuning.
-
Updated
Apr 23, 2026 - Python
Análise Avançada de Dados com Causalidade e Aprendizado por Reforço
-
Updated
Feb 27, 2025 - Jupyter Notebook
Fine-tune Qwen3.8-27B to validate Form I-9 compliance (M-274/8 CFR 274a.2), citing violations. Synthetic-only data.
-
Updated
Sep 6, 2026 - Python
Source-tracked catalog of abliterated, low-refusal, and community-reported open-weight and hosted language models
-
Updated
Oct 5, 2026
The course teaches how to fine-tune LLMs using Group Relative Policy Optimization (GRPO)—a reinforcement learning method that improves model reasoning with minimal data. Learn RFT concepts, reward design, LLM-as-a-judge evaluation, and deploy jobs on the Predibase platform.
-
Updated
Jun 13, 2025 - Jupyter Notebook
Experiments in Latin dactylic hexameter generation with transformers: A hybrid post hoc feedback framework
-
Updated
Dec 26, 2025 - Jupyter Notebook
No-code desktop app for generating high-quality synthetic datasets to fine-tune LLMs — plan-then-execute pipeline, LLM-as-judge, HuggingFace upload.
-
Updated
May 20, 2026 - Python
FlowerTune LLM on Coding Dataset
-
Updated
Feb 18, 2025 - Python
Unified Amazon SageMaker AI plugin for Codex, Kiro, and Claude Code — SDK v3, training, inference, HyperPod, monitoring, and warm pools.
-
Updated
Sep 18, 2026 - Python
Flipper Zero Sub-GHz RF dataset (280-1100 MHz, 9 countries) with 1500 Q&A pairs for LLM fine-tuning, fact-checked allocations, and a GPU-accelerated validation pipeline (Ollama Qwen 32B + DeBERTa NLI).
-
Updated
Sep 29, 2026 - Python
Add this topic to your repo
To associate your repository with the llm-fine-tuning topic, visit your repo's landing page and select "manage topics."