llm-fine-tuning
Here are 86 public repositories matching this topic...
Collection of resources for finetuning Large Language Models (LLMs).
-
Updated
Jan 12, 2025
synthetic dataset generation workflow using local file resources for finetuning llms.
-
Updated
Oct 9, 2025 - Python
Sustain-LC is a benchmarking environment for traditional and reinforcement learning based controls as well as LLM based control
-
Updated
Aug 7, 2025 - Jupyter Notebook
Distributed Reinforcement Learning for LLM Fine-Tuning with multi-GPU utilization
-
Updated
Mar 12, 2025 - Python
The Personal Knowledge Graph You Didn’t Know You Already Wrote
-
Updated
Apr 20, 2026 - Python
Fine tune LLM with HuggingFace
-
Updated
Jun 18, 2026 - Jupyter Notebook
This repository contains code associated with Neuro-LIFT: A Neuromorphic, LLM-based Interactive Framework for Autonomous Drone FlighT at the Edge
-
Updated
Apr 25, 2025 - Python
A sacred space for heartfelt conversations, where wisdom flows freely and memories gently fade like whispers at sunset.
-
Updated
Dec 11, 2025 - HTML
Análise Avançada de Dados com Causalidade e Aprendizado por Reforço
-
Updated
Feb 27, 2025 - Jupyter Notebook
Experiments in Latin dactylic hexameter generation with transformers: A hybrid post hoc feedback framework
-
Updated
Dec 26, 2025 - Jupyter Notebook
The course teaches how to fine-tune LLMs using Group Relative Policy Optimization (GRPO)—a reinforcement learning method that improves model reasoning with minimal data. Learn RFT concepts, reward design, LLM-as-a-judge evaluation, and deploy jobs on the Predibase platform.
-
Updated
Jun 13, 2025 - Jupyter Notebook
CLI tool for generating high-quality synthetic datasets for LLM fine-tuning.
-
Updated
Apr 23, 2026 - Python
Flipper Zero Sub-GHz RF dataset (280-1100 MHz, 9 countries) with 1500 Q&A pairs for LLM fine-tuning, fact-checked allocations, and a GPU-accelerated validation pipeline (Ollama Qwen 32B + DeBERTa NLI).
-
Updated
Jun 27, 2026 - Python
Fully Connected Neural Networks, Multilayer Neural Networks, MAdaline, CNNs, Segmentation, Detection, RNNs, CNN-LSTM, LSTM, Bi-LSTM, GRU, Transformers, Huber Loss, ViT, DGMs, Triplet VAE, AdvGAN, Image Caption Generation, attention, LLM Fine-Tuning, Soft Prompting, LoRA, Layer Freezing, SlimOrca
-
Updated
Dec 25, 2025 - Jupyter Notebook
No-code desktop app for generating high-quality synthetic datasets to fine-tune LLMs — plan-then-execute pipeline, LLM-as-judge, HuggingFace upload.
-
Updated
May 20, 2026 - Python
FlowerTune LLM on Coding Dataset
-
Updated
Feb 18, 2025 - Python
High-performance Rust extensions for Axolotl (no OOM for large datasets) - drop-in acceleration for existing installations.
-
Updated
Jul 2, 2026 - Python
Process-supervised RL for a multi-step reasoning agent — DAPO + a learned Process Reward Model (PRM) training a Qwen3-8B Planner. A modern, from-scratch rebuild of the AgentFlow paper (ICLR 2026).
-
Updated
Jun 2, 2026 - Python
ARC-Test-Time-Training (ARC-TTT)
-
Updated
Jan 15, 2025 - Python
Improve this page
Add a description, image, and links to the llm-fine-tuning topic page so that developers can more easily learn about it.
Add this topic to your repo
To associate your repository with the llm-fine-tuning topic, visit your repo's landing page and select "manage topics."