An open API service for producing an overview of a list of open source projects.

Collections: awesome-llama

https://github.com/modelscope/ms-swift

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).

deepseek-r1 embedding grpo internvl liger llama llama4 llm lora megatron moe multimodal open-r1 peft qwen3 qwen3-6 qwen3-omni qwen3-vl reranker sft

Last synced: 02 Sep 2026

https://github.com/predibase/lorax

Multi-LoRA inference server that scales to 1000s of fine-tuned LLMs

fine-tuning gpt llama llm llm-inference llm-serving llmops lora model-serving pytorch transformers

Last synced: 02 Sep 2026

https://github.com/hiyouga/LlamaFactory

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

agent ai deepseek fine-tuning gemma gpt instruction-tuning large-language-models llama llama3 llm lora moe nlp peft qlora quantization qwen rlhf transformers

Last synced: 02 Sep 2026

https://github.com/stochasticai/xTuring

Build, personalize and control your own LLMs. From data pre-processing to fine-tuning, xTuring provides an easy way to personalize open-source LLMs. Join our discord community: https://discord.gg/TgHXuSJEk6

adapter deep-learning fine-tuning finetuning gen-ai generative-ai gpt-2 gpt-j language-model llama llm lora mistral mixed-precision peft quantization

Last synced: 02 Sep 2026

https://github.com/stochasticai/xturing

Build, personalize and control your own LLMs. From data pre-processing to fine-tuning, xTuring provides an easy way to personalize open-source LLMs. Join our discord community: https://discord.gg/TgHXuSJEk6

adapter deep-learning fine-tuning finetuning gen-ai generative-ai gpt-2 gpt-j language-model llama llm lora mistral mixed-precision peft quantization

Last synced: 02 Sep 2026

https://github.com/datawhalechina/self-llm

《开源大模型食用指南》针对中国宝宝量身打造的基于Linux环境快速微调(全参数/Lora)、部署国内外开源大模型(LLM)/多模态大模型(MLLM)教程

chatglm chatglm3 gemma-2b-it glm-4 internlm2 llama3 llm lora minicpm q-wen qwen qwen1-5 qwen2

Last synced: 02 Sep 2026

https://github.com/georgian-io/LLM-Finetuning-Toolkit

Toolkit for fine-tuning, ablating and unit-testing open-source LLMs.

ablation-study classification falcon fine-tuning finetuning flan-t5 large-language-models llama2 llm-test lora mistral-7b nlp nlp-machine-learning qlora redpajama summarization unit-testing zephyr

Last synced: 02 Sep 2026

https://github.com/LianjiaTech/BELLE

BELLE: Be Everyone's Large Language model Engine(开源中文对话大模型)

bloom chinese-nlp gpt-evaluation gpt-q instruct-finetune instruct-gpt instruction-set llama lora open-models

Last synced: 02 Sep 2026

https://github.com/ymcui/Chinese-LLaMA-Alpaca

中文LLaMA&Alpaca大语言模型+本地CPU/GPU训练部署 (Chinese LLaMA & Alpaca LLMs)

alpaca alpaca-2 large-language-models llama llama-2 llm lora nlp plm pre-trained-language-models quantization

Last synced: 02 Sep 2026

https://github.com/TUDB-Labs/mLoRA

An Efficient "Factory" to Build Multiple LoRA Adapters

baichuan chatglm dpo finetune gpu llama llama2 llm lora mlora peft rlhf

Last synced: 02 Sep 2026

https://github.com/PhoebusSi/Alpaca-CoT

We unified the interfaces of instruction-tuning data (e.g., CoT data), multiple LLMs and parameter-efficient methods (e.g., lora, p-tuning) together for easy use. We welcome open-source enthusiasts to initiate any meaningful PR on this repo and integrate as many LLM related technologies as possible. 我们打造了方便研究人员上手和使用大模型等微调平台,我们欢迎开源爱好者发起任何有意义的pr!

alpaca chatglm chatgpt cot instruction-tuning llama llm lora moss p-tuning parameter-efficient pytorch tabul tabular-data tabular-model

Last synced: 02 Sep 2026

https://github.com/zjunlp/KnowLM

An Open-sourced Knowledgable Large Language Model Framework.

bilingual chinese deep-learning deepspeed english gpt-3 instructie instruction-following instruction-tuning instructions knowlm language-model large-language-models llama lora models pre-trained-language-models pre-trained-model pre-training reasoning

Last synced: 02 Sep 2026

https://github.com/chenking2020/FindTheChatGPTer

ChatGPT爆火,开启了通往AGI的关键一步,本项目旨在汇总那些ChatGPT的开源平替们,包括文本大模型、多模态大模型等,为大家提供一些便利

agi alpaca autogpt baichuan belle ceval chatglm chatgpt codi guanaco learderboard linly llama llama2 llava lora minigpt4 self-instruct vicuna wizadlm

Last synced: 02 Sep 2026

https://github.com/ashishpatel26/LLM-Finetuning

LLM Finetuning with peft

falcon fine-tuning huggingface llama llama2 llm llms lora peft pytorch text-generation

Last synced: 02 Sep 2026

https://github.com/wenge-research/YAYI

雅意大模型:为客户打造安全可靠的专属大模型,基于大规模中英文多领域指令数据训练的 LlaMA 2 & BLOOM 系列模型,由中科闻歌算法团队研发。(Repo for YaYi Chinese LLMs based on LlaMA2 & BLOOM)

bloom chat chinese llama llama2 llm lora yayi

Last synced: 02 Sep 2026

https://github.com/soulteary/llama-docker-playground

Quick Start LLaMA models with multiple methods, and fine-tune 7B/65B with One-Click.

alpaca-lora docker llama llm lora

Last synced: 02 Sep 2026

https://github.com/ddzipp/AutoAudit

AutoAudit—— the LLM for Cyber Security 网络安全大语言模型

cyber-security fine-tuning gpt llama lora qlora security-tools

Last synced: 02 Sep 2026

https://github.com/airaria/Visual-Chinese-LLaMA-Alpaca

多模态中文LLaMA&Alpaca大语言模型(VisualCLA)

alpaca chinese llama llm lora multimodal nlp vision-language

Last synced: 02 Sep 2026

https://github.com/WangRongsheng/ChatGenTitle

🌟 ChatGenTitle:使用百万arXiv论文信息在LLaMA模型上进行微调的论文题目生成模型

arxiv large-language-models llama llm llms lora

Last synced: 02 Sep 2026

https://github.com/Chongjie-Si/Subspace-Tuning

A generalized framework for subspace tuning methods in parameter efficient fine-tuning.

adapter commonsense-reasoning glue llama llama2-7b llama3-8b lora lora-dash low-rank-adaptation natural-language-generation natural-language-processing natural-language-understanding parameter-efficient-fine-tuning pretrained-models soft-prompt-tuning subject-driven-generation subspace-tuning

Last synced: 02 Sep 2026

https://github.com/misonsky/HiFT

memory-efficient fine-tuning; support 24G GPU memory fine-tuning 7B

chinese-llama chinese-llama-65b huggingface-transformers large-language-models llama2 llama3 lora memory-efficient-tuning peft-fine-tuning-llm pytorch-implementation transformers

Last synced: 02 Sep 2026

https://github.com/GURPREETKAURJETHRA/END-TO-END-GENERATIVE-AI-PROJECTS

End to End Generative AI Industry Projects on LLM Models with Deployment_Awesome LLM Projects

chainlit finetuning-llms gemini generative-ai gpt4o gradio-python-llm huggingface langchain large-language-models llama llama-index llama3 llama3-meta-ai llm llmops lora mergekit mistral openai-api qlora

Last synced: 02 Sep 2026

https://github.com/ukairia777/tensorflow-nlp-tutorial

tensorflow를 사용하여 텍스트 전처리부터, Topic Models, BERT, GPT, LLM과 같은 최신 모델의 다운스트림 태스크들을 정리한 Deep Learning NLP 저장소입니다.

bert bert-ner dpo huggingface keras-tutorial llama llm lora named-entity-recognition natural-language-processing nlp nlp-tutorial question-answering sft tensorflow trainer transformers

Last synced: 02 Sep 2026

https://github.com/Longyichen/Alpaca-family-library

Summarize all open source Large Languages Models and low-cost replication methods for Chatgpt.

adapter alpaca chatgpt datasets large-language-models llama lora

Last synced: 02 Sep 2026

https://github.com/jasonvanf/llama-trl

LLaMA-TRL: Fine-tuning LLaMA with PPO and LoRA

adapter chatgpt gpt gpt-4 llama lora peft ppo rlhf transformer trl

Last synced: 02 Sep 2026

https://github.com/Gunale0926/SORSA

SORSA: Singular Values and Orthonormal Regularized Singular Vectors Adaptation of Large Language Models

deep-learning fine-tuning llama lora machine-learning nlp peft python pytorch rwkv sorsa svd transformer

Last synced: 02 Sep 2026

https://github.com/Joyce94/LLM-RLHF-Tuning

LLM Tuning with PEFT (SFT+RM+PPO+DPO with LoRA)

fine-tuning language-model llama llm lora peft ppo reinforcement-learning rlhf

Last synced: 02 Sep 2026

https://github.com/yangjianxin1/Firefly-LLaMA2-Chinese

Firefly中文LLaMA-2大模型,支持增量预训练Baichuan2、Llama2、Llama、Falcon、Qwen、Baichuan、InternLM、Bloom等大模型

baichaun2 baichuan baichuan-13b bloom chatglm falcon firefly internlm llama llama-2 llama2 llm lora pretrain qlora qwen xverse

Last synced: 02 Sep 2026

https://github.com/abhi-arya1/tuna

fine tuning, reimagined. welcome to tuna 🎣 - we're simplifying cloud compute architecture, datasets, and more, to get your specialized AI from 0->100 asap

ai cli code-generation fine-tune generative-ai generative-code llama lora python

Last synced: 02 Sep 2026

https://github.com/git-cloner/llama-lora-fine-tuning

llama fine-tuning with lora

finetuning llama lora

Last synced: 02 Sep 2026

https://github.com/mlpc-ucsd/BLIVA

(AAAI 2024) BLIVA: A Simple Multimodal LLM for Better Handling of Text-rich Visual Questions

blip2 bliva chatbot instruction-tuning llama llm lora multimodal visual-language-learning

Last synced: 02 Sep 2026

https://github.com/jackaduma/Vicuna-LoRA-RLHF-PyTorch

A full pipeline to finetune Vicuna LLM with LoRA and RLHF on consumer hardware. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the Vicuna architecture. Basically ChatGPT but with Vicuna

chatgpt finetune gpt llama llm lora peft ppo pytorch reward-models rlhf vicuna vicuna-7b

Last synced: 02 Sep 2026

https://github.com/taishan1994/Llama3.1-Finetuning

对llama3进行全参微调、lora微调以及qlora微调。

llama3 lora qlora qwen

Last synced: 02 Sep 2026

https://github.com/ziwang-com/AGM

AGM阿格姆:AI基因图谱模型,从token-weight权重微粒角度,探索AI模型,GPT\LLM大模型的内在运作机制。

agi agm gene genemap gpt llama2 llm lora model tensor token weight

Last synced: 02 Sep 2026

https://github.com/ssbuild/llm_finetuning

Large language Model fintuning bloom , opt , gpt, gpt2 ,llama,llama-2,cpmant and so on

adalora bloom cpmant gpt gpt2 llama llama2 lora mistral opt qlora

Last synced: 02 Sep 2026

https://github.com/git-cloner/llama2-lora-fine-tuning

llama2 finetuning with deepspeed and lora

deepspeed finetuning llama2 lora

Last synced: 02 Sep 2026

https://github.com/monk1337/auto-ollama

run ollama & gguf easily with a single command

autogguf autoollama gguf inference llama llm llm-inference lora mergelora mistral ollama openai

Last synced: 02 Sep 2026

https://github.com/ASSERT-KTH/repairllama

RepairLLaMA: Efficient Representations and Fine-Tuned Adapters for Program Repair http://arxiv.org/pdf/2312.15698

apr codellama llama llms lora repair

Last synced: 02 Sep 2026

https://github.com/eliahuhorwitz/Spectral-DeTuning

Official PyTorch Implementation for the "Recovering the Pre-Fine-Tuning Weights of Generative Models" paper (ICML 2024).

deep-learning jailbreak llama2 llm lora machine-learning mistral stable-diffusion weight-space-learning

Last synced: 02 Sep 2026

https://github.com/A-baoYang/alpaca-7b-chinese

Finetune LLaMA-7B with Chinese instruction datasets

alpaca chatgpt deep-learning fine-tuning instruction-following llm lora nlp pytorch

Last synced: 02 Sep 2026

https://github.com/jackaduma/ChatGLM-LoRA-RLHF-PyTorch

A full pipeline to finetune ChatGLM LLM with LoRA and RLHF on consumer hardware. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the ChatGLM architecture. Basically ChatGPT but with ChatGLM

chatglm chatglm-6b chatgpt deepspeed finetune gpt llama llm lora peft ppo pytorch reward-models rlhf

Last synced: 02 Sep 2026

https://github.com/Abbey4799/CuteGPT

An open-source conversational language model developed by the Knowledge Works Research Laboratory at Fudan University.

chatgpt chinese-nlp deep-learning dialogue-systems instruction-tuning large-language-models llama llm lora natual-language-processing nlp text-generation

Last synced: 02 Sep 2026

https://github.com/leehanchung/SMILE-factory

Finetune Falcon, LLaMA, MPT, and RedPajama on consumer hardware using PEFT LoRA

agi falcon gpt llama llm lora mpt nlp redpajama

Last synced: 02 Sep 2026

https://github.com/ziwang-com/zero-lora

zero零训练llm调参

gpt gptq llama llm lora

Last synced: 02 Sep 2026

https://github.com/jwliao1209/TWLLM-Tutor

📘 Taiwan-LLM Tutor: Large Language Models for Taiwanese Secondary Education

bert fine-tuning llama llm lora

Last synced: 02 Sep 2026

https://github.com/taishan1994/qlora-chinese-LLM

使用qlora对中文大语言模型进行微调,包含ChatGLM、Chinese-LLaMA-Alpaca、BELLE

alpaca belle bloomz chatglm llama lora qlora

Last synced: 02 Sep 2026

https://github.com/poteminr/instruct-ner

Instruct LLMs for flat and nested NER. Fine-tuning Llama and Mistral models for instruction named entity recognition. (Instruction NER)

alpaca flat-ner llama llama2 llamacpp lora mistral-7b named-entity-recognition ner

Last synced: 02 Sep 2026

https://github.com/xyjigsaw/LLM-Pretrain-SFT

Scripts of LLM pre-training and fine-tuning (w/wo LoRA, DeepSpeed)

baichuan2 deepspeed large-language-models llama lora mistral

Last synced: 02 Sep 2026

https://github.com/billvsme/train_law_llm

✏️0成本LLM微调上手项目,⚡️一步一步使用colab训练法律LLM,基于microsoft/phi-1_5、chatglm3,包含lora微调,全参微调

ai deepspeed law llama2 llm lora python

Last synced: 02 Sep 2026

https://github.com/5663015/LLMs_train

一套代码指令微调大模型

baichuan bloom chatglm-6b deepspeed language-model llama llm-training llms lora pythia

Last synced: 02 Sep 2026

https://github.com/ziwang-com/zwPython

## zw-GPT-stuido智王GPT研发平台,全球首个集成式GPT开发平台。 zwPython-AI优化升级完成。 采用苹果MAC电脑开箱即用模式,无需安装,解压即用。

gpt llama llm lora python

Last synced: 02 Sep 2026

https://github.com/jackaduma/Alpaca-LoRA-RLHF-PyTorch

A full pipeline to finetune Alpaca LLM with LoRA and RLHF on consumer hardware. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the Alpaca architecture. Basically ChatGPT but with Alpaca

alpaca chatgpt deepspeed finetune gpt llama llm lora peft ppo pytorch reward-models rlhf

Last synced: 02 Sep 2026

https://github.com/adithya-s-k/CompanionLLM

CompanionLLM - A framework to finetune LLMs to be your own sentient conversational companion

fine-tuning finetuning hacktoberfest hacktoberfest-accepted hacktoberfest2023 huggingface llama llama2 llamacpp llm llm-inference llm-training lora mit-license open-source peft

Last synced: 02 Sep 2026

https://github.com/WangRongsheng/Chinese-LLaMA-Alpaca-Usage

📔 对Chinese-LLaMA-Alpaca进行使用说明和核心代码注解

alpaca fine-tuning large-language-models llama llm llms lora pre-trained-language-models webui

Last synced: 02 Sep 2026

https://github.com/lxe/llama-peft-tuner

Tune LLaMa-7B on Alpaca Dataset using PEFT / LORA Based on @zphang's https://github.com/zphang/minimal-llama scripts.

llama llm lora machine-learning pytorch

Last synced: 02 Sep 2026

https://github.com/chuangchuangtan/LLaVA-NeXT-Image-Llama3-Lora

LLaVA-NeXT-Image-Llama3-Lora, Modified from https://github.com/arielnlee/LLaVA-1.6-ft

finetuning llama3 llava-next lora

Last synced: 02 Sep 2026

https://github.com/ziwang-com/mini-AGI

GPT+神器,简单实用的一站式AGI架构,内置本地化,LLM模型,agent,矢量数据库,智能链chain

agi chatglm gpt llama lora vicuan vicuna

Last synced: 02 Sep 2026

https://github.com/Tommy-s-Online-Courses/LLM

大模型LLM系列 课程资料

llama llama3 llm lora

Last synced: 02 Sep 2026

https://github.com/RangiLyu/llama.mmengine

Training LLaMA language model with MMEngine! It supports LoRA fine-tuning!

alpaca fine-tuning language-model llama lora nlp

Last synced: 02 Sep 2026

https://github.com/Jiacheng-Zhu-AIML/AsymmetryLoRA

Preprint: Asymmetry in Low-Rank Adapters of Foundation Models

bert fine-tuning foundation-models llama llm lora peft

Last synced: 02 Sep 2026

https://github.com/remixer-dec/botality-ii

telegram bot for self-hosted local inference of stable diffusion, text-to-speech and large language models, such as llama3

ai alpaca gpt-2 gpt-j llama llama3 llamacpp lora m1-mac mps multimodal self-hosted stable-diffusion stt telegram-bot tta tts

Last synced: 02 Sep 2026

https://github.com/wangermeng2021/llm-webui

A Gradio web UI for Large Language Models. Supports LoRA/QLoRA finetuning,RAG(Retrieval-augmented generation) and Chat

finetune-llms finetuning-large-language-models finetuning-llms large-language-models llama2 llm-web-ui llms lora mistral-7b qlora rag retrieval-augmented-generation webui zaphyr

Last synced: 02 Sep 2026

https://github.com/avocardio/Zicklein

Finetuning instruct-LLaMA on german datasets.

alpaca finetuning german ggml language-model llama llama2 llm lora meta

Last synced: 02 Sep 2026

https://github.com/LeVuMinhHuy/brocode

a bro who codes with you

code-generation code-summarization huggingface instruction-tuning llama2 lora peft transformer typescript

Last synced: 02 Sep 2026

https://github.com/gauss5930/AlpaGasus2-QLoRA

This is AlpaGasus2-QLoRA based on LLaMA2 with AlpaGasus mechanism using QLoRA!

alpaca huggingface huggingface-transformers llama2 lora parameter-efficient-tuning peft qlora transformer

Last synced: 02 Sep 2026

https://github.com/camenduru/alpaca-lora-colab

Alpaca Lora

ai alpaca colab colab-notebook colaboratory llama llm lora

Last synced: 02 Sep 2026

https://github.com/Miraclemarvel55/LLaMA-MOSS-RLHF-LoRA

用RLHF可选LoRA对LLaMA和MOSS进行训练|Training LLaMA or MOSS with RLHF [LoRA]

chinese llama lora moss ppo reward rl rlhf similarity

Last synced: 02 Sep 2026

https://github.com/dasdristanta13/LLM-Lora-PEFT_accumulate

LLM-Lora-PEFT_accumulate explores optimizations for Large Language Models (LLMs) using PEFT, LORA, and QLORA. Contribute experiments and implementations to enhance LLM efficiency. Join discussions and push the boundaries of LLM optimization. Let's make LLMs more efficient together!

alpaca bitsandbytes falcon int8 llama llm lora peft qlora

Last synced: 02 Sep 2026

https://github.com/louisc-s/QLoRA-Fine-tuning-for-Film-Character-Styled-Responses-from-LLM

Code for fine-tuning Llama2 LLM with custom text dataset to produce film character styled responses

chatbot deep-learning finetuning-llms generative-ai llama2 lora parameter-efficient-tuning peft qlora

Last synced: 02 Sep 2026

https://github.com/aman-17/MediSOAP

FineTuning LLMs on conversational medical dataset.

fine-tuning generative-ai llama llama-2 llm-training lora medical peft peft-fine-tuning-llm qlora summarization

Last synced: 02 Sep 2026

https://github.com/monk1337/NanoPeft

The simplest repository & Neat implementation of different Lora methods for training/fine-tuning Transformer-based models (i.e., BERT, GPTs). [ Research purpose ]

huggingface llama llm lora low-rank-adaptation mistral peft qlora quantization

Last synced: 02 Sep 2026

https://github.com/Pawandeep-prog/finetune-llm-lora-guide

llama llm lora peft transformers

Last synced: 02 Sep 2026

https://github.com/soniawmeyer/WanderChat

A Comparison of LLM Chat Bot Implementation Methods with Travel Use Case

ai-engineering chatbot fine-tuning llama llm-training llms lora machine-learning mistral qlora rag rlhf sjsu travel

Last synced: 02 Sep 2026

https://github.com/HEMANGANI/LLM-Recommendation-Systems

This project fine-tunes large language models (LLMs) for text-based recommendations, using a novel prompt mechanism to improve accuracy and user satisfaction. It demonstrates efficient model adaptation with diverse datasets, leveraging advanced libraries and techniques for optimal performance.

llama llm llm-recommendation lora mistral qlora recommendation-system

Last synced: 02 Sep 2026

https://github.com/hululuzhu/llama-lora-chinese-couplet

llama-lora e2e example to demo a Chinese Couplet AI in 10 mins. some thoughts on high quality chat AI

chinese couplet llama llm lora

Last synced: 02 Sep 2026

https://github.com/tien02/llm-math

Fine tune Large Language Model on Mathematic dataset

huggingface llama llama2 llm lora mathematics supervised-finetuning transformer

Last synced: 02 Sep 2026

https://github.com/EvilFreelancer/MoDA

Is a framework designed to enhance the performance and flexibility of large language models by dynamically selecting and integrating specialized LoRA adapters based on the input query.

adapters gpt llama2 lora router transformers

Last synced: 02 Sep 2026

https://github.com/Coldwave96/llama-honeypot

A honeypot backend API based on fine-tuning LLaMA via LoRA.

fine-tuning llama lora

Last synced: 02 Sep 2026

https://github.com/minggnim/fine-tune-llms

Fine tuning approaches for LLMs

fine-tuning llama2 llms lora

Last synced: 02 Sep 2026

https://github.com/heisemind/qlora-llama

This project enables fine-tuning using QLoRA for question-answering Llama 2 model.

huggingface llama llama2 lora python

Last synced: 02 Sep 2026

https://github.com/KalbeDigitalLab/ALPARA-TUTORIAL-PRICAI-2023

This repository contains hands on code for tutorials on PRICAI 2023 with the topics Developing Open Source Large Language Model using LLaMA and Alpaca

alpaca large-language-models llama lora

Last synced: 02 Sep 2026

https://github.com/zetavg/llama-lora-tuner

UI tool for fine-tuning and testing your own LoRA models base on LLaMA, GPT-J and more. One-click run on Google Colab. + A Gradio ChatGPT-like Chat UI to demonstrate your language models.

ai alpaca alpaca-lora google-colab gpt gpt-j language-model llama lora machine-learning peft

Last synced: 02 Sep 2026

https://github.com/wxjiao/ParroT

The ParroT framework to enhance and regulate the Translation Abilities during Chat based on open-sourced LLMs (e.g., LLaMA-7b, Bloomz-7b1-mt) and human written translation and evaluation data.

bloomz chatgpt contrastive error-guided gpt-4 human-feedback instruction-tuning llama lora machine-translation

Last synced: 02 Sep 2026

https://github.com/l294265421/chat-sentiment-analysis

Solve all sentiment analysis tasks in chat-style by finetuning LLaMA with lora

absa aspect-based-sentiment-analysis aspect-opinion-pair-extraction aspect-sentiment-opinion-triplet aspect-term-extraction chatgpt llama lora sentiment-analysis

Last synced: 02 Sep 2026

https://github.com/nopperl/Zicklein-GGML

German alpaca finetune converted to GGML format (compatible with llama.cpp).

alpaca finetune german ggml llama lora

Last synced: 02 Sep 2026

https://github.com/manthan410/finetune-llama2-for-docker_command

llama-2 model finetuned to generate docker commands

fine-tuning large-language-models llama2 lora peft qlora

Last synced: 02 Sep 2026

https://github.com/oldgrev/coco-LoRA

evaluating LoRA for training small amounts of data dynamically

llama lora

Last synced: 02 Sep 2026

https://github.com/justdepie/MSc-Thesis-From-Tables-to-Natural-Language-Summaries

Thesis scope: Train and Develop a Table-to-Text Transformer-based model for contextual summarization of tabular data. To achieve this T5-small , T5-base, Bart-base and Llama2 7B chat were finetuned on ToTTo and QTSumm. Regarding ToTTo, the models outperformed the benchmark.

bart-base huggingface-transformers llama2 lora peft seq2seq-model summarization t5-base t5-small table-to-text transformer wikipedia

Last synced: 02 Sep 2026

https://github.com/yangjianxin1/Firefly

Firefly: 大模型训练工具,支持训练Qwen2.5、Qwen2、Yi1.5、Phi-3、Llama3、Gemma、MiniCPM、Yi、Deepseek、Orion、Xverse、Mixtral-8x7B、Zephyr、Mistral、Baichuan2、Llma2、Llama、Qwen、Baichuan、ChatGLM2、InternLM、Ziya2、Vicuna、Bloom等大模型

alpaca aquila baichuan chatglm gemma gpt internlm llama llama2 llama3 llm lora minicpm mistral mixtral peft qlora qwen qwen2 zephyr

Last synced: 02 Sep 2026

https://github.com/yangjianxin1/firefly

Firefly: 大模型训练工具,支持训练Qwen2.5、Qwen2、Yi1.5、Phi-3、Llama3、Gemma、MiniCPM、Yi、Deepseek、Orion、Xverse、Mixtral-8x7B、Zephyr、Mistral、Baichuan2、Llma2、Llama、Qwen、Baichuan、ChatGLM2、InternLM、Ziya2、Vicuna、Bloom等大模型

alpaca aquila baichuan chatglm gemma gpt internlm llama llama2 llama3 llm lora minicpm mistral mixtral peft qlora qwen qwen2 zephyr

Last synced: 02 Sep 2026

https://github.com/zetavg/LLaMA-LoRA-Tuner

UI tool for fine-tuning and testing your own LoRA models base on LLaMA, GPT-J and more. One-click run on Google Colab. + A Gradio ChatGPT-like Chat UI to demonstrate your language models.

ai alpaca alpaca-lora google-colab gpt gpt-j language-model llama lora machine-learning peft

Last synced: 02 Sep 2026

Statistics

  • Projects: 2,239
  • Last updated: about 2 years ago