Collections: awesome-llama
https://github.com/mallorbc/Finetune_LLMs
Repo for fine-tuning Casual LLMs
docker falcon gpt gpt-3 gpt-35-turbo gpt-4 gpt-j-6b llama llama2 llm llm-training mpt
Last synced: 12 Aug 2026
https://github.com/riccardomusmeci/mlx-llm
Large Language Models (LLMs) applications and tools running on Apple Silicon in real-time with Apple MLX.
llama llm mistral mlx phi transformers
Last synced: 12 Aug 2026
https://github.com/SeungyounShin/Llama2-Code-Interpreter
Make Llama2 use Code Execution, Debug, Save Code, Reuse it, Access to Internet
codeinterpreter codellama llama llm
Last synced: 12 Aug 2026
https://github.com/bigsk1/voice-chat-ai
🎙️ Speak with AI - Run locally using Ollama, OpenAI, Anthropic or xAI - Speech uses SparkTTS, OpenAI, ElevenLabs, Kokoro, Typecast or xAI
ai-speech ai-voice ai-voice-agent anthropic-claude conversational-ai elevenlabs-api fastapi ollama selfhosted tts typecast voice-ai webrtc whisper-ai xai xai-tts
Last synced: 12 Aug 2026
https://github.com/galatolofederico/vanilla-llama
Plain pytorch implementation of LLaMA
llama llama-inference-server llama-pytorch
Last synced: 12 Aug 2026
https://github.com/interestingLSY/swiftLLM
A tiny yet powerful LLM inference system tailored for researching purpose. vLLM-equivalent performance with only 2k lines of code (2% of vLLM).
cuda gpt inference inference-engine llama llm llm-inference llm-serving llmops mlops model-serving pytorch transformer transformers
Last synced: 12 Aug 2026
https://github.com/calcuis/gguf-core
a simple way to interact llama with gguf
cli core gguf gui llama metadata pdf reader vision wav
Last synced: 13 Aug 2026
https://github.com/RLHF-V/RLHF-V
[CVPR'24] RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback
chatbot gpt-4 llama multi-modality multimodal rlhf-v visual-language-learning
Last synced: 12 Aug 2026
https://github.com/misonsky/HiFT
memory-efficient fine-tuning; support 24G GPU memory fine-tuning 7B
chinese-llama chinese-llama-65b huggingface-transformers large-language-models llama2 llama3 lora memory-efficient-tuning peft-fine-tuning-llm pytorch-implementation transformers
Last synced: 12 Aug 2026
https://github.com/neph1/LlamaTale
Giving the power of LLM's to a MUD lib.
generative-ai interactive-fiction large-language-models llama llm mud roleplaying
Last synced: 12 Aug 2026
https://github.com/jankais3r/LLaMA_MPS
Run LLaMA (and Stanford-Alpaca) inference on Apple Silicon GPUs.
alpaca apple-silicon chat chatbot chatgpt llama llms macos metal ml mps stanford-alpaca torch
Last synced: 12 Aug 2026
https://github.com/xNul/code-llama-for-vscode
Use Code Llama with Visual Studio Code and the Continue extension. A local LLM alternative to GitHub Copilot.
assistant code code-llama codellama continue continuedev copilot llama llama2 llamacpp llm local meta ollama studio visual vscode
Last synced: 12 Aug 2026
https://github.com/kaarthik108/snowChat
Chat snowflake - Text to SQL
agents chatgpt langchain langgraph llama snowflake snowpark streamlit supabase
Last synced: 12 Aug 2026
https://github.com/janelu9/EasyLLM
Running Large Language Model easily.
deepseek deepspeed fine-tuning llama megatron npu pretrain qwen qwen-vl rlhf vllm
Last synced: 12 Aug 2026
https://github.com/ChuloAI/BrainChulo
Harnessing the Memory Power of the Camelids
chromadb fastapi langchain llama llm microsoft-guidance python retrieval-augmented sqlmodel vector-store
Last synced: 12 Aug 2026
https://github.com/sae-llm-coconut/coconut-ai
Python library that ease the installation process of Stable Diffusion, and allows to genrate images with a nice to use API.
llama2 python-library stable-diffusion
Last synced: 12 Aug 2026
https://github.com/WangRongsheng/CareGPT
🌞 CareGPT (关怀GPT)是一个医疗大语言模型,同时它集合了数十个公开可用的医疗微调数据集和开放可用的医疗大语言模型,包含LLM的训练、测评、部署等以促进医疗LLM快速发展。Medical LLM, Open Source Driven for a Healthy Future.
baichuan gpt large-language-models llama llama2 medical-llm
Last synced: 12 Aug 2026
https://github.com/kyegomez/M2PT
Implementation of M2PT in PyTorch from the paper: "Multimodal Pathway: Improve Transformers with Irrelevant Data from Other Modalities"
ai attention attention-is-all-you-need gpt4 gpt5 llama ml models mulit-modality multi-modal
Last synced: 12 Aug 2026
https://github.com/hpcaitech/SwiftInfer
Efficient AI Inference & Serving
artificial-intelligence deep-learning gpt inference llama llama2 llm-inference llm-serving
Last synced: 12 Aug 2026
https://github.com/alexeichhorn/typegpt
Make GPT safe for production
gpt llama llm openai prompt-engineering
Last synced: 12 Aug 2026
https://github.com/airaria/Visual-Chinese-LLaMA-Alpaca
多模态中文LLaMA&Alpaca大语言模型(VisualCLA)
alpaca chinese llama llm lora multimodal nlp vision-language
Last synced: 12 Aug 2026
https://github.com/mbzuai-oryx/VideoGPT-plus
Official Repository of paper VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding
chatbot clip dual-encoder gpt4 gpt4o image-encoder llama3 llava multimodal phi-3-mini vicuna video-chatbot video-conversation video-encoder vision-language vision-language-pretraining
Last synced: 12 Aug 2026
https://github.com/HenryNdubuaku/nanodl
JAX library for training sub-4B foundation models for edge
attention attention-mechanism deep-learning distributed-training flax gpt jax llama machine-learning mistral nlp transformer
Last synced: 12 Aug 2026
https://github.com/stanleylsx/llms_tool
一个基于HuggingFace开发的大语言模型训练、测试工具。支持各模型的webui、终端预测,低参数量及全参数模型训练(预训练、SFT、RM、PPO、DPO)和融合、量化。
aquila aquila2 baichuan baichuan2 bloom chatglm chatglm2 chatglm3 deepspeed falcon internlm llama llama2 mistral moss pytorch qwen xverse
Last synced: 12 Aug 2026
https://github.com/okuvshynov/slowllama
Finetune llama2-70b and codellama on MacBook Air without quantization
apple-silicon fine-tuning llama llama2
Last synced: 12 Aug 2026
https://github.com/kyegomez/GATS
Implementation of GATS from the paper: "GATS: Gather-Attend-Scatter" in pytorch and zeta
ai attention attention-is-all-you-need attention-mechanism gpt4 llama ml multi-modal multi-modality multimodal open-source
Last synced: 12 Aug 2026
https://github.com/YJ-20/auto-subtitle-translate
Automatically generate, translate, and overlay subtitles for any video.
ai ai-subtitle automatic-subtitle deep-learning ffmpeg llama llama2 python subtitle-generator subtitles subtitles-generator translates translator whisper
Last synced: 12 Aug 2026
https://github.com/josStorer/selfhostedAI
A collection of one-click self-hosted AI
chatglm chatgpt llama llm stable-diffusion
Last synced: 13 Aug 2026
https://github.com/WangRongsheng/ChatGenTitle
🌟 ChatGenTitle:使用百万arXiv论文信息在LLaMA模型上进行微调的论文题目生成模型
arxiv large-language-models llama llm llms lora
Last synced: 12 Aug 2026
https://github.com/Gary3410/TaPA
[arXiv 2023] Embodied Task Planning with Large Language Models
ai2thor embodied-agent llama robotics
Last synced: 13 Aug 2026
https://github.com/luogen1996/LLaVA-HR
[ICLR2025] LLaVA-HR: High-Resolution Large Language-Vision Assistant
high-resolution-mllms llama2 llms multimodal-chatbot
Last synced: 12 Aug 2026
https://github.com/holoviz-topics/panel-chat-examples
Examples of Chat Bots using Panels chat features: Traditional, LLMs, AI Agents, LangChain, OpenAI etc
chat gpt llama mistral openai panel python
Last synced: 12 Aug 2026
https://github.com/alexfazio/OpenPlexity-Pages
SearchGPT / Perplexity Pages clone, but personalised for you.
crewai groq llama3 search-engine streamlit
Last synced: 12 Aug 2026
https://github.com/taesiri/ArXivQA
WIP - Automated Question Answering for ArXiv Papers with Large Language Models (https://arxiv.taesiri.xyz/)
arxiv arxiv-daily arxiv-dataset arxiv-papers arxiv-preprint automated-qa claude claude2 gpt gpt-4 llama llama2 llm question-answering
Last synced: 12 Aug 2026
https://github.com/OneInterface/realtime-bakllava
llama.cpp with BakLLaVA model describes what does it see
bakllavva cpp demo-application inference llama llamacpp llm
Last synced: 12 Aug 2026
https://github.com/Chongjie-Si/Subspace-Tuning
A generalized framework for subspace tuning methods in parameter efficient fine-tuning.
adapter commonsense-reasoning glue llama llama2-7b llama3-8b lora lora-dash low-rank-adaptation natural-language-generation natural-language-processing natural-language-understanding parameter-efficient-fine-tuning pretrained-models soft-prompt-tuning subject-driven-generation subspace-tuning
Last synced: 12 Aug 2026
https://github.com/innightwolfsleep/llm_telegram_bot
LLM telegram bot
aiogram bot exllama large-language-models llamacpp llm telegram telegram-bot text-generation-webui
Last synced: 12 Aug 2026
https://github.com/simulatrex/simulatrex-engine
Enable decision-making based on simulations
chatgpt generative-ai gpt-4 llama2 pypi python simaas simulations
Last synced: 12 Aug 2026
https://github.com/JosemyDuarte/gpt-researcher-ollama
Based on assafelovic/gpt-researcher - Modified to support local Ollama models
gpt gpt-researcher llama llama3 ollama
Last synced: 12 Aug 2026
https://github.com/BIDS-Xu-Lab/Me-LLaMA
A novel medical large language model family with 13/70B parameters, which have SOTA performances on various medical tasks
biomedical-nlp biomedical-text-mining chatgpt clinical-nlp gpt-4 large-language-models llama llama2 llm medical- medical-application medical-large-language-models medical-question-answering
Last synced: 12 Aug 2026
https://github.com/jerry1993-tech/Cornucopia-LLaMA-Fin-Chinese
聚宝盆(Cornucopia): 中文金融系列开源可商用大模型,并提供一套高效轻量化的垂直领域LLM训练框架(Pretraining、SFT、RLHF、Quantize等)
chinese finance large-language-models llama nlp qa rlhf sft text-generation transformers
Last synced: 12 Aug 2026
https://github.com/amithkoujalgi/ollama-pdf-bot
A bot that accepts PDF docs and lets you ask questions on it.
bot chat-bot llama llama2 llm ollama pdf pdf-bot
Last synced: 13 Aug 2026
https://github.com/SteveKGYang/MentalLLaMA
This repository introduces MentaLLaMA, the first open-source instruction following large language model for interpretable mental health analysis.
chatgpt gpt4 interpretability language-model large-language-models llama2 mental-health natural-language-processing natural-language-understanding social-media
Last synced: 12 Aug 2026
https://github.com/BayLing-Models/BayLing
“百聆”是一个基于LLaMA的语言对齐增强的英语/中文大语言模型,具有优越的英语/中文能力,在多语言和通用任务等多项测试中取得ChatGPT 90%的性能。BayLing is an English/Chinese LLM equipped with advanced language alignment, showing superior capability in English/Chinese generation, instruction following and multi-turn interaction.⚠️ This project has been moved to: https://github.com/BayLing-Models/BayLing
aigc bayling chatgpt chinese cross-lingual general-language-model gpt4 human-performance instruction-tuning interactive large-language-models llama machine-translation multilingual-translation translation
Last synced: 12 Aug 2026
https://github.com/jianzhnie/LLamaTuner
Easy and Efficient Finetuning LLMs. (Supported LLama, LLama2, LLama3, Qwen, Baichuan, GLM , Falcon) 大模型高效量化训练+部署.
chatgpt dpo llama llama3 mixtral ppo qlora qwen rlhf
Last synced: 12 Aug 2026
https://github.com/maclandrol/molfeat-hype
Can ChatGPT generate molecular features ?
Last synced: 12 Aug 2026
https://github.com/Gunale0926/SORSA
SORSA: Singular Values and Orthonormal Regularized Singular Vectors Adaptation of Large Language Models
deep-learning fine-tuning llama lora machine-learning nlp peft python pytorch rwkv sorsa svd transformer
Last synced: 12 Aug 2026
https://github.com/laelhalawani/glai
glai - GGUF LLAMA AI - Package for simplified model handling and text generation with Llama models quantized to GGUF format. APIs for downloading and loading models automatically, includes a db with models of various scale and quantizations. With this high level API you need one line to load the model and one to generate text completions.
ai chatgpt generative-ai gguf llama llm quantization
Last synced: 12 Aug 2026
https://github.com/OFA-Sys/ExpertLLaMA
An opensource ChatBot built with ExpertPrompting which achieves 96% of ChatGPT's capability.
alignment alpaca chatgpt llama vicuna
Last synced: 12 Aug 2026
https://github.com/remixer-dec/llama-mps
Experimental fork of Facebooks LLaMa model which runs it with GPU acceleration on Apple Silicon M1/M2
llama llm macos mps
Last synced: 13 Aug 2026
https://github.com/minosvasilias/godot-dodo
Finetuning large language models for GDScript generation.
ai finetuning gdscript godot llama
Last synced: 12 Aug 2026
https://github.com/soulteary/docker-llama2-chat
Play LLaMA2 (official / 中文版 / INT4 / llama2.cpp) Together! ONLY 3 STEPS! ( non GPU / 5GB vRAM / 8~14GB vRAM)
llama llama2 llama2-docker llama2-playground llm
Last synced: 12 Aug 2026
https://github.com/FSoft-AI4Code/CodeCapybara
Open-source Self-Instruction Tuning Code LLM
ai4code alpaca codellm instruction-tuning llama
Last synced: 12 Aug 2026
https://github.com/SteelPh0enix/unreasonable-llama
Python API for llama.cpp webserver
Last synced: 12 Aug 2026
https://github.com/jasonvanf/llama-trl
LLaMA-TRL: Fine-tuning LLaMA with PPO and LoRA
adapter chatgpt gpt gpt-4 llama lora peft ppo rlhf transformer trl
Last synced: 13 Aug 2026
https://github.com/declare-lab/flacuna
Flacuna was developed by fine-tuning Vicuna on Flan-mini, a comprehensive instruction collection encompassing various tasks. Vicuna is already an excellent writing assistant, and the intention behind Flacuna was to enhance Vicuna's problem-solving capabilities. To achieve this, we curated a dedicated instruction dataset called Flan-mini.
large-language-models llama transformer
Last synced: 12 Aug 2026
https://github.com/ayaka14732/llama-2-jax
JAX implementation of the Llama 2 model
jax llama llama2 natural-language-processing nlp
Last synced: 12 Aug 2026
https://github.com/Joyce94/LLM-RLHF-Tuning
LLM Tuning with PEFT (SFT+RM+PPO+DPO with LoRA)
fine-tuning language-model llama llm lora peft ppo reinforcement-learning rlhf
Last synced: 12 Aug 2026
https://github.com/SqueezeAILab/KVQuant
[NeurIPS 2024] KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
compression efficient-inference efficient-model large-language-models llama llm localllama localllm mistral model-compression natural-language-processing quantization small-models text-generation transformer
Last synced: 12 Aug 2026
https://github.com/kyegomez/AttnWithConvolutions
Interleaved Attention's with convolutions for text modeling
artificial-intelligence attention attention-mechanism convolution convolutional-neural-networks gpt4 llama machine-learning machine-learning-algorithms
Last synced: 12 Aug 2026
https://github.com/Riccorl/llama-trainer
Llama Trainer Utility
huggingface llama llm llm-inference llm-training llms transformer
Last synced: 12 Aug 2026
https://github.com/vinjn/llm-metahuman
An open solution for AI-powered photorealistic digital humans.
ai audio2face chatbot chatgpt digital-human gpt3 llama llama2 llm metahuman nvidia omniverse openai
Last synced: 12 Aug 2026
https://github.com/yangjianxin1/Firefly-LLaMA2-Chinese
Firefly中文LLaMA-2大模型,支持增量预训练Baichuan2、Llama2、Llama、Falcon、Qwen、Baichuan、InternLM、Bloom等大模型
baichaun2 baichuan baichuan-13b bloom chatglm falcon firefly internlm llama llama-2 llama2 llm lora pretrain qlora qwen xverse
Last synced: 12 Aug 2026
https://github.com/e-lab/SyntaxShaper
Powering Agent Chains by Constraining LLM Outputs
agent ai ai-eng llama llama-cpp llm llm-agent
Last synced: 12 Aug 2026
https://github.com/tsangwailam/langchain-runpod-llm
Python library for using RunPod API endpoint as LangChain LLM
Last synced: 12 Aug 2026
https://github.com/zorazrw/filco
[Preprint] Learning to Filter Context for Retrieval-Augmented Generaton
dialog-generation fact-verification flan-t5 llama2 question-answering retrieval-augmented-generation
Last synced: 12 Aug 2026
https://github.com/nlp-uoregon/Okapi
Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
bloom chatbot dataset instruction-tuning language-model large-language-models llama multilingual natural-language-processing nlp question-answering reinforcement-learning reinforcement-learning-from-human-feedback rlhf
Last synced: 12 Aug 2026
https://github.com/dataprofessor/llama2
This chatbot app is built using the Llama 2 open source LLM from Meta.
large-language-models llama2 llm meta python streamlit
Last synced: 12 Aug 2026
https://github.com/c0sogi/llama-api
An OpenAI-like LLaMA inference API
api exllama fastapi llama llamacpp
Last synced: 12 Aug 2026
https://github.com/hkproj/pytorch-llama
LLaMA 2 implemented from scratch in PyTorch
language-model llama llama2 paper-implementations pytorch
Last synced: 12 Aug 2026
https://github.com/abhi-arya1/tuna
fine tuning, reimagined. welcome to tuna 🎣 - we're simplifying cloud compute architecture, datasets, and more, to get your specialized AI from 0->100 asap
ai cli code-generation fine-tune generative-ai generative-code llama lora python
Last synced: 13 Aug 2026
https://github.com/tpoisonooo/llama.onnx
LLaMa/RWKV onnx models, quantization and testcase
alpaca llama llm onnx onnxruntime quantization rwkv transformer
Last synced: 12 Aug 2026
https://github.com/cycneuramus/signal-aichat
An AI chatbot for Signal powered by Google Bard, Bing Chat, ChatGPT, HuggingChat, and llama.cpp
ai-bot bard bing-chat chatgpt chatgpt-bot google-bard huggingchat llama llamacpp signal-bot signal-messenger
Last synced: 12 Aug 2026
https://github.com/IAAR-Shanghai/Grimoire
Grimoire is All You Need for Enhancing Large Language Models
baichuan chatgpt datasets gpt-4 grimoire icl in-context-learning llama llm phi2
Last synced: 12 Aug 2026
https://github.com/thomas-yanxin/KarmaVLM
🧘🏻♂️KarmaVLM (相生):A family of high efficiency and powerful visual language model.
llama2 llava multimodel qwen2 vision-language-model visual-language-learning vlm
Last synced: 12 Aug 2026
https://github.com/declare-lab/red-instruct
Codes and datasets of the paper Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
huggingface-transformers llama llama2 llm llms
Last synced: 12 Aug 2026
https://github.com/4AI/LS-LLaMA
A Simple but Powerful SOTA NER Model | Official Code For Label Supervised LLaMA Finetuning
conll2003 llama llama2 llms named-entity-recognition ontonotes sequence-classification token-classification
Last synced: 12 Aug 2026
https://github.com/calcuis/llama-core
solo connector core built on llama.cpp
core gguf llama
Last synced: 12 Aug 2026
https://github.com/kaymen99/Upwork-AI-jobs-applier
AI tool for automating Upwork job applications using AI agents to find and qualify jobs, write personalized cover letters, and prepare for interviews based on your skills and experience.
ai-agents ai-automation langchain langgraph llm-agent llm-scraper playwright scraping upwork upwork-automation upwork-jobs-scraping upwork-scraper
Last synced: 12 Aug 2026
https://github.com/photomz/BabyDoctor
The AI Radiologist You Can Chat With
clip huggingface llama2
Last synced: 12 Aug 2026
https://github.com/ictnlp/TruthX
Code for ACL 2024 paper "TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space"
baichuan chatglm chatgpt explainable-ai gpt-4 hallucination hallucinations language-model llama llama2 llama3 llm llm-inference llms mistral representation safety truthfulness
Last synced: 12 Aug 2026
https://github.com/luchangli03/export_llama_to_onnx
export llama to onnx
llama llm llm-inference onnx onnxruntime pytorch
Last synced: 13 Aug 2026
https://github.com/FuxiaoLiu/LRV-Instruction
[ICLR'24] Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning
chatgpt evaluation evaluation-metrics foundation-models gpt gpt-4 hallucination iclr iclr2024 llama llava multimodal object-detection prompt-engineering vicuna vision vision-and-language vqa
Last synced: 12 Aug 2026
https://github.com/git-cloner/llama-lora-fine-tuning
llama fine-tuning with lora
finetuning llama lora
Last synced: 12 Aug 2026
https://github.com/allenai/CommonGen-Eval
Evaluating LLMs with CommonGen-Lite
chatgpt evaluation gpt-evaluation llama2 llm llm-evaluation text-generation
Last synced: 13 Aug 2026
https://github.com/sazonovanton/SirChatalot
SirChatalot is a Telegram bot leveraging ChatGPT, Claude or YandexGPT. It uses Whisper for speech-to-text and DALL-E, Stability AI or YandexART for image creation. It can use vision capabilities, tools and semantic search in vector DB.
agentic-ai anthropic chatgpt claude claude-api dall-e function-calling openai openai-api python-telegram-bot rag semantic-search stability-ai telegram-bot tool-use web-search whisper yandex-gpt yandexart yandexgpt
Last synced: 12 Aug 2026
https://github.com/mlpc-ucsd/BLIVA
(AAAI 2024) BLIVA: A Simple Multimodal LLM for Better Handling of Text-rich Visual Questions
blip2 bliva chatbot instruction-tuning llama llm lora multimodal visual-language-learning
Last synced: 12 Aug 2026
https://github.com/jianzhnie/Open-R1
The open source implementation of DeepSeek-R1. 开源复现 DeepSeek-R1
deepseek-r1 deepseek-v3 grpo llm rlhf
Last synced: 12 Aug 2026
https://github.com/kaymen99/langgraph-email-automation
Multi AI agents for customer support email automation built with Langchain & Langgraph
ai-agents ai-automation ai-customer-service ai-customer-support customer-support customer-support-automation email-automation gmail-api langchain langgraph llama3 rag rag-agents rag-application
Last synced: 12 Aug 2026
https://github.com/starmpcc/CAMEL
Clinically Adapted Model Enhanced from LLaMA
alpaca camel clinical gpt large-language-model llama llm self-instruct
Last synced: 12 Aug 2026
https://github.com/boostcampaitech5/LawBot-Online-Legal-Advice-LLM-Service
사용자가 채팅웹을 통해 자신이 처한 법률적 상황을 제시하면, 입력에 대한 문맥을 모델이 이해하여 가이드라인을 제시하고, 유사한 상황의 판례를 제공하는 웹 서비스입니다. (2023.08.18 서비스 종료)
airflow docker llama2 reactjs selenium tailwindcss
Last synced: 12 Aug 2026
https://github.com/ant4g0nist/polar
A LLDB plugin which brings LLMs to LLDB
langchain llama2 lldb llm ollama reverse-engineering
Last synced: 12 Aug 2026
https://github.com/li-plus/chat4u
用微信聊天记录训练一个你专属的聊天机器人
alpaca chatbot llama wechat wechaty
Last synced: 13 Aug 2026
https://github.com/xNul/chat-llama-discord-bot
A Discord Bot for chatting with LLaMA, Vicuna, Alpaca, MPT, or any other Large Language Model (LLM) supported by text-generation-webui or llama.cpp.
alpaca bot chat chat-bot chatbot chatgpt chatllama discord gpt-4 gpt4 large-language-model large-language-models llama llamacpp llm text-generation-webui vicuna
Last synced: 12 Aug 2026
https://github.com/l294265421/alpaca-rlhf
Finetuning LLaMA with RLHF (Reinforcement Learning with Human Feedback) based on DeepSpeed Chat
alpaca chatgpt language-model large-language-models llama llm reinforcement-learning rlhf
Last synced: 12 Aug 2026
https://github.com/harleyszhang/llm_counts
llm theoretical performance analysis tools and support params, flops, memory and latency analysis.
gpu-performance llama llm llm-inference profiler python3 transformer
Last synced: 12 Aug 2026
https://github.com/quack-ai/companion
VSCode coding companion for software teams 🦆 Turn your team insights into a portable plug-and-play context for code generation. Alternative to GitHub Copilot & OpenAI GPT powered by OSS LLMs (Phi 3, Llama 3, CodeQwen, Mistral, etc.), made with ❤️ using FastAPI & Ollama.
ai api code-generation code-quality deep-learning developer-tools docker fastapi gpt4o groq llama llm mistral ollama open-source openai python self-hosted visual-studio-code vscode
Last synced: 12 Aug 2026
https://github.com/jackaduma/Vicuna-LoRA-RLHF-PyTorch
A full pipeline to finetune Vicuna LLM with LoRA and RLHF on consumer hardware. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the Vicuna architecture. Basically ChatGPT but with Vicuna
chatgpt finetune gpt llama llm lora peft ppo pytorch reward-models rlhf vicuna vicuna-7b
Last synced: 12 Aug 2026
https://github.com/taishan1994/Llama3.1-Finetuning
对llama3进行全参微调、lora微调以及qlora微调。
llama3 lora qlora qwen
Last synced: 12 Aug 2026
https://github.com/mkellerman/gpt4all-ui
Simple Docker Compose to load gpt4all (Llama.cpp) as an API and chatbot-ui for the web interface. This mimics OpenAI's ChatGPT but as a local instance (offline).
api gpt4all llama python ui web
Last synced: 13 Aug 2026
Statistics
- Projects: 2,239
- Last updated: about 2 years ago