An open API service for producing an overview of a list of open source projects.

Collections: awesome-llama

https://github.com/mallorbc/Finetune_LLMs

Repo for fine-tuning Casual LLMs

docker falcon gpt gpt-3 gpt-35-turbo gpt-4 gpt-j-6b llama llama2 llm llm-training mpt

Last synced: 12 Aug 2026

https://github.com/riccardomusmeci/mlx-llm

Large Language Models (LLMs) applications and tools running on Apple Silicon in real-time with Apple MLX.

llama llm mistral mlx phi transformers

Last synced: 12 Aug 2026

https://github.com/SeungyounShin/Llama2-Code-Interpreter

Make Llama2 use Code Execution, Debug, Save Code, Reuse it, Access to Internet

codeinterpreter codellama llama llm

Last synced: 12 Aug 2026

https://github.com/bigsk1/voice-chat-ai

🎙️ Speak with AI - Run locally using Ollama, OpenAI, Anthropic or xAI - Speech uses SparkTTS, OpenAI, ElevenLabs, Kokoro, Typecast or xAI

ai-speech ai-voice ai-voice-agent anthropic-claude conversational-ai elevenlabs-api fastapi ollama selfhosted tts typecast voice-ai webrtc whisper-ai xai xai-tts

Last synced: 12 Aug 2026

https://github.com/galatolofederico/vanilla-llama

Plain pytorch implementation of LLaMA

llama llama-inference-server llama-pytorch

Last synced: 12 Aug 2026

https://github.com/interestingLSY/swiftLLM

A tiny yet powerful LLM inference system tailored for researching purpose. vLLM-equivalent performance with only 2k lines of code (2% of vLLM).

cuda gpt inference inference-engine llama llm llm-inference llm-serving llmops mlops model-serving pytorch transformer transformers

Last synced: 12 Aug 2026

https://github.com/calcuis/gguf-core

a simple way to interact llama with gguf

cli core gguf gui llama metadata pdf reader vision wav

Last synced: 13 Aug 2026

https://github.com/RLHF-V/RLHF-V

[CVPR'24] RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

chatbot gpt-4 llama multi-modality multimodal rlhf-v visual-language-learning

Last synced: 12 Aug 2026

https://github.com/misonsky/HiFT

memory-efficient fine-tuning; support 24G GPU memory fine-tuning 7B

chinese-llama chinese-llama-65b huggingface-transformers large-language-models llama2 llama3 lora memory-efficient-tuning peft-fine-tuning-llm pytorch-implementation transformers

Last synced: 12 Aug 2026

https://github.com/neph1/LlamaTale

Giving the power of LLM's to a MUD lib.

generative-ai interactive-fiction large-language-models llama llm mud roleplaying

Last synced: 12 Aug 2026

https://github.com/jankais3r/LLaMA_MPS

Run LLaMA (and Stanford-Alpaca) inference on Apple Silicon GPUs.

alpaca apple-silicon chat chatbot chatgpt llama llms macos metal ml mps stanford-alpaca torch

Last synced: 12 Aug 2026

https://github.com/xNul/code-llama-for-vscode

Use Code Llama with Visual Studio Code and the Continue extension. A local LLM alternative to GitHub Copilot.

assistant code code-llama codellama continue continuedev copilot llama llama2 llamacpp llm local meta ollama studio visual vscode

Last synced: 12 Aug 2026

https://github.com/kaarthik108/snowChat

Chat snowflake - Text to SQL

agents chatgpt langchain langgraph llama snowflake snowpark streamlit supabase

Last synced: 12 Aug 2026

https://github.com/janelu9/EasyLLM

Running Large Language Model easily.

deepseek deepspeed fine-tuning llama megatron npu pretrain qwen qwen-vl rlhf vllm

Last synced: 12 Aug 2026

https://github.com/ChuloAI/BrainChulo

Harnessing the Memory Power of the Camelids

chromadb fastapi langchain llama llm microsoft-guidance python retrieval-augmented sqlmodel vector-store

Last synced: 12 Aug 2026

https://github.com/sae-llm-coconut/coconut-ai

Python library that ease the installation process of Stable Diffusion, and allows to genrate images with a nice to use API.

llama2 python-library stable-diffusion

Last synced: 12 Aug 2026

https://github.com/WangRongsheng/CareGPT

🌞 CareGPT (关怀GPT)是一个医疗大语言模型,同时它集合了数十个公开可用的医疗微调数据集和开放可用的医疗大语言模型,包含LLM的训练、测评、部署等以促进医疗LLM快速发展。Medical LLM, Open Source Driven for a Healthy Future.

baichuan gpt large-language-models llama llama2 medical-llm

Last synced: 12 Aug 2026

https://github.com/kyegomez/M2PT

Implementation of M2PT in PyTorch from the paper: "Multimodal Pathway: Improve Transformers with Irrelevant Data from Other Modalities"

ai attention attention-is-all-you-need gpt4 gpt5 llama ml models mulit-modality multi-modal

Last synced: 12 Aug 2026

https://github.com/hpcaitech/SwiftInfer

Efficient AI Inference & Serving

artificial-intelligence deep-learning gpt inference llama llama2 llm-inference llm-serving

Last synced: 12 Aug 2026

https://github.com/alexeichhorn/typegpt

Make GPT safe for production

gpt llama llm openai prompt-engineering

Last synced: 12 Aug 2026

https://github.com/airaria/Visual-Chinese-LLaMA-Alpaca

多模态中文LLaMA&Alpaca大语言模型(VisualCLA)

alpaca chinese llama llm lora multimodal nlp vision-language

Last synced: 12 Aug 2026

https://github.com/mbzuai-oryx/VideoGPT-plus

Official Repository of paper VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding

chatbot clip dual-encoder gpt4 gpt4o image-encoder llama3 llava multimodal phi-3-mini vicuna video-chatbot video-conversation video-encoder vision-language vision-language-pretraining

Last synced: 12 Aug 2026

https://github.com/HenryNdubuaku/nanodl

JAX library for training sub-4B foundation models for edge

attention attention-mechanism deep-learning distributed-training flax gpt jax llama machine-learning mistral nlp transformer

Last synced: 12 Aug 2026

https://github.com/stanleylsx/llms_tool

一个基于HuggingFace开发的大语言模型训练、测试工具。支持各模型的webui、终端预测,低参数量及全参数模型训练(预训练、SFT、RM、PPO、DPO)和融合、量化。

aquila aquila2 baichuan baichuan2 bloom chatglm chatglm2 chatglm3 deepspeed falcon internlm llama llama2 mistral moss pytorch qwen xverse

Last synced: 12 Aug 2026

https://github.com/okuvshynov/slowllama

Finetune llama2-70b and codellama on MacBook Air without quantization

apple-silicon fine-tuning llama llama2

Last synced: 12 Aug 2026

https://github.com/kyegomez/GATS

Implementation of GATS from the paper: "GATS: Gather-Attend-Scatter" in pytorch and zeta

ai attention attention-is-all-you-need attention-mechanism gpt4 llama ml multi-modal multi-modality multimodal open-source

Last synced: 12 Aug 2026

https://github.com/YJ-20/auto-subtitle-translate

Automatically generate, translate, and overlay subtitles for any video.

ai ai-subtitle automatic-subtitle deep-learning ffmpeg llama llama2 python subtitle-generator subtitles subtitles-generator translates translator whisper

Last synced: 12 Aug 2026

https://github.com/josStorer/selfhostedAI

A collection of one-click self-hosted AI

chatglm chatgpt llama llm stable-diffusion

Last synced: 13 Aug 2026

https://github.com/WangRongsheng/ChatGenTitle

🌟 ChatGenTitle:使用百万arXiv论文信息在LLaMA模型上进行微调的论文题目生成模型

arxiv large-language-models llama llm llms lora

Last synced: 12 Aug 2026

https://github.com/Gary3410/TaPA

[arXiv 2023] Embodied Task Planning with Large Language Models

ai2thor embodied-agent llama robotics

Last synced: 13 Aug 2026

https://github.com/luogen1996/LLaVA-HR

[ICLR2025] LLaVA-HR: High-Resolution Large Language-Vision Assistant

high-resolution-mllms llama2 llms multimodal-chatbot

Last synced: 12 Aug 2026

https://github.com/holoviz-topics/panel-chat-examples

Examples of Chat Bots using Panels chat features: Traditional, LLMs, AI Agents, LangChain, OpenAI etc

chat gpt llama mistral openai panel python

Last synced: 12 Aug 2026

https://github.com/alexfazio/OpenPlexity-Pages

SearchGPT / Perplexity Pages clone, but personalised for you.

crewai groq llama3 search-engine streamlit

Last synced: 12 Aug 2026

https://github.com/taesiri/ArXivQA

WIP - Automated Question Answering for ArXiv Papers with Large Language Models (https://arxiv.taesiri.xyz/)

arxiv arxiv-daily arxiv-dataset arxiv-papers arxiv-preprint automated-qa claude claude2 gpt gpt-4 llama llama2 llm question-answering

Last synced: 12 Aug 2026

https://github.com/OneInterface/realtime-bakllava

llama.cpp with BakLLaVA model describes what does it see

bakllavva cpp demo-application inference llama llamacpp llm

Last synced: 12 Aug 2026

https://github.com/Chongjie-Si/Subspace-Tuning

A generalized framework for subspace tuning methods in parameter efficient fine-tuning.

adapter commonsense-reasoning glue llama llama2-7b llama3-8b lora lora-dash low-rank-adaptation natural-language-generation natural-language-processing natural-language-understanding parameter-efficient-fine-tuning pretrained-models soft-prompt-tuning subject-driven-generation subspace-tuning

Last synced: 12 Aug 2026

https://github.com/innightwolfsleep/llm_telegram_bot

LLM telegram bot

aiogram bot exllama large-language-models llamacpp llm telegram telegram-bot text-generation-webui

Last synced: 12 Aug 2026

https://github.com/simulatrex/simulatrex-engine

Enable decision-making based on simulations

chatgpt generative-ai gpt-4 llama2 pypi python simaas simulations

Last synced: 12 Aug 2026

https://github.com/JosemyDuarte/gpt-researcher-ollama

Based on assafelovic/gpt-researcher - Modified to support local Ollama models

gpt gpt-researcher llama llama3 ollama

Last synced: 12 Aug 2026

https://github.com/BIDS-Xu-Lab/Me-LLaMA

A novel medical large language model family with 13/70B parameters, which have SOTA performances on various medical tasks

biomedical-nlp biomedical-text-mining chatgpt clinical-nlp gpt-4 large-language-models llama llama2 llm medical- medical-application medical-large-language-models medical-question-answering

Last synced: 12 Aug 2026

https://github.com/jerry1993-tech/Cornucopia-LLaMA-Fin-Chinese

聚宝盆(Cornucopia): 中文金融系列开源可商用大模型,并提供一套高效轻量化的垂直领域LLM训练框架(Pretraining、SFT、RLHF、Quantize等)

chinese finance large-language-models llama nlp qa rlhf sft text-generation transformers

Last synced: 12 Aug 2026

https://github.com/amithkoujalgi/ollama-pdf-bot

A bot that accepts PDF docs and lets you ask questions on it.

bot chat-bot llama llama2 llm ollama pdf pdf-bot

Last synced: 13 Aug 2026

https://github.com/SteveKGYang/MentalLLaMA

This repository introduces MentaLLaMA, the first open-source instruction following large language model for interpretable mental health analysis.

chatgpt gpt4 interpretability language-model large-language-models llama2 mental-health natural-language-processing natural-language-understanding social-media

Last synced: 12 Aug 2026

https://github.com/BayLing-Models/BayLing

“百聆”是一个基于LLaMA的语言对齐增强的英语/中文大语言模型,具有优越的英语/中文能力,在多语言和通用任务等多项测试中取得ChatGPT 90%的性能。BayLing is an English/Chinese LLM equipped with advanced language alignment, showing superior capability in English/Chinese generation, instruction following and multi-turn interaction.⚠️ This project has been moved to: https://github.com/BayLing-Models/BayLing

aigc bayling chatgpt chinese cross-lingual general-language-model gpt4 human-performance instruction-tuning interactive large-language-models llama machine-translation multilingual-translation translation

Last synced: 12 Aug 2026

https://github.com/jianzhnie/LLamaTuner

Easy and Efficient Finetuning LLMs. (Supported LLama, LLama2, LLama3, Qwen, Baichuan, GLM , Falcon) 大模型高效量化训练+部署.

chatgpt dpo llama llama3 mixtral ppo qlora qwen rlhf

Last synced: 12 Aug 2026

https://github.com/maclandrol/molfeat-hype

Can ChatGPT generate molecular features ?

Last synced: 12 Aug 2026

https://github.com/Gunale0926/SORSA

SORSA: Singular Values and Orthonormal Regularized Singular Vectors Adaptation of Large Language Models

deep-learning fine-tuning llama lora machine-learning nlp peft python pytorch rwkv sorsa svd transformer

Last synced: 12 Aug 2026

https://github.com/laelhalawani/glai

glai - GGUF LLAMA AI - Package for simplified model handling and text generation with Llama models quantized to GGUF format. APIs for downloading and loading models automatically, includes a db with models of various scale and quantizations. With this high level API you need one line to load the model and one to generate text completions.

ai chatgpt generative-ai gguf llama llm quantization

Last synced: 12 Aug 2026

https://github.com/OFA-Sys/ExpertLLaMA

An opensource ChatBot built with ExpertPrompting which achieves 96% of ChatGPT's capability.

alignment alpaca chatgpt llama vicuna

Last synced: 12 Aug 2026

https://github.com/remixer-dec/llama-mps

Experimental fork of Facebooks LLaMa model which runs it with GPU acceleration on Apple Silicon M1/M2

llama llm macos mps

Last synced: 13 Aug 2026

https://github.com/minosvasilias/godot-dodo

Finetuning large language models for GDScript generation.

ai finetuning gdscript godot llama

Last synced: 12 Aug 2026

https://github.com/soulteary/docker-llama2-chat

Play LLaMA2 (official / 中文版 / INT4 / llama2.cpp) Together! ONLY 3 STEPS! ( non GPU / 5GB vRAM / 8~14GB vRAM)

llama llama2 llama2-docker llama2-playground llm

Last synced: 12 Aug 2026

https://github.com/FSoft-AI4Code/CodeCapybara

Open-source Self-Instruction Tuning Code LLM

ai4code alpaca codellm instruction-tuning llama

Last synced: 12 Aug 2026

https://github.com/SteelPh0enix/unreasonable-llama

Python API for llama.cpp webserver

Last synced: 12 Aug 2026

https://github.com/jasonvanf/llama-trl

LLaMA-TRL: Fine-tuning LLaMA with PPO and LoRA

adapter chatgpt gpt gpt-4 llama lora peft ppo rlhf transformer trl

Last synced: 13 Aug 2026

https://github.com/declare-lab/flacuna

Flacuna was developed by fine-tuning Vicuna on Flan-mini, a comprehensive instruction collection encompassing various tasks. Vicuna is already an excellent writing assistant, and the intention behind Flacuna was to enhance Vicuna's problem-solving capabilities. To achieve this, we curated a dedicated instruction dataset called Flan-mini.

large-language-models llama transformer

Last synced: 12 Aug 2026

https://github.com/ayaka14732/llama-2-jax

JAX implementation of the Llama 2 model

jax llama llama2 natural-language-processing nlp

Last synced: 12 Aug 2026

https://github.com/Joyce94/LLM-RLHF-Tuning

LLM Tuning with PEFT (SFT+RM+PPO+DPO with LoRA)

fine-tuning language-model llama llm lora peft ppo reinforcement-learning rlhf

Last synced: 12 Aug 2026

https://github.com/SqueezeAILab/KVQuant

[NeurIPS 2024] KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization

compression efficient-inference efficient-model large-language-models llama llm localllama localllm mistral model-compression natural-language-processing quantization small-models text-generation transformer

Last synced: 12 Aug 2026

https://github.com/kyegomez/AttnWithConvolutions

Interleaved Attention's with convolutions for text modeling

artificial-intelligence attention attention-mechanism convolution convolutional-neural-networks gpt4 llama machine-learning machine-learning-algorithms

Last synced: 12 Aug 2026

https://github.com/Riccorl/llama-trainer

Llama Trainer Utility

huggingface llama llm llm-inference llm-training llms transformer

Last synced: 12 Aug 2026

https://github.com/vinjn/llm-metahuman

An open solution for AI-powered photorealistic digital humans.

ai audio2face chatbot chatgpt digital-human gpt3 llama llama2 llm metahuman nvidia omniverse openai

Last synced: 12 Aug 2026

https://github.com/yangjianxin1/Firefly-LLaMA2-Chinese

Firefly中文LLaMA-2大模型,支持增量预训练Baichuan2、Llama2、Llama、Falcon、Qwen、Baichuan、InternLM、Bloom等大模型

baichaun2 baichuan baichuan-13b bloom chatglm falcon firefly internlm llama llama-2 llama2 llm lora pretrain qlora qwen xverse

Last synced: 12 Aug 2026

https://github.com/e-lab/SyntaxShaper

Powering Agent Chains by Constraining LLM Outputs

agent ai ai-eng llama llama-cpp llm llm-agent

Last synced: 12 Aug 2026

https://github.com/tsangwailam/langchain-runpod-llm

Python library for using RunPod API endpoint as LangChain LLM

Last synced: 12 Aug 2026

https://github.com/zorazrw/filco

[Preprint] Learning to Filter Context for Retrieval-Augmented Generaton

dialog-generation fact-verification flan-t5 llama2 question-answering retrieval-augmented-generation

Last synced: 12 Aug 2026

https://github.com/nlp-uoregon/Okapi

Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback

bloom chatbot dataset instruction-tuning language-model large-language-models llama multilingual natural-language-processing nlp question-answering reinforcement-learning reinforcement-learning-from-human-feedback rlhf

Last synced: 12 Aug 2026

https://github.com/dataprofessor/llama2

This chatbot app is built using the Llama 2 open source LLM from Meta.

large-language-models llama2 llm meta python streamlit

Last synced: 12 Aug 2026

https://github.com/c0sogi/llama-api

An OpenAI-like LLaMA inference API

api exllama fastapi llama llamacpp

Last synced: 12 Aug 2026

https://github.com/hkproj/pytorch-llama

LLaMA 2 implemented from scratch in PyTorch

language-model llama llama2 paper-implementations pytorch

Last synced: 12 Aug 2026

https://github.com/abhi-arya1/tuna

fine tuning, reimagined. welcome to tuna 🎣 - we're simplifying cloud compute architecture, datasets, and more, to get your specialized AI from 0->100 asap

ai cli code-generation fine-tune generative-ai generative-code llama lora python

Last synced: 13 Aug 2026

https://github.com/tpoisonooo/llama.onnx

LLaMa/RWKV onnx models, quantization and testcase

alpaca llama llm onnx onnxruntime quantization rwkv transformer

Last synced: 12 Aug 2026

https://github.com/cycneuramus/signal-aichat

An AI chatbot for Signal powered by Google Bard, Bing Chat, ChatGPT, HuggingChat, and llama.cpp

ai-bot bard bing-chat chatgpt chatgpt-bot google-bard huggingchat llama llamacpp signal-bot signal-messenger

Last synced: 12 Aug 2026

https://github.com/IAAR-Shanghai/Grimoire

Grimoire is All You Need for Enhancing Large Language Models

baichuan chatgpt datasets gpt-4 grimoire icl in-context-learning llama llm phi2

Last synced: 12 Aug 2026

https://github.com/thomas-yanxin/KarmaVLM

🧘🏻‍♂️KarmaVLM (相生):A family of high efficiency and powerful visual language model.

llama2 llava multimodel qwen2 vision-language-model visual-language-learning vlm

Last synced: 12 Aug 2026

https://github.com/declare-lab/red-instruct

Codes and datasets of the paper Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment

huggingface-transformers llama llama2 llm llms

Last synced: 12 Aug 2026

https://github.com/4AI/LS-LLaMA

A Simple but Powerful SOTA NER Model | Official Code For Label Supervised LLaMA Finetuning

conll2003 llama llama2 llms named-entity-recognition ontonotes sequence-classification token-classification

Last synced: 12 Aug 2026

https://github.com/calcuis/llama-core

solo connector core built on llama.cpp

core gguf llama

Last synced: 12 Aug 2026

https://github.com/kaymen99/Upwork-AI-jobs-applier

AI tool for automating Upwork job applications using AI agents to find and qualify jobs, write personalized cover letters, and prepare for interviews based on your skills and experience.

ai-agents ai-automation langchain langgraph llm-agent llm-scraper playwright scraping upwork upwork-automation upwork-jobs-scraping upwork-scraper

Last synced: 12 Aug 2026

https://github.com/photomz/BabyDoctor

The AI Radiologist You Can Chat With

clip huggingface llama2

Last synced: 12 Aug 2026

https://github.com/ictnlp/TruthX

Code for ACL 2024 paper "TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space"

baichuan chatglm chatgpt explainable-ai gpt-4 hallucination hallucinations language-model llama llama2 llama3 llm llm-inference llms mistral representation safety truthfulness

Last synced: 12 Aug 2026

https://github.com/luchangli03/export_llama_to_onnx

export llama to onnx

llama llm llm-inference onnx onnxruntime pytorch

Last synced: 13 Aug 2026

https://github.com/FuxiaoLiu/LRV-Instruction

[ICLR'24] Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

chatgpt evaluation evaluation-metrics foundation-models gpt gpt-4 hallucination iclr iclr2024 llama llava multimodal object-detection prompt-engineering vicuna vision vision-and-language vqa

Last synced: 12 Aug 2026

https://github.com/git-cloner/llama-lora-fine-tuning

llama fine-tuning with lora

finetuning llama lora

Last synced: 12 Aug 2026

https://github.com/allenai/CommonGen-Eval

Evaluating LLMs with CommonGen-Lite

chatgpt evaluation gpt-evaluation llama2 llm llm-evaluation text-generation

Last synced: 13 Aug 2026

https://github.com/sazonovanton/SirChatalot

SirChatalot is a Telegram bot leveraging ChatGPT, Claude or YandexGPT. It uses Whisper for speech-to-text and DALL-E, Stability AI or YandexART for image creation. It can use vision capabilities, tools and semantic search in vector DB.

agentic-ai anthropic chatgpt claude claude-api dall-e function-calling openai openai-api python-telegram-bot rag semantic-search stability-ai telegram-bot tool-use web-search whisper yandex-gpt yandexart yandexgpt

Last synced: 12 Aug 2026

https://github.com/mlpc-ucsd/BLIVA

(AAAI 2024) BLIVA: A Simple Multimodal LLM for Better Handling of Text-rich Visual Questions

blip2 bliva chatbot instruction-tuning llama llm lora multimodal visual-language-learning

Last synced: 12 Aug 2026

https://github.com/jianzhnie/Open-R1

The open source implementation of DeepSeek-R1. 开源复现 DeepSeek-R1

deepseek-r1 deepseek-v3 grpo llm rlhf

Last synced: 12 Aug 2026

https://github.com/kaymen99/langgraph-email-automation

Multi AI agents for customer support email automation built with Langchain & Langgraph

ai-agents ai-automation ai-customer-service ai-customer-support customer-support customer-support-automation email-automation gmail-api langchain langgraph llama3 rag rag-agents rag-application

Last synced: 12 Aug 2026

https://github.com/starmpcc/CAMEL

Clinically Adapted Model Enhanced from LLaMA

alpaca camel clinical gpt large-language-model llama llm self-instruct

Last synced: 12 Aug 2026

https://github.com/boostcampaitech5/LawBot-Online-Legal-Advice-LLM-Service

사용자가 채팅웹을 통해 자신이 처한 법률적 상황을 제시하면, 입력에 대한 문맥을 모델이 이해하여 가이드라인을 제시하고, 유사한 상황의 판례를 제공하는 웹 서비스입니다. (2023.08.18 서비스 종료)

airflow docker llama2 reactjs selenium tailwindcss

Last synced: 12 Aug 2026

https://github.com/ant4g0nist/polar

A LLDB plugin which brings LLMs to LLDB

langchain llama2 lldb llm ollama reverse-engineering

Last synced: 12 Aug 2026

https://github.com/li-plus/chat4u

用微信聊天记录训练一个你专属的聊天机器人

alpaca chatbot llama wechat wechaty

Last synced: 13 Aug 2026

https://github.com/xNul/chat-llama-discord-bot

A Discord Bot for chatting with LLaMA, Vicuna, Alpaca, MPT, or any other Large Language Model (LLM) supported by text-generation-webui or llama.cpp.

alpaca bot chat chat-bot chatbot chatgpt chatllama discord gpt-4 gpt4 large-language-model large-language-models llama llamacpp llm text-generation-webui vicuna

Last synced: 12 Aug 2026

https://github.com/l294265421/alpaca-rlhf

Finetuning LLaMA with RLHF (Reinforcement Learning with Human Feedback) based on DeepSpeed Chat

alpaca chatgpt language-model large-language-models llama llm reinforcement-learning rlhf

Last synced: 12 Aug 2026

https://github.com/harleyszhang/llm_counts

llm theoretical performance analysis tools and support params, flops, memory and latency analysis.

gpu-performance llama llm llm-inference profiler python3 transformer

Last synced: 12 Aug 2026

https://github.com/quack-ai/companion

VSCode coding companion for software teams 🦆 Turn your team insights into a portable plug-and-play context for code generation. Alternative to GitHub Copilot & OpenAI GPT powered by OSS LLMs (Phi 3, Llama 3, CodeQwen, Mistral, etc.), made with ❤️ using FastAPI & Ollama.

ai api code-generation code-quality deep-learning developer-tools docker fastapi gpt4o groq llama llm mistral ollama open-source openai python self-hosted visual-studio-code vscode

Last synced: 12 Aug 2026

https://github.com/jackaduma/Vicuna-LoRA-RLHF-PyTorch

A full pipeline to finetune Vicuna LLM with LoRA and RLHF on consumer hardware. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the Vicuna architecture. Basically ChatGPT but with Vicuna

chatgpt finetune gpt llama llm lora peft ppo pytorch reward-models rlhf vicuna vicuna-7b

Last synced: 12 Aug 2026

https://github.com/taishan1994/Llama3.1-Finetuning

对llama3进行全参微调、lora微调以及qlora微调。

llama3 lora qlora qwen

Last synced: 12 Aug 2026

https://github.com/mkellerman/gpt4all-ui

Simple Docker Compose to load gpt4all (Llama.cpp) as an API and chatbot-ui for the web interface. This mimics OpenAI's ChatGPT but as a local instance (offline).

api gpt4all llama python ui web

Last synced: 13 Aug 2026

Statistics

  • Projects: 2,239
  • Last updated: about 2 years ago