Collections: awesome-llama
https://github.com/thejasmeetsingh/moody-llm
A LLM whose mood keeps changing.
asynchronous-programming fastapi genai highlightjs langchain llama3 llm ollama pydantic python3 reactjs restful-api supabase tailwindcss websockets
Last synced: 27 Aug 2026
https://github.com/X-PLUG/mPLUG-Owl
mPLUG-Owl: The Powerful Multi-modal Large Language Model Family
alpaca chatbot chatgpt damo dialogue gpt gpt4 gpt4-api huggingface instruction-tuning large-language-models llama mplug mplug-owl multimodal pretraining pytorch transformer video visual-recognition
Last synced: 27 Aug 2026
https://github.com/WangRongsheng/CareGPT
🌞 CareGPT (关怀GPT)是一个医疗大语言模型,同时它集合了数十个公开可用的医疗微调数据集和开放可用的医疗大语言模型,包含LLM的训练、测评、部署等以促进医疗LLM快速发展。Medical LLM, Open Source Driven for a Healthy Future.
baichuan gpt large-language-models llama llama2 medical-llm
Last synced: 27 Aug 2026
https://github.com/declare-lab/red-instruct
Codes and datasets of the paper Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
huggingface-transformers llama llama2 llm llms
Last synced: 27 Aug 2026
https://github.com/zetavg/LLaMA-LoRA-Tuner
UI tool for fine-tuning and testing your own LoRA models base on LLaMA, GPT-J and more. One-click run on Google Colab. + A Gradio ChatGPT-like Chat UI to demonstrate your language models.
ai alpaca alpaca-lora google-colab gpt gpt-j language-model llama lora machine-learning peft
Last synced: 27 Aug 2026
https://github.com/ATH-MaaS/Ovis
A novel Multimodal Large Language Model (MLLM) architecture, designed to structurally align visual and textual embeddings.
chatbot llama3 multimodal multimodal-large-language-models multimodality qwen vision-language-learning vision-language-model
Last synced: 27 Aug 2026
https://github.com/Azure-Samples/miyagi
Sample to envision intelligent apps with Microsoft's Copilot stack for AI-infused product experiences.
agents aks assistants azure azure-openai azureai copilot gpt-4 guidance langchain llama-index llama2 openai phi-2 prompt-engineering promptflow semantic-kernel taskweaver typechat
Last synced: 27 Aug 2026
https://github.com/kennethleungty/Llama-2-Open-Source-LLM-CPU-Inference
Running Llama 2 and other Open-Source LLMs on CPU Inference Locally for Document Q&A
c-transformers chatgpt cpu cpu-inference deep-learning document-qa faiss langchain language-models large-language-models llama llama-2 llm machine-learning natural-language-processing nlp open-source-llm python sentence-transformers transformers
Last synced: 27 Aug 2026
https://github.com/NVIDIA/nim-anywhere
Accelerate your Gen AI with NVIDIA NIM and NVIDIA AI Workbench
genai langchain lcel llama llama3 llm nim nvidia nvwb-project rag
Last synced: 27 Aug 2026
https://github.com/MetaGLM/langchain-zhipuai
基于 Langchain,快速集成GLM-4 AllTools 功能的插件
chatbot glm gpt gpt-4 langchain llama ollama rag
Last synced: 27 Aug 2026
https://github.com/henrywoo/pyllama
LLaMA: Open and Efficient Foundation Language Models
Last synced: 27 Aug 2026
https://github.com/EmbeddedLLM/embeddedllm
EmbeddedLLM: API server for Embedded Device Deployment. Currently support CUDA/OpenVINO/IpexLLM/DirectML/CPU
aipc cpu directml directx-12 gemma ipexllm llama llm llm-inference llm-serving mistral model-inference npu open-source-llm openvino openvino-inference-engine phi-3 windows
Last synced: 27 Aug 2026
https://github.com/flojud/DocsChat
The chatbot utilizes a conversational retrieval chain to answer user queries based on the content of embedded documents. It leverages various NLP techniques, including language models and embeddings, to provide relevant responses.
Last synced: 27 Aug 2026
https://github.com/ictnlp/TruthX
Code for ACL 2024 paper "TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space"
baichuan chatglm chatgpt explainable-ai gpt-4 hallucination hallucinations language-model llama llama2 llama3 llm llm-inference llms mistral representation safety truthfulness
Last synced: 27 Aug 2026
https://github.com/Dino-Kupinic/blackrose
fastapi llama3 meta-ai ollama python3
Last synced: 27 Aug 2026
https://github.com/FuxiaoLiu/LRV-Instruction
[ICLR'24] Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning
chatgpt evaluation evaluation-metrics foundation-models gpt gpt-4 hallucination iclr iclr2024 llama llava multimodal object-detection prompt-engineering vicuna vision vision-and-language vqa
Last synced: 27 Aug 2026
https://github.com/aj-archipelago/cortex
Open-source AI backend control plane for model routing, agent tools, OpenAI-compatible APIs, and private workspaces.
ai ai-agent ai-workspace anthropic-compatible chatgpt claude entities gemini graphql llama llm llm-router mcp model-router multimodal openai openai-compatible rest-api router vertex-ai
Last synced: 26 Aug 2026
https://github.com/nandxorandor/chatbot_VS_chatbot
This project showcases engaging interactions between two AI chatbots.
ai-automation chat chatbot chatbot-development chatgpt conversational-ai flask llama model python-chatbot zephyr
Last synced: 27 Aug 2026
https://github.com/Yxxxb/VoCo-LLaMA
[CVPR'2025] VoCo-LLaMA: This repo is the official implementation of "VoCo-LLaMA: Towards Vision Compression with Large Language Models".
image-compression llama llava
Last synced: 27 Aug 2026
https://github.com/titanml/takeoff-community
TitanML Takeoff Server is an optimization, compression and deployment platform that makes state of the art machine learning models accessible to everyone.
deployment llama llm python quantization
Last synced: 27 Aug 2026
https://github.com/helixml/helix
♾️ Private Agent Fleet with Spec Coding. Each agent gets their own GPU-accelerated desktop. Run Claude, Codex, Gemini and open models on a full private AI Stack ♾️
agents api genai glm golang helm k8s kimi llm llm-agent llm-serving openai openapi qwen rag self-hosted swagger swarm
Last synced: 27 Aug 2026
https://github.com/iSiddharth20/LLM-Chatbot
Enables users to interact with the LLM via Ollama by implementating a client-server architecture utilizing FastAPI as server-side framework and Streamlit for user interface.
client-server client-server-architecture fastapi hacktoberfest large-language-models linux-server llama-index llama3 natural-language-processing ollama ollama-gui streamlit
Last synced: 27 Aug 2026
https://github.com/momegas/megabots
🤖 State-of-the-art, production ready LLM apps made mega-easy, so you don't have to build them from scratch 🤯 Create a bot, now 🫵
chatbot faiss fastapi gpt-35-turbo gpt-4 information-retrieval langchain llama natural-language-processing nlp pinecone prompt-engineering python question-answering s3
Last synced: 27 Aug 2026
https://github.com/MajidRaimi/Chat-With-PDF
Small project where you can chat with pdf using langchain framework.
chroma langchain llama2 llm python
Last synced: 27 Aug 2026
https://github.com/Anirudh1905/Databot
LLM based RAG pipeline
chromadb fastapi langchain llama2 llm openai-api rag streamlit vectordb
Last synced: 27 Aug 2026
https://github.com/YeonwooSung/ai_book
AI book for everyone
ai ai-ml cv deep-learning gpt-4 knowledge-distillation llama llm llmops machine-learning machinelearning mlops mlops-workflow nlp pytorch sagemaker tutorial xai
Last synced: 27 Aug 2026
https://github.com/CMKRG/QiZhenGPT
QiZhenGPT: An Open Source Chinese Medical Large Language Model|一个开源的中文医疗大语言模型
alpaca chatgpt chinese-medical disease gpt llama llm medicine
Last synced: 27 Aug 2026
https://github.com/albertstarfield/project-zephyrine
Zephyrine: An augmented Agentic Assistant GNC framework system for experimental Aircraft. Engineered for processing of navigation loops and sensor fusion and adaptation to bridge the gap between strategic virtual path and action planning and real-world kinematic control.
adacore assistant aviation-tools control-systems flightsystem ifcs openai-api realtime uav
Last synced: 27 Aug 2026
https://github.com/neel1996/lagoon
Llama2 + Supabase powered tool to ingest and chat with github documents
generative-ai llama2 llms supabase
Last synced: 27 Aug 2026
https://github.com/candle-org/step_into_llm
MindSpore online courses: Step into LLM
bert chatglm chatglm2 chatgpt codegeex gpt gpt2 instruction-tuning large-language-models llama llama2 llm mindspore moe natural-language-processing nlp parallel-computing peft prompt-tuning rlhf
Last synced: 27 Aug 2026
https://github.com/kwaroran/RisuAI
Make your own story. User-friendly software for LLM roleplaying
ai characters chat chatbot claude gemini gpt llama llm mcp mcp-client mistral roleplay tauri
Last synced: 27 Aug 2026
https://github.com/minggnim/nlp-models
A repository for training transformer based models
chatbot chatbots ctransformers deeplearning falcon fine-tuning gpt-2 langchain llama2 llms multi-label-classification multi-task-learning nlp pytorch qdrant-vector-database transformers
Last synced: 27 Aug 2026
https://github.com/duongnguyen-dev/WhisBot
Your Voice Assistant
langchain llama speech-reco text-generation voice-assistant
Last synced: 27 Aug 2026
https://github.com/eidolon-ai/eidolon
The first AI Agent Server, Eidolon is a pluggable Agent SDK and enterprise ready, deployment server for Agentic applications
agents generative-ai langchain llama llm openai python services
Last synced: 27 Aug 2026
https://github.com/ModelTC/lightllm
LightLLM is a Python-based LLM (Large Language Model) inference and serving framework, notable for its lightweight design, easy scalability, and high-speed performance.
deep-learning gpt llama llm model-serving nlp openai-triton
Last synced: 27 Aug 2026
https://github.com/mirpo/fastapi-gen
Build LLM-enabled FastAPI applications without build configuration.
cli fastapi gemma-2b huggingface langchain langchain-python llama llama2 llamacpp llm named-entity-recognition ner nlp smollm summarization text-generation
Last synced: 27 Aug 2026
https://github.com/Dicklesworthstone/swiss_army_llama
A FastAPI service for semantic text search using precomputed embeddings and advanced similarity measures, with built-in support for various file types through textract.
embedding-similarity embedding-vectors embeddings llama2 llamacpp semantic-search
Last synced: 27 Aug 2026
https://github.com/mbzuai-oryx/VideoGPT-plus
Official Repository of paper VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding
chatbot clip dual-encoder gpt4 gpt4o image-encoder llama3 llava multimodal phi-3-mini vicuna video-chatbot video-conversation video-encoder vision-language vision-language-pretraining
Last synced: 27 Aug 2026
https://github.com/thomas-yanxin/KarmaVLM
🧘🏻♂️KarmaVLM (相生):A family of high efficiency and powerful visual language model.
llama2 llava multimodel qwen2 vision-language-model visual-language-learning vlm
Last synced: 27 Aug 2026
https://github.com/icip-cas/ChatAlpaca
A Multi-Turn Dialogue Corpus based on Alpaca Instructions
Last synced: 27 Aug 2026
https://github.com/lenML/Speech-AI-Forge
🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.
agent asr chattts chattts-forge chinese colab cosy-voice cosyvoice english firered fireredtts fish-speech gpt llama llm ssml stt text-to-speech tts whisper
Last synced: 27 Aug 2026
https://github.com/Mahesh3394/query_response_generation_llama
In this project we hosted LLAMA model with 7B parameter for response generation. Here we created a rest api which can generate a response when provided a query text.
ctransformers fastapi langchain llama2 llm nlp pytorch rest-api
Last synced: 27 Aug 2026
https://github.com/mlc-ai/mlc-llm
Universal LLM Deployment Engine with ML Compilation
language-model llm machine-learning-compilation tvm
Last synced: 27 Aug 2026
https://github.com/nomic-ai/gpt4all
GPT4All: Run Local LLMs on Any Device. Open-source and available for commercial use.
ai-chat llm-inference
Last synced: 27 Aug 2026
https://github.com/chatchat-space/Langchain-Chatchat
Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain-ChatGLM), local knowledge based LLM (like ChatGLM, Qwen and Llama) RAG and Agent app with langchain
chatbot chatchat chatglm chatgpt embedding faiss fastchat gpt knowledge-base langchain langchain-chatglm llama llm milvus ollama qwen rag retrieval-augmented-generation streamlit xinference
Last synced: 27 Aug 2026
https://github.com/AkiRusProd/llm-agent
LLM using long-term memory through vector database
chromadb gpt gpt4all intent-classification large-language-models llama llm machine-learning nlp rag
Last synced: 27 Aug 2026
https://github.com/ParthSareen/ducky
Natural language to bash commands. Run, understand, copy bash commands generated by an LLM
agent ai-agent bash llm ollama
Last synced: 27 Aug 2026
https://github.com/InternRobotics/PointLLM
[ECCV 2024 Best Paper Candidate & TPAMI 2025] PointLLM: Empowering Large Language Models to Understand Point Clouds
3d chatbot foundation-models gpt-4 large-language-models llama multimodal objaverse point-cloud pointllm representation-learning vision-and-language
Last synced: 27 Aug 2026
https://github.com/sinanazem/llm-apps
This repository showcases projects leveraging advanced language models.
chatbots llama2 ollama rag
Last synced: 27 Aug 2026
https://github.com/LLM-Semantic-Router/vllm-router
vLLM Router
huggingface kubernetes llama2 llm llm-inference vllm
Last synced: 27 Aug 2026
https://github.com/FoundationVision/Groma
[ECCV2024] Grounded Multimodal Large Language Model with Localized Visual Tokenization
foundation-models grounding large-language-models llama llama2 llm mllm multimodal vision-language-model
Last synced: 27 Aug 2026
https://github.com/vtuber-plan/langport
Langport is a language model inference service
api chatgpt chatgpt-api fauxpilot langchain language-model llama llama-cpp llm openai tabby
Last synced: 27 Aug 2026
https://github.com/jorge-armando-navarro-flores/chat_with_your_docs
Discover and converse with advanced AI models like Mistral, LLAMA2, and GPT-3.5 from leading sources like OLLAMA, Hugging Face, and OpenAI. Easily extract insights from PDFs, web pages, and YouTube videos with our intuitive interface. Unlock the power of knowledge with seamless chat interactions.
chatbot docs faiss gemma gpt-3-5-turbo gpt-4 gradio huggingface langchain llama2 llms mistral ollama openai pdf python retrieval-chatbot vectorstore web youtube
Last synced: 27 Aug 2026
https://github.com/c0sogi/llama-api
An OpenAI-like LLaMA inference API
api exllama fastapi llama llamacpp
Last synced: 27 Aug 2026
https://github.com/oobabooga/textgen
Open-source desktop app for local LLMs. Text, vision, tool-calling, OpenAI/Anthropic-compatible API. 100% private.
Last synced: 27 Aug 2026
https://github.com/nikolamilosevic86/local-genAI-search
Local-GenAI-Search is a generative search engine based on Llama 3, langchain and qdrant that answers questions based on your local files
generative-ai langchain large-language-models llama3 local msmarco python3 qdrant-client search-engine sentence-embeddings sentence-transformers
Last synced: 27 Aug 2026
https://github.com/the-open-agent/openagent
⚡️next-generation personal AI assistant powered by LLM, RAG and agent loops, supporting computer-use, browser-use and coding agent, demo: https://demo.openagentai.org
agent agentic agentic-ai agi chatbot chatgpt gpt harness hermes-agent knowledge-base langchain llm mcp model-context-protocol multi-agent openagent openai openclaw rag
Last synced: 27 Aug 2026
https://github.com/yueying-teng/llama-streamlit
Hashtags recommendation based on item titles using LLaMA
langchain llama2 llamacpp recommendation streamlit
Last synced: 27 Aug 2026
https://github.com/vemonet/libre-chat
🦙 Free and Open Source Large Language Model (LLM) chatbot web UI and API. Self-hosted, offline capable and easy to setup.
chatbot chatgpt langchain large-language-models llm llm-inference openapi self-hosted
Last synced: 27 Aug 2026
https://github.com/joshaustintech/python-assistant
Python development AI assistant built on CodeLlama
gradio-interface llama2 python
Last synced: 27 Aug 2026
https://github.com/milinddeore/ner-anon-mode
This model facilitates both data anonymization and the incorporation of synthetic data.
anonymization llama2 llm named-entity-recognition ner synthetic-data
Last synced: 27 Aug 2026
https://github.com/Ramseths/app-llama2
Generative AI - LLaMA 2 7B & LangChain, to generate stories based on a genre.
artificial-intelligence generative-ai github gradio langchain large-language-models learn llama2 meta natural-language-processing python student-vscode transformers
Last synced: 27 Aug 2026
https://github.com/chatchat-space/langchain-chatchat
Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain-ChatGLM), local knowledge based LLM (like ChatGLM, Qwen and Llama) RAG and Agent app with langchain
chatbot chatchat chatglm chatgpt embedding faiss fastchat gpt knowledge-base langchain langchain-chatglm llama llm milvus ollama qwen rag retrieval-augmented-generation streamlit xinference
Last synced: 27 Aug 2026
https://github.com/kyegomez/Exa
Unleash the full potential of exascale LLMs on consumer-class GPUs, proven by extensive benchmarks, with no long-term adjustments and minimal learning curve.
inference-engine llama2 llama2-7b llamacpp llamas llm-inference llms opensource
Last synced: 27 Aug 2026
https://github.com/HROlive/Poland-End-To-End-LLM-Bootcamp
This bootcamp is designed to give NLP researchers an end-to-end overview on the fundamentals of NVIDIA NeMo framework, complete solution for building large language models. It will also have hands-on exercises complimented by tutorials, code snippets, and presentations to help researchers kick-start with NeMo LLM Service and Guardrails.
gpt llama2 llm llm-inference llm-training nemo-guardrails nvidia nvidia-nemo p-tuning prompt-tuning tensorrt triton
Last synced: 27 Aug 2026
https://github.com/apocas/restai
RESTai is an AIaaS (AI as a Service) open-source platform. Supports many public and local LLM suported by Ollama/vLLM/etc. Precise embeddings usage, tuning, analytics etc. Built-in image/audio generation with dynamic loading generators. Live chat deployment. Built-in block based graphical language. Prompt versioning and much more...
blocky embeddings fastapi langchain llama llamaindex llm ollama openai openaiapi python rag stable-diffusion transformers
Last synced: 27 Aug 2026
https://github.com/emilsharkov/ChatGPTAPIClone
This is a scalable and fault tolerant API cluster that answers prompts similarly to ChatGPT, built with Terraform and Python
aws docker-image kubernetes llama2 python terraform
Last synced: 27 Aug 2026
https://github.com/areebahmed575/Learn-Generative-AI
This repo contains almost everything about Generative Ai
chromadb docker fastapi gpts huggingface kafka langchain llama2 llm oauth2 objectbox openai pinecone poetry python sqlachemy sqlalchemy sqlmodel streamlit vector-database
Last synced: 27 Aug 2026
https://github.com/aws-samples/generative-ai-use-cases
Application implementation with business use cases for safely utilizing generative AI in business operations
aws bedrock chatbot claude claude3 claude4 command-r deepseek-r1 generative-ai image-generation lambda llama3 llm mistral nova rag react sagemaker typescript
Last synced: 27 Aug 2026
https://github.com/alibaba/rtp-llm
RTP-LLM: Alibaba's high-performance LLM inference engine for diverse applications.
gpt inference llama llm llm-serving llmops model-serving
Last synced: 27 Aug 2026
https://github.com/expectedparrot/edsl
Design, conduct and analyze results of AI-powered surveys and experiments. Simulate social science and market research with large numbers of AI agents and LLMs.
anthropic data-labeling deepinfra domain-specific-language experiments llama2 llm llm-agent llm-framework llm-inference market-research mixtral open-source openai python social-science surveys synthetic-data
Last synced: 27 Aug 2026
https://github.com/meta-llama/llama-cookbook
Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama model family and using them on various provider services
ai finetuning langchain llama llama2 llm machine-learning python pytorch vllm
Last synced: 27 Aug 2026
https://github.com/tien02/llm-math
Fine tune Large Language Model on Mathematic dataset
huggingface llama llama2 llm lora mathematics supervised-finetuning transformer
Last synced: 27 Aug 2026
https://github.com/zetavg/llama-lora-tuner
UI tool for fine-tuning and testing your own LoRA models base on LLaMA, GPT-J and more. One-click run on Google Colab. + A Gradio ChatGPT-like Chat UI to demonstrate your language models.
ai alpaca alpaca-lora google-colab gpt gpt-j language-model llama lora machine-learning peft
Last synced: 27 Aug 2026
https://github.com/silvanmelchior/IncognitoPilot
An AI code interpreter for sensitive data, powered by GPT-4 or Code Llama / Llama 2.
ai chatgpt chatgpt-code-interpreter codellama copilot gpt-4 llama2 llm python
Last synced: 27 Aug 2026
https://github.com/ramchaik/facts
Facts is a simple RAG using Langchain, Ollama, and Chroma to answer queries based on provided data.
chromadb embeddings langchain llama3 ollama rag vector-database
Last synced: 27 Aug 2026
https://github.com/theodo-group/GenossGPT
One API for all LLMs either Private or Public (Anthropic, Llama V2, GPT 3.5/4, Vertex, GPT4ALL, HuggingFace ...) 🌈🐂 Replace OpenAI GPT with any LLMs in your app with one line.
api gpt gpt4all huggingface inference llama llm openai private public
Last synced: 27 Aug 2026
https://github.com/georgian-io/LLM-Finetuning-Toolkit
Toolkit for fine-tuning, ablating and unit-testing open-source LLMs.
ablation-study classification falcon fine-tuning finetuning flan-t5 large-language-models llama2 llm-test lora mistral-7b nlp nlp-machine-learning qlora redpajama summarization unit-testing zephyr
Last synced: 27 Aug 2026
https://github.com/vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
amd blackwell cuda deepseek deepseek-v3 gpt gpt-oss inference kimi llama llm llm-serving model-serving moe openai pytorch qwen qwen3 tpu transformer
Last synced: 27 Aug 2026
https://github.com/xusenlinzy/api-for-open-llm
Openai style api for open large language models, using LLMs just as chatgpt! Support for LLaMA, LLaMA-2, BLOOM, Falcon, Baichuan, Qwen, Xverse, SqlCoder, CodeLLaMA, ChatGLM, ChatGLM2, ChatGLM3 etc. 开源大模型的统一后端接口
baichuan chatglm code-llama docker internlm langchain llama llama2 llms nlp openai qwen sqlcoder xverse
Last synced: 27 Aug 2026
https://github.com/JuliusHaring/chatbot-template
A comprehensive chatbot system with integrated LLM querying and Messenger Bot interfacing capabilities. To be used as a template for implementation.
chatbot chatgpt chatgpt-api llama llamaindex telegram telegram-bot vectorstore
Last synced: 26 Aug 2026
https://github.com/jindalAnuj/LLama3-AIEmailAgent
LLama3 agent using crewai , detect spam email and auto responsd if important
agent crewai llama3 python
Last synced: 27 Aug 2026
https://github.com/LennardZuendorf/thesis-webapp
Webapp/Application implemention of my thesis about XAI and Interpretability of Transformer Models.
bertviz gradio huggingface interpretable-ai llama2 mistral shap xai
Last synced: 27 Aug 2026
https://github.com/ray-project/llm-applications
A comprehensive guide to building RAG-based LLM applications for production.
anyscale fine-tuning llama2 llms machine-learning openai ray serving
Last synced: 27 Aug 2026
https://github.com/davzoku/cria
An end-to-end LLM app prototype based on Llama 2
artificial-intelligence chatbot llama2 llm nextjs transformers
Last synced: 27 Aug 2026
https://github.com/ddh0/easy-llama
Python package wrapping llama.cpp for on-device LLM inference
llama llamacpp llm llms text-generation
Last synced: 27 Aug 2026
https://github.com/OwlAIProject/Owl
A personal wearable AI that runs locally
ai ble bluetooth esp32 llama2 mistral nrf52840 ollama wearable whisper
Last synced: 27 Aug 2026
https://github.com/sooraj12/ragchatbot_llama3
rag
agent langchain llama3 rag
Last synced: 27 Aug 2026
https://github.com/anslin-raj/anil_ai_chat_llama_2
Anil AI is an innovative chat application that harnesses the robust capabilities of Generative AI, utilizing the advanced LLAMA 2 model for inference.
automation chatbot devops docker generative-ai gpt llama2 llm ml mlops
Last synced: 27 Aug 2026
https://github.com/luogen1996/LLaVA-HR
[ICLR2025] LLaVA-HR: High-Resolution Large Language-Vision Assistant
high-resolution-mllms llama2 llms multimodal-chatbot
Last synced: 27 Aug 2026
https://github.com/bolna-ai/bolna
Conversational voice AI agents
agentic-ai agents ai-agents cartesia conversational-ai deepgram deepseek deepseek-chat elevenlabs function-calling gpt-4 llama openai plivo twilio voice-agents voice-ai-agents voice-assistant whisper
Last synced: 27 Aug 2026
https://github.com/boostcampaitech5/LawBot-Online-Legal-Advice-LLM-Service
사용자가 채팅웹을 통해 자신이 처한 법률적 상황을 제시하면, 입력에 대한 문맥을 모델이 이해하여 가이드라인을 제시하고, 유사한 상황의 판례를 제공하는 웹 서비스입니다. (2023.08.18 서비스 종료)
airflow docker llama2 reactjs selenium tailwindcss
Last synced: 27 Aug 2026
https://github.com/UbiquitousLearning/mllm
Fast Multimodal LLM on Mobile Devices
ai llama llm mobile multimodal
Last synced: 27 Aug 2026
https://github.com/modular/modular
The Modular Platform (includes MAX & Mojo)
ai language machine-learning max modular mojo programming-language
Last synced: 27 Aug 2026
https://github.com/xorbitsai/inference
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.
artificial-intelligence deployment diffusers gemma glm glm-5-3 inference kimi kimi-k3 llama-cpp llamacpp llm machine-learning openai-api pytorch qwen sglang transformers vllm whisper
Last synced: 27 Aug 2026
https://github.com/PaddlePaddle/PaddleNLP
Easy-to-use and powerful LLM and SLM library with awesome model zoo.
bert compression distributed-training document-intelligence embedding ernie information-extraction llama llm neural-search nlp paddlenlp pretrained-models question-answering search-engine semantic-analysis sentiment-analysis transformers uie
Last synced: 27 Aug 2026
https://github.com/nptt9/illama
A fast, lightweight, parallel inference server for Llama LLMs.
exllama exllamav2 flash-attention-2 inference llama llama2 llama3 llm-inference paged-attention server
Last synced: 27 Aug 2026
https://github.com/TinyLLaVA/TinyLLaVA_Factory
A Framework of Small-scale Large Multimodal Models
large-multimodal-models llama llava nlp tinyllama transformers vision-language
Last synced: 27 Aug 2026
https://github.com/FunnySaltyFish/best_llm
Vote the Best LLM by yourself! 票选你最喜欢的大语言模型
ai artificial-intelligence chatgpt gpt-4 llama llm natural-language-generation nature-language-processing nature-language-understanding nlp openai transformer
Last synced: 27 Aug 2026
Statistics
- Projects: 2,239
- Last updated: about 2 years ago