An open API service for producing an overview of a list of open source projects.

Collections: awesome-llama

https://github.com/thejasmeetsingh/moody-llm

A LLM whose mood keeps changing.

asynchronous-programming fastapi genai highlightjs langchain llama3 llm ollama pydantic python3 reactjs restful-api supabase tailwindcss websockets

Last synced: 27 Aug 2026

https://github.com/X-PLUG/mPLUG-Owl

mPLUG-Owl: The Powerful Multi-modal Large Language Model Family

alpaca chatbot chatgpt damo dialogue gpt gpt4 gpt4-api huggingface instruction-tuning large-language-models llama mplug mplug-owl multimodal pretraining pytorch transformer video visual-recognition

Last synced: 27 Aug 2026

https://github.com/WangRongsheng/CareGPT

🌞 CareGPT (关怀GPT)是一个医疗大语言模型,同时它集合了数十个公开可用的医疗微调数据集和开放可用的医疗大语言模型,包含LLM的训练、测评、部署等以促进医疗LLM快速发展。Medical LLM, Open Source Driven for a Healthy Future.

baichuan gpt large-language-models llama llama2 medical-llm

Last synced: 27 Aug 2026

https://github.com/declare-lab/red-instruct

Codes and datasets of the paper Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment

huggingface-transformers llama llama2 llm llms

Last synced: 27 Aug 2026

https://github.com/zetavg/LLaMA-LoRA-Tuner

UI tool for fine-tuning and testing your own LoRA models base on LLaMA, GPT-J and more. One-click run on Google Colab. + A Gradio ChatGPT-like Chat UI to demonstrate your language models.

ai alpaca alpaca-lora google-colab gpt gpt-j language-model llama lora machine-learning peft

Last synced: 27 Aug 2026

https://github.com/ATH-MaaS/Ovis

A novel Multimodal Large Language Model (MLLM) architecture, designed to structurally align visual and textual embeddings.

chatbot llama3 multimodal multimodal-large-language-models multimodality qwen vision-language-learning vision-language-model

Last synced: 27 Aug 2026

https://github.com/Azure-Samples/miyagi

Sample to envision intelligent apps with Microsoft's Copilot stack for AI-infused product experiences.

agents aks assistants azure azure-openai azureai copilot gpt-4 guidance langchain llama-index llama2 openai phi-2 prompt-engineering promptflow semantic-kernel taskweaver typechat

Last synced: 27 Aug 2026

https://github.com/kennethleungty/Llama-2-Open-Source-LLM-CPU-Inference

Running Llama 2 and other Open-Source LLMs on CPU Inference Locally for Document Q&A

c-transformers chatgpt cpu cpu-inference deep-learning document-qa faiss langchain language-models large-language-models llama llama-2 llm machine-learning natural-language-processing nlp open-source-llm python sentence-transformers transformers

Last synced: 27 Aug 2026

https://github.com/NVIDIA/nim-anywhere

Accelerate your Gen AI with NVIDIA NIM and NVIDIA AI Workbench

genai langchain lcel llama llama3 llm nim nvidia nvwb-project rag

Last synced: 27 Aug 2026

https://github.com/MetaGLM/langchain-zhipuai

基于 Langchain,快速集成GLM-4 AllTools 功能的插件

chatbot glm gpt gpt-4 langchain llama ollama rag

Last synced: 27 Aug 2026

https://github.com/henrywoo/pyllama

LLaMA: Open and Efficient Foundation Language Models

Last synced: 27 Aug 2026

https://github.com/EmbeddedLLM/embeddedllm

EmbeddedLLM: API server for Embedded Device Deployment. Currently support CUDA/OpenVINO/IpexLLM/DirectML/CPU

aipc cpu directml directx-12 gemma ipexllm llama llm llm-inference llm-serving mistral model-inference npu open-source-llm openvino openvino-inference-engine phi-3 windows

Last synced: 27 Aug 2026

https://github.com/flojud/DocsChat

The chatbot utilizes a conversational retrieval chain to answer user queries based on the content of embedded documents. It leverages various NLP techniques, including language models and embeddings, to provide relevant responses.

Last synced: 27 Aug 2026

https://github.com/ictnlp/TruthX

Code for ACL 2024 paper "TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space"

baichuan chatglm chatgpt explainable-ai gpt-4 hallucination hallucinations language-model llama llama2 llama3 llm llm-inference llms mistral representation safety truthfulness

Last synced: 27 Aug 2026

https://github.com/Dino-Kupinic/blackrose

fastapi llama3 meta-ai ollama python3

Last synced: 27 Aug 2026

https://github.com/FuxiaoLiu/LRV-Instruction

[ICLR'24] Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

chatgpt evaluation evaluation-metrics foundation-models gpt gpt-4 hallucination iclr iclr2024 llama llava multimodal object-detection prompt-engineering vicuna vision vision-and-language vqa

Last synced: 27 Aug 2026

https://github.com/aj-archipelago/cortex

Open-source AI backend control plane for model routing, agent tools, OpenAI-compatible APIs, and private workspaces.

ai ai-agent ai-workspace anthropic-compatible chatgpt claude entities gemini graphql llama llm llm-router mcp model-router multimodal openai openai-compatible rest-api router vertex-ai

Last synced: 26 Aug 2026

https://github.com/nandxorandor/chatbot_VS_chatbot

This project showcases engaging interactions between two AI chatbots.

ai-automation chat chatbot chatbot-development chatgpt conversational-ai flask llama model python-chatbot zephyr

Last synced: 27 Aug 2026

https://github.com/Yxxxb/VoCo-LLaMA

[CVPR'2025] VoCo-LLaMA: This repo is the official implementation of "VoCo-LLaMA: Towards Vision Compression with Large Language Models".

image-compression llama llava

Last synced: 27 Aug 2026

https://github.com/titanml/takeoff-community

TitanML Takeoff Server is an optimization, compression and deployment platform that makes state of the art machine learning models accessible to everyone.

deployment llama llm python quantization

Last synced: 27 Aug 2026

https://github.com/helixml/helix

♾️ Private Agent Fleet with Spec Coding. Each agent gets their own GPU-accelerated desktop. Run Claude, Codex, Gemini and open models on a full private AI Stack ♾️

agents api genai glm golang helm k8s kimi llm llm-agent llm-serving openai openapi qwen rag self-hosted swagger swarm

Last synced: 27 Aug 2026

https://github.com/iSiddharth20/LLM-Chatbot

Enables users to interact with the LLM via Ollama by implementating a client-server architecture utilizing FastAPI as server-side framework and Streamlit for user interface.

client-server client-server-architecture fastapi hacktoberfest large-language-models linux-server llama-index llama3 natural-language-processing ollama ollama-gui streamlit

Last synced: 27 Aug 2026

https://github.com/momegas/megabots

🤖 State-of-the-art, production ready LLM apps made mega-easy, so you don't have to build them from scratch 🤯 Create a bot, now 🫵

chatbot faiss fastapi gpt-35-turbo gpt-4 information-retrieval langchain llama natural-language-processing nlp pinecone prompt-engineering python question-answering s3

Last synced: 27 Aug 2026

https://github.com/MajidRaimi/Chat-With-PDF

Small project where you can chat with pdf using langchain framework.

chroma langchain llama2 llm python

Last synced: 27 Aug 2026

https://github.com/Anirudh1905/Databot

LLM based RAG pipeline

chromadb fastapi langchain llama2 llm openai-api rag streamlit vectordb

Last synced: 27 Aug 2026

https://github.com/YeonwooSung/ai_book

AI book for everyone

ai ai-ml cv deep-learning gpt-4 knowledge-distillation llama llm llmops machine-learning machinelearning mlops mlops-workflow nlp pytorch sagemaker tutorial xai

Last synced: 27 Aug 2026

https://github.com/CMKRG/QiZhenGPT

QiZhenGPT: An Open Source Chinese Medical Large Language Model|一个开源的中文医疗大语言模型

alpaca chatgpt chinese-medical disease gpt llama llm medicine

Last synced: 27 Aug 2026

https://github.com/albertstarfield/project-zephyrine

Zephyrine: An augmented Agentic Assistant GNC framework system for experimental Aircraft. Engineered for processing of navigation loops and sensor fusion and adaptation to bridge the gap between strategic virtual path and action planning and real-world kinematic control.

adacore assistant aviation-tools control-systems flightsystem ifcs openai-api realtime uav

Last synced: 27 Aug 2026

https://github.com/neel1996/lagoon

Llama2 + Supabase powered tool to ingest and chat with github documents

generative-ai llama2 llms supabase

Last synced: 27 Aug 2026

https://github.com/candle-org/step_into_llm

MindSpore online courses: Step into LLM

bert chatglm chatglm2 chatgpt codegeex gpt gpt2 instruction-tuning large-language-models llama llama2 llm mindspore moe natural-language-processing nlp parallel-computing peft prompt-tuning rlhf

Last synced: 27 Aug 2026

https://github.com/kwaroran/RisuAI

Make your own story. User-friendly software for LLM roleplaying

ai characters chat chatbot claude gemini gpt llama llm mcp mcp-client mistral roleplay tauri

Last synced: 27 Aug 2026

https://github.com/minggnim/nlp-models

A repository for training transformer based models

chatbot chatbots ctransformers deeplearning falcon fine-tuning gpt-2 langchain llama2 llms multi-label-classification multi-task-learning nlp pytorch qdrant-vector-database transformers

Last synced: 27 Aug 2026

https://github.com/duongnguyen-dev/WhisBot

Your Voice Assistant

langchain llama speech-reco text-generation voice-assistant

Last synced: 27 Aug 2026

https://github.com/eidolon-ai/eidolon

The first AI Agent Server, Eidolon is a pluggable Agent SDK and enterprise ready, deployment server for Agentic applications

agents generative-ai langchain llama llm openai python services

Last synced: 27 Aug 2026

https://github.com/ModelTC/lightllm

LightLLM is a Python-based LLM (Large Language Model) inference and serving framework, notable for its lightweight design, easy scalability, and high-speed performance.

deep-learning gpt llama llm model-serving nlp openai-triton

Last synced: 27 Aug 2026

https://github.com/mirpo/fastapi-gen

Build LLM-enabled FastAPI applications without build configuration.

cli fastapi gemma-2b huggingface langchain langchain-python llama llama2 llamacpp llm named-entity-recognition ner nlp smollm summarization text-generation

Last synced: 27 Aug 2026

https://github.com/Dicklesworthstone/swiss_army_llama

A FastAPI service for semantic text search using precomputed embeddings and advanced similarity measures, with built-in support for various file types through textract.

embedding-similarity embedding-vectors embeddings llama2 llamacpp semantic-search

Last synced: 27 Aug 2026

https://github.com/mbzuai-oryx/VideoGPT-plus

Official Repository of paper VideoGPT+: Integrating Image and Video Encoders for Enhanced Video Understanding

chatbot clip dual-encoder gpt4 gpt4o image-encoder llama3 llava multimodal phi-3-mini vicuna video-chatbot video-conversation video-encoder vision-language vision-language-pretraining

Last synced: 27 Aug 2026

https://github.com/thomas-yanxin/KarmaVLM

🧘🏻‍♂️KarmaVLM (相生):A family of high efficiency and powerful visual language model.

llama2 llava multimodel qwen2 vision-language-model visual-language-learning vlm

Last synced: 27 Aug 2026

https://github.com/icip-cas/ChatAlpaca

A Multi-Turn Dialogue Corpus based on Alpaca Instructions

Last synced: 27 Aug 2026

https://github.com/lenML/Speech-AI-Forge

🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.

agent asr chattts chattts-forge chinese colab cosy-voice cosyvoice english firered fireredtts fish-speech gpt llama llm ssml stt text-to-speech tts whisper

Last synced: 27 Aug 2026

https://github.com/Mahesh3394/query_response_generation_llama

In this project we hosted LLAMA model with 7B parameter for response generation. Here we created a rest api which can generate a response when provided a query text.

ctransformers fastapi langchain llama2 llm nlp pytorch rest-api

Last synced: 27 Aug 2026

https://github.com/mlc-ai/mlc-llm

Universal LLM Deployment Engine with ML Compilation

language-model llm machine-learning-compilation tvm

Last synced: 27 Aug 2026

https://github.com/nomic-ai/gpt4all

GPT4All: Run Local LLMs on Any Device. Open-source and available for commercial use.

ai-chat llm-inference

Last synced: 27 Aug 2026

https://github.com/chatchat-space/Langchain-Chatchat

Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain-ChatGLM), local knowledge based LLM (like ChatGLM, Qwen and Llama) RAG and Agent app with langchain

chatbot chatchat chatglm chatgpt embedding faiss fastchat gpt knowledge-base langchain langchain-chatglm llama llm milvus ollama qwen rag retrieval-augmented-generation streamlit xinference

Last synced: 27 Aug 2026

https://github.com/AkiRusProd/llm-agent

LLM using long-term memory through vector database

chromadb gpt gpt4all intent-classification large-language-models llama llm machine-learning nlp rag

Last synced: 27 Aug 2026

https://github.com/ParthSareen/ducky

Natural language to bash commands. Run, understand, copy bash commands generated by an LLM

agent ai-agent bash llm ollama

Last synced: 27 Aug 2026

https://github.com/InternRobotics/PointLLM

[ECCV 2024 Best Paper Candidate & TPAMI 2025] PointLLM: Empowering Large Language Models to Understand Point Clouds

3d chatbot foundation-models gpt-4 large-language-models llama multimodal objaverse point-cloud pointllm representation-learning vision-and-language

Last synced: 27 Aug 2026

https://github.com/sinanazem/llm-apps

This repository showcases projects leveraging advanced language models.

chatbots llama2 ollama rag

Last synced: 27 Aug 2026

https://github.com/LLM-Semantic-Router/vllm-router

vLLM Router

huggingface kubernetes llama2 llm llm-inference vllm

Last synced: 27 Aug 2026

https://github.com/FoundationVision/Groma

[ECCV2024] Grounded Multimodal Large Language Model with Localized Visual Tokenization

foundation-models grounding large-language-models llama llama2 llm mllm multimodal vision-language-model

Last synced: 27 Aug 2026

https://github.com/vtuber-plan/langport

Langport is a language model inference service

api chatgpt chatgpt-api fauxpilot langchain language-model llama llama-cpp llm openai tabby

Last synced: 27 Aug 2026

https://github.com/jorge-armando-navarro-flores/chat_with_your_docs

Discover and converse with advanced AI models like Mistral, LLAMA2, and GPT-3.5 from leading sources like OLLAMA, Hugging Face, and OpenAI. Easily extract insights from PDFs, web pages, and YouTube videos with our intuitive interface. Unlock the power of knowledge with seamless chat interactions.

chatbot docs faiss gemma gpt-3-5-turbo gpt-4 gradio huggingface langchain llama2 llms mistral ollama openai pdf python retrieval-chatbot vectorstore web youtube

Last synced: 27 Aug 2026

https://github.com/c0sogi/llama-api

An OpenAI-like LLaMA inference API

api exllama fastapi llama llamacpp

Last synced: 27 Aug 2026

https://github.com/oobabooga/textgen

Open-source desktop app for local LLMs. Text, vision, tool-calling, OpenAI/Anthropic-compatible API. 100% private.

Last synced: 27 Aug 2026

https://github.com/nikolamilosevic86/local-genAI-search

Local-GenAI-Search is a generative search engine based on Llama 3, langchain and qdrant that answers questions based on your local files

generative-ai langchain large-language-models llama3 local msmarco python3 qdrant-client search-engine sentence-embeddings sentence-transformers

Last synced: 27 Aug 2026

https://github.com/the-open-agent/openagent

⚡️next-generation personal AI assistant powered by LLM, RAG and agent loops, supporting computer-use, browser-use and coding agent, demo: https://demo.openagentai.org

agent agentic agentic-ai agi chatbot chatgpt gpt harness hermes-agent knowledge-base langchain llm mcp model-context-protocol multi-agent openagent openai openclaw rag

Last synced: 27 Aug 2026

https://github.com/yueying-teng/llama-streamlit

Hashtags recommendation based on item titles using LLaMA

langchain llama2 llamacpp recommendation streamlit

Last synced: 27 Aug 2026

https://github.com/vemonet/libre-chat

🦙 Free and Open Source Large Language Model (LLM) chatbot web UI and API. Self-hosted, offline capable and easy to setup.

chatbot chatgpt langchain large-language-models llm llm-inference openapi self-hosted

Last synced: 27 Aug 2026

https://github.com/joshaustintech/python-assistant

Python development AI assistant built on CodeLlama

gradio-interface llama2 python

Last synced: 27 Aug 2026

https://github.com/milinddeore/ner-anon-mode

This model facilitates both data anonymization and the incorporation of synthetic data.

anonymization llama2 llm named-entity-recognition ner synthetic-data

Last synced: 27 Aug 2026

https://github.com/Ramseths/app-llama2

Generative AI - LLaMA 2 7B & LangChain, to generate stories based on a genre.

artificial-intelligence generative-ai github gradio langchain large-language-models learn llama2 meta natural-language-processing python student-vscode transformers

Last synced: 27 Aug 2026

https://github.com/chatchat-space/langchain-chatchat

Langchain-Chatchat(原Langchain-ChatGLM)基于 Langchain 与 ChatGLM, Qwen 与 Llama 等语言模型的 RAG 与 Agent 应用 | Langchain-Chatchat (formerly langchain-ChatGLM), local knowledge based LLM (like ChatGLM, Qwen and Llama) RAG and Agent app with langchain

chatbot chatchat chatglm chatgpt embedding faiss fastchat gpt knowledge-base langchain langchain-chatglm llama llm milvus ollama qwen rag retrieval-augmented-generation streamlit xinference

Last synced: 27 Aug 2026

https://github.com/kyegomez/Exa

Unleash the full potential of exascale LLMs on consumer-class GPUs, proven by extensive benchmarks, with no long-term adjustments and minimal learning curve.

inference-engine llama2 llama2-7b llamacpp llamas llm-inference llms opensource

Last synced: 27 Aug 2026

https://github.com/HROlive/Poland-End-To-End-LLM-Bootcamp

This bootcamp is designed to give NLP researchers an end-to-end overview on the fundamentals of NVIDIA NeMo framework, complete solution for building large language models. It will also have hands-on exercises complimented by tutorials, code snippets, and presentations to help researchers kick-start with NeMo LLM Service and Guardrails.

gpt llama2 llm llm-inference llm-training nemo-guardrails nvidia nvidia-nemo p-tuning prompt-tuning tensorrt triton

Last synced: 27 Aug 2026

https://github.com/apocas/restai

RESTai is an AIaaS (AI as a Service) open-source platform. Supports many public and local LLM suported by Ollama/vLLM/etc. Precise embeddings usage, tuning, analytics etc. Built-in image/audio generation with dynamic loading generators. Live chat deployment. Built-in block based graphical language. Prompt versioning and much more...

blocky embeddings fastapi langchain llama llamaindex llm ollama openai openaiapi python rag stable-diffusion transformers

Last synced: 27 Aug 2026

https://github.com/emilsharkov/ChatGPTAPIClone

This is a scalable and fault tolerant API cluster that answers prompts similarly to ChatGPT, built with Terraform and Python

aws docker-image kubernetes llama2 python terraform

Last synced: 27 Aug 2026

https://github.com/areebahmed575/Learn-Generative-AI

This repo contains almost everything about Generative Ai

chromadb docker fastapi gpts huggingface kafka langchain llama2 llm oauth2 objectbox openai pinecone poetry python sqlachemy sqlalchemy sqlmodel streamlit vector-database

Last synced: 27 Aug 2026

https://github.com/aws-samples/generative-ai-use-cases

Application implementation with business use cases for safely utilizing generative AI in business operations

aws bedrock chatbot claude claude3 claude4 command-r deepseek-r1 generative-ai image-generation lambda llama3 llm mistral nova rag react sagemaker typescript

Last synced: 27 Aug 2026

https://github.com/alibaba/rtp-llm

RTP-LLM: Alibaba's high-performance LLM inference engine for diverse applications.

gpt inference llama llm llm-serving llmops model-serving

Last synced: 27 Aug 2026

https://github.com/expectedparrot/edsl

Design, conduct and analyze results of AI-powered surveys and experiments. Simulate social science and market research with large numbers of AI agents and LLMs.

anthropic data-labeling deepinfra domain-specific-language experiments llama2 llm llm-agent llm-framework llm-inference market-research mixtral open-source openai python social-science surveys synthetic-data

Last synced: 27 Aug 2026

https://github.com/meta-llama/llama-cookbook

Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama model family and using them on various provider services

ai finetuning langchain llama llama2 llm machine-learning python pytorch vllm

Last synced: 27 Aug 2026

https://github.com/tien02/llm-math

Fine tune Large Language Model on Mathematic dataset

huggingface llama llama2 llm lora mathematics supervised-finetuning transformer

Last synced: 27 Aug 2026

https://github.com/zetavg/llama-lora-tuner

UI tool for fine-tuning and testing your own LoRA models base on LLaMA, GPT-J and more. One-click run on Google Colab. + A Gradio ChatGPT-like Chat UI to demonstrate your language models.

ai alpaca alpaca-lora google-colab gpt gpt-j language-model llama lora machine-learning peft

Last synced: 27 Aug 2026

https://github.com/silvanmelchior/IncognitoPilot

An AI code interpreter for sensitive data, powered by GPT-4 or Code Llama / Llama 2.

ai chatgpt chatgpt-code-interpreter codellama copilot gpt-4 llama2 llm python

Last synced: 27 Aug 2026

https://github.com/ramchaik/facts

Facts is a simple RAG using Langchain, Ollama, and Chroma to answer queries based on provided data.

chromadb embeddings langchain llama3 ollama rag vector-database

Last synced: 27 Aug 2026

https://github.com/theodo-group/GenossGPT

One API for all LLMs either Private or Public (Anthropic, Llama V2, GPT 3.5/4, Vertex, GPT4ALL, HuggingFace ...) 🌈🐂 Replace OpenAI GPT with any LLMs in your app with one line.

api gpt gpt4all huggingface inference llama llm openai private public

Last synced: 27 Aug 2026

https://github.com/georgian-io/LLM-Finetuning-Toolkit

Toolkit for fine-tuning, ablating and unit-testing open-source LLMs.

ablation-study classification falcon fine-tuning finetuning flan-t5 large-language-models llama2 llm-test lora mistral-7b nlp nlp-machine-learning qlora redpajama summarization unit-testing zephyr

Last synced: 27 Aug 2026

https://github.com/vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

amd blackwell cuda deepseek deepseek-v3 gpt gpt-oss inference kimi llama llm llm-serving model-serving moe openai pytorch qwen qwen3 tpu transformer

Last synced: 27 Aug 2026

https://github.com/xusenlinzy/api-for-open-llm

Openai style api for open large language models, using LLMs just as chatgpt! Support for LLaMA, LLaMA-2, BLOOM, Falcon, Baichuan, Qwen, Xverse, SqlCoder, CodeLLaMA, ChatGLM, ChatGLM2, ChatGLM3 etc. 开源大模型的统一后端接口

baichuan chatglm code-llama docker internlm langchain llama llama2 llms nlp openai qwen sqlcoder xverse

Last synced: 27 Aug 2026

https://github.com/JuliusHaring/chatbot-template

A comprehensive chatbot system with integrated LLM querying and Messenger Bot interfacing capabilities. To be used as a template for implementation.

chatbot chatgpt chatgpt-api llama llamaindex telegram telegram-bot vectorstore

Last synced: 26 Aug 2026

https://github.com/jindalAnuj/LLama3-AIEmailAgent

LLama3 agent using crewai , detect spam email and auto responsd if important

agent crewai llama3 python

Last synced: 27 Aug 2026

https://github.com/LennardZuendorf/thesis-webapp

Webapp/Application implemention of my thesis about XAI and Interpretability of Transformer Models.

bertviz gradio huggingface interpretable-ai llama2 mistral shap xai

Last synced: 27 Aug 2026

https://github.com/ray-project/llm-applications

A comprehensive guide to building RAG-based LLM applications for production.

anyscale fine-tuning llama2 llms machine-learning openai ray serving

Last synced: 27 Aug 2026

https://github.com/davzoku/cria

An end-to-end LLM app prototype based on Llama 2

artificial-intelligence chatbot llama2 llm nextjs transformers

Last synced: 27 Aug 2026

https://github.com/ddh0/easy-llama

Python package wrapping llama.cpp for on-device LLM inference

llama llamacpp llm llms text-generation

Last synced: 27 Aug 2026

https://github.com/OwlAIProject/Owl

A personal wearable AI that runs locally

ai ble bluetooth esp32 llama2 mistral nrf52840 ollama wearable whisper

Last synced: 27 Aug 2026

https://github.com/sooraj12/ragchatbot_llama3

rag

agent langchain llama3 rag

Last synced: 27 Aug 2026

https://github.com/anslin-raj/anil_ai_chat_llama_2

Anil AI is an innovative chat application that harnesses the robust capabilities of Generative AI, utilizing the advanced LLAMA 2 model for inference.

automation chatbot devops docker generative-ai gpt llama2 llm ml mlops

Last synced: 27 Aug 2026

https://github.com/luogen1996/LLaVA-HR

[ICLR2025] LLaVA-HR: High-Resolution Large Language-Vision Assistant

high-resolution-mllms llama2 llms multimodal-chatbot

Last synced: 27 Aug 2026

https://github.com/bolna-ai/bolna

Conversational voice AI agents

agentic-ai agents ai-agents cartesia conversational-ai deepgram deepseek deepseek-chat elevenlabs function-calling gpt-4 llama openai plivo twilio voice-agents voice-ai-agents voice-assistant whisper

Last synced: 27 Aug 2026

https://github.com/boostcampaitech5/LawBot-Online-Legal-Advice-LLM-Service

사용자가 채팅웹을 통해 자신이 처한 법률적 상황을 제시하면, 입력에 대한 문맥을 모델이 이해하여 가이드라인을 제시하고, 유사한 상황의 판례를 제공하는 웹 서비스입니다. (2023.08.18 서비스 종료)

airflow docker llama2 reactjs selenium tailwindcss

Last synced: 27 Aug 2026

https://github.com/zhangleinice/chatbot

llama2聊天机器人

langchain llama2

Last synced: 27 Aug 2026

https://github.com/UbiquitousLearning/mllm

Fast Multimodal LLM on Mobile Devices

ai llama llm mobile multimodal

Last synced: 27 Aug 2026

https://github.com/modular/modular

The Modular Platform (includes MAX & Mojo)

ai language machine-learning max modular mojo programming-language

Last synced: 27 Aug 2026

https://github.com/xorbitsai/inference

Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.

artificial-intelligence deployment diffusers gemma glm glm-5-3 inference kimi kimi-k3 llama-cpp llamacpp llm machine-learning openai-api pytorch qwen sglang transformers vllm whisper

Last synced: 27 Aug 2026

https://github.com/PaddlePaddle/PaddleNLP

Easy-to-use and powerful LLM and SLM library with awesome model zoo.

bert compression distributed-training document-intelligence embedding ernie information-extraction llama llm neural-search nlp paddlenlp pretrained-models question-answering search-engine semantic-analysis sentiment-analysis transformers uie

Last synced: 27 Aug 2026

https://github.com/nptt9/illama

A fast, lightweight, parallel inference server for Llama LLMs.

exllama exllamav2 flash-attention-2 inference llama llama2 llama3 llm-inference paged-attention server

Last synced: 27 Aug 2026

https://github.com/TinyLLaVA/TinyLLaVA_Factory

A Framework of Small-scale Large Multimodal Models

large-multimodal-models llama llava nlp tinyllama transformers vision-language

Last synced: 27 Aug 2026

https://github.com/FunnySaltyFish/best_llm

Vote the Best LLM by yourself! 票选你最喜欢的大语言模型

ai artificial-intelligence chatgpt gpt-4 llama llm natural-language-generation nature-language-processing nature-language-understanding nlp openai transformer

Last synced: 27 Aug 2026

Statistics

  • Projects: 2,239
  • Last updated: about 2 years ago