FazBrowse GitHub Viewer | Trending |
URL:
| Home
Tools: [Download Repo ZIP]   [Original HTTPS Page]

llama-cpp-python ยท GitHub Topics ยท GitHub

#

llama-cpp-python

Here are 85 public repositories matching this topic...

Setup and run a local LLM and Chatbot using consumer grade hardware.

  • Updated Nov 23, 2025
  • JavaScript

Gradio based tool to run opensource LLM models directly from Huggingface

  • Updated Jun 27, 2024
  • Python

Information on optimizing python libraries specifically for oobabooga to take advantage of Apple Silicon and Accelerate Framework.

  • Updated Feb 12, 2025
  • Python

An open source, Gradio-based chatbot app that combines the best of retrieval augmented generation and prompt engineering into an intelligent assistant for modern professionals.

  • Updated Aug 1, 2024
  • Python

GPU-accelerated LLaMA inference wrapper for legacy Vulkan-capable systems a Pythonic way to run AI with knowledge (Ilm) on fire (Vulkan).

  • Updated Oct 14, 2025
  • Python

Local character AI chatbot with chroma vector store memory and some scripts to process documents for Chroma

  • Updated Oct 7, 2024
  • Python

This repository is a CUA (computer use agent) system that, using the Qwen3-VL model on Ubuntu computers, aims to perform tasks on your behalf using the keyboard and mouse in a local Sandbox environment in GGUF format, based on the commands you provide.

  • Updated Mar 3, 2026
  • Python

Experimental interface environment for open source LLM, designed to democratize the use of AI. Powered by llama-cpp, llama-cpp-python and Gradio.

  • Updated Oct 11, 2025
  • Python

Tool for test diferents large language models without code.

  • Updated Oct 18, 2025
  • Python

UnOfficial Gradio Repo for ICML 2024 paper "Executable Code Actions Elicit Better LLM Agents" by Xingyao Wang, Yangyi Chen, Lifan Yuan, Yizhe Zhang, Yunzhu Li, Hao Peng, Heng Ji.

  • Updated Sep 30, 2024
  • Jupyter Notebook

A quick and optimized solution to manage llama based gguf quantized models, download gguf files, retreive messege formatting, add more models from hf repos and more. It's super easy to use and comes prepacked with best preconfigured open source models: dolphin phi-2 2.7b, mistral 7b v0.2, mixtral 8x7b v0.1, solar 10.7b and zephyr 3b

  • Updated Jan 13, 2024
  • Python

Multimodal prompt generator nodes for ComfyUI, designed to generate prompts for QwenImageEdit and Wan2.2. Supports local LLM / local GGUF models (Qwen2.5-VL, Qwen3-VL, Qwen3.5 and Qwen3.6) and Qwen API for image and video prompt generation and enhancement.

  • Updated Jul 30, 2026
  • Python

A financial chatbot powered by an LLM and retrieval-augmented generation.

  • Updated Oct 2, 2023
  • Jupyter Notebook

Simple chat interface for local AI using llama-cpp-python and llama-cpp-agent

  • Updated Jul 27, 2024
  • Python

A modular, local AI companion featuring a RAG pipeline with FAISS and SentenceTransformers for semantic long-term memory. Powered by GGUF models via llama-cpp-python

  • Updated Jul 31, 2026
  • Python

A comprehensive toolkit for training and running lightweight adapters for GGUF-based language models (ERNIE, Llama, Mistral, Phi-3, etc.) without modifying the base model.

  • Updated Feb 23, 2026
  • Python

Improve this page

Add a description, image, and links to the llama-cpp-python topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the llama-cpp-python topic, visit your repo's landing page and select "manage topics."

Learn more


Back | FazBrowse Home | New Git URL