FazBrowse GitHub Viewer | Trending |
URL:
| Home
Tools: [Download Repo ZIP]   [Original HTTPS Page]

gpu-memory · GitHub Topics · GitHub

#

gpu-memory

Here are 42 public repositories matching this topic...

A fast GPU memory copy library based on NVIDIA GPUDirect RDMA technology

  • Updated Jul 14, 2026
  • C

Training neural networks in TensorFlow 2.0 with 5x less memory

  • Updated Feb 21, 2022
  • Python

A Toolkit for Training, Tracking, Saving Models and Syncing Results

  • Updated Mar 12, 2020
  • Python

A memory profiler for NVIDIA GPUs to explore memory inefficiencies in GPU-accelerated applications.

  • Updated May 30, 2026
  • Python

OpenCV & Spout C++ library. Shared GPU memory and processing at reach.

  • Updated May 1, 2020
  • C++

Rust embedded things running on the seL4 microkernel for the Raspberry Pi 3

  • Updated Dec 8, 2018
  • Rust

A simple tool to find out GPU VRAM requirements for running LLMs

  • Updated Mar 23, 2026
  • HTML

A tiny, useful command-line tool to show each user gpu usage, pid under each gpu, provide more details than nvidia-smi/gpustat

  • Updated Sep 21, 2019
  • Shell

Demonstration of generating mini-batches in Tensorlfow from GPU memory.

  • Updated Apr 20, 2017
  • Python

Accurate VRAM calculator for Local LLMs (Llama 4, DeepSeek V3, Qwen 2.5). Calculates GGUF quantization, GQA context overhead, and offloading limits

  • Updated Nov 27, 2025
  • HTML

A prefix-cache advisor for LLM serving infrastructure that recommends KV-cache capacity and eviction policies from your request traces/logs.

  • Updated Aug 7, 2026
  • Python
  • Updated Jun 24, 2022
  • C++

Dynamic GPU Layer Swapping: Train large models on consumer GPUs with intelligent memory management

  • Updated Sep 12, 2025
  • Python

A fork of Kubernetes with support of schedulable resource of NVIDIA GPU memory

  • Updated Nov 17, 2018
  • Go

Detailed VRAM profiler for transformer inference with per-layer breakdown, activation analysis, and a predictive memory model that predicts VRAM with <1.2% error. Shows that FFN layers dominate static memory and that measured runtime VRAM exceeds KV-cache estimates by 2-4x.

  • Updated Jul 14, 2026
  • Python

Event-driven benchmark of adaptive batch composition policies for LLM serving, measuring how prefill and decode interference affects TTFT, TPOT, and throughput under different memory pressure regimes.

  • Updated Jul 24, 2026
  • Python

A CLI tool for estimating GPU VRAM requirements for Hugging Face models, supporting various data types, parallelization strategies, and fine-tuning scenarios like LoRA.

  • Updated Oct 22, 2025
  • Python

Research harness for evaluating query-time bounded elimination of reconstructable KV-cache witnesses in long-context transformer inference workloads. Related provisional filing: IN 202641062451.

  • Updated May 18, 2026
  • Python

Improve this page

Add a description, image, and links to the gpu-memory topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the gpu-memory topic, visit your repo's landing page and select "manage topics."

Learn more


Back | FazBrowse Home | New Git URL