FazBrowse GitHub Viewer | Trending |
URL:
| Home
Tools: [Download Repo ZIP]   [Original HTTPS Page]

memory-optimization · GitHub Topics · GitHub

#

memory-optimization

Here are 184 public repositories matching this topic...

This free RAM cleaner uses native Windows features to optimize memory areas. It's a compact, portable, and smart application.

  • Updated Dec 19, 2025
  • C#

Run larger LLMs with longer contexts on Apple Silicon by using differentiated precision for KV cache quantization. KVSplit enables 8-bit keys & 4-bit values, reducing memory by 59% with <1% quality loss. Includes benchmarking, visualization, and one-command setup. Optimized for M1/M2/M3 Macs with Metal support.

  • Updated May 21, 2025
  • Python

VS Code extension: Go struct layout, padding, reorder

  • Updated Jul 16, 2026
  • TypeScript

A fast and memory-optimized string library for heavy-text manipulation in Python

  • Updated Apr 22, 2020
  • Python

Keep ChatGPT fast — Firefox & Chrome extension that trims DOM in long conversations

  • Updated Apr 9, 2026
  • TypeScript

Materials about memory optimization and zero-allocation samples.

  • Updated Jan 6, 2026
  • C#

Flash weight streaming for MLX: run massive models larger than your RAM on Apple Silicon.

  • Updated Jun 13, 2026
  • Python

NVIDIA Sol-Attn for ComfyUI / Triton kernel on SM89 - SM121, with zero-copy MiniMax H3 nodes: memory-efficient attention, scheduled tau with graph preview, and feed-forward chunking. Measured 1.14–1.44× vs SageAttention and −37% MLP peak VRAM on H3

  • Updated Aug 13, 2026
  • Python

This code repository contains the code used for my "Optimizing Memory Usage for Training LLMs and Vision Transformers in PyTorch" blog post.

  • Updated Jul 14, 2023
  • Python

First open-source implementation of Google TurboQuant (ICLR 2026) -- near-optimal KV cache compression for LLM inference. 5x compression with near-zero quality loss.

  • Updated May 25, 2026
  • Python

Lightweight Python lazy imports that defer module loading to reduce startup time and initial memory use.

  • Updated Aug 15, 2026
  • Python

A performant and memory efficient storage for immutable strings with C++17. Supports all standard char types: char, wchar_t, char16_t, char32_t and C++20's char8_t.

  • Updated May 2, 2022
  • C++

🚀 Zero-config automatic memory allocator for Rust - just add one line and get up to 1.6x faster allocation performance across all platforms

  • Updated Jul 1, 2025
  • Rust

A curated list of awesome optimizations that you can do to improve your redis deployment (both client side and server side).

  • Updated Apr 26, 2021

PRISM: O(1) Photonic Block Selection for Long-Context LLM Inference — eliminates the O(N) KV cache scan via photonic broadcast-and-weight similarity engine on TFLN

  • Updated Apr 28, 2026
  • Python

An extension of micro mouse on WEBOTS using the flood filled algorithm, A star, Dijkstra’s and Breadth first search algorithm for moving the E-puck robot from start to goal in an NxM sized maze whose map was unknown to the robot (mapping and path planning). Further, leveraged Error Correction for accurate turning and recursive Backtracking algor…

  • Updated Jun 22, 2021
  • C++

Unified KV-cache compression for LLM inference: 12 Python-native methods, Debian-tested isolated add-ons, Godzilla KVarN/TriAttention, exact Godzilla/Gigatoken profiles, CUDA weight sharing, and multi-GPU planning.

  • Updated Aug 3, 2026
  • Python

Improve this page

Add a description, image, and links to the memory-optimization topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the memory-optimization topic, visit your repo's landing page and select "manage topics."

Learn more


Back | FazBrowse Home | New Git URL