FazBrowse GitHub Viewer | Trending |
URL:
| Home
Tools: [Download Repo ZIP]   [Original HTTPS Page]

token-optimization · GitHub Topics · GitHub

#

token-optimization

Here are 1,239 public repositories matching this topic...

CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies

  • Updated Aug 20, 2026
  • Rust

Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

  • Updated Aug 19, 2026
  • Python

Control what your AI can see. LeanCTX (Lean Context) is the context intelligence layer for AI agents — one local Rust binary that decides what they read, remembers what they learn, guards what they touch, and proves what they save. 60–90% fewer tokens as the receipt. 76 MCP tools, 30+ agents, local-first.

  • Updated Aug 19, 2026
  • Rust

Cut AI token costs 95%+ on code exploration. The leading MCP server for precise, symbol-level GitHub code retrieval via tree-sitter AST. Works with Claude Code, Cursor & any MCP client. 313B+ tokens saved.

  • Updated Aug 20, 2026
  • Python

Sharper context. Fewer tokens. Open-source middleware for Claude Code.

  • Updated Aug 19, 2026
  • TypeScript

Find the ghost tokens. Fix them. Survive compaction. Avoid context quality decay.

  • Updated Aug 19, 2026
  • Python

Non-destructive compression gateway for AI coding agents. Cuts token bills 25% on turn 1 to past 85% in long or saturated sessions, and fits ~3× more turns in the same context window. Powered by our open-source code-native 4B model. Drop-in for Claude Code, Cursor, Codex, OpenHands, and any BASE_URL agent.

  • Updated Aug 20, 2026
  • Python

Up to 71.5x fewer tokens per session on Claude Code with Obsidian + Graphify. Persistent memory, codebase knowledge graphs, and chat import pipeline. 🇧🇷 PT-BR included.

  • Updated Jun 1, 2026
  • Python

Compress LLM context to save tokens and reduce costs

  • Updated Aug 19, 2026
  • Rust

Build agent that uses 80% less token and delivers better results.

  • Updated Aug 17, 2026
  • Rust

lowfat - slim your command output. strips noise, saves tokens.

  • Updated Aug 17, 2026
  • Rust

Optimize token usage for Claude API calls

  • Updated Aug 17, 2026
  • JavaScript

Headroom for macOS — cut Claude Code and Codex token costs by ~50%

  • Updated Aug 19, 2026
  • Rust

Measure token savings per AI coding agent, optimize context, and share a live local knowledge graph across 16 CLI clients.

  • Updated Aug 16, 2026
  • JavaScript

Working memory for Claude Code - persistent context and multi-instance coordination

  • Updated Jan 17, 2026
  • Python

AI context optimization platform for files, conversations, RAG, code, logs and AI agents with context compression, evidence preservation, content-addressed recovery and auditable Context Receipts.

  • Updated Aug 19, 2026
  • Python

Stop Claude Code from burning through your quota in 20 minutes. Auto-rotates oversized sessions and preserves context.

  • Updated Apr 16, 2026
  • TypeScript

Context engineering for AI agents. ~80% fewer tokens. Fix tool overload. Skills and memory with in-process BM25 and semantic retrieval. Progressive Disclosure. No vector DB.

  • Updated Aug 19, 2026
  • TypeScript

CLI proxy that reduces LLM token usage by 60-90%. Declarative YAML filters for Claude Code, Cursor, Copilot, Gemini. rtk alternative in Go.

  • Updated Aug 19, 2026
  • Go

Generate a compact codebase index for AI assistants — saves 50K+ tokens per conversation

  • Updated May 30, 2026
  • TypeScript

Improve this page

Add a description, image, and links to the token-optimization topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the token-optimization topic, visit your repo's landing page and select "manage topics."

Learn more


Back | FazBrowse Home | New Git URL