| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.
You must be logged in to block users.
Contact GitHub support about this user’s behavior. Learn more about reporting abuse.
Report abuse4-5x faster Qwen3.5 on ASUS GX10 / DGX Spark — Hybrid INT4+FP8 + MTP via one shell script
Qwen3.6 27B × DFlash — 30-35 tok/s on NVIDIA DGX Spark (GB10) - LLama.Cpp
Forked from ggml-org/llama.cpp
llama.cpp optimized for DeepSeek V4 Flash on NVIDIA DGX Spark / ASUS GX10 (128GB unified, 273GB/s). Fixes graph buffer overflow ▎ for 8K+ context. Achieves 6-7 tok/s on 284B IQ2_XXS.
C++ 7
Transform Your AI Coding Experience: Claude/Antigravity as the Brains, Roo Code as the Brawn, and Telegram as a Bonus Remote Control!
JavaScript 2
A smart Layer-7 Auto-Scaler & Lifecycle Proxy for llama.cpp. Save 100% VRAM when idle, auto-spawn multiple instances on demand, and prevent AI Agent bottlenecks with zero-config.
Python 1
Deterministic solver fleet that cracked 9,500/9,500 NVIDIA Nemotron reasoning puzzles (Private 0.86, Silver). Solvers + CoT generators + quality gates.
Python
| Back | FazBrowse Home | New Git URL |