| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.
You must be logged in to block users.
Contact GitHub support about this user’s behavior. Learn more about reporting abuse.
Report abuseReproducible llama.cpp CPU inference profiling and a deterministic LLM serving simulator with continuous batching, KV cache, prefix caching, and workload-driven latency analysis.
A production-oriented LLM engineering platform with OpenAI-compatible serving, streaming, observability, evaluation, reproducible experiments, and deterministic RAG.
Python 17
Evidence-backed structural validation of Kimi K3 UD-IQ1_M and UD-Q4_K_XL split GGUF releases using OMIV.
Shell 14
Offline-first evidence and verification framework for AI model artifacts, transformations, runtime identity, and provenance.
Python 2
Forked from Liquid4All/pipette-clients
Client-side harnesses and apps for running measurements on target devices, including dedicated iOS and Android apps.
Rust 1
| Back | FazBrowse Home | New Git URL |