| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
Extensible Python SDK for developing Flyte tasks and workflows. Simple to get started and learn and highly extensible.
Community maintained hardware plugin for vLLM on Tensordyne systems
Ray is a unified framework for scaling AI and Python applications. Ray consists of a core distributed runtime and a toolkit of libraries (Ray AIR) for accelerating ML workloads.
Allows reading from cloud based storage.
Production-grade client-side tracing, profiling, and analysis for complex software systems.
A high-throughput and memory-efficient inference and serving engine for LLMs
[MLSys'24] Atom: Low-bit Quantization for Efficient and Accurate LLM Serving
🚀 Collection of components for development, training, tuning, and inference of foundation models leveraging PyTorch native components.
Loading…
Loading…
| Back | FazBrowse Home | New Git URL |