| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.
You must be logged in to block users.
Contact GitHub support about this user’s behavior. Learn more about reporting abuse.
Report abuseMaking large AI models cheaper, faster and more accessible
xDiT: A Scalable Inference Engine for Diffusion Transformers (DiTs) with Massive Parallelism
a fast and user-friendly runtime for transformer inference (Bert, Albert, GPT2, Decoders, etc) on CPU and GPU.
Fast inference from large lauguage models via speculative decoding
PatrickStar enables Larger, Faster, Greener Pretrained Models for NLP and democratizes AI for everyone.
USP: Unified (a.k.a. Hybrid, 2D) Sequence Parallel Attention for Long Context Transformers Model Training and Inference
| Back | FazBrowse Home | New Git URL |