| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
Forked from NVIDIA/Megatron-LM
Ongoing research training transformer language models at scale, including: BERT & GPT-2
MII makes low-latency and high-throughput inference possible, powered by DeepSpeed.
Forked from EleutherAI/gpt-neox
An implementation of model parallel autoregressive transformers on GPUs, based on the DeepSpeed library.
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
Ongoing research training transformer language models at scale, including: BERT & GPT-2
MII makes low-latency and high-throughput inference possible, powered by DeepSpeed.
An implementation of model parallel autoregressive transformers on GPUs, based on the DeepSpeed library.
This organization has no public members. You must be a member to see who’s a part of this organization.
Loading…
Loading…
| Back | FazBrowse Home | New Git URL |