- NVIDIA
- Bay Area, CA, USA
- https://ericrxw.github.io/xiaoweiren/
Popular repositories Loading
- apex
apex PublicForked from NVIDIA/apex
A PyTorch Extension: Tools for easy mixed precision and distributed training in Pytorch
Python
- NeMo-Megatron-Launcher
NeMo-Megatron-Launcher PublicForked from NVIDIA/NeMo-Framework-Launcher
NeMo Megatron launcher and tools
Python
- flash-attention
flash-attention PublicForked from Dao-AILab/flash-attention
Fast and memory-efficient exact attention
Python
- collective_matmul
collective_matmul PublicPython
- TransformerEngine
TransformerEngine PublicForked from NVIDIA/TransformerEngine
A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit floating point (FP8) precision on Hopper GPUs, to provide better performance with lower memory utilization in bot…
Python
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.
Uh oh!
There was an error while loading. Please reload this page.

