Skip to content
#

gpudirect

Here are 18 public repositories matching this topic...

A hybrid testbed for evaluating top open-source LLMs (like gpt-oss-20b and Llama 3.3) on local, cloud GPUs, and AWS Inferentia2/Trainium instances, focusing on vLLM optimization, capacity management, kernel bypass, hardware-software co-design, as well as supporting infrastructure such as NCCL, RDMA, NVMeoF.

  • Updated Apr 21, 2026
  • Python

Add this topic to your repo

To associate your repository with the gpudirect topic, visit your repo's landing page and select "manage topics."

Learn more