Hands-on GPU/HPC infrastructure operations: K8s GPU scheduling, HAMi sharing, Slurm, observability & vLLM inference. Learn it free on a laptop; validate on one cheap GPU.
-
Updated
Jul 22, 2026 - Shell
Hands-on GPU/HPC infrastructure operations: K8s GPU scheduling, HAMi sharing, Slurm, observability & vLLM inference. Learn it free on a laptop; validate on one cheap GPU.
Doctor consultation app designed using flutter.
A science inference cluster — every model, from protein folding to LLMs, behind one OpenAI/Anthropic-compatible endpoint and a single key. Built to run beside HPC so agentic AI researchers can drive simulations and model inference together. Self-deploying on RKE2 with KServe/Knative + HAMi GPU.
Simulate Moore Threads MUSA and NVIDIA CUDA GPUs in Kubernetes — fake MUSA/MTML/CUDA/NVML user-space libs, mthreads-gmi/nvidia-smi tools, and a device-plugin for HAMi-style scheduling without physical GPUs.
Add a description, image, and links to the hami topic page so that developers can more easily learn about it.
To associate your repository with the hami topic, visit your repo's landing page and select "manage topics."