Rent ready-to-use cloud GPUs in seconds. Lium CLI makes it easy to launch, manage, and scale GPU compute directly from your terminal. Fast, cost-optimized, and built for AI & ML developers.
-
Updated
Sep 6, 2026 - Python
Rent ready-to-use cloud GPUs in seconds. Lium CLI makes it easy to launch, manage, and scale GPU compute directly from your terminal. Fast, cost-optimized, and built for AI & ML developers.
A Flexible and High-Performance Inference Serving Engine for Diffusion Language Models
⚡ TIMTEH Model Forge — Uncensored, abliterated & reasoning-distilled GGUFs. Forged on 8×H200 SXM5 | 1.1TB VRAM
Monitor low-utilization time, idle-state episodes, and workload starvation signals on NVIDIA datacenter GPUs.
Thermal-aware batch controller for vLLM/TensorRT-LLM. Prevents HBM thermal throttling from killing p99 latency on H100/H200. Monitors nvidia-smi, auto-cuts batch size at 85°C, migrates cold KV to DRAM. Prometheus + Grafana included. 4.2s -> 2.1s p99 at 128K context.
Daily-verified cloud GPU rental prices across 18+ providers (H100, H200, B200, A100, RTX 4090...). Open dataset, updated daily, CC BY 4.0. Live at gpurentalprices.com
dd-ready fully packed FreeDOS disk image for crossfalshing SAS2008 like Dell H200 to HBA
GLM-5.2 744B at 4-bit on Modal 4x H200 via vLLM, plus a static streaming chat UI.
Reproducible vLLM-Omni serving and benchmark reference for NVIDIA Cosmos3-Super on H200 and B200 GPUs
AMD MI300X vs NVIDIA H200: a controlled single-GPU benchmark of inference and training. AMD's 1.32x paper FLOPs advantage inverts on achieved efficiency, but its 1.36x memory becomes a 1.84x KV-cache advantage that wins large-model serving outright.
Voltage Park — independent third-party profile of a public API surface, by API Evangelist. Voltage Park is a GPU cloud offering on-demand and reserved NVIDIA H100 and H200 clusters as bare metal and virtual machines. Its On-Demand API (served at cloud-api.voltagepark.com, running on TensorDock infrastructure) lets developers deploy and manage insta
SGPU — Simple GPU monitor for the SGVR H200 lab (MLXP / NAVER Cloud NKS). Zero-install kubectl dashboards, per-pod GPU process attribution, in-pod TUI, and per-user usage accounting.
To associate your repository with the h200 topic, visit your repo's landing page and select "manage topics."