A high-throughput and memory-efficient inference and serving engine for LLMs
-
Updated
Oct 3, 2026 - Python
A high-throughput and memory-efficient inference and serving engine for LLMs
ODS V3 Pre-Release: Public testing and refinement ahead of the official V3 launch. Turn your PC, Mac, or Linux box into a private AI server.
QualityScaler - image/video AI upscaler app
Stable Diffusion web UI
A unified orchestration layer for heterogeneous AI compute. It standardizes how to manage compute and run training and inference on GPU clouds, Kubernetes, VMs, or bare-metal clusters.
Open Source Inference Research Platform Standard / 开源推理研究平台
AMD-SHARK Studio -- Web UI for SHARK+IREE High Performance Machine Learning Distribution
OpenCL integration for Python, plus shiny features
RealScaler - image/video AI upscaler app (Real-ESRGAN)
Evidence-backed AMD Strix Halo local-AI setup and benchmarks: Qwen3.8, Ollama, llama.cpp, Vulkan/ROCm, large GGUFs, and cross-OEM results.
Nabla: High-Performance Scientific Computing
FluidFrames | video AI frame-generation app
[DEPRECATED] Moved to ROCm/rocm-libraries repo
To associate your repository with the amd topic, visit your repo's landing page and select "manage topics."