Open Machine Learning Compiler Framework
-
Updated
Oct 3, 2026 - Python
Open Machine Learning Compiler Framework
Low-Latency Accelerated Web Remote Desktop Streaming Platform for Self-Hosting, Containers, Kubernetes, or Cloud/HPC
Next-gen fast plotting library running on WGPU using the pygfx rendering engine
Building a Neural Network-Native Engine (3N). Python authoring, C++ runtime, Vulkan/WebGPU. Windows, Linux, Android and Web.
A front-end to glxinfo, vulkaninfo, clinfo and es2_info - Linux
Evidence-backed AMD Strix Halo local-AI setup and benchmarks: Qwen3.8, Ollama, llama.cpp, Vulkan/ROCm, large GGUFs, and cross-OEM results.
Qwen 3.8 27B ROCmFP4 on AMD Strix Halo (Ryzen AI Max+ 395). Up to 36 tok/s via MTP Speculation, TurboQuant & Mesa RADV Wave64.
Performance-tuned llama.cpp for AMD Strix Halo (gfx1151): FA + MoE-prefill fixes with a bundled current Mesa driver. Vulkan and HIP; portable dir, Docker, and distrobox.
Scriptable CLI for RenderDoc captures — built for terminal workflows, CI pipelines, and AI agents
Stable Diffusion UI: Diffusers (CUDA/ONNX)
😎 A curated list of awesome GPGPU (CUDA/OpenCL/Vulkan) resources
xllamacpp - a Python wrapper of llama.cpp
High-performance Lemonade alternative for AMD Strix Halo & Radeon — latest MoE models (Ling-3.0-Flash, Qwen 3.8 Flash Next, Ornith 1.5), bleeding-edge RDNA 3.5 Wave64/ROCmFP4 kernels, and silicon-tuned profiles.
To associate your repository with the vulkan topic, visit your repo's landing page and select "manage topics."