
hipfire-models/qwen3.8-27b · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
hipfire is a Rust and HIP inference engine for AMD RDNA and CDNA GPUs, providing an Ollama-style CLI and an OpenAI-compatible API. v0.4.0 runs quantized Qwen models with speculative decode and image generation on RDNA1 through RDNA4 via a kernel dispatch layer called Redline.
Weights ship through Hugging Face; the daemon binds to loopback port 11435 by default.