Distribute and run LLMs with a single file.
cross-platform
gguf
llama-cpp
local-ai
local-inference
local-llm
open-source-ai
single-file-executable
speech-to-text
Updated 2026-10-08 19:53:38 +00:00
🔥 Automated nightly & release builds of ROCmFPX, Ciru ROCmFPX (DualView Q7/Q8), and q38rocm with AMD ROCm 7 for Windows & Linux. Portable standalone binaries with bundled runtime libraries for Strix Halo (gfx1151), RDNA4, RDNA3, and Homebrew tap.
ciru-rocmfpx
dualview
gfx1151
gguf
hip
homebrew
llama-cpp
llamacpp
portable-binaries
q38rocm
quantization
radeon-8060s
rdna3
rdna4
rocm
rocm-7
rocmfpx
strix-halo
ubuntu-rocm
windows-rocm
Updated 2026-09-05 19:17:42 +00:00
⚡ Automated nightly builds & portable ROCm 7 releases of Ember for DeepSeek-V4-Flash on AMD Strix Halo (gfx1151 / Radeon 8060S). Standalone C inference server with bundled ROCm runtime, continuous batching, and Homebrew support.
ai-inference
amd-apu
c-inference
continuous-batching
deepseek-v4
dflash
ember
gfx1151
gguf
hip
homebrew
llm-inference
nightly-builds
portable-binaries
radeon-8060s
rdna3-5
rocm
rocm-7
strix-halo
the-rock
Updated 2026-08-31 03:00:27 +00:00