Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
AMD Strix Halo / Ryzen AI Halo local LLM setup and benchmark guide for Ryzen AI MAX+ 395 and Radeon 8060S: Ollama, llama.cpp Vulkan/RADV, ROCm, 101 t/s Qwen3-Coder, CHADROCK MTP, 120B GGUF, and raw evidence.
| Date | Stars |
|---|---|
| 2026-07-31 | 251 |
| 2026-08-01 | 251 |
| 2026-08-02 | 255 |
| 2026-08-04 | 256 |
| 2026-08-05 | 257 |
| 2026-08-06 | 260 |
Today
+3 stars today
This week
— stars this week
This month
— stars this month
Momentum
27.0
growth rate 0.00%/day
     [](COMMUNITY_RESULTS.md)    [](https://github.com/hogeheer499-commits/strix-halo-guide/actions/workflows/validate.yml) # AMD Strix Halo Local LLM Guide for Ryzen AI MAX+ 395 / Radeon 8060S (gfx1151) A complete, practical guide to running large language models locally on AMD Strix Halo / Ryzen AI MAX+ 395 systems with Radeon 8060S (`gfx1151`) and 96GB/128GB unified memory. Covers BIOS config, Ubuntu 24.04/kernel setup, Ollama, `llama.cpp` Vulkan/RADV, ROCm/HIP experiments, vLLM notes, 70B/120B and selected 284B GGUF capacity evidence, benchmarks, raw logs, and reproducibility checks. AMD now publicly frames Ryzen AI Halo-class systems as a local-AI and developer-platform direction. This repository is the independent practical layer: copyable setup, measured rows, raw evidence, failures, and community reproductions. It is not official AMD or OEM endorsement. See [`RYZEN_AI_HALO_CONTEXT.md`](RYZEN_AI_HALO_CONTEXT.md). Project website: <https://strixhaloguide.com/>. This GitHub repository remains the source of truth for setup commands, benchmark claims, and raw evidence. Maintainer credibility is also public and reviewable: accepted upstream contributions include code and validation in [`llama.cpp`](https://github.com/ggml-org/llama.cpp/pull/25643), LocalAI, Qwen Code, OpenTelemetry GenAI, and NVIDIA AICR, plus tested-coverage documentation in the vLLM GGUF plugin. See [`UPSTREAM_CONTRIBUTIONS.md`](UPSTREAM_CONTRIBUTIONS.md) for exact PR links, scope, and honest boundaries. Upstream acceptance strengthens confidence in the engineering process; it does not replace the raw evidence required for each benchmark claim. What you get: - Copyable Ubuntu + Vulkan/RADV setup for Ollama and `llama.cpp`. - Practical model/backend choices for a local AI PC. - Direct local results: Qwen3-Coder 30B at 101.0 t/s on the official `llama.cpp` b9851 Vulkan release binary, Qwen3-30B-A3B-Instruct-2507 IQ4_XS at 100.0 t/s with a b9544 control at 103.2 t/s, LFM2.5 8B-A1B at 170.0 t/s with a b9544 control at 176.5 t/s, Nemotron 3 Super 120B-A12B at 18.4-18.9 t/s, and a 90.86GB DeepSeek V4 Flash 284B `UD-IQ2_XXS` capacity scout at 13.27 t/s direct on Vulkan/RADV. - Experimental server routes: Qwen3.6 MTP at 101.1 t/s, Gemma 4 26B-A4B QAT MTP up to 110.0 t/s best-repeat, and CHADROCK ACE/SABER 35B ROCmFP4 at 141.37 t/s across three repeats on one exact high-acceptance reference profile with `llama-server` speculative decoding. The CHADROCK number is prompt-shape-specific, not a universal server speed. - Multi-user evidence: b9979 stock, opt-in AMD/RADV density, dense16, and Lemonade ROCm concurrency matrices for 30B 128-expert/top-8 and 80B 512-expert/top-10 MoE models. - Raw CSVs, logs, charts, and reproducibility notes for headline claims. - Community validation from Beelink, Corsair, GMKtec, MS-S1-Max, Nimo, NixOS, NPU, ROCmFP4, and other Strix Halo owner stacks. > Measured primarily on one Beelink GTR9 Pro. Community
Excerpt of 222,001 characters
Read on GitHubWould you bet a product on this? Bounded 0–100 and slow moving.
matched fp:98f15485360b8830, llm:description: 'AMD Strix Halo / Ryzen AI Halo local LLM setup and benchmark guide... Ollama, llama.cpp Vulkan/RADV, ROCm, ... 120B GGUF' ; topics include: llama-cpp, llm, local-ai, local-llm, gguf, benchmark, roc m, vulkan
matched fp:98f15485360b8830, llm:description: 'AMD Strix Halo / Ryzen AI Halo local LLM setup and benchmark guide... Ollama, llama.cpp Vulkan/RADV, ROCm, ... 120B GGUF' ; topics include: llama-cpp, llm, local-ai, local-llm, gguf, benchmark, roc m, vulkan
matched fp:98f15485360b8830, llm:description: 'AMD Strix Halo / Ryzen AI Halo local LLM setup and benchmark guide... Ollama, llama.cpp Vulkan/RADV, ROCm, ... 120B GGUF' ; topics include: llama-cpp, llm, local-ai, local-llm, gguf, benchmark, roc m, vulkan
matched fp:98f15485360b8830, llm:description: 'AMD Strix Halo / Ryzen AI Halo local LLM setup and benchmark guide... Ollama, llama.cpp Vulkan/RADV, ROCm, ... 120B GGUF' ; topics include: llama-cpp, llm, local-ai, local-llm, gguf, benchmark, roc m, vulkan