Recommendations for people using Strix Halo (AI MAX)

183.157.***.***
96

https://github.com/Nathanw1014/strix-halo-llamacpp

This is a version where the developer fixed some inefficient operations between Strix Halo and llamacpp.

I heard that a merge request was made to llamacpp, but it seems that it hasn't been approved yet because there wasn't enough proof,

but I think it's significantly improved.

Based on Qwen 3.6 27B Q6K, with 180K Prefill, it achieves 130t/s and about 10tps in normal operation.

로그인한 회원만 댓글 등록이 가능합니다.

개발한당

KR | ID | EN
  • IDR
  • KOR
7.83 0.01

2026.08.24 KEB 하나은행 고시회차 482회

다가오는 한인 행사일정

  • 등록 된 일정이 없어요!