GLM and I created a llama.cpp fork optimized for AMD GFX906 (Mi50, Mi60, Radeon VII, GCN HIP)

I, in fact, saw your post on Reddit and then noticed the URL and came here. I will surely try it out and report back with numbers. I do have several MI50 that currently run the default llama.cpp.

1 Like