A set of performant RDNA/HIP patches against llama.cpp focusing on high speed inferencing with numerics assurance
↑ +17 today
Discover the most starred and trending open source tools tagged with #rocm.
A set of performant RDNA/HIP patches against llama.cpp focusing on high speed inferencing with numerics assurance
Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk
LLM speculative inference server for heterogeneous hardware & consumer GPUs
UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms