August 4, 2026BUILDING
I turned an ASUS GX10 into our own local model provider
We turned a conference-gift ASUS GX10 into an OpenAI-compatible model provider the whole team can hit with the client code they already use — LiteLLM in front, llama.cpp and vLLM behind, three access tiers, and one genuinely annoying Docker networking wall.
infrastructurelocal modelsgx10