Hyper-Inference Router: AI Routing on MI300X
A routing layer that sends every AI request to the cheapest capable model. Uses Gemma 4 E4B on a dedicated AMD Instinct MI300X deployment for reasoning and creative tasks. Already running in production inside Kronos AI, a SaaS platform.