Hi everyone,
If I understand correctly, AI Gateway currently supports Anthropic, OpenAI, and Google models running locally on infrastructure (thanks to the ESXi integration).
Is there a solution to intercept traffic from a wider range of models—such as Qwen, Llama, DeepSeek, or others—running locally in our data center on a vLLM or LM Studio instance?
Thanks !



