Add gpuos models to LiteLLM
model_list:
- model_name: qwen3-32b
litellm_params:
model: openai/qwen3-32b
api_base: https://gpuos.si/v1
api_key: os.environ/GPUOS_API_KEY
- model_name: bge-m3
litellm_params:
model: openai/bge-m3
api_base: https://gpuos.si/v1
api_key: os.environ/GPUOS_API_KEYLiteLLM then routes these model names to your GPUs and everything else to the providers you already use.
Official documentation: docs.litellm.ai