This model always redirects to the latest model in the OpenAI GPT Mini family.
多个平台托管同一个模型。请求将根据优先级和健康状况进行分配,如果本地 GPU 可用,则始终优先使用。