This model always redirects to the latest Grok model from xAI.
多个平台托管同一个模型。请求将根据优先级和健康状况进行分配,如果本地 GPU 可用,则始终优先使用。