Tab | Managed Object |
Models | Various large models supported by the platform (language, vision, multimodal understanding, vector, speech models), managing the pay-as-you-go billing for their free trial packages and default model services. |
Tools | Various tool capabilities provided by the platform, managing their enablement status and pay-as-you-go billing. |
Column | Description | |
Model Name | | Click a model name to go to the corresponding default inference service details. If the free package has not been claimed for the model, the default inference service endpoint has not been created, and the jump cannot be performed. |
Model Rate Limiting | | Displays the TPM / RPM quota of the model. |
Status | Running | Indicates that the default model service is running normally. |
| Not Enabled | Indicates that a model with free quota has not claimed the free trial package, or a model without free quota has not enabled pay-as-you-go billing. |
| Stopped | Indicates that the default model service has been stopped. This may occur when the free quota is exhausted but pay-as-you-go billing is not enabled, or when the account has an overdue payment. |
| Creating | In the process of claiming the free package or enabling pay-as-you-go billing. |
| Running (Model Pending Deprecation) | Indicates that the default model service is running normally, but the model is about to be taken offline. |
| Running (Model Under Maintenance) | Indicates that the default model service is running normally, but the model is under maintenance. |
Column | Description |
Claim Status | Unclaimed: You can click the button to claim the free trial package for this model. Claimed: The button is grayed out and cannot be used to claim again. Not Supported: This model does not support free trials. To use it, enable pay-as-you-go billing. |
Free quota balance | Displays the remaining free quota for claimed models in the form of remaining percentage + progress bar. Hover over the progress bar to view details: free quota tokens, consumed tokens, remaining tokens, and remaining percentage %. Models that are not claimed / not supported are not displayed. |
Free quota expiration time | Displays the expiration time of the free quota. Not displayed for models that are not claimed / not supported. |
Column | Description |
Enabling Status | Not Enabled: You can click the button to enable pay-as-you-go billing for this model. Enabled: You can click the button to disable pay-as-you-go billing for this model. A secondary confirmation dialog box pops up for both enabling and disabling. |
Billing Mode | Currently supports token-based billing. For models that do not have pay-as-you-go billing enabled, this column does not display information. |
Page | Responsibilities |
Activation | Model-level overview entry for viewing model lists, claiming free quotas, and enabling/disabling pay-as-you-go billing, helping users perform lightweight activation. |
Online Inference Service | Service-level management entry. A model can have multiple inference services (1 default service with the same name + n custom services). You can create custom services, configure rate limiting and billing, and make API calls as needed. |
Apakah halaman ini membantu?
Anda juga dapat Menghubungi Penjualan atau Mengirimkan Tiket untuk meminta bantuan.
masukan