tencent cloud

LLM Service TokenHub

Billing Mode

Unduh
Mode fokus
Ukuran font
Terakhir diperbarui: 2026-09-10 21:19:10
Diterjemahkan oleh AI
On TokenHub, different models may have different billing methods. The following introduces the billing methods for each type of model. You can also go to the model details page from Model Gallery to view the billing rules for each model.

Language Models

Billable Item
Billing Methods
Billing Unit
Description
Input tokens
Pay-as-you-go
USD / million tokens
Token consumption of user input text (including system prompt)
Output tokens
Pay-as-you-go
USD / million tokens
Token consumption of model-generated text
Cached Input
Pay-as-you-go
USD / million tokens
Cached input Token consumption
Billing Notes:
Input Tokens and Output Tokens are billed separately, and different models have different unit prices. See Model Pricing (Language Models).
Some models support tiered pricing (such as different unit prices for different input length ranges).
Some reasoning models' thinking process (Reasoning) and regular output are priced separately.

Prerequisites

Each time you claim a free resource package for a model or enable postpaid service for a model, USD 1 will be automatically frozen in your account. If your account balance or credit limit is insufficient, the claim or enablement will fail.
The pre-frozen amount can be viewed in the Billing Center.
The frozen amount will be returned to your account after the resources are terminated. If you want to get your pre-frozen amount back, delete all model endpoints on the Online Inference page. To delete the default endpoint, contact customer service.

References

For detailed pricing information of each model, see Model Pricing.
For details on the impact of billing arrears and recovery mechanisms, see Overdue Payments.


Bantuan dan Dukungan

Apakah halaman ini membantu?

masukan