PPIO 企业 Token Plan
主流模型统一接入,调用低至 4.5 折,支持 200 团队席位
模型列表
精选高能模型,赋能多领域 AI 应用建设
企业私有化部署
企业级性能,弹性扩展,零运维负担
热门模型
NEW
PPIO 智能模型网关提供的多模型融合推理能力价格按实际调用模型计算
NEW
Qwen3.8 Flash
¥0.8/Mt输入
¥2.7/Mt输出
1M上下文
¥0.1/MtCache Read
NEW
GLM 5.3 Flash
¥0.8/Mt输入
¥2.8/Mt输出
1M上下文
¥0.23/MtCache Read
NEW
DeepSeek V4 Flash Vision Exp
¥3/Mt输入
¥9/Mt输出
1M上下文
¥1/MtCache Read
NEW
GLM 5.3
¥8/Mt输入
¥28/Mt输出
1M上下文
¥2/MtCache Read
NEW
DeepSeek V4 Pro 0813
¥9/Mt输入
¥27/Mt输出
1M上下文
¥3/MtCache Read
NEW
DeepSeek V4 Flash 0731
¥3/Mt输入
¥9/Mt输出
1M上下文
¥1/MtCache Read
NEW
Kimi K3
¥20/Mt输入
¥100/Mt输出
1M上下文
¥2/MtCache Read
NEW
GLM 5.2
¥8/Mt输入
¥28/Mt输出
1M上下文
¥2/MtCache Read
NEW
Kimi K2.7 Code
¥6.5/Mt输入
¥27/Mt输出
262.1K上下文
¥1.3/MtCache Read
NEW
MiniMax-M3
¥4.2/Mt输入
¥16.8/Mt输出
1M上下文
¥0.84/MtCache Read
NEW
DeepSeek V4 Flash
¥3/Mt输入
¥9/Mt输出
1M上下文
¥1/MtCache Read
NEW
DeepSeek V4 Pro
¥9/Mt输入
¥27/Mt输出
1M上下文
¥3/MtCache Read
NEWFeatured
Qwen3.8 Max
¥12/Mt输入
¥36/Mt输出
1M上下文
¥1.5/MtCache Read
NEW
Xiaomi MiMo V2.5
¥1/Mt输入
¥2/Mt输出
1M上下文
¥0.02/MtCache Read
限时5折
Qwen3.7 Max
¥6/Mt输入
¥18/Mt输出
1M上下文
¥1.2/MtCache Read
NEW
Xiaomi MiMo V2.5 Pro
¥3/Mt输入
¥6/Mt输出
1M上下文
¥0.025/MtCache Read
NEW
Kimi K2.6
¥6.5/Mt输入
¥27/Mt输出
262.1K上下文
¥1.1/MtCache Read
NEW
GLM 5.1
¥8/Mt输入
¥28/Mt输出
204.8K上下文
¥2/MtCache Read
NEW
MiniMax M2.7-highspeed
¥4.2/Mt输入
¥16.8/Mt输出
204.8K上下文
¥0.42/MtCache Read
¥2.625/MtCache Write(5m)
HOT
GLM-5-Turbo
¥7/Mt输入
¥26/Mt输出
202.8K上下文
¥1.8/MtCache Read
NEW
MiniMax M2.7
¥2.1/Mt输入
¥8.4/Mt输出
204.8K上下文
¥0.42/MtCache Read
¥2.625/MtCache Write(5m)
HOT
MiniMax M2.5 Highspeed
¥4.2/Mt输入
¥16.8/Mt输出
204.8K上下文
¥0.21/MtCache Read
¥2.625/MtCache Write(5m)
HOT
Qwen3.5 397B A17B
¥3/Mt输入
¥18/Mt输出
262.1K上下文
HOT
Qwen3.5-Plus
¥4/Mt输入
¥24/Mt输出
1M上下文
¥0.4/MtCache Read
¥5/MtCache Write(5m)
HOT
MiniMax M2.5
¥2.1/Mt输入
¥8.4/Mt输出
204.8K上下文
¥0.21/MtCache Read
¥2.625/MtCache Write(5m)
HOT
GLM 5
¥6/Mt输入
¥22/Mt输出
202.8K上下文
¥1.5/MtCache Read
NEW
Qwen3 Coder Next
¥1.4/Mt输入
¥10.5/Mt输出
262.1K上下文
DeepSeek OCR 2
¥0.216/Mt输入
¥0.216/Mt输出
8.2K上下文
HOT
Kimi K2.5
¥4/Mt输入
¥21/Mt输出
262.1K上下文
¥0.7/MtCache Read
HOT
MiniMax M2.1
¥2.1/Mt输入
¥8.4/Mt输出
204.8K上下文
¥0.21/MtCache Read
¥2.625/MtCache Write(5m)
HOT
GLM 4.7
¥4/Mt输入
¥16/Mt输出
204.8K上下文
¥0.8/MtCache Read
¥0.8/MtCache Write(5m)
HOT
DeepSeek V3.2
¥2/Mt输入
¥3/Mt输出
163.8K上下文
¥0.2/MtCache Read
Kimi K2 Thinking
¥4/Mt输入
¥16/Mt输出
262.1K上下文
MiniMax M2
¥2.1/Mt输入
¥8.4/Mt输出
204.8K上下文
¥0.21/MtCache Read
¥2.625/MtCache Write(5m)
HOT
DeepSeek V3.2 Exp
¥2/Mt输入
¥3/Mt输出
163.8K上下文
NEW
Qwen3.6 35B A3B
¥1.8/Mt输入
¥10.8/Mt输出
262.1K上下文
GLM 4.6v
¥2/Mt输入
¥6/Mt输出
131.1K上下文
¥0.4/MtCache Read
GLM 4.6
¥4/Mt输入
¥16/Mt输出
204.8K上下文
¥0.8/MtCache Read
¥0.8/MtCache Write(5m)
NEW
Qwen3.6-Plus
¥8/Mt输入
¥48/Mt输出
1M上下文
¥0.8/MtCache Read
¥10/MtCache Write(5m)
HOT
DeepSeek V3 0324
¥2/Mt输入
¥8/Mt输出
163.8K上下文
¥0.6/MtCache Read
简单易用只需一行代码,开发者即可快速使用派欧云的模型服务。
from openai import OpenAI
client = OpenAI(
base_url='https://api.ppio.com/openai',
api_key='<你的 API KEY>',
# 获取 API Key 请参考:https://ppio.com/docs/support/api-key
)
completion_res = client.completions.create(
model='deepseek/deepseek-v3-0324',
prompt='大语言模型会给我们的生活带来什么改变?',
stream=True,
max_tokens=512,
)
大语言模型 API
为您提供企业级大语言模型服务,比您自行部署 AI Infra,更可靠、更快、更经济、更具扩展性。
您可将精力集中在应用增长和客户服务上,而大型语言模型基础设施可放心交给 PPIO
可靠稳定
超高性价比
快速扩容

私有化部署,企业级定制化模型服务
如果您的企业需要更高性能保障、定制服务等级协议(SLA)或私有化部署能力,我们提供专属解决方案,满足您的模型定制化需求。
定制化定价方案
在线率与响应延迟保障
无限扩展能力
专属计算集群
典型应用场景

AI 情感陪伴机器人

AI 小说生成器

AI 总结摘要

AI 代码生成