Inkling (batch)

Inkling · 批量推理档(Batch)

异步批量调用档:以折扣价处理离线批量请求,吞吐高、延迟不敏感,适合大规模非实时任务。

适用场景:大规模离线推理、成本敏感批处理、跑批与回填

本页为 Inkling 的批量推理档,与其实时标准档共享底层模型能力,主要差异在计费模式与调用方式。标准实时档详见 Inkling 标准实时档

厂商:Thinking Machines

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

关键参数

Token 定价

输入输出模态

文本、图像、音频

典型用例

什么时候选这个模型

基于公开定价与能力字段派生的选型信号(非实测benchmark),最终选型请结合官方文档与你的评测。

相关问题

同厂商其他模型

查看完整信息:Inkling (batch)