Token导航 LogoToken导航TokenDH.com
Arcee AI / Trinity logo
官方核验于 2026-04-16

Arcee AI / Trinity

Trinity 系列兼顾开放权重和生产 API,适合做 Agent、结构化输出和私有化部署。它比较适合开发团队深用,不是最偏消费端聊天体验的路线。

官方入口原生官方接口付费入口语言模型

主分类:语言模型

适合:Agent / 模型微调 / 私有化部署

付费:开源下载 / 企业合作

地区:国际可用

接入判断

按原生接口接

Arcee AI / Trinity 适合作为官方原生入口,接入前先按文档核对鉴权方式、模型 ID 和计费口径。

协议:官方原生 API

价格:以官方价格页为准

模型:已收录 4 个

模型能力

4 个模型
语言模型端侧部署低延迟轻量 Agent

Trinity Nano 6B

分类:语言模型 / 上下文:128K / 调用名:trinity-nano-6b

Pricing Integration List chevron-right Language Models chevron-right Trinity-Nano (6B) Trinity-Mini (26B) Trinity-Large-Preview Trinity-Large-Thinking API Reference chevron-right Your First API Call Chat Completion Usage Models Capabilities chevron-right Streaming Messages Multi-Turn Conversations Function Calling Structured Outputs Reasoning Traces Quick Deploys chevron-right Download Models arrow-up-right Hardware Prerequisites Consumer Hardware chevron-right Inference Engines chevron-right Policies chevron-right Deprecation Policy chevron-up chevron-down gitbook Powered by GitBook gitbook xmark block-quote On this page xmark xmark copy Copy chevron-down block-quote On this page block-quote Language Models Trinity-Nano (6B) Overview Trinity Nano is a 6B-parameter (1B active) sparse mixture-of-experts language model, optimized for high-efficiency inference in real-time, on-device, and embedded AI applications.

函数调用长上下文
语言模型函数调用Agent 工作流长上下文

Trinity Mini

分类:语言模型 / 上下文:128K / 调用名:trinity-mini

Pricing Integration List chevron-right Language Models chevron-right Trinity-Nano (6B) Trinity-Mini (26B) Trinity-Large-Preview Trinity-Large-Thinking API Reference chevron-right Your First API Call Chat Completion Usage Models Capabilities chevron-right Streaming Messages Multi-Turn Conversations Function Calling Structured Outputs Reasoning Traces Quick Deploys chevron-right Download Models arrow-up-right Hardware Prerequisites Consumer Hardware chevron-right Inference Engines chevron-right Policies chevron-right Deprecation Policy chevron-up chevron-down gitbook Powered by GitBook gitbook xmark block-quote On this page xmark xmark copy Copy chevron-down block-quote On this page block-quote Language Models Trinity-Mini (26B) Overview Trinity Mini is a 26B-parameter (3B active) sparse mixture-of-experts language model, engineered for efficient inference over long contexts with robust function calling and multi-step agent workflows.

函数调用结构化输出长上下文
语言模型复杂任务编码推理高级 Agent

Trinity Large Preview

分类:语言模型 / 上下文:128K / 调用名:trinity-large-preview

Pricing Integration List chevron-right Language Models chevron-right Trinity-Nano (6B) Trinity-Mini (26B) Trinity-Large-Preview Trinity-Large-Thinking API Reference chevron-right Your First API Call Chat Completion Usage Models Capabilities chevron-right Streaming Messages Multi-Turn Conversations Function Calling Structured Outputs Reasoning Traces Quick Deploys chevron-right Download Models arrow-up-right Hardware Prerequisites Consumer Hardware chevron-right Inference Engines chevron-right Policies chevron-right Deprecation Policy chevron-up chevron-down gitbook Powered by GitBook gitbook xmark block-quote On this page xmark xmark copy Copy chevron-down block-quote On this page block-quote Language Models Trinity-Large-Preview Overview Trinity Large (Preview) is a 400B-parameter (13B active) sparse mixture-of-experts language model, engineered to scale model capacity while maintaining inference efficiency over long contexts, with strong performance in reasoning-heavy workloads including math, coding-related tasks, and multi-step agent workflows.

函数调用结构化输出长上下文
语言模型多步推理Reasoning TraceAgent 工作流

Trinity Large Thinking

分类:语言模型 / 上下文:512K / 调用名:trinity-large-thinking

Pricing Integration List chevron-right Language Models chevron-right Trinity-Nano (6B) Trinity-Mini (26B) Trinity-Large-Preview Trinity-Large-Thinking API Reference chevron-right Your First API Call Chat Completion Usage Models Capabilities chevron-right Streaming Messages Multi-Turn Conversations Function Calling Structured Outputs Reasoning Traces Quick Deploys chevron-right Download Models arrow-up-right Hardware Prerequisites Consumer Hardware chevron-right Inference Engines chevron-right Policies chevron-right Deprecation Policy chevron-up chevron-down gitbook Powered by GitBook gitbook xmark block-quote On this page xmark xmark copy Copy chevron-down block-quote On this page block-quote Language Models Trinity-Large-Thinking Overview Trinity-Large-Thinking is a reasoning-optimized variant of Arcee AI's Trinity-Large family — a 398B-parameter sparse Mixture-of-Experts (MoE) model with approximately 13B active parameters per token. Built on Trinity-Large-Base and post-trained with extended chain-of-thought reasoning and agentic RL, Trinity-Large-Thinking delivers state-of-the-art performance on agentic benchmarks while maintaining strong general capabilities. Trinity-Large-Thinking generates explicit reasoning traces wrapped in <think>...</think> blocks before producing its final response. This thinking process is critical to the model's performance — thinking tokens must be kept in context for multi-turn conversations and agentic loops to function correctly.

函数调用推理结构化输出长上下文