ModelBeat
Models/Trinity Large Thinking
All models
A

Arcee AI: Trinity Large Thinking

arcee-ai/trinity-large-thinking

Trinity Large Thinking is Arcee AI’s AI model, released Apr 1, 2026. It costs $0.25/M input and $0.80/M output tokens, and handles a 262K-token context window.

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks.

Context
262K
Input
$0.25 / 1M
Output
$0.80 / 1M
Released
Apr 1, 2026
Modalities
texttext

Benchmarks

Scores on standardized evaluations. Higher is better — and rank shows where Trinity Large Thinking lands among all models tracked on Model Beat.

32.0
Intelligence Index
Epoch AI
32nd percentile of tracked models
18.0
Coding Index
Epoch AI
18th percentile of tracked models
58.0
Agentic Index
Epoch AI
58th percentile of tracked models

Reasoning

2 evals
GPQA Diamond75.2%

Graduate-level scientific reasoning

Humanity's Last Exam15.8%

Frontier of human expert knowledge

Coding

2 evals
SciCode40.6%

Scientific research coding

WebDev Arena1238

Human-rated web development

Agentic & Tools

1 evals
τ²-bench90.1%

Tool-agent-user reliability

Changelog

2
  • Trinity Large Thinking: reported SciCode score rose from 36.1% to 40.6% (+12%)benchmark · Sep 7, 2026
  • Trinity Large Thinking: Humanity's Last Exam improved from 14.7% to 15.8% (+7%)benchmark · Aug 6, 2026

Compare Trinity Large Thinking with

Frequently asked questions

What is Trinity Large Thinking?

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks.

Who created Trinity Large Thinking?

Trinity Large Thinking is an AI model developed by Arcee AI, released Apr 1, 2026.

How much does Trinity Large Thinking cost?

Trinity Large Thinking's list price from Arcee AI is $0.25 per million input tokens and $0.80 per million output tokens.

What is Trinity Large Thinking's context window?

Trinity Large Thinking supports a context window of up to 262K tokens.

How does Trinity Large Thinking perform on benchmarks?

On standardized evaluations tracked by Epoch AI, Trinity Large Thinking scores 90.1% on τ²-bench, 40.6% on SciCode, 15.8% on Humanity's Last Exam.

Benchmark data from Epoch AI (CC BY) and Artificial Analysis; pricing, specs & descriptions from OpenRouter. Data updated Sep 10, 2026.