ModelBeat
Models/Trinity Large Thinking
All models
A

Arcee AI: Trinity Large Thinking

arcee-ai/trinity-large-thinking

Trinity Large Thinking is Arcee AI’s AI model, released Apr 1, 2026. It costs $0.25/M input and $0.80/M output tokens, and handles a 262K-token context window. Its strongest use case is agentic & tools, where it ranks in the 60th percentile of the models Model Beat tracks.

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks.

Context
262K
Input
$0.25 / 1M
Output
$0.80 / 1M
Released
Apr 1, 2026
Modalities
texttext
Strongest atAgentic & Tools top 40%

Benchmarks

Scores on standardized evaluations. Higher is better — and rank shows where Trinity Large Thinking lands among all models tracked on Model Beat.

32.0
Intelligence Index
Epoch AI
32nd percentile of tracked models
12.0
Coding Index
Epoch AI
12th percentile of tracked models
60.0
Agentic Index
Epoch AI
60th percentile of tracked models

Reasoning

2 evals
GPQA Diamond75.2%

Graduate-level scientific reasoning

Humanity's Last Exam14.7%

Frontier of human expert knowledge

Coding

2 evals
SciCode36.1%

Scientific research coding

WebDev Arena1239

Human-rated web development

Agentic & Tools

1 evals
τ²-bench90.1%

Tool-agent-user reliability

Compare Trinity Large Thinking with

Frequently asked questions

What is Trinity Large Thinking?

Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks.

Who created Trinity Large Thinking?

Trinity Large Thinking is an AI model developed by Arcee AI, released Apr 1, 2026.

How much does Trinity Large Thinking cost?

Trinity Large Thinking's list price from Arcee AI is $0.25 per million input tokens and $0.80 per million output tokens.

What is Trinity Large Thinking's context window?

Trinity Large Thinking supports a context window of up to 262K tokens.

How does Trinity Large Thinking perform on benchmarks?

On standardized evaluations tracked by Epoch AI, Trinity Large Thinking scores 90.1% on τ²-bench, 14.7% on Humanity's Last Exam, 75.2% on GPQA Diamond.

Benchmark data from Epoch AI (CC BY) and Artificial Analysis; pricing, specs & descriptions from OpenRouter. Data updated Jul 30, 2026.