halcyon-3 ultra
Most capableDeep reasoning, long documents and complex agents.
- context
- 1M
- input /1M
- £2.40
- output /1M
- £9.60
relative speed
halcyon-3 is now generally available
Frontier-level accuracy, private by default and priced for production. Build agents, copilots and document workflows on infrastructure hosted in the UK and EU.
▸ read_documents(3 files, 84 pages)
Two conflicts found. Northgate Ltd §7.2 requires payment within 14 days, and Arden Supply §11 adds a 4% late fee after 21 days. The third contract matches your terms.
1.2B
tokens served daily
99.99%
API uptime, last 12 months
<180ms
median time to first token
SOC 2 · ISO 27001
independently audited
01 / models
Deep reasoning, long documents and complex agents.
relative speed
The everyday workhorse for chat, RAG and coding.
relative speed
Sub-200ms answers for high-volume classification.
relative speed
02 / research
Scores on public benchmarks, run by our evaluation team with published prompts and seeds. Full methodology in the technical report.
read the report| benchmark | halcyon-3 | halcyon-2 | open-weights avg. |
|---|---|---|---|
| MMLU-Pro | 84.1 | 76.3 | 71.0 |
| GPQA Diamond | 71.8 | 60.2 | 55.4 |
| SWE-bench Verified | 63.5 | 49.0 | 41.7 |
| HumanEval+ | 92.4 | 85.1 | 80.6 |
03 / platform
from halcyon import Halcyon
client = Halcyon()
reply = client.chat.create(
model="halcyon-3",
messages=[{"role": "user", "content": prompt}],
region="uk-south",
)
print(reply.text)