Hermes · Preliminary Findings · Not decision-grade

Frontier Model Selection per Constellation

Date 2026-07-15 Author Hermes Claude facts live-verified (claude-api SSOT) Mapping inferred, not benchmarked GPT axis not researched

Methodology declared up front (rule #05 research-depth tiering): the model facts below are verified against the Anthropic model catalog (cached 2026-06-24, newer than training cutoff). The per-constellation mapping is reasoned from documented characteristics + known roles — not head-to-head benchmarked. This report deliberately stops short of a decision-grade arc.

§0Premise corrections

The originating question named "Fable 5 and Opus 5.6" and "GPT 5.6." Three corrections gate everything downstream:

  1. There is no Opus 5.6. The Opus line tops out at Opus 4.8 (claude-opus-4-8). The current "5"-generation flagship is Fable 5 (claude-fable-5) — a distinct top tier, not an Opus point-release.
  2. GPT is a different vendor and architecture. The whole Pantheon runs on Claude Code. Swapping a constellation to an OpenAI GPT model is a harness/architecture change, not a config swap — and current GPT specs can't be asserted from a stale (Jan 2026) prior without live research. Out of scope here.
  3. Both Claude frontier tiers carry a 1M context window. "Smaller context window" and "Fable is context-limited" don't apply — Fable 5 and Opus 4.8 are both 1M.

§1Fable 5 vs Opus 4.8 — the real comparison

DimensionFable 5Opus 4.8
PositionMost capable widely-releasedMost capable Opus-tier
Context / max output1M / 128K1M / 128K
Price (in / out per MTok)$10 / $50$5 / $25 (half)
ThinkingAlways on; can't disableAdaptive; can disable
Turn latencyMinutes on hard tasksFaster; fast mode ≈2.5× throughput
Data retention30-day required (no ZDR)Standard
Named strengthsLong-horizon autonomous runs; first-shot well-specified systems; enterprise deliverables (financial analysis, spreadsheets, docs); repo-history search; ambiguity; parallel sub-agent + async peer commsSOTA agentic execution; knowledge work; memory; one-shot bug fixes; catches flakes; mid-session system messages; warmer writing
Security caveatflag Cyber/bio classifiers false-positive on legit security tooling; bug-finding gains exclude security analysisNone

§2The key reframe

The intuition "Fable = planning, weak at engineering → switch to Opus for engineering" is inverted. Fable 5's largest gains are in long-horizon autonomous coding and first-shot system builds — it's the top tier at both planning and hard engineering. The real decision variable is not planning-vs-engineering. It is:

stakes × latency-tolerance × volume ÷ cost
TierEconomicsUse for
Fable 52× price, minutes/turnRare, high-stakes, long-horizon, latency-tolerant work where a better outcome clearly justifies cost + wait
Opus 4.8Half cost, fast modeThe fleet default workhorse — near-top capability, interactive-friendly
Sonnet 5$3 / $15High-volume lanes; near-Opus on coding/agentic
Haiku 4.5$1 / $5, 200K ctxSubagents, sweeps, simple tasks

§3Preliminary per-constellation mapping

ConstellationPrimary roleTierRationale
AthenaGovernance, cross-fleet rulingsFable 5Rare, load-bearing verdicts; deepest reasoning wins, cost irrelevant at low volume
ProteusSystems architecture, schema/DDL, roadmapsFable 5"First-shot well-specified systems" is Fable's named sweet spot (Opus for routine DDL)
PlutusTrading, contracts, financial analysisFable 5"Enterprise deliverables — financial analysis, spreadsheets" is a named strength (Opus for ops)
HephaistosPR review / merge gateOpus 4.8Strong bug-finding + fast-mode throughput + cost; escalate hardest reviews to Fable
HadesCredentials, vault, security scansOpus 4.8 not FableFable's cyber classifier false-positives on legit security work; bug-finding excludes security
AtlasInfra/ops, skills & rules authoringOpus 4.8Careful critical-path authoring, high volume
MetisCRM/scheduling architecture + buildOpus 4.8Mixed design/build (Fable for large architecture arcs)
HermesResearch + intelligence pipelineOpus 4.8 + Fable + SonnetOpus controller; Fable for T-deep arcs; Sonnet subagents; Haiku sweeps. Volume drives cost efficiency
ApolloFrontend / UI / visualOpus 4.8 / SonnetDesign instinct adequate; cheaper
MnemosyneMemory / knowledge graphOpus 4.8Pipeline/build work
pantheon-opsn8n / hub automationSonnet 5 / OpusOps volume
Shape of it: Fable 5 for the 3–4 constellations whose output is rare, high-stakes, and long-horizon (Athena, Proteus, Plutus-analysis). Opus 4.8 as fleet default everywhere else. Sonnet 5 for high-volume/ops. Secondary lever — Opus 4.8 fast mode for interactive coordinator sessions, Fable 5 for overnight autonomous runs where latency is free.
Hard carve-out: keep Hades and all security-adjacent work off Fable 5 — cyber-classifier friction plus documented exclusion of security bug-finding.

§4Open before this becomes decision-grade

Promote to a T-deep arc (parallel per-model sub-agent lanes, live-sourced benchmarks, cited verdict, cost model) only if a fleet-wide model-assignment ruling will hang on it.