Efficiency First
We select silicon by tokens-per-watt and workload fit, not by brand logo. Galaxy, Cerebras, Helios, Trainium, and photonic paths compete on benchmarks.
§ 01 Sovereign AI Stack · 1,801 planned
As of August 2026 the Vasilikos campus is pre-construction. Phase 1 CapEx is Tenstorrent Galaxy RISC-V inference. Later halls evaluate AMD Helios, Cerebras CS-3/CS-4, IBM z17, Trainium 4, and photonic fabrics — mixed ISA, vendor-agnostic, no exclusive lock-in.
§ 02 Philosophy · Open architecture
The AI silicon market is in rapid design churn — new XPUs, 3D-stacked inference chips, rack-scale open standards, and lightwave fabrics ship on 12–24 month cycles. AGICY evaluates and integrates the most efficient open-weight-compatible hardware per workload class. We do not sign exclusive, decade-long single-vendor infrastructure marriages: SRAs reserve capacity; the technology committee re-scores silicon each generation.
We select silicon by tokens-per-watt and workload fit, not by brand logo. Galaxy, Cerebras, Helios, Trainium, and photonic paths compete on benchmarks.
Every hardware decision is supported by published benchmarks, vendor specifications, and internal validation testing.
Open fabrics (OCP ORW, Ethernet, UALoE) and modular halls let us adopt Trainium 4, MI450, or lightwave switching when diligence clears — without rewriting the campus.
§ 03 Zone Architecture · Two planned zones
Token Factory is planned RISC-V (Galaxy). Training Zone is mixed ISA on evaluation — Helios GPUs, Cerebras wafers, Trainium XPUs. Not two RISC-V halls.
Dense model training, fine-tuning, and research workloads.
Real-time inference serving at scale. Planned output is tokens — a token factory, not a storage warehouse.
§ 04 Heterogeneous Fleet · 8 silicon lanes
Phase 1 and Phase 2 anchors receive equal architectural weight. Additional lanes (AMD Helios, Taalas HC1, AWS Trainium 4, photonic fabrics) stay on the technology committee watchlist until independent EU workload validation completes.
§ 04b Footprint · Galaxy · NVL72
A Tenstorrent Galaxy is a 32-chip Blackhole mesh in 6U — approximately 19,200 mm² of silicon at ~600 mm² per die (roadmap neighbourhood; die area, not core area), drawing 8–10 kW operating and air-cooled. Seven occupy 42U of rack space at $770,000 public list and roughly 63 kW IT, before top-of-rack switching. One NVIDIA GB200 NVL72 is a purpose-built liquid-cooled rack at ~$3,000,000 analyst mid (band $2.0–3.4M) and 120+ kW, occupying a comparable floor footprint. Campus is PRE-CONSTRUCTION. Executed offtake is 0.
$110,000
Public list · per 6U server · 8–10 kW operating · ~12 kW max
~$3,000,000
Analyst mid · per rack · band $2.0–3.4M
Hardware list price for that footprint: $770,000 versus $2.0–3.4M — a 2.6× to 4.4× spread on purchase price alone. That is not a claim that either produces the same number of tokens. Cost-per-token is the thesis; same-model, same-precision benchmarks are in progress. Planned output is tokens — a token factory, not a storage warehouse. Logging and provenance are designed to support customers' EU AI Act Article 12 record-keeping obligations. Hosting compute does not make AGICY the provider of a customer's high-risk AI system. Standalone Annex III high-risk obligations apply from 2 December 2027 (Digital Omnibus).
§ 05 Performance · Labeled bases · Aug 2026
Four figures, four bases — do not mix them. The retired 400B dense ~2,500 tok/s (2.5× unnamed GPU) row is gone: it collided with vendor-class DeepSeek-R1 671B (~350–400 tok/s/user). Campus is pre-construction. Nothing here is a live meter.
| Figure | Value | Basis |
|---|---|---|
| CY meter / server | 300 tok/s | MODELED commercial meter: 1 Galaxy treated as 1 CY. Same spine as 17T/yr. |
| CY meter / Phase-1 fleet | 540K tok/s | 1,801 × 300 = 540,300. MODELED — not a live meter. |
| SC-36 blended / Galaxy | ~6,649 tok/s | MODELED theoretical blended. Fleet ~378T/yr on the homepage. Not 100% utilization. Not the CY meter. |
| DeepSeek-R1 671B / user | ~350–400 tok/s | Vendor-class / research citation — not AGICY lab. Not a 400B dense class row. |
| Operator class | How compute is paid for |
|---|---|
| AGICY | Plans to own the campus iron (pre-construction). Executed offtake is zero. Return math stays under NDA. |
| Frontier labs | Typically lease hyperscaler GPU. Public margin figures are not AGICY-measured and are omitted here. |
| GPU clouds / colo | Rent or colo GPU systems. Different stack than planned Galaxy air-cooled halls. |
§ 06 Sovereign Infrastructure · Pre-construction
Renders and energy cards are the planned Vasilikos campus as of August 2026 — not a live hall. Phase 1 Galaxy racks are air-cooled (zero water at the rack). Campus PUE 1.20 TARGET is the facility envelope — a seawater-plant design target on the homepage, not a claim that Galaxy servers are water-cooled. Behind-the-meter generation is a design target — not ordered. Phase 2 Helios / Cerebras halls would be liquid-cooled if procured.
Vasilikos energy zone, Limassol — planned sovereign AI campus. No live halls, turbines, or meters.
Air-cooled RISC-V inference — Phase 1 CapEx anchor. Helios, Cerebras, and Taalas remain Phase 2+ evaluation.
Phase 1 Galaxy racks: air-cooled, zero water at the rack. Campus PUE 1.20 TARGET is the facility envelope (seawater plant design target) — not rack water cooling. Phase 2 liquid halls only if procured.
Behind-the-meter generation is modeled for investors — not ordered and not operational. Facility MW, fuel, and turbine diligence: investor briefing / data room.
Facility MW, fuel €/MWh, turbine slot, and HRSG diligence are in the investor energy model — not repeated on this public page. Investor relations →
§ 07 Phase 2 Planning · Vendor-agnostic roadmap
Phase 1 deploys 1,801 Tenstorrent Galaxy servers for air-cooled sovereign inference. Phase 2 opens liquid-cooled training halls with AMD Helios / MI450 open racks first on the evaluation queue, then Taalas HC1 (model-hardwired inference · AMD acquisition agreement 6 Aug 2026 · not ordered · AMD newsroom), Cerebras CS-3 / CS-4 (CS-4 announced 18 Aug 2026 · not ordered), IBM z17 / LinuxONE 5 (on Watchlist), plus AWS Trainium 4 (~2027), d-Matrix Corsair (on Watchlist), and photonic fabrics. Nothing is procured on faith or exclusive vendor contracts.
Phase 1 · Tenstorrent Galaxy briefing
Planned Phase 1 CapEx — 1,801 Galaxy RISC-V servers. Campus is pre-construction. Cerebras CS-4 is announced, not ordered. We do not OCR or invent financials. This is Aphrodite, an AI agent — not a human spokesperson.
Cerebras CS-3 · Aphrodite Video Briefing (CS-3 footage · CS-4 on the hardware page)
Phase 2+ · Taalas HC1 product brief (vendor media · watchlist)
Model-hardwired Llama 3.1 8B demonstrator — vendor-claimed ~17k tok/s/user. AMD definitive acquisition agreement announced 6 Aug 2026 (subject to closing). AMD press release ↗. Not Phase 1 CapEx. Not AGICY lab measurement.
Phase 2 · Evaluation
OCP ORW open rack · 64–72 MI450 GPUs · H2 2026 ramp · no exclusive AMD commit.
Helios hardware page →Phase 2 · CS-4 announced · not ordered
CS-3 WSE-3 plus CS-4 Nexus (three WSE-3 Turbo wafers, 18 Aug 2026). Not WSE-4. Not owned Phase 1 fleet.
Cerebras hardware page →Phase 2+ · Watchlist
Telum II + Spyre · AI next to the ledger · GA 12 Aug 2026 · not Galaxy CapEx.
IBM z17 hardware page →Phase 2+ Evaluation Queue
Model-hardwired Llama 3.1 8B · ~17k tok/s/user (vendor) · AMD deal announced · not CapEx
DIMC decode lane · sub-2 ms · Phase 2–3 · on Watchlist
Custom XPU · FP4 native · ~2027 GA · hybrid NVLink Fusion (reported)
Lightwave fabrics + 3D memory stacking for tokens-per-watt
Chip & system builders — propose accelerators for the Vasilikos review queue
Cerebras CS-3 · Wafer-Scale Engine (CS-4: see hardware page)
Cerebras vs Groq · vendor claims, not AGICY lab
Compiled from Cerebras and Groq public marketing (70B/8B open-weight class, 2024–2026 product pages). Not same-model, same-precision, same-date AGICY measurements. Do not treat as a ranking.
1,801 Tenstorrent Galaxy RISC-V (planned CapEx)
Inference · Fine-tuning · Marketplace
CS-3 / CS-4 + Helios eval + liquid halls
Training · Foundation models
Hybrid multi-vendor + photonic fabric
Full-stack sovereign AI
Phase 1 · Galaxy thermal profile (air-cooled racks · planned · not a comparison)
| Metric | Phase 1 Galaxy (air, planned) |
|---|---|
| Chip TDP | 300W (Blackhole) |
| Server power | 8–10 kW operating · ~12 kW max |
| Cooling method | Forced air (standard HVAC) |
| Plumbing | None for Phase 1 Galaxy |
| Water for cooling | Zero |
§ 08 Sovereign Co-Design · Alliance intent
Through the Sovereign Compute Alliance, top-tier reserved-capacity clients (SRA L5 — the highest published reservation band) are intended to join architectural design, fine-tuning, and deployment of models once the campus is live. Campus is pre-construction. Images below are concept, not a current lab.
§ 09 Developer Experience · Galaxy SDK · CS-3 vendor DX
Phase 1 developers use Tenstorrent TT-Forge / TT-Metal. The table below is Cerebras CS-3 vendor-claimed DX for a possible Phase 2 wafer-scale lane — not the Galaxy compiler, and not a live AGICY cluster.
| Metric | Cerebras CS-3 (Cerebras vendor marketing) | Multi-node GPU (Cerebras vendor comparison) |
|---|---|---|
| Code for 175B-class training | ~565 lines (Cerebras vendor) | ~20,000 lines (Cerebras vendor GPU comparison) |
| Engineering team required | 3–5 ML engineers (Cerebras vendor) | 35+ systems engineers (Cerebras vendor GPU comparison) |
| Cluster management | Single wafer-scale device (Cerebras vendor) | Multi-node GPU orchestration (Cerebras vendor comparison) |
| Time to first inference | Hours (Cerebras SDK-native, vendor) | Weeks (Cerebras vendor GPU comparison) |
High-level ML compiler — converts PyTorch/JAX/TF models to Tensix instructions automatically
Low-level hardware interface — direct kernel programming for custom compute operations
Runtime scheduler — manages chip-level execution, memory allocation, and data movement
Information Security
Target: Go-LiveTrust Service Criteria
Target: Y1Data Protection
By DesignNetwork Security
Target: Go-LiveAI Regulation
By DesignDC Uptime Institute
Target§ 10 Hardware Review · Chip & system builders
AGICY runs a vendor-agnostic evaluation queue for the Vasilikos campus. Phase 1 CapEx stays on Tenstorrent Galaxy. Phase 2 priority today: AMD Helios / MI450, then Cerebras CS-3 / CS-4 (CS-4 announced 18 Aug 2026 · not ordered), IBM z17 / LinuxONE 5 (on Watchlist), with Trainium 4, d-Matrix Corsair (on Watchlist), and photonic fabric lanes behind. Propose rack-scale, wafer-scale, decode, or fabric systems for committee review — no exclusive lock-in implied by a listing.
§ 11 Ready? · Deploy on sovereign compute
Whether you're evaluating compute for your enterprise, planning an AI research programme, or investing in sovereign infrastructure.