Provable, not pitchable.
Every artifact below links to something you can audit, run, or verify without trusting us. Pick any one — there's no slide deck behind it.
Live track record
These artifacts update over time and accumulate evidence. Hard to fake without leaving traces.
Forecast integrity chain
Every forecast is hashed, sealed daily into a Merkle tree, and committed to a public GitHub repo. Verify any forecast against a published root.
Live anchored forecasting tournament
System forecasts on real Manifold and Polymarket questions, each cryptographically anchored at creation. System Brier vs. market Brier shown for every resolved question.
Calibration leaderboard
Public rankings by Brier score and calibration error. Filterable by tier, timeframe, and domain.
Source credibility leaderboard
Which news outlets, think tanks, and analysts actually predict accurately? Rankings derived from the predictive value of sources cited in resolved forecasts.
Fed-Watch
Live FOMC forecasts vs CME Fed Funds futures-implied probabilities. System Brier vs market Brier per meeting. Each forecast cryptographically anchored at creation.
Earnings-Watch
Pre-earnings strategic briefs on top US public companies. EPS direction + price reaction forecasts. ICD 203-graded prior-quarter forward-looking commentary.
Pre-registered predictions
Every public forecast hashed (SHA-256), sealed daily into a Merkle tree, and committed to GitHub BEFORE resolution. Cherry-picking is mechanically impossible.
Research
Open standards and benchmarks. Submitted to or in submission for top venues.
Strategic Cognition Benchmark leaderboard
Frontier-model evaluations on the SCB benchmark — 205 tasks across 5 dimensions of strategic cognition, grounded in ICD 203.
Strategic Cognition Benchmark — paper
v15 candidate specification (~35,000 words). 205 tasks, JCS non-compensatory gating, ICD 203 grounding, BetterBench self-assessment 91/100.
Strategic Cognition Ontology — paper
v21 candidate specification (~28,000 words). 35 OWL classes, 18 SKOS schemes, 8 Exchange Profiles, A2A protocol extension.
Capability Cards
Per-AI-system capability summaries (SCI / SCI-H, archetype, dimension breakdown, strengths, limitations). Citable at /capability-cards/<system>/<date>.
Interactive demos
Try the platform on your own input. No account needed for a sample run.
Strategic overlay for Claude Tag
A simulated Slack channel showing the judgment layer respond on top of Anthropic's Claude Tag: calibrated forecasts, ambient contradiction catches over logged assumptions, groupthink checks, and Merkle-sealed decisions captured in-thread.
SCO conformance demo
Enter a strategic topic. Watch the system generate validated SCO artifacts (Landscape, Drivers, Predictions, Scenarios) in real-time, with conformance violations surfaced.
Brief Critic — ICD 203 public lab
Paste any analytical document. Get an ICD 203 tradecraft scorecard with paragraph-level findings. Share results via permanent link.
OKR strategic twin
Paste a Google Sheets URL or upload a CSV of your OKRs. See them as a strategic landscape with forecast probabilities and external-signal binding.
WhatsApp bot demo
Daily intelligence briefs over WhatsApp. Voice messages transcribed, audio briefs delivered, war-game scenarios you can play through interactively.
Multi-persona memo critic
Paste a strategic memo. Run it through 4 personas in parallel: McKinsey partner, Sequoia GP, Pentagon J5 strategist, Chief Risk Officer.
Storyline arc tracker
8 industry-shaping narratives (Fed pivot, China-Taiwan, AI Act, energy transition, etc.) tracked from emergence → rising → climax → resolution. Posture, signposts, and assumption decay.
Watch the AI think
ADW Delphi decomposition trace step-by-step. Semantic parse → decomposition → base-rate anchor → panel vote → adversarial round → ensemble synth → ICD 203 check → signposts.
Counterfactual sandbox
Pick a forecast. Toggle off individual sources / signals / assumptions. See the probability shift live, with attribution math per element.
Scenario room
Pick a strategic question. Get 3-5 distinct futures with conditional probabilities, decision implications, pre-mortem signals, and key drivers.
Source vs source
Pick any two news sources. Head-to-head predictive value (Brier improvement vs uninformed base rate) with 95% CIs and per-category breakdown.
Posture matrix
For each strategic situation, system recommends a posture (Act / Probe / Watch / Hedge / Exit) with explicit rationale, the regret accepted, and 3 leading indicators that would shift the call.
Information gaps
Paste a thesis. System ranks the load-bearing unknowns by relevance × knowability ÷ cost. The research priority list — 'if you only have time to find out 3 things, here they are.'
Coherence check
Paste a strategy. System identifies internal contradictions, tensions, unstated assumptions, reasoning gaps, circular reasoning — with paragraph-level citations and suggested fixes.
Assumption audit
Paste a memo. System extracts every load-bearing belief, scores initial vs. current confidence, names the kill-criterion, and tracks decay across accumulated evidence.
Decision frame
Pose a strategic question. System emits the operational decision frame: who decides, options with cost/yield/reversibility, recommended option with rationale, change-of-mind criteria, deadline.
Pre-mortem
Paste a thesis. System runs adversarial passes producing failure modes — each with leading-indicator signposts and the load-bearing assumption that breaks. 'If I see X, I'm wrong.'
Perspective pack
Two persona-trained reasoners attack the same question. System aligns claims, identifies agreement zones, and surfaces divergences with reconciliation paths. SCO §6 Exchange Profile in action.
Landscape composer
Pose a strategic question. Compose the landscape graph: actors, drivers, dependencies, decision-points, actions, scenarios, signals, constraints. Live SCO SHACL validation as you toggle nodes.
Governing one AI agent, end to end
One agent's strategic call passes through five governance surfaces: an ICD 203 tradecraft contract that blocks the sloppy draft, a critic whose own reliability is measured and whose verdict escalates when it is low, a hash-chained decision ledger, outcome scoring against what actually happened, and an offline auditor check.
Developer surfaces
Install, integrate, and run yourself. Code is on GitHub or npm.
Connect to Claude Tag
Add the LUU strategic tools to your Claude Tag in Slack (Mode ①). Generate a key, add an MCP connector, and @Claude a verb. LUU holds no Slack token — it receives only task-scoped inputs.
AI agent framework showcase
Three runnable agents — daily-briefing, enterprise-risk-monitor, geopolitical-advisor — built on CrewAI, LangChain, Semantic Kernel, and Google ADK.
@luu/attestation-verifier↗
Offline verifier for LUU evidence packs. Zero network I/O, native crypto, timing-safe comparisons. Verifies content hashes, Merkle proofs, and Ed25519 signatures.
MCP server (live)↗
Model Context Protocol server exposing strategic-intelligence tools to AI agents over JSON-RPC 2.0. Live endpoint; Tier 0 anonymous access available (5 queries/day per IP).
Infrastructure
The plumbing behind everything else. Verifiable on its own merits.
CloudEvents subscriber
Subscribe to platform events (decision-created, action-executed, outcome-observed, seal-anchored). HMAC-SHA256 signed, RFC-compliant CloudEvents 1.0.
Public anchors repo↗
GitHub repository of daily Merkle-root commits. Append-only, signed commits, branch-protected. Walk the chain offline with the bundled examples/walk-chain.mjs.
Want this for your team?
Drop your email and we'll be in touch when access opens for your vertical.