Institutional & sovereign reasoning
Models investigating, orchestrating, cross-evaluating, and synthesizing across local, private, public-cloud, and sovereign environments.
A clear line between vision and evidence
NORTHSTAR is in active research and development toward institutional and sovereign-scale synthetic reasoning. This site explains that vision; it is not a live release dashboard or a declaration of general availability.
Public vision · September 27, 2026What the website represents
Models investigating, orchestrating, cross-evaluating, and synthesizing across local, private, public-cloud, and sovereign environments.
Locally rendered stars, connections, and example workflows. The display is not connected to a fleet and does not report real resource counts or model performance.
Runtime behavior, installed-package acceptance, provider interoperability, security, and reasoning quality must be evaluated for the actual version and environment.
How progress should be measured
A credible result states the source revision, artifacts, model identities, environment, expected behavior, observed outcome, and remaining limitations. Fixture, simulation, loopback, installed-package, and physical-network evidence are different proof levels.
Compare with strong single-model and simpler multi-model baselines. Examine evidence coverage, contradiction detection, calibration, reproducibility, and human usefulness under comparable cost and time constraints.
Show actual, authorized composition across the claimed workers and providers. Test admission, capacity, quota, billing, cancellation, revocation, restart, and uncertain outcomes.
Dozens and hundreds of nodes are design objectives. Thousands remain a future horizon as control-plane hardware and orchestration capacity advance. Demonstrate scheduler throughput, bounded memory, coordination overhead, evidence volume, and safe recovery at each claimed size. Simulation does not certify production fleet capacity.
Qualify each supported routing, switching, firewall, NAT, and VPN backend in an isolated lab. Prove effective traffic behavior and recovery—not merely successful command execution.
Demonstrate the specific deployment’s data-flow restrictions, identity controls, tenant separation, auditability, and operational ownership. Do not assume geography alone establishes sovereignty.
Verify installation, accessibility, first-run setup, upgrade, interruption, and recovery on the actual supported platforms. Interface messages must match current execution evidence.
Missing credentials, labs, independent assessment, or release evidence are reported as remaining gates. They do not become passes, and they do not shrink the vision.
Follow the work, not the hype
Follow Mainely Code for development announcements. Any public release should carry its own supported scope, installation guidance, and qualification record.