Workbench Lite
For trying the AIGIS engineering workflow
- 1 local repo
- 1 active feature lane
- Manual lane workflow
- Basic patch receipts & apply queue
- Bring your own provider
AIGIS SOLUTIONS
View Tool Pricing
AIGIS Workbench
AIGIS Workbench is the governed core of AIGIS software engineering — a local-first workbench where your engineering standards run as machine-enforced gates on autonomous coding agents, not prompts they are politely asked to honor. Two provider families run every role, and can be pitted against each other to reason and to QC each other's work.
>_ in plain termsNew to AI coding tools? Think of it as a safe control room: the AI writes the code, and Workbench checks it against your rules before anything is kept.
>_ Bring your own provider — Codex/GPT and Claude. AIGIS sells the engineering control layer, not model credits.
Work flows in from two providers, the gate decides, and only clean work stacks into history — held work never slips through. Illustrative loop.
Why it is different
Most AI coding tools hand the agent a style guide and hope. AIGIS Workbench compiles your standards into gates that run against the work, block a bad apply, and stay separated from advisory rules. That is the whole thesis: governed autonomy, honest by construction.
shared_architecture
>_foundation......shared_with_flagship_adsp1 >_role............stands_alone_as_a_product >_second_role.....proves_architecture_for_flagships
Workbench is built on the same architecture that powers our flagship ADSP1. It stands on its own as a product — and it is where AIGIS exercises and validates the engineering quality that flows into everything we ship.
>_ in plain termsYou set the rules once, and the workbench actually enforces them — instead of hoping the AI remembers. It is the same engine behind our flagship product, offered on its own.
Enforced rules run as deterministic, zero-cost checks that hold a patch back — not text an agent can quietly skip. Advisory rules stay clearly labeled as advice.
Claude and Codex/GPT run every role. Route each role independently, or pit them against each other to reason and to review each other's work.
Reviewers are read-only, cost is real or labeled an estimate, and no score, benchmark, or metric on any surface is fabricated. If a number appears, the product measured it.
The differentiator
One primitive, two surfaces: two different provider families reason independently on the same artifact, their disagreement is surfaced structurally, and a neutral lighter model synthesizes the result — never a fabricated agreement score.
>_ in plain termsTwo different AIs answer the same thing on their own, then a third neutral one writes up where they agreed and where they still disagree — so you are never trusting a single opinion.
Debate · artifact = a prompt
One prompt. Claude and Codex each answer first with no peeking, then run configurable structured rebuttal rounds. A separate neutral model — the Documenter role — organizes one combined answer plus an explicit "still contested" list. Watch divergence, rebuttal, and synthesis live, side by side, over SSE.
Review · artifact = a finished patch
Click Review on a patch or lane and the opposite provider and the same provider each QC the work, read-only, run as a task with a brief. Same-provider review catches execution slips; opposite-provider catches blind spots. When the two reviewers disagree with each other, that disagreement is elevated — not averaged away.
Debate and Review are committed day-one launch capabilities, currently being finalized. We describe what they do — we do not fake it with invented transcripts, scores, or benchmarks.
The workspace
Register your repositories, then work the way the task wants — unattended, hands-on, or conversational — over one governed core. Runs and apply are scoped per repo. Premium light and dark themes.
>_ in plain termsPick how hands-on you want to be: let it run on its own, guide it step by step, or just chat with it — all in one place.
The unattended governed pipeline: compose, lint, queue, run, review, apply.
Ad-hoc cockpit: generate a grounded prompt, run one lane, gate the apply.
Fluid single-agent conversation with git as the safety net.
Drop a list, or let an agent propose ranked, evidence-cited work.
Token-budgeted context retrieval and operator-approved memory cards.
Enforced gates and advisory rules, honestly separated, with a drift readout.
Per-role model routing, concurrency and burn caps, default review policy.
Register N repos with provider and scan depth; everything scopes per repo.
operational_status_bar
>_repo...........scoped_per_registered_repo >_providers......codex + claude live_status >_state..........applied / ready / idle >_infra..........poll_cadence · dedup · sse · in_flight
Batch Process
Author a task, lint it against a hard gate, and route it into the pipeline. One task is a trajectory; several become a multi-task trajectory, with automatic lane routing and grounding pulled from the repo. Then the scheduler runs it — and every action publishes to one live governance stream.
>_ in plain termsLine up the work, press go, and walk away. If something needs you, it asks one clear question instead of quietly getting stuck.
Scheduler controls
Real AIGIS Workbench UI — Batch Process running the Agent Patch Scheduler · scroll to explore
The front door: author a task and lint it against a hard gate before it can enter the pipeline. Lint-green in, or it does not run.
Live concurrency and burn-safety caps, a burn-fuse, repairs per session, rework ratio, and a session cap — with real token capture: cost per patch and burn rate in actual dollars, projected against the cap.
A live strip per lane (task / phase / elapsed / policy), and a per-task drawer with the assessment ribbon, file-by-file colorized diff, clickable changed-file paths, and the exact agent prompt sent.
Every action publishes to one live SSE stream: the deterministic zero-cost gates, operator decisions, and inline live validation screenshots — one place to watch it all.
Distinguishes connection-lost, provider-error, and usage-limit, then offers the matched action — Continue, Rerun-after-reset, or Resolve — so unattended runs stay trustworthy.
A skipped or orphaned gate surfaces exactly one operator decision — never a silent hang. Copy-to-clipboard on every brief and debrief.
Deterministic, zero-cost gates
Real dollar token capture in the Burn Ledger is a committed day-one launch feature, being finalized. Where a figure is an estimate, it is labeled an estimate. No burn number on this site is fabricated.
Hands-on modes
Not every task belongs in an unattended queue. Two more surfaces put you in the seat — with the same gates underneath.
Manual · the ad-hoc cockpit
A Prompt Generator turns a goal into a grounded prompt and runs a single lane. Lane placeholders show live state — Active or Ready — with Prepare, Run, and Remove. The Apply Queue uses tri-state "main clean / dirty" gating: apply is blocked while the working tree is dirty. Batch submissions and job queue are one glance away, per repo.
Repo · git as the safety net
A repositories registry with typed repo maps — index, architecture, source-of-truth, changelog — compiles into an AIGIS.md with a managed pointer block. Open a live conversational session: choose Codex or Claude, model and reasoning level, read-only or workspace-write. Work on a branch with checkpoint, discard, and change-set review. Both providers share one durable transcript, with live token streaming and honest stream diagnostics.
Context & queue
Memory
A context compiler previews a token-budgeted retrieval with a why-reason per slice — non-destructive, showing exactly what an agent would receive. Memory cards cover user preference, project, continuity, rule, and decision; pinned cards are always included, and secrets are stripped on save. The capture queue is operator-approved and never auto-promotes.
Backlog
The Backlog Composer turns each line of a list into an independent lint-green task in its own lane, auto-queued, with a persistent "N tasks queued → open Batch" readout that survives refresh. The Backlog Proposer reads the repo, maps, and memory and proposes ranked, evidence-cited work; approve routes it through the same pipeline, dismiss drops it.
Governance controls
Standards
Two standard sets — General good engineering and AIGIS engineering — as a per-rule corpus with toggles; toggling a rule off omits it from the compiled AIGIS.md. Enforced gates surface live and stay separated from advisory rules, with a drift readout for what has changed since the maps were written. The Trust Ledger tracks per-lane trust level and record (applies / holds / fails), proposes graduation to the review policy, and supports drop-hard-on-failure and pin-to-enforce.
Settings
Per-role model routing across Tier 1/2/3, Bugfix, Reviewer, Prompt generator, Documenter, and Backlog Proposer — each Codex or Claude, each with its own reasoning and effort. Set concurrency and burn-safety caps, and a default review policy: enforce, agent, or bypass.
What makes it different
AIGIS Workbench does not replace Codex, Claude, or your editor. It governs the engineering process around them — and shows its work.
Bring your own provider
AIGIS Workbench does not include third-party model usage. Use your own Codex, Claude, OpenAI, Anthropic, or supported provider access where available. AIGIS sells the engineering control layer, not model credits.
Provider availability, API access, subscriptions, and usage costs depend on each third-party provider. AIGIS Workbench pricing does not include those costs unless explicitly stated in the future.
Coming next
Directions we are building toward. These are not part of the launch feature set and are not implied as available.
Planned tool pricing
AIGIS Tools pricing is still being shaped. The goal is a light free version, paid yearly editions users can keep, and optional subscriptions for people who want continuous updates.
For trying the AIGIS engineering workflow
Solo devs & founders running multiple AI lanes
Power users, technical founders & agencies
Small teams that want shared AI-dev standards