● OPERATOR OF RECORD · MICHAEL MOFFETT · accountable on every commit team@caliperforge.com
AI RESEARCH STUDIO · SECURITY & INVARIANTS

Verifiable, auditable, correctable.

An AI research studio specializing in security and invariants. We build autonomous systems you can check, and the delivery is the proof.

"We build autonomous systems you can check. The proof is in the receipts, not in a claim about the receipts."
HUMAN IN THE LOOP
OPERATOR OF RECORD
HOW THE STUDIO OPERATES AT SCALE
39
SPECIALIZED AGENTS
scoped roles, cross-checked
4,793
LOGGED AGENT DECISIONS
every classifier + gate row
1,638
AGENT DISPATCHES
each one produces receipts
49
DAYS RUNNING
since 2026-05-27
Operational totals as of 2026-07-14. The adjudicated catch record — 35 defects caught at 83%, 0 false positives — follows below.  Jump to the record ↓
01 / WHAT CALIPERFORGE IS

An agent org that ships receipts, not claims.

An AI research studio specializing in security and invariants. The studio runs as an agent org: specialized roles, invariant tooling, and a human operator who reviews every deliverable before it ships.

We build autonomous systems you can check. Stateful invariant testing. Planted-twin CI on every harness. A published record of what our own gates caught, and what they missed. The proof is in the receipts.

02 / THE METHOD
01 · EXPRESS

Express the invariant

We express a protocol's safety rule as a machine-checkable invariant, one line that captures what must always hold.

02 · BUILD TWIN

Ship a planted-bug twin

A clean reference where the invariant holds, and a planted-bug twin where it fires. Both run in CI on every push.

03 · PUBLISH

Receipts, not claims

Every deliverable is verifiable. The CI record is public. The self-correction log is public.

03 / WHAT WE CAUGHT

The adjudicated record — the proof.

Full gate record →

Scale alone is not the claim. Of 86 adjudicated §4a/§4b decisions, here is what the gates caught, and what they missed — every one logged.

35
DEFECTS CAUGHT
across 86 decisions
83%
CATCH RATE
of defects present
0
FALSE POSITIVES
zero, to date
7
DOCUMENTED MISSES
published in full
THE FULL-WINDOW LEDGER
35 caught 7 missed 28 clean, 0 false positives
04 / WHAT WE SHIP

Open-source tooling we maintain.

Source on GitHub, license on the card. AI involvement disclosed at point of use.
Latest public repo · 2026-07-02

uniswap-v4-invariants

Defender-side invariant harness for Uniswap v4 hooks. Recurring bug classes are expressed as stateful invariants on real v4-core, each shipped as a clean and planted-bug twin pair that both run in CI on every push.

New · flagship EVM · Uniswap v4 · Foundry Apache-2.0 github.com/caliperforge/uniswap-v4-invariants →
Exploit→Invariant Atlas
CI ✓

Seven real-world hacks across four VMs, each with a runnable invariant that would have caught the bug class on the pre-exploit code. First defender-side, pre-deploy CI benchmark across Cairo, Move, Solana, and EVM.

cf-invariants
CI ✓

Open-source snforge sidecar that adds stateful invariant testing and AI-suggested invariants to Cairo 2.x. Twelve reference contracts, Voyager-verified on Starknet Sepolia.

hyperevm-safety
CI ✓

Invariants and CI-runnable property tests for HyperEVM lending protocols that consume HyperCore oracle reads. Six HyperCore-boundary invariants in v0.1; planted/incident twin fires on the same run.

cf-modeleval
CI ✓

Planted-twin discrimination-power harness for AI safety properties, prompt-injection and sycophancy resistance, run across Anthropic, OpenAI, and Groq. Every CLEAN claim paired to a PLANTED control.

Also in the public portfolio

Complete public listing · synced 2026-07-12
Defender-side stateful invariant harness for Euler Earn allocator and vault share accounting, running against real Euler Earn contracts (pinned submodule). One case in the initial cut, encoded as a same-source clean and planted twin pair against the Pashov Audit Group M-01 finding and Euler Labs' fix commit. Regression fixture; not an audit.
EVM · GPL-2.0-or-later
Reusable Foundry and Recon Chimera scaffold for EVM build-to-win contest entries. CI-verified.
EVM · MIT carveout
CI-verified cross-side conservation invariant reference for lock/mint bridges, anchored on the Verus-Ethereum 2026 case. Clean and planted-bug twin on every push.
EVM · Apache-2.0
Invariant testing and planted-twin CI for BSC DeFi protocols. PancakeSwap v3 harness live; Venus and Stargate planned. Foundry.
BSC · Apache-2.0
Language-conditioned detection-rate eval harness for AI code auditors. EN/ES/PT/CS variant corpus built on Atlas planted-bug twins. Apart Global South AI Safety Hackathon 2026.
Eval · Apache-2.0
Differential equivalence tests for Taiko's Type-1 Ethereum-equivalence guarantee, encoded as clean and planted twin Foundry projects. AI-augmented, hard-disclosed.
EVM · Apache-2.0
Soroban (Stellar) stateful invariant atlas. Clean and planted twin fixtures over already-public Soroban findings. AI-augmented, hard-disclosed.
Stellar · Apache-2.0 OR MIT
github.com/caliperforge  ·  18 public repos, all listed above  ·  AI-disclosed at point of use
05 / CHAINS WE COVER

We work where contracts are shipping.

STARKNET · CAIRO
Cairo 2.x
Contract work and developer tooling on snforge.
SOLANA · ANCHOR
Anchor / SPL
Programs, tooling, Foundation-funded work.
HYPEREVM · FOUNDRY
Solidity
HyperCore-boundary invariants and property tests for lending.
ETHEREUM · EVM
Invariant forensics
Bridge-conservation references and exploit-to-invariant reproductions.
Base and the Optimism Superchain come online as ecosystem work lands; Move (Sui, Aptos) and Go (Cosmos, Celestia) when contract volume warrants.
06 / ON-CHAIN & VERIFIABLE

Deployed on-chain, not just described.

We publish reference contracts on-chain so anyone can run our invariants against a live target, each shipped as a clean reference and a planted-bug twin, every deployment independently verified. Coverage expands chain by chain as the work lands.

12
REFERENCE CONTRACTS LIVE
100%
INDEPENDENTLY VERIFIED
2
LEGS EACH · CLEAN + PLANTED TWIN
INVARIANT CLASSES COVERED
Supply accounting Governance state AMM constant-product Vault share / asset Lending solvency Oracle monotonicity Staking conservation Timelock delay Multisig threshold Vesting cap
Live today on Starknet Sepolia · Cairo 2.x, more chains as coverage lands. Read the findings report →
07 / HOW WE'RE ORGANIZED

Why the receipts are trustworthy.

Independent checks at every load-bearing step, and every one of them logs to the published catch record.

INDEPENDENT REVIEWERS

§4a content and §4b code. The author never reviews their own work, every public claim and every code change passes a reviewer who didn't write it.

The 35 defects on record were caught here.
AI OPS INTEGRITY SWEEPS

A nightly pass over the task queue, role files, decision log, and gate verdicts, looking for process drift, missed handoffs, and stale state.

Operational hygiene, it keeps the records honest.
AI HR ROLE AUDITS

Every role is audited against its actual behavior. When a role acts outside scope, it surfaces in the audit log, not after a customer notices.

Reviewer stays separate from author.
DREAMING CONSOLIDATION LOOP EMERGING

A daily pass consolidating yesterday's gate outcomes and reviewer feedback into the standing playbook. Just coming online.

Reported as it matures, not load-bearing yet.
QA REVIEWERS

The review discipline applied to code, applied to public copy. Anything on this site, in a grant, or customer-facing passes a content QA reviewer first.

Every verdict is logged.

Every check above logs to the published gate record.

Read what we caught →
08 / CONTACT

Operated by Michael Moffett.

Accountable for every commit, PR, grant application, and bounty claim made under this org. For grant collaboration, engagement inquiries, or security-tooling questions:

Start an engagement →
X · FARCASTER@caliperforge
ENScaliperforge.eth