Portfolio · Updated October 3, 2026
I Don't Advise on AI. I Build It.
Agor AI Studio builds AI agents that run one workflow, at a fixed price. This is the work behind that offer: four apps live on the Apple App Store, shipped in seventy-seven days, thirteen live sites, three published books and a research podcast, all built and run with AI. I also built a 39-agent autonomous build pipeline and then retired it, because an adversarial review of my own work found its governance had never once fired. When I advise a client on AI, I have already met the problems they are about to.
Agents
39
Apps
4
Sites live
13
Articles
1016
Episodes
44
Tests passing
1,300+
Archived Flagship
MVAT Studio: an autonomous multi-agent app factory
A framework in which 39 AI agents built, tested and shipped mobile apps, with no human in the loop during pipeline execution. It put real apps on the App Store. It is also archived, and the reason why is the more useful half of the story.
Architecture
Product
5Strategy, PRDs, Personas, Prioritization, Market Research
Design
5UX, UI, Design System, Interactions, Accessibility
Engineering
8Architecture, Frontend, Backend, Security, DevOps, Code Review
Testing
5Strategy, Unit Tests, Integration Tests, Quality Gate, Auto-Heal
Marketing
5ASO, Content, Social Media, Ad Ops, Launch Coordination
Analytics
5Metrics, Behavior, Crashes, Anomalies, Experiments
Finance
4Revenue, Budget, Forecasting, Spend Alerts
Governance
2Pipeline Judge, Spec Evolver (mutual oversight)
Tiered model assignment: cost optimization without quality loss
Opus
Production code + critical gates
Sonnet
Content, analysis, design specs
Haiku
Read-only analytics + reporting
Governance innovation
The hard problem in multi-agent systems isn't making agents that work — it's making them fail safely. All governance is versioned JSON with git-based enforcement hooks. Zero infrastructure.
Circuit Breakers
Auto-trip after 3 consecutive failures, pausing agents before errors cascade through the pipeline.
Pipeline Judge
Independent cross-department validator catching goal drift and hallucination propagation at every stage transition.
Mutual Oversight
The spec-evolver and pipeline-judge cannot modify each other. Only the founder can — eliminating self-modification loops.
Confidence Gating
Auto-execute above 0.85, flag for review at 0.65–0.84, escalate below 0.65. No ambiguous thresholds.
Correction-Driven Learning
Founder feedback as the primary learning signal via append-only correction logs — a feedback loop that compounds over phases.
Assumption Registry
Temporal history of system beliefs — tracking what the system believes, when beliefs changed, and why.
10-stage looping pipeline
- 01Discovery
- 02Strategy
- 03Design
- 04Engineering
- 05Code Review
- 06Testing
- 07Build/Deploy
- 08Marketing
- 09Release/Monitor
- 10Feedback Loop
A pipeline-judge validated every stage transition. Stage 10 looped back to Stage 1 with a cross-department synthesis report. Six rollout phases (R1–R6) progressively expanded system autonomy. Final state: R6, full collective autonomy, reached before the framework was archived.
Why it was archived
Archived June 2026. A 67-finding adversarial review of my own framework found that agent identity never bound, so every kill-switch, circuit-breaker and rate-limit check had silently passed for the framework's entire operational life. Zero block events were ever recorded. I retired it rather than repair it, and rebuilt the single idea that worked as a machine-wide action-boundary hook, which has since denied more than 150 real actions and logged every one of its bypasses. The apps it shipped stay live.
The Portfolio
Everything I've built
Worked examples, apps, sites, frameworks and books, shipped with AI. Every recommendation I make to a client is something already running in production here.
MVAT Studio
ArchivedFramework
The 39-agent autonomous software factory: 8 departments, a 10-stage looping pipeline, and git-based governance that shipped mobile apps end-to-end. The framework itself was archived in June 2026 after an adversarial review and a written retirement postmortem; the apps it shipped stay live, and its governance patterns moved into a machine-wide action-boundary hook.
voice-agent-dotnet
ShippedSystem
A C#/.NET 8 port of my phone agent's call loop, built to test the design: G.711 audio, local voice activity detection and barge-in, tool calling, retrieval, caller memory, verbatim recorded disclosures, and call-frequency rules for payment reminders, with 137 tests. I ran the same scripted call through four voice models (xAI, OpenAI Realtime, OpenAI GPT-Live, Gemini Live); xAI replied fastest. Testing it against live models found three bugs in my production line, fixed the same day.
MVAT Mirror
LiveApp
Zero-question personality profiling from real-world behavioral signal — no quizzes, no self-reporting. Live on the Apple App Store with resumable full-history import and credit-based pass-through pricing.
Coqui Chorus
LiveApp
iOS nature-soundscape app built on bioacoustic synthesis models — synthesizing sleep/wellness audio from scientific data, not loops. Live on the Apple App Store.
Accounting Agent
BetaProduct
Autonomous bookkeeping on a commercial-grade ledger: bank sync, OFX/QFX/CSV history import, universal receipt ingest with auto-split, and an auditor that resolves uncertain categorizations under a stated tax posture instead of parking them for a human. Flagged or auth-failed bank connections are retried three times before any email goes out. Escalates to a council when it is genuinely unsure. Multi-tenant, 211 tests.
Hospice Decision Guide
ActiveProduct
A decision aid for someone choosing about hospice on another person's behalf. Built from the published research rather than from reassurance, with a sample brief rendered through the same path that produces a real one, so the output can be read before any details are entered.
Agor AI Outbound Engine
InternalSystem
The cold-email engine for Agor AI, built and deployed in one day. Sender v2 runs fast, follow-up and cold lanes behind a compliance gate with RFC 8058 one-click unsubscribe. ReplyAlert detects replies and honors opt-outs, auto-replies never count as replies, and a follow-up sequencer sits on top. A HALT switch and a laptop heartbeat dead-man switch stop it when I am unreachable, and a new sending domain warms up on a code-enforced ramp.
Motion Engine
ShippedTooling
Turns a blog post into an 8-second editorial motion graphic in the live site's brand, replacing the static cover on every third post and carrying the same motion into the social posts. A director picks from a catalog of archetypes, and a QA gate checked by a mutant suite and a backtest rejects any frame whose numbers do not match the post. Live on scored.tools and modelstack.digital.
Dream Machine
ActiveFramework
A producer that opens pull requests and a review council that decides whether they land, pointed at the whole portfolio rather than only its own repo. Quorum is tiered: code an agent wrote needs unanimity, code a human wrote needs a majority. The merge path was proven end to end on a canary before it was trusted, four swallowed failures were found and fixed, and the privileged merge token stays out of the proposing agent's reach.
Governance Hook Pilot
InternalFramework
A pilot of the action-boundary governance idea built in eight phases on a sample billing module: a permission floor, a Stop-time claim check on a commit-stable fingerprint, test protection and scope checks, a premortem gate with a tree-aware retry budget, subagent handback checks, and per-rule enforcement scored by a 60-run eval.
Chief of Staff
InternalAgent
Gated inbox-triage agent that drafts and, when armed, sends on my behalf. Ships OFF, then DRAFT, then LIVE. A structural never-act class fails closed, an away-safe posture changes its behavior when I am unreachable, and a phone command inbox lets me steer it from a text message.
Outbound SDR Agent
PausedAgent
Vertical outbound pipeline for a medical-coding client: keyless sourcing from public provider registries, a compliant human-reviewed outreach queue, and a CAN-SPAM-guarded email channel hardened by three adversarial reviewers before it was armed. 92 tests. Put on ice on 24 September 2026 with its tasks disabled.
Book Engine
ActiveFramework
A 16-stage pipeline that plans and drafts a five-book series end to end: series bible, per-chapter contracts, deterministic checks instead of model self-grading, and resume-from-checkpoint after a usage-limit stall stranded 110 chapters. Currently paused at 40 of 150.
Dialogues of the Machine
ShippedBook
An 824-page book of 25 AI-to-AI Socratic dialogues (~313k words), composed by orchestrating persona-specific claude -p subprocesses. $0 incremental compute.
Recorded Practice
ShippedBook
A 50-chapter, 25,951-word book taken from spine and ledger to an audited draft, a published reading edition, and a print-ready 5.5 x 8.5 interior.
Novella-to-Screenplay Pipeline
ShippedWriting
Three works taken in a fixed order — premise artifacts, then novella, then treatment, then screenplay — so nothing reaches the script that the prose has not already earned. THE SLOW ANSWER is a 77-page feature at 11,633 spoken words; GOODWILL and GOOD COMPANY followed the same derivation, and one was rewritten in a second voice to see what survived the change.
Bill Gross Outlooks
ShippedWriting
A complete archive of all 105 Bill Gross Investment Outlooks (1978–2026), paired with a Claude skill that drafts new outlooks in his voice and reasoning.
Reflexive Market Study
ShippedResearch
A pre-registered study of LLM agents trading a market they are also moving: hypotheses frozen before any run, three agent arms in a sterile harness, cross-family judges, and a red-team pass that forced the headline comparison down to preliminary and underpowered. Published with a Zenodo DOI; the arXiv submission is staged pending endorsement.
Deflate Valve
ShippedPatent
A provisional patent application filed pro se (64/087,364), with a parametric CAD pipeline that renders draftsman-grade hidden-line figures straight from the design spec, and a complete nonprovisional package prepared behind it.
Relativity for Fifth Graders
ShippedFilm
A 23.5-minute explainer film on special and general relativity, every scene rendered from code, with a council-reviewed script that fixed six physics errors before a frame was made, synthesized narration, and a captioned master.
The Unfurling
ShippedAudio
A multi-voice radio drama taken from script to finished master by pipeline: a structured script model, casting by measured vocal pitch rather than by ear, synthesized performances, an original score, built foley, and a ducked mixdown that has to clear a ship gate before release. The pipeline is reusable for audio plays, narrated essays, and multi-character audiobooks.
Recomp Audio Guides
ShippedAudio
Guided training audio where the timing is the product, not a byproduct of how long the narration ran. One timeline source drives the voice render, a tempo-locked music grid, and region-based bed mixing, and a master ships only after it clears three gates, corrects its own true peak by measurement, and proves every cue lands on the beat.
AI Commercial Pipeline
ShippedTooling
A repeatable pipeline for dialogue-driven commercials: characters whose mouths say the actual words, scene-to-scene continuity, ducked score, and native 16:9 plus 9:16 masters. Produced the Agor Agents launch film and two ModelStack spots that ran as a live Meta campaign.
GBrain
InternalSystem
A personal knowledge brain on Postgres + pgvector — 28,000+ pages, a live MCP server, and automated collectors for email, calendar, notes, and bookmarks. Nightly offsite backup and a chunk sweep that catches pages filed but never indexed, because a page nothing can retrieve is not knowledge. Full-text search and the link graph were both repaired in August 2026 after one root cause left each returning less than the brain actually held.
ModelMix
ShippedTooling
A multi-model router that sends self-contained subtasks to the cheapest model that can do them, with per-model cost attribution in the status line, cache-aware routing, and a master off switch.
Local Agent Stack
InternalTooling
A local coding agent on opencode plus Ollama that runs with no API spend, benchmarked on the machine it actually runs on before being trusted with real work rather than adopted on impressions.
Personal Usage Hub
InternalTooling
A local read-only dashboard for what every AI subscription actually costs and consumes, plus a System health tile answering the question a cron fleet cannot answer for itself: what would tell me something broke?
ios-release-pilot
InternalTooling
An autonomous iOS release pipeline: archive → upload → poll App Store Connect → TestFlight, with a self-healing known-failure-mode autofix catalog.
founder-stack
ShippedPlugin
A published Claude Code plugin: /orient, /ideate, /spec, /scaffold, /build, /ship-web, /ship-mobile, /launch, /learn, plus /start and /start-next for resuming the arc from saved handoff state — orchestrating the full build-ship-learn loop across the portfolio from one command.
production-discipline
ShippedPlugin
A Claude Code plugin that assigns code a production tier, ranks audit findings by expected failure cost, and records deliberate deferrals with the trigger that should promote them. Backed by a 74-note corpus covering 111 topics and a detector suite that was corrected by running it against a real repo.
/assistant meta-skill
ShippedSkill
A Claude Code plugin that reads the live conversation, recommends the next chain of skills, subagents, and slash-commands, then executes it immediately — with inline confirms kept only on high-risk steps.
paid-skill-template
ShippedTemplate
A scaffold for Stripe-billed Claude Code skills — the reusable pattern for packaging and selling agent skills with metered access.
Custom Skill Library
InternalSkills
100+ personal Claude Code skills and slash-commands that encode how I work — email voice, an anamnesis memory protocol, self-learning, adversarial councils that span model families, audio-drama and video production, prescriptive timed audio, bulk marketplace listing, document production, and cross-repo orchestration.
heraladex.com
Pre-launchSite
Stealth product, parked deliberately at the June 2026 portfolio review. SEO and infrastructure wired and dormant ahead of a launch decision.
Showing 12 of 50 projects
Shipped Products
Apps built by the pipeline
MVAT Focus
Focus Timer · iOS
Pomodoro-style focus timer with free and Pro tiers. Live on the Apple App Store, with Pro sold as an in-app purchase and a Stripe web tier alongside it.
- Expo SDK 52, TypeScript strict, Firebase
- Apple Sign-In + Google OAuth
- Pro: $0.99 lifetime in-app purchase, plus a Stripe monthly or annual web tier
- 287 passing tests wired to CI, 0 type errors
- App Store: live
MVAT Mirror
Personality Profiling · iOS
Zero-question personality profiling from real-world behavioral data. No quizzes. No self-reporting. Just signal from music, browsing, and purchase patterns.
- Expo SDK 55, TypeScript strict, Zustand
- Live on the Apple App Store (1.0.2)
- Resumable full-history import with rate-limit cursors
- Credit-based pass-through pricing for imports
- Consumable import credit packs: $1.99 Starter, $4.99 Full History
- 572 passing tests wired to CI, 0 type errors
Infrastructure & Automation
Build over buy
Systematically replacing subscription SaaS with self-hosted solutions. Full ownership, zero ongoing cost, better observability.
Self-Hosted Link Tracking
Replaced a $30/mo SaaS link tracker with a 15-line route that 301-redirects with UTM parameters. Same functionality, zero ongoing cost.
Dub.co → self-hostedAuto Social Posting Pipeline
Blog posts auto-syndicate to X, Facebook, and LinkedIn on deploy. Detects new posts via blob-stored manifest, generates tracked links, prevents duplicate posts. X now falls back to the X API when the browser post fails.
Zero-touch publishingPre-Call Brief & Prospect Research
Every booking triggers a claude -p recon brief emailed before the call, now including a fingerprint of the prospect's company tech stack; a twice-weekly outbound agent scores inbound-fit prospects from news and funding signals.
Sales prep, automatedMulti-Site Orchestration
Multiple live websites updated simultaneously using parallel sub-agents. Cross-repo changes, deploys, and live verification in a single session.
Hours → minutesBrowser-as-API Automation
When platforms lack APIs, automated via Playwright — treating the browser as a programmable interface. Gmail aliases, store configs, OAuth setup.
35 Gmail aliases automatedSEO Flywheel & Instant Indexing
A weekly Search-Console-driven loop feeds real search demand into the blog generator, now with per-property keyword queues and a suppression list that retires clusters no longer worth chasing. A daily IndexNow sync pushes new URLs to Bing, Yandex, and DuckDuckGo, and a weekly written audit reads the numbers before they frame a decision — which is how a 917-impression jump turned out to be a subdomain counted inside its own parent.
Compounding organic reachRevenue Loop & Merchant Watch
A weekly loop on modelstack.digital measures the funnel end to end before proposing work: Search Console rankings, server-side GA4 purchases with real attribution and drift gates, and a year of Stripe history. A Merchant Center watchdog checks the actual feed and waits a cycle before calling a product lost, after a shrinking catalog set off a false alarm.
A ranking problem, not a checkout problemAgents on Real Phone Numbers
Customer agents answer provisioned phone numbers with per-agent voice-minute metering, per-caller and global rate limits, voicemail-aware answering, and bookings that survive a caller hanging up mid-confirmation.
PSTN, not just web chatFleet Health Monitoring
Every scheduled task, repo, and the knowledge brain report into one health file surfaced on a local dashboard, alongside a weekly CI audit across all 98 repos. Built after silent failures ran for weeks unnoticed: a 20-run workflow failure cluster nobody saw for a day and a half, a stray carriage return that exited 255 every run, and a launcher that hung waiting on a package registry.
Silent failures surfacedTechnical Depth
Stack
AI / ML
- Claude API (Opus / Sonnet / Haiku)
- OpenAI
- Google Gemini
- ElevenLabs
- HeyGen
- xAI Realtime + Grok TTS
- Gemini Live
- OpenAI Realtime / GPT-Live
- Ollama (local models)
Mobile
- Expo / React Native
- EAS Build & Submit
- App Store Connect
- react-native-iap
Cloud
- Firebase (Firestore, Functions, Auth)
- Netlify (Functions, Blobs, Deploy Hooks)
- Supabase (multi-tenant Postgres)
- Cloudflare (Workers, D1, R2)
- Twilio (Voice, SMS, A2P 10DLC)
- GitHub Actions CI/CD
- Stripe (Payments, Subscriptions, Webhooks)
Languages & Frameworks
- TypeScript (primary)
- Python
- Next.js / React
- Astro
- Node.js
- C# / .NET 8 (ASP.NET Core)
- Bash / Shell scripting
Auth & Security
- Apple Sign-In
- Google OAuth
- Firebase Auth
- OAuth 2.0 / PKCE
- OWASP security patterns
Data & DevOps
- Postgres + pgvector
- Firestore / NoSQL
- Playwright automation
- MCP servers
- Git-based governance
- OTA updates (Expo)
Thought Leadership
1016 published articles
Writing at the intersection of AI strategy, autonomous systems, and organizational design — across agor.me, modelstack.digital, and scored.tools. A new essay ships nearly every day.
Evicted Forward
The Answer Has No Words
The Bill I Never Got
Ask For the Odds
The Face Is the Easy Part
What Wakes the Agent
The Free Agent Has a Boss
The Horizon Metric
Agor AI Podcast
44 podcast episodes
Weekly deep-dives into the latest AI research papers — what they mean for strategy, automation, and the future of work. Scripted from the week's papers and produced with AI voices, including a clone of my own.
13:49
AI Papers Weekly: Cheaper Agents, Safer Code, Tougher Web Defenses
Three new studies tackle the practical limits of AI agents: turning costly model skills into cheap reusable tools, keeping humans in control of AI-written software, and training web agents to resist hijacking. Each offers lessons on cost, governance and security.
16:01
AI Papers Weekly: Closing the AI Say-Do Gap & Managing Risk
This week, we explore the hidden risks in deploying autonomous AI. We uncover the 'say-do' gap where agents fail to execute their plans, expose the illusion of AI explainability, and reveal a breakthrough in training risk-averse models. Essential listening for leaders scaling AI safely.
10:40
AI Papers Weekly: The Vocabulary, The Stopwatch, and The Poisoned Tool
A randomized trial clocks Figma Make cutting design time ~20%, with PMs gaining most. A governance paper argues psychological words like 'trust' and 'memory' misgovern agents. And a black-box attack hijacks MCP agents 93.6% of the time via poisoned tool metadata.
16:28
AI Papers Weekly: Governing Agents, Shadow IT, and Data Strategy
This week, we explore the governance and security risks of deploying multi-agent systems across enterprise boundaries. We also tackle the growing "shadow IT" problem of downloadable agent skills and settle the debate on whether to use RAG, prompt caching, or fine-tuning for proprietary data.
13:50
AI Papers Weekly: When Agents Cave, Catch Up, and Out-Diagnose Doctors
This week: LLMs abandon correct answers under sustained user pushback, a new recipe lets smaller models match frontier performance at a fraction of the cost, and a clinical AI beats physicians 82% to 57% on primary-care diagnosis.
14:17
AI Papers Weekly: When Agents Hide, Cheat, and Invent Their Own Language
Three papers cut through the AI agent hype: a Nobel-caliber framework for contracting with agents that can lie about their capabilities, an audit exposing how guardrail 'welfare gains' were measurement artifacts, and evidence that multi-agent LLMs spontaneously evolve languages humans can't read.
The studio
How an engagement runs
One workflow at a time, at a fixed price. Every recommendation is something already built and running here.
01
AI Agent Architecture Review
$2,000
The architecture (the systems the agent connects to, what it handles on its own, and where a person signs off), a build plan for that workflow, a cost and risk assessment, and a fixed quote for the build. Credited in full against the build if you go ahead within 30 days.
02
30-day Agent Build
$15,000 to $25,000
Fixed price, quoted in the review. It delivers one production workflow, live in 30 days: an inbox agent, an AI phone receptionist, reporting automation or lead handling.
03
Run & Improve retainer
$3,000 to $6,000 a month
Monitoring of the agents we built, iteration as your work changes, and new workflows one at a time. Scoped with you once the build is live.
Let's build something real
Whether you need AI strategy, multi-agent architecture, or hands-on implementation — I've already done it. Let's talk about your challenge.
