Portfolio· Updated September 1, 2026

I Don't Advise on AI. I Build It.

Four apps live on the Apple App Store, shipped in seventy-seven days. Seven live sites, a 313,000-word book, and a research podcast, all built and run solo with AI. I also built a 39-agent autonomous build pipeline and then retired it, because an adversarial review of my own work found its governance had never once fired. When I advise clients on AI strategy, I have already met the problems they are about to.

39

AI Agents Built

4

Mobile Apps

12

Sites Live

923

Articles Published

40

Podcast Episodes

1,300+

Tests Passing

Archived Flagship

MVAT Studio: Autonomous Multi-Agent App Factory

A framework in which 39 AI agents built, tested and shipped mobile apps, with no human in the loop during pipeline execution. It put real apps on the App Store. It is also archived, and the reason why is the more useful half of the story.

Architecture

Product

5

Strategy, PRDs, Personas, Prioritization, Market Research

Design

5

UX, UI, Design System, Interactions, Accessibility

Engineering

8

Architecture, Frontend, Backend, Security, DevOps, Code Review

Testing

5

Strategy, Unit Tests, Integration Tests, Quality Gate, Auto-Heal

Marketing

5

ASO, Content, Social Media, Ad Ops, Launch Coordination

Analytics

5

Metrics, Behavior, Crashes, Anomalies, Experiments

Finance

4

Revenue, Budget, Forecasting, Spend Alerts

Governance

2

Pipeline Judge, Spec Evolver (mutual oversight)

Tiered Model Assignment — Cost Optimization Without Quality Loss

7

Opus

Production code + critical gates

19

Sonnet

Content, analysis, design specs

13

Haiku

Read-only analytics + reporting

Governance Innovation

The hard problem in multi-agent systems isn't making agents that work — it's making them fail safely. All governance is versioned JSON with git-based enforcement hooks. Zero infrastructure.

Circuit Breakers

Auto-trip after 3 consecutive failures, pausing agents before errors cascade through the pipeline.

Pipeline Judge

Independent cross-department validator catching goal drift and hallucination propagation at every stage transition.

Mutual Oversight

The spec-evolver and pipeline-judge cannot modify each other. Only the founder can — eliminating self-modification loops.

Confidence Gating

Auto-execute above 0.85, flag for review at 0.65–0.84, escalate below 0.65. No ambiguous thresholds.

Correction-Driven Learning

Founder feedback as the primary learning signal via append-only correction logs — a feedback loop that compounds over phases.

Assumption Registry

Temporal history of system beliefs — tracking what the system believes, when beliefs changed, and why.

10-Stage Looping Pipeline

1Discovery
2Strategy
3Design
4Engineering
5Code Review
6Testing
7Build/Deploy
8Marketing
9Release/Monitor
10Feedback Loop
↺ 1

A pipeline-judge validated every stage transition. Stage 10 looped back to Stage 1 with a cross-department synthesis report. Six rollout phases (R1–R6) progressively expanded system autonomy. Final state: R6, full collective autonomy, reached before the framework was archived.

Why it was archived

Archived June 2026. A 67-finding adversarial review of my own framework found that agent identity never bound, so every kill-switch, circuit-breaker and rate-limit check had silently passed for the framework's entire operational life. Zero block events were ever recorded. I retired it rather than repair it, and rebuilt the single idea that worked as a machine-wide action-boundary hook, which has since denied more than 150 real actions and logged every one of its bypasses. The apps it shipped stay live.

The Portfolio

Everything I've Built

Apps, sites, frameworks, and a book — shipped solo with AI. Every recommendation I make to a client is something already running in production here.

Shipped

MVAT Studio

Framework

The 39-agent autonomous software factory: 8 departments, a 10-stage looping pipeline, and git-based governance that shipped mobile apps end-to-end. The framework itself was archived in June 2026 after an adversarial review and a written retirement postmortem; the apps it shipped stay live, and its governance patterns moved into a machine-wide action-boundary hook.

Live

Agor Agents

Product

Self-serve AI agents for businesses, live and taking real payments. Pick a pre-built agent, connect your Google Calendar, and it builds itself with no sales call. It answers from your own documents, checks real availability, and books appointments — in web chat, in a one-tag embeddable widget, and on a real phone number with provisioned Twilio routing, customer voice choice, per-agent voice-minute metering, and caller rate limits. Every published agent clears a quality gate first.

Live

MVAT Focus

App

Pomodoro-style focus timer for iOS. Live on the Apple App Store with free/Pro tiers and live Stripe subscriptions. Built by the pipeline.

Live

MVAT Mirror

App

Zero-question personality profiling from real-world behavioral signal — no quizzes, no self-reporting. Live on the Apple App Store with resumable full-history import and credit-based pass-through pricing.

Live

GifLoop

App

Native iOS 17+ universal app for creating, editing, and exporting animated GIFs. Live on the Apple App Store, shipped through a fully automated release pipeline.

Live

Coqui Chorus

App

iOS nature-soundscape app built on bioacoustic synthesis models — synthesizing sleep/wellness audio from scientific data, not loops. Live on the Apple App Store.

Live

agor.me

Site

This site. AI consulting with a dual-modal assistant (Gemini text chat + xAI realtime voice), calendar booking, and a daily-generated blog with AI-narrated audio and a word-timed reading highlight.

Live

modelstack.digital

Site

Finance, M&A, and investment-banking model templates on live Stripe checkout, wrapped in a daily blog pipeline. Every post prerendered to static HTML, products and bundles fed to Google Merchant Center behind a watchdog that checks the feed before calling anything lost, server-side GA4 purchases with real attribution, and the full catalog listed for distribution on Flevy and eFinancialModels. Download expiry and review requests handled automatically.

Live

scored.tools

Site

An AI-tool scoring and comparison directory — browse, compare, and search tools with honest verdicts. New tools are auto-discovered from Hacker News and other sources, reviewed, and merged by pipeline. Nightly bundled deploys to control cost, and a search audit that stopped 296 tool pages from competing with their own reviews.

Live

aphor.me

Site

Aphor Subconscious Studio — Ericksonian hypnotherapy guided audio sessions for sleep, anxiety relief, and personal transformation.

Live

Hansum

Site

Merch store on live Stripe checkout with shipping baked into the displayed price, per-variant Printful fulfillment behind a design allowlist, branded confirmation email, and Rex, a Gemini receptionist widget on every page.

Live

OffWeekend

Site

Availability matching for separated parents dating on the apps they already use. No accounts: two people answer a week each, the overlap reveals, and the pairing is recoverable by email. Built in one overnight session, then corrected by a cross-family council that caught three false copy claims.

Live

agortherapy.com

Site

Marketing site for a compliance-first practice-operations layer. The runtime behind it — meeting poller, session distill, pre-session brief, reminders, daily digest — is built and held in a deliberately human-gated Phase 0.

Beta

The Origination Engine

Product

Deal-origination signal engine for middle-market capital markets: a recurring brief built from filings and market signals, a prose verifier that refuses to let stale cash imply runway urgency, a 12-month lead-time backtest, and a public explorer. Sending is stopped while the engagement decision is pending.

Beta

Accounting Agent

Product

Autonomous bookkeeping on a commercial-grade ledger: bank sync, OFX/QFX/CSV history import, universal receipt ingest with auto-split, and an auditor that resolves uncertain categorizations under a stated tax posture instead of parking them for a human. Escalates to a council when it is genuinely unsure. Multi-tenant, 211 tests.

Live

MLR Guard

Product

Claims-grounded generation for regulated promotional copy: the model may only assemble text from a pre-approved claims library, never author or reword a claim. Vue and Hono on Cloudflare Workers with D1 and an optional R2 snapshot store, 28 tests, deployed live and kept public deliberately as a readable work sample.

Active

Hospice Decision Guide

Product

A decision aid for someone choosing about hospice on another person's behalf. Built from the published research rather than from reassurance, with a sample brief rendered through the same path that produces a real one, so the output can be read before any details are entered.

Active

Dream Machine

Framework

A producer that opens pull requests and a review council that decides whether they land, pointed at the whole portfolio rather than only its own repo. Quorum is tiered: code an agent wrote needs unanimity, code a human wrote needs a majority. The merge path was proven end to end on a canary before it was trusted, four swallowed failures were found and fixed, and the privileged merge token stays out of the proposing agent's reach.

Internal

Chief of Staff

Agent

Gated inbox-triage agent that drafts and, when armed, sends on my behalf. Ships OFF, then DRAFT, then LIVE. A structural never-act class fails closed, an away-safe posture changes its behavior when I am unreachable, and a phone command inbox lets me steer it from a text message.

Internal

Outbound SDR Agent

Agent

Vertical outbound pipeline for a medical-coding client: keyless sourcing from public provider registries, a compliant human-reviewed outreach queue, and a CAN-SPAM-guarded email channel hardened by three adversarial reviewers before it was armed. 92 tests.

Active

Book Engine

Framework

A 16-stage pipeline that plans and drafts a five-book series end to end: series bible, per-chapter contracts, deterministic checks instead of model self-grading, and resume-from-checkpoint after a usage-limit stall stranded 110 chapters. Currently paused at 40 of 150.

Shipped

Dialogues of the Machine

Book

An 824-page book of 25 AI-to-AI Socratic dialogues (~313k words), composed by orchestrating persona-specific claude -p subprocesses. $0 incremental compute.

Shipped

Recorded Practice

Book

A 50-chapter, 25,951-word book taken from spine and ledger to an audited draft, a published reading edition, and a print-ready 5.5 x 8.5 interior.

Shipped

Novella-to-Screenplay Pipeline

Writing

Three works taken in a fixed order — premise artifacts, then novella, then treatment, then screenplay — so nothing reaches the script that the prose has not already earned. THE SLOW ANSWER is a 77-page feature at 11,633 spoken words; GOODWILL and GOOD COMPANY followed the same derivation, and one was rewritten in a second voice to see what survived the change.

Shipped

Bill Gross Outlooks

Writing

A complete archive of all 105 Bill Gross Investment Outlooks (1978–2026), paired with a Claude skill that drafts new outlooks in his voice and reasoning.

Shipped

Reflexive Market Study

Research

A pre-registered study of LLM agents trading a market they are also moving: hypotheses frozen before any run, three agent arms in a sterile harness, cross-family judges, and a red-team pass that forced the headline comparison down to preliminary and underpowered. Published with a Zenodo DOI; the arXiv submission is staged pending endorsement.

Shipped

Deflate Valve

Patent

A provisional patent application filed pro se (64/087,364), with a parametric CAD pipeline that renders draftsman-grade hidden-line figures straight from the design spec, and a complete nonprovisional package prepared behind it.

Shipped

Relativity for Fifth Graders

Film

A 23.5-minute explainer film on special and general relativity, every scene rendered from code, with a council-reviewed script that fixed six physics errors before a frame was made, synthesized narration, and a captioned master.

Shipped

The Unfurling

Audio

A multi-voice radio drama taken from script to finished master by pipeline: a structured script model, casting by measured vocal pitch rather than by ear, synthesized performances, an original score, built foley, and a ducked mixdown that has to clear a ship gate before release. The pipeline is reusable for audio plays, narrated essays, and multi-character audiobooks.

Shipped

Recomp Audio Guides

Audio

Guided training audio where the timing is the product, not a byproduct of how long the narration ran. One timeline source drives the voice render, a tempo-locked music grid, and region-based bed mixing, and a master ships only after it clears three gates, corrects its own true peak by measurement, and proves every cue lands on the beat.

Shipped

AI Commercial Pipeline

Tooling

A repeatable pipeline for dialogue-driven commercials: characters whose mouths say the actual words, scene-to-scene continuity, ducked score, and native 16:9 plus 9:16 masters. Produced the Agor Agents launch film and two ModelStack spots that ran as a live Meta campaign.

Internal

GBrain

System

A personal knowledge brain on Postgres + pgvector — 28,000+ pages, a live MCP server, and automated collectors for email, calendar, notes, and bookmarks. Nightly offsite backup and a chunk sweep that catches pages filed but never indexed, because a page nothing can retrieve is not knowledge. Full-text search and the link graph were both repaired in August 2026 after one root cause left each returning less than the brain actually held.

Shipped

ModelMix

Tooling

A multi-model router that sends self-contained subtasks to the cheapest model that can do them, with per-model cost attribution in the status line, cache-aware routing, and a master off switch.

Internal

Local Agent Stack

Tooling

A local coding agent on opencode plus Ollama that runs with no API spend, benchmarked on the machine it actually runs on before being trusted with real work rather than adopted on impressions.

Internal

Personal Usage Hub

Tooling

A local read-only dashboard for what every AI subscription actually costs and consumes, plus a System health tile answering the question a cron fleet cannot answer for itself: what would tell me something broke?

Internal

ios-release-pilot

Tooling

An autonomous iOS release pipeline: archive → upload → poll App Store Connect → TestFlight, with a self-healing known-failure-mode autofix catalog.

Shipped

founder-stack

Plugin

A published Claude Code plugin: /orient, /ideate, /spec, /scaffold, /build, /ship-web, /ship-mobile, /launch, /learn, plus /start and /start-next for resuming the arc from saved handoff state — orchestrating the full build-ship-learn loop across the portfolio from one command.

Shipped

production-discipline

Plugin

A Claude Code plugin that assigns code a production tier, ranks audit findings by expected failure cost, and records deliberate deferrals with the trigger that should promote them. Backed by a 74-note corpus covering 111 topics and a detector suite that was corrected by running it against a real repo.

Shipped

/assistant meta-skill

Skill

A Claude Code plugin that reads the live conversation, recommends the next chain of skills, subagents, and slash-commands, then executes it immediately — with inline confirms kept only on high-risk steps.

Shipped

paid-skill-template

Template

A scaffold for Stripe-billed Claude Code skills — the reusable pattern for packaging and selling agent skills with metered access.

Internal

Custom Skill Library

Skills

100+ personal Claude Code skills and slash-commands that encode how I work — email voice, an anamnesis memory protocol, self-learning, adversarial councils that span model families, audio-drama and video production, prescriptive timed audio, bulk marketplace listing, document production, and cross-repo orchestration.

Pre-launch

heraladex.com

Site

Stealth product, parked deliberately at the June 2026 portfolio review. SEO and infrastructure wired and dormant ahead of a launch decision.

Shipped Products

Apps Built by the Pipeline

MVAT Focus

Focus Timer · iOS

Pomodoro-style focus timer with free/Pro tiers. Live on the Apple App Store. TestFlight live with 15 testers. Stripe subscriptions processing live payments.

Expo SDK 52, TypeScript strict, Firebase
Apple Sign-In + Google OAuth
Stripe: $4.99/mo, $39.99/yr (live)
287 passing tests wired to CI, 0 type errors
App Store: live

MVAT Mirror

Personality Profiling · iOS

Zero-question personality profiling from real-world behavioral data. No quizzes. No self-reporting. Just signal from music, browsing, and purchase patterns.

Expo SDK 55, TypeScript strict, Zustand
Live on the Apple App Store (1.0.2)
Resumable full-history import with rate-limit cursors
Credit-based pass-through pricing for imports
Free tier + $9.99/mo + $49.99 lifetime
572 passing tests wired to CI, 0 type errors
Infrastructure & Automation

Build Over Buy

Systematically replacing subscription SaaS with self-hosted solutions. Full ownership, zero ongoing cost, better observability.

Self-Hosted Link Tracking

Replaced a $30/mo SaaS link tracker with a 15-line route that 301-redirects with UTM parameters. Same functionality, zero ongoing cost.

Dub.co → self-hosted

Auto Social Posting Pipeline

Blog posts auto-syndicate to X, Facebook, and LinkedIn on deploy. Detects new posts via blob-stored manifest, generates tracked links, prevents duplicate posts.

Zero-touch publishing

Pre-Call Brief & Prospect Research

Every booking triggers a claude -p recon brief emailed before the call; a twice-weekly outbound agent scores inbound-fit prospects from news and funding signals.

Sales prep, automated

Multi-Site Orchestration

Multiple live websites updated simultaneously using parallel sub-agents. Cross-repo changes, deploys, and live verification in a single session.

Hours → minutes

Browser-as-API Automation

When platforms lack APIs, automated via Playwright — treating the browser as a programmable interface. Gmail aliases, store configs, OAuth setup.

35 Gmail aliases automated

SEO Flywheel & Instant Indexing

A weekly Search-Console-driven loop feeds real search demand into the blog generator, now with per-property keyword queues and a suppression list that retires clusters no longer worth chasing. A daily IndexNow sync pushes new URLs to Bing, Yandex, and DuckDuckGo, and a weekly written audit reads the numbers before they frame a decision — which is how a 917-impression jump turned out to be a subdomain counted inside its own parent.

Compounding organic reach

Revenue Loop & Merchant Watch

A weekly loop on modelstack.digital measures the funnel end to end before proposing work: Search Console rankings, server-side GA4 purchases with real attribution and drift gates, and a year of Stripe history. A Merchant Center watchdog checks the actual feed and waits a cycle before calling a product lost, after a shrinking catalog set off a false alarm.

A ranking problem, not a checkout problem

Agents on Real Phone Numbers

Customer agents answer provisioned phone numbers with per-agent voice-minute metering, per-caller and global rate limits, voicemail-aware answering, and bookings that survive a caller hanging up mid-confirmation.

PSTN, not just web chat

Fleet Health Monitoring

Every scheduled task, repo, and the knowledge brain report into one health file surfaced on a local dashboard, alongside a weekly CI audit across all 98 repos. Built after silent failures ran for weeks unnoticed: a 20-run workflow failure cluster nobody saw for a day and a half, a stray carriage return that exited 255 every run, and a launcher that hung waiting on a package registry.

Silent failures surfaced
Technical Depth

Stack

AI / ML

  • Claude API (Opus / Sonnet / Haiku)
  • OpenAI
  • Google Gemini
  • ElevenLabs
  • HeyGen
  • xAI Realtime + Grok TTS
  • Ollama (local models)

Mobile

  • Expo / React Native
  • EAS Build & Submit
  • App Store Connect
  • react-native-iap

Cloud

  • Firebase (Firestore, Functions, Auth)
  • Netlify (Functions, Blobs, Deploy Hooks)
  • Supabase (multi-tenant Postgres)
  • Cloudflare (Workers, D1, R2)
  • Twilio (Voice, SMS, A2P 10DLC)
  • GitHub Actions CI/CD
  • Stripe (Payments, Subscriptions, Webhooks)

Languages & Frameworks

  • TypeScript (primary)
  • Python
  • Next.js / React
  • Astro
  • Node.js
  • Bash / Shell scripting

Auth & Security

  • Apple Sign-In
  • Google OAuth
  • Firebase Auth
  • OAuth 2.0 / PKCE
  • OWASP security patterns

Data & DevOps

  • Postgres + pgvector
  • Firestore / NoSQL
  • Playwright automation
  • MCP servers
  • Git-based governance
  • OTA updates (Expo)
Thought Leadership

923 Published Articles

Writing at the intersection of AI strategy, autonomous systems, and organizational design — across agor.me, modelstack.digital, and scored.tools. A new essay ships nearly every day.

Read all articles
Agor AI Podcast

40 Podcast Episodes

Weekly deep-dives into the latest AI research papers — what they mean for strategy, automation, and the future of work. Generated end-to-end with Google NotebookLM.

13:50

AI Papers Weekly: When Agents Cave, Catch Up, and Out-Diagnose Doctors

This week: LLMs abandon correct answers under sustained user pushback, a new recipe lets smaller models match frontier performance at a fraction of the cost, and a clinical AI beats physicians 82% to 57% on primary-care diagnosis.

14:17

AI Papers Weekly: When Agents Hide, Cheat, and Invent Their Own Language

Three papers cut through the AI agent hype: a Nobel-caliber framework for contracting with agents that can lie about their capabilities, an audit exposing how guardrail 'welfare gains' were measurement artifacts, and evidence that multi-agent LLMs spontaneously evolve languages humans can't read.

45:00

Could You Stop It Completely?

A 45-minute guided self-hypnosis in Ariel's own voice. It settles the body, gladdens the mind into absorption, then turns that stillness on the architecture of self with three tiers of honest inquiry, ending on a single question: could you stop it completely? Sit down for it. Not while driving.

14:03

AI Papers Weekly: When Retrieval Isn't Reading, and the Invisible Layer Between Weights and Words

Three papers dismantle comfortable assumptions about deploying LLMs: retrieved information often fails to shape judgments, society's AI vocabulary is still being fought over, and an undisclosed 'editorial layer' can steer outputs after training ends.

15:14

AI Papers Weekly: When Agents Get Fragile, Delegated, and Graded

This week: self-improving agents crack under task-order noise, dating-app users want to send AI agents but not receive them, and small cheap models grade as reliably as frontier ones when handed a rubric. Three papers, three levers for anyone deploying LLMs at scale.

14:10

AI Papers Weekly: Why Your Agent's Memory Is a Liability

This week: why CLAUDE.md files grow forever and what to do about it, how to prove an AI's probability estimates are internally honest, and a case study of humans and AI jointly cracking a famous math constant.

Advisory Services

What I Bring to Engagements

Every recommendation is something I've already built. No slide decks without substance. No theory without production evidence.

1

AI Strategy & Roadmap

Assess where AI creates real value in your business, not where it's hype. Practical roadmaps with measurable milestones.

2

Multi-Agent System Design

Architecture, governance, and rollout planning for autonomous AI systems. How to make agents that fail safely and improve over time.

3

AI-Native Product Development

From concept through app store submission. Full-stack technical advisory for teams building AI-first products.

4

Automation Audit

Identify manual processes that can be automated with AI. Prioritized by ROI, implemented with your existing stack.

5

AI Governance & Safety

Circuit breakers, confidence gating, oversight patterns, and audit trails. Making autonomous systems that executives can trust.

Let's Build Something Real

Whether you need AI strategy, multi-agent architecture, or hands-on implementation — I've already done it. Let's talk about your challenge.