I Don't Advise on AI. I Build It.
Four apps live on the Apple App Store, shipped in seventy-seven days. Seven live sites, a 313,000-word book, and a research podcast, all built and run solo with AI. I also built a 39-agent autonomous build pipeline and then retired it, because an adversarial review of my own work found its governance had never once fired. When I advise clients on AI strategy, I have already met the problems they are about to.
39
AI Agents Built
4
Mobile Apps
12
Sites Live
923
Articles Published
40
Podcast Episodes
1,300+
Tests Passing
MVAT Studio: Autonomous Multi-Agent App Factory
A framework in which 39 AI agents built, tested and shipped mobile apps, with no human in the loop during pipeline execution. It put real apps on the App Store. It is also archived, and the reason why is the more useful half of the story.
Architecture
Product
5Strategy, PRDs, Personas, Prioritization, Market Research
Design
5UX, UI, Design System, Interactions, Accessibility
Engineering
8Architecture, Frontend, Backend, Security, DevOps, Code Review
Testing
5Strategy, Unit Tests, Integration Tests, Quality Gate, Auto-Heal
Marketing
5ASO, Content, Social Media, Ad Ops, Launch Coordination
Analytics
5Metrics, Behavior, Crashes, Anomalies, Experiments
Finance
4Revenue, Budget, Forecasting, Spend Alerts
Governance
2Pipeline Judge, Spec Evolver (mutual oversight)
Tiered Model Assignment — Cost Optimization Without Quality Loss
Opus
Production code + critical gates
Sonnet
Content, analysis, design specs
Haiku
Read-only analytics + reporting
Governance Innovation
The hard problem in multi-agent systems isn't making agents that work — it's making them fail safely. All governance is versioned JSON with git-based enforcement hooks. Zero infrastructure.
Circuit Breakers
Auto-trip after 3 consecutive failures, pausing agents before errors cascade through the pipeline.
Pipeline Judge
Independent cross-department validator catching goal drift and hallucination propagation at every stage transition.
Mutual Oversight
The spec-evolver and pipeline-judge cannot modify each other. Only the founder can — eliminating self-modification loops.
Confidence Gating
Auto-execute above 0.85, flag for review at 0.65–0.84, escalate below 0.65. No ambiguous thresholds.
Correction-Driven Learning
Founder feedback as the primary learning signal via append-only correction logs — a feedback loop that compounds over phases.
Assumption Registry
Temporal history of system beliefs — tracking what the system believes, when beliefs changed, and why.
10-Stage Looping Pipeline
A pipeline-judge validated every stage transition. Stage 10 looped back to Stage 1 with a cross-department synthesis report. Six rollout phases (R1–R6) progressively expanded system autonomy. Final state: R6, full collective autonomy, reached before the framework was archived.
Why it was archived
Archived June 2026. A 67-finding adversarial review of my own framework found that agent identity never bound, so every kill-switch, circuit-breaker and rate-limit check had silently passed for the framework's entire operational life. Zero block events were ever recorded. I retired it rather than repair it, and rebuilt the single idea that worked as a machine-wide action-boundary hook, which has since denied more than 150 real actions and logged every one of its bypasses. The apps it shipped stay live.
Everything I've Built
Apps, sites, frameworks, and a book — shipped solo with AI. Every recommendation I make to a client is something already running in production here.
MVAT Studio
Framework
The 39-agent autonomous software factory: 8 departments, a 10-stage looping pipeline, and git-based governance that shipped mobile apps end-to-end. The framework itself was archived in June 2026 after an adversarial review and a written retirement postmortem; the apps it shipped stay live, and its governance patterns moved into a machine-wide action-boundary hook.
Agor Agents
Product
Self-serve AI agents for businesses, live and taking real payments. Pick a pre-built agent, connect your Google Calendar, and it builds itself with no sales call. It answers from your own documents, checks real availability, and books appointments — in web chat, in a one-tag embeddable widget, and on a real phone number with provisioned Twilio routing, customer voice choice, per-agent voice-minute metering, and caller rate limits. Every published agent clears a quality gate first.
MVAT Focus
App
Pomodoro-style focus timer for iOS. Live on the Apple App Store with free/Pro tiers and live Stripe subscriptions. Built by the pipeline.
MVAT Mirror
App
Zero-question personality profiling from real-world behavioral signal — no quizzes, no self-reporting. Live on the Apple App Store with resumable full-history import and credit-based pass-through pricing.
GifLoop
App
Native iOS 17+ universal app for creating, editing, and exporting animated GIFs. Live on the Apple App Store, shipped through a fully automated release pipeline.
Coqui Chorus
App
iOS nature-soundscape app built on bioacoustic synthesis models — synthesizing sleep/wellness audio from scientific data, not loops. Live on the Apple App Store.
agor.me
Site
This site. AI consulting with a dual-modal assistant (Gemini text chat + xAI realtime voice), calendar booking, and a daily-generated blog with AI-narrated audio and a word-timed reading highlight.
modelstack.digital
Site
Finance, M&A, and investment-banking model templates on live Stripe checkout, wrapped in a daily blog pipeline. Every post prerendered to static HTML, products and bundles fed to Google Merchant Center behind a watchdog that checks the feed before calling anything lost, server-side GA4 purchases with real attribution, and the full catalog listed for distribution on Flevy and eFinancialModels. Download expiry and review requests handled automatically.
scored.tools
Site
An AI-tool scoring and comparison directory — browse, compare, and search tools with honest verdicts. New tools are auto-discovered from Hacker News and other sources, reviewed, and merged by pipeline. Nightly bundled deploys to control cost, and a search audit that stopped 296 tool pages from competing with their own reviews.
aphor.me
Site
Aphor Subconscious Studio — Ericksonian hypnotherapy guided audio sessions for sleep, anxiety relief, and personal transformation.
Hansum
Site
Merch store on live Stripe checkout with shipping baked into the displayed price, per-variant Printful fulfillment behind a design allowlist, branded confirmation email, and Rex, a Gemini receptionist widget on every page.
OffWeekend
Site
Availability matching for separated parents dating on the apps they already use. No accounts: two people answer a week each, the overlap reveals, and the pairing is recoverable by email. Built in one overnight session, then corrected by a cross-family council that caught three false copy claims.
agortherapy.com
Site
Marketing site for a compliance-first practice-operations layer. The runtime behind it — meeting poller, session distill, pre-session brief, reminders, daily digest — is built and held in a deliberately human-gated Phase 0.
The Origination Engine
Product
Deal-origination signal engine for middle-market capital markets: a recurring brief built from filings and market signals, a prose verifier that refuses to let stale cash imply runway urgency, a 12-month lead-time backtest, and a public explorer. Sending is stopped while the engagement decision is pending.
Accounting Agent
Product
Autonomous bookkeeping on a commercial-grade ledger: bank sync, OFX/QFX/CSV history import, universal receipt ingest with auto-split, and an auditor that resolves uncertain categorizations under a stated tax posture instead of parking them for a human. Escalates to a council when it is genuinely unsure. Multi-tenant, 211 tests.
MLR Guard
Product
Claims-grounded generation for regulated promotional copy: the model may only assemble text from a pre-approved claims library, never author or reword a claim. Vue and Hono on Cloudflare Workers with D1 and an optional R2 snapshot store, 28 tests, deployed live and kept public deliberately as a readable work sample.
Hospice Decision Guide
Product
A decision aid for someone choosing about hospice on another person's behalf. Built from the published research rather than from reassurance, with a sample brief rendered through the same path that produces a real one, so the output can be read before any details are entered.
Dream Machine
Framework
A producer that opens pull requests and a review council that decides whether they land, pointed at the whole portfolio rather than only its own repo. Quorum is tiered: code an agent wrote needs unanimity, code a human wrote needs a majority. The merge path was proven end to end on a canary before it was trusted, four swallowed failures were found and fixed, and the privileged merge token stays out of the proposing agent's reach.
Chief of Staff
Agent
Gated inbox-triage agent that drafts and, when armed, sends on my behalf. Ships OFF, then DRAFT, then LIVE. A structural never-act class fails closed, an away-safe posture changes its behavior when I am unreachable, and a phone command inbox lets me steer it from a text message.
Outbound SDR Agent
Agent
Vertical outbound pipeline for a medical-coding client: keyless sourcing from public provider registries, a compliant human-reviewed outreach queue, and a CAN-SPAM-guarded email channel hardened by three adversarial reviewers before it was armed. 92 tests.
Book Engine
Framework
A 16-stage pipeline that plans and drafts a five-book series end to end: series bible, per-chapter contracts, deterministic checks instead of model self-grading, and resume-from-checkpoint after a usage-limit stall stranded 110 chapters. Currently paused at 40 of 150.
Dialogues of the Machine
Book
An 824-page book of 25 AI-to-AI Socratic dialogues (~313k words), composed by orchestrating persona-specific claude -p subprocesses. $0 incremental compute.
Recorded Practice
Book
A 50-chapter, 25,951-word book taken from spine and ledger to an audited draft, a published reading edition, and a print-ready 5.5 x 8.5 interior.
Novella-to-Screenplay Pipeline
Writing
Three works taken in a fixed order — premise artifacts, then novella, then treatment, then screenplay — so nothing reaches the script that the prose has not already earned. THE SLOW ANSWER is a 77-page feature at 11,633 spoken words; GOODWILL and GOOD COMPANY followed the same derivation, and one was rewritten in a second voice to see what survived the change.
Bill Gross Outlooks
Writing
A complete archive of all 105 Bill Gross Investment Outlooks (1978–2026), paired with a Claude skill that drafts new outlooks in his voice and reasoning.
Reflexive Market Study
Research
A pre-registered study of LLM agents trading a market they are also moving: hypotheses frozen before any run, three agent arms in a sterile harness, cross-family judges, and a red-team pass that forced the headline comparison down to preliminary and underpowered. Published with a Zenodo DOI; the arXiv submission is staged pending endorsement.
Deflate Valve
Patent
A provisional patent application filed pro se (64/087,364), with a parametric CAD pipeline that renders draftsman-grade hidden-line figures straight from the design spec, and a complete nonprovisional package prepared behind it.
Relativity for Fifth Graders
Film
A 23.5-minute explainer film on special and general relativity, every scene rendered from code, with a council-reviewed script that fixed six physics errors before a frame was made, synthesized narration, and a captioned master.
The Unfurling
Audio
A multi-voice radio drama taken from script to finished master by pipeline: a structured script model, casting by measured vocal pitch rather than by ear, synthesized performances, an original score, built foley, and a ducked mixdown that has to clear a ship gate before release. The pipeline is reusable for audio plays, narrated essays, and multi-character audiobooks.
Recomp Audio Guides
Audio
Guided training audio where the timing is the product, not a byproduct of how long the narration ran. One timeline source drives the voice render, a tempo-locked music grid, and region-based bed mixing, and a master ships only after it clears three gates, corrects its own true peak by measurement, and proves every cue lands on the beat.
AI Commercial Pipeline
Tooling
A repeatable pipeline for dialogue-driven commercials: characters whose mouths say the actual words, scene-to-scene continuity, ducked score, and native 16:9 plus 9:16 masters. Produced the Agor Agents launch film and two ModelStack spots that ran as a live Meta campaign.
GBrain
System
A personal knowledge brain on Postgres + pgvector — 28,000+ pages, a live MCP server, and automated collectors for email, calendar, notes, and bookmarks. Nightly offsite backup and a chunk sweep that catches pages filed but never indexed, because a page nothing can retrieve is not knowledge. Full-text search and the link graph were both repaired in August 2026 after one root cause left each returning less than the brain actually held.
ModelMix
Tooling
A multi-model router that sends self-contained subtasks to the cheapest model that can do them, with per-model cost attribution in the status line, cache-aware routing, and a master off switch.
Local Agent Stack
Tooling
A local coding agent on opencode plus Ollama that runs with no API spend, benchmarked on the machine it actually runs on before being trusted with real work rather than adopted on impressions.
Personal Usage Hub
Tooling
A local read-only dashboard for what every AI subscription actually costs and consumes, plus a System health tile answering the question a cron fleet cannot answer for itself: what would tell me something broke?
ios-release-pilot
Tooling
An autonomous iOS release pipeline: archive → upload → poll App Store Connect → TestFlight, with a self-healing known-failure-mode autofix catalog.
founder-stack
Plugin
A published Claude Code plugin: /orient, /ideate, /spec, /scaffold, /build, /ship-web, /ship-mobile, /launch, /learn, plus /start and /start-next for resuming the arc from saved handoff state — orchestrating the full build-ship-learn loop across the portfolio from one command.
production-discipline
Plugin
A Claude Code plugin that assigns code a production tier, ranks audit findings by expected failure cost, and records deliberate deferrals with the trigger that should promote them. Backed by a 74-note corpus covering 111 topics and a detector suite that was corrected by running it against a real repo.
/assistant meta-skill
Skill
A Claude Code plugin that reads the live conversation, recommends the next chain of skills, subagents, and slash-commands, then executes it immediately — with inline confirms kept only on high-risk steps.
paid-skill-template
Template
A scaffold for Stripe-billed Claude Code skills — the reusable pattern for packaging and selling agent skills with metered access.
Custom Skill Library
Skills
100+ personal Claude Code skills and slash-commands that encode how I work — email voice, an anamnesis memory protocol, self-learning, adversarial councils that span model families, audio-drama and video production, prescriptive timed audio, bulk marketplace listing, document production, and cross-repo orchestration.
heraladex.com
Site
Stealth product, parked deliberately at the June 2026 portfolio review. SEO and infrastructure wired and dormant ahead of a launch decision.
Apps Built by the Pipeline
MVAT Focus
Focus Timer · iOS
Pomodoro-style focus timer with free/Pro tiers. Live on the Apple App Store. TestFlight live with 15 testers. Stripe subscriptions processing live payments.
MVAT Mirror
Personality Profiling · iOS
Zero-question personality profiling from real-world behavioral data. No quizzes. No self-reporting. Just signal from music, browsing, and purchase patterns.
Build Over Buy
Systematically replacing subscription SaaS with self-hosted solutions. Full ownership, zero ongoing cost, better observability.
Self-Hosted Link Tracking
Replaced a $30/mo SaaS link tracker with a 15-line route that 301-redirects with UTM parameters. Same functionality, zero ongoing cost.
Auto Social Posting Pipeline
Blog posts auto-syndicate to X, Facebook, and LinkedIn on deploy. Detects new posts via blob-stored manifest, generates tracked links, prevents duplicate posts.
Pre-Call Brief & Prospect Research
Every booking triggers a claude -p recon brief emailed before the call; a twice-weekly outbound agent scores inbound-fit prospects from news and funding signals.
Multi-Site Orchestration
Multiple live websites updated simultaneously using parallel sub-agents. Cross-repo changes, deploys, and live verification in a single session.
Browser-as-API Automation
When platforms lack APIs, automated via Playwright — treating the browser as a programmable interface. Gmail aliases, store configs, OAuth setup.
SEO Flywheel & Instant Indexing
A weekly Search-Console-driven loop feeds real search demand into the blog generator, now with per-property keyword queues and a suppression list that retires clusters no longer worth chasing. A daily IndexNow sync pushes new URLs to Bing, Yandex, and DuckDuckGo, and a weekly written audit reads the numbers before they frame a decision — which is how a 917-impression jump turned out to be a subdomain counted inside its own parent.
Revenue Loop & Merchant Watch
A weekly loop on modelstack.digital measures the funnel end to end before proposing work: Search Console rankings, server-side GA4 purchases with real attribution and drift gates, and a year of Stripe history. A Merchant Center watchdog checks the actual feed and waits a cycle before calling a product lost, after a shrinking catalog set off a false alarm.
Agents on Real Phone Numbers
Customer agents answer provisioned phone numbers with per-agent voice-minute metering, per-caller and global rate limits, voicemail-aware answering, and bookings that survive a caller hanging up mid-confirmation.
Fleet Health Monitoring
Every scheduled task, repo, and the knowledge brain report into one health file surfaced on a local dashboard, alongside a weekly CI audit across all 98 repos. Built after silent failures ran for weeks unnoticed: a 20-run workflow failure cluster nobody saw for a day and a half, a stray carriage return that exited 255 every run, and a launcher that hung waiting on a package registry.
Stack
AI / ML
- Claude API (Opus / Sonnet / Haiku)
- OpenAI
- Google Gemini
- ElevenLabs
- HeyGen
- xAI Realtime + Grok TTS
- Ollama (local models)
Mobile
- Expo / React Native
- EAS Build & Submit
- App Store Connect
- react-native-iap
Cloud
- Firebase (Firestore, Functions, Auth)
- Netlify (Functions, Blobs, Deploy Hooks)
- Supabase (multi-tenant Postgres)
- Cloudflare (Workers, D1, R2)
- Twilio (Voice, SMS, A2P 10DLC)
- GitHub Actions CI/CD
- Stripe (Payments, Subscriptions, Webhooks)
Languages & Frameworks
- TypeScript (primary)
- Python
- Next.js / React
- Astro
- Node.js
- Bash / Shell scripting
Auth & Security
- Apple Sign-In
- Google OAuth
- Firebase Auth
- OAuth 2.0 / PKCE
- OWASP security patterns
Data & DevOps
- Postgres + pgvector
- Firestore / NoSQL
- Playwright automation
- MCP servers
- Git-based governance
- OTA updates (Expo)
923 Published Articles
Writing at the intersection of AI strategy, autonomous systems, and organizational design — across agor.me, modelstack.digital, and scored.tools. A new essay ships nearly every day.
The Wrong Line Went Up
The Rung You Can Reach
The Ledger Was Wrong
Approval Was A Placebo
The Agent Runs As You
The Denial Reads Back
The Fallback Never Passed
The Wait Was The Business
40 Podcast Episodes
Weekly deep-dives into the latest AI research papers — what they mean for strategy, automation, and the future of work. Generated end-to-end with Google NotebookLM.
AI Papers Weekly: When Agents Cave, Catch Up, and Out-Diagnose Doctors
This week: LLMs abandon correct answers under sustained user pushback, a new recipe lets smaller models match frontier performance at a fraction of the cost, and a clinical AI beats physicians 82% to 57% on primary-care diagnosis.
AI Papers Weekly: When Agents Hide, Cheat, and Invent Their Own Language
Three papers cut through the AI agent hype: a Nobel-caliber framework for contracting with agents that can lie about their capabilities, an audit exposing how guardrail 'welfare gains' were measurement artifacts, and evidence that multi-agent LLMs spontaneously evolve languages humans can't read.
Could You Stop It Completely?
A 45-minute guided self-hypnosis in Ariel's own voice. It settles the body, gladdens the mind into absorption, then turns that stillness on the architecture of self with three tiers of honest inquiry, ending on a single question: could you stop it completely? Sit down for it. Not while driving.
AI Papers Weekly: When Retrieval Isn't Reading, and the Invisible Layer Between Weights and Words
Three papers dismantle comfortable assumptions about deploying LLMs: retrieved information often fails to shape judgments, society's AI vocabulary is still being fought over, and an undisclosed 'editorial layer' can steer outputs after training ends.
AI Papers Weekly: When Agents Get Fragile, Delegated, and Graded
This week: self-improving agents crack under task-order noise, dating-app users want to send AI agents but not receive them, and small cheap models grade as reliably as frontier ones when handed a rubric. Three papers, three levers for anyone deploying LLMs at scale.
AI Papers Weekly: Why Your Agent's Memory Is a Liability
This week: why CLAUDE.md files grow forever and what to do about it, how to prove an AI's probability estimates are internally honest, and a case study of humans and AI jointly cracking a famous math constant.
What I Bring to Engagements
Every recommendation is something I've already built. No slide decks without substance. No theory without production evidence.
AI Strategy & Roadmap
Assess where AI creates real value in your business, not where it's hype. Practical roadmaps with measurable milestones.
Multi-Agent System Design
Architecture, governance, and rollout planning for autonomous AI systems. How to make agents that fail safely and improve over time.
AI-Native Product Development
From concept through app store submission. Full-stack technical advisory for teams building AI-first products.
Automation Audit
Identify manual processes that can be automated with AI. Prioritized by ROI, implemented with your existing stack.
AI Governance & Safety
Circuit breakers, confidence gating, oversight patterns, and audit trails. Making autonomous systems that executives can trust.
Let's Build Something Real
Whether you need AI strategy, multi-agent architecture, or hands-on implementation — I've already done it. Let's talk about your challenge.