Available for contract or full-time work

AI & Document-Automation Engineer

I turn messy PDFs and documents into clean, validated, structured data — and build the AI systems and SaaS products around them.

I ship real products — then make the results provably correct: tested code does the math, not the model.

Ohio, USA · Remote Python · TypeScript · C#/.NET LinkedIn ↗ GitHub ↗
production document-automation tools shipped for a CPA firm
5
live B2B SaaS in production (OBBBA Tracker)
1
tie-out checks · 0 exceptions on a published report
25
verification gauntlet behind AgentA
7-tier

What I do

Three things, done to an audit standard

A finance background, an engineer's hands, and a hard rule: tested code proves the answer — the model never gets the last word on a number.

Document extraction & validation

PDFs into clean, structured JSON — then every total re-derived (foot, crossfoot, articulate) and locked with golden-file tests. No OCR for born-digital text.

AI systems that stay honest

Multi-agent systems, RAG, and LLM pipelines built behind verification — sealed holdouts and fresh-context critics — so improvements are earned, not hallucinated.

Full-stack SaaS

Production B2B products end to end — auth, multi-tenant data, billing, payroll-format exports, audit trails, and dashboards. Real users, real subscriptions.

Selected work

Products and systems I built and shipped

Live SaaS, document intelligence, autonomous AI, and a strength-per-byte ML benchmark. See all six, including a native C++ Pac-Man →

OBBBA Tracker dashboard showing tipped-wage and overtime deduction tracking
LiveB2B SaaS · I build & operate it

OBBBA Tracker — live tax-compliance SaaS

Helps tipped-industry employers track and document the “no tax on tips & overtime” deductions under the One Big Beautiful Bill Act: automatic Treasury Tipped Occupation Code assignment, FLSA overtime, W-2 Box 14 exports (ADP/Gusto/QuickBooks), an audit trail, multi-tenant role-based access, and an analytics dashboard. Subscriptions plus a free trial.

full-stack SaaSauthmulti-tenantpayroll exportsbilling
Visit the live product
Multi-agent system

AgentA — self-improving engineering harness

An autonomous system that improves its own code behind a 7-tier verification gauntlet — parse, unit, property, mutation, benchmark, sealed holdout, fresh-context critic — so improvements are earned, not hallucinated.

Improvements gated by a sealed holdout + fresh-context critic

Inner research loop ported from Udit Goenka’s autoresearch (MIT, after Karpathy); the meta-improvement architecture and verification stack are my own.

multi-agentverificationPython
PDF → validated structured data

Financial-Statement Extractor

Turns a government audit statement (PDF) into clean structured JSON, then re-derives every total to prove the extraction is correct, with a golden-file regression test. The model never does the arithmetic — tested code does.

25 checks, 0 exceptions — inject one wrong figure and it’s caught

Pythonpdfplumberpytestaudit-grade
Federal grant compliance SaaS

GrantLedger

Full-stack B2B SaaS that auto-categorizes nonprofit grant spending into 2 CFR 200 budget categories, tracks budget-to-actual per grant, and generates audit-ready compliance reports, with QuickBooks/Xero integration.

Audit-ready compliance reports, budget-to-actual per grant

Next.jsTypeScriptSupabaseStripeOpenAIaudit-ready
ML benchmark · search vs. learning

NeuroFour — Connect 4 strength-per-byte

Connect 4 is a solved game, so perfect play is a known, fixed target. NeuroFour ranks 19 Connect-4 agents by strength per byte under a hard 5M-FLOP-per-move budget, scored against an exact solver.

Zero leads optimality, ladder Elo, and NeuroFour Score among the 19 under-budget agents, at 0 bytes, ahead of all 14 learned-net entries (6 distinct weight files, largest ~24 KB): ladder Elo 754 vs 631 and NeuroFour Score 96.45 vs 74.25 — a score that divides strength by a size penalty — optimality 0.960 vs 0.957 by a single move in 300. Each “vs” is the strongest learned entry on that axis (optimality is a tie); the four entries behind them load just two weight files. The cost: Zero spends 4,999,028 of the board’s 5M-FLOP budget — 99.98% of it, 18th of 19 on FLOPs/move (and 4th on soundness, 15th on latency).

Free-tier API — first move may take ~30–60s to wake.

benchmark designPython / FastAPIReact 19

About

Finance brain, engineer’s hands

I’m an AI & document-automation engineer with a finance background. I build document-extraction pipelines, AI systems, and full-stack SaaS — with a focus on output that is verifiable, not just plausible.

Over the past year I designed and shipped five production document-automation tools for a public-accounting firm (Millhuff-Stang, CPA), built a live tax-compliance SaaS (OBBBA Tracker), and built a self-improving multi-agent engineering harness. Earlier I built a Retrieval-Augmented Generation (RAG) document system for a law firm and a case-management web application for a legal practice.

More about me, the full stack & credentials

PythonC#/.NETTypeScript / React FastAPIpdfplumber / PyMuPDFRAG / vector search Claude / Gemini / OpenAINext.jsDocker
Education
B.B.A., Finance — University of Cincinnati, Lindner College of Business (Cum Laude)
Certifications
CFI — Financial Modeling & Valuation Analyst (FMVA); Business Intelligence & Data Analyst (BIDA)
Based in
Ohio, USA — working remote
Open to
Contract or full-time AI / automation / document-intelligence work

Writing & research

Notes on building systems you can trust

Short, sanitized pieces on verifiable AI. Read them all →

Contact

Let’s build something verifiable

I’m available for contract or full-time AI, automation, and document-intelligence work. The fastest way to reach me is email — no forms, no backend, just say hello.