Projects & experiments
Things I’m building.
Case studies
AI Reliability Layer
A shared layer for monitoring, diagnosis, and automated recovery across production agents and automations.
Immigrant Invest · Production reliabilityEngineering Coordination Layer
A multi-agent system that helps me maintain a high standard of project coordination by connecting discussions, decisions, and project updates to clear tasks, owners, and next steps.
Immigrant Invest · Engineering operationsCompany Brain
Shared company context from 17 internal systems, available to employee agents and internal products through permission-aware access.
Immigrant Invest · Context infrastructurePII Sanitization Pipeline
A shared pipeline for preparing confidential client and partner data for AI workflows, with field-level policies and local processing before downstream agents receive it.
Immigrant Invest · Data privacy
Show 3 more case studiesShow fewer case studies
Presales Automation
A multi-agent presales system that qualifies leads, answers questions using a knowledge base, and schedules sales calls, with shared context across voice, WhatsApp, email, and SMS.
Immigrant Invest · Presales automationDocument Intelligence Platform
An agent-routed platform that prepares 100K+ confidential documents for downstream systems through local OCR and LLM processing, selective Azure fallback, and legal review.
Immigrant Invest · Document intelligenceQuote Agent
A multi-agent quoting system that interprets site assessments, compares purchasing options, and prepares costed proposals with code-based checks and human approval.
Paul Kick · Renovation operations · 2024
Personal projects
pgwarden
Give AI assistants Postgres access with per-person permissions, PII masking, human-approved writes, and an auditable query history.
MCP infrastructure · Database access controlproof-of-done
Check a coding agent’s verification claims against command results after its latest edits, with a Claude Code stop hook and a transcript audit CLI.
Developer tools · Verification evidenceagent-claimcheck
Check AI agent success claims against tool receipts and state evidence, with calibrated decisions and a review queue for uncertain cases.
Agent evaluation · Claim verificationTern
Search a video archive by what was said, shown, or written on screen. Built to run locally on Apple Silicon.
Multimodal search · Desktop apptaskdistill
Turn a recurring LLM task into a small local model, with measured quality and automatic escalation to the original API.
Model distillation · Local inferencebooking-truth
Check whether an AI booking agent did what it promised, using calendar-state evaluation, injected failures, and a guarded reference agent.
Agent evaluation · Booking reliabilitywellbrief
Ask questions across drilling reports and build pre-drill risk briefs, with traceable figures and source quotes. Runs offline by default.
Offline retrieval · Decision support