Read more about the article AI in Healthcare: A Builder’s Guide to What Actually Works
AI running healthcare operations

AI in Healthcare: A Builder’s Guide to What Actually Works

AI in healthcare is not really about diagnosis. After building AI for real practices, here is where it actually delivers — operations, communication, and access — and the principles that separate what works from what just demos well.

Continue ReadingAI in Healthcare: A Builder’s Guide to What Actually Works

LLM-as-a-Judge for Voice Agents: Testing Non-Deterministic AI with Simulated Callers

You cannot unit-test a conversation. The testing playbook for production voice agents: a four-layer test pyramid, simulated callers over real audio, LLM-as-a-judge scoring calibrated to design intent, the transcript-integrity trap, and the 2-of-3 flake rule.

Continue ReadingLLM-as-a-Judge for Voice Agents: Testing Non-Deterministic AI with Simulated Callers

AI Voice Agent Architecture: What I Learned Building the Same Agent Three Times

I built the same production voice agent three times. The orchestrator collapsed under coupling, server-gated turns created dead air, and the third architecture — where the realtime model owns the conversation — is the one that survived. Pros, cons, and diagrams of all three.

Continue ReadingAI Voice Agent Architecture: What I Learned Building the Same Agent Three Times