Read more about the article AI in Healthcare: A Builder’s Guide to What Actually Works
AI running healthcare operations

AI in Healthcare: A Builder’s Guide to What Actually Works

AI in healthcare is not really about diagnosis. After building AI for real practices, here is where it actually delivers — operations, communication, and access — and the principles that separate what works from what just demos well.

Continue ReadingAI in Healthcare: A Builder’s Guide to What Actually Works

LLM-as-a-Judge for Voice Agents: Testing Non-Deterministic AI with Simulated Callers

You cannot unit-test a conversation. The testing playbook for production voice agents: a four-layer test pyramid, simulated callers over real audio, LLM-as-a-judge scoring calibrated to design intent, the transcript-integrity trap, and the 2-of-3 flake rule.

Continue ReadingLLM-as-a-Judge for Voice Agents: Testing Non-Deterministic AI with Simulated Callers