Building safe generative AI with human touch

Subhabrata (Subho) Mukherjee, PhD

Co-Founder & Chief Scientific Officer, Hippocratic AI

At Hippocratic AI, I lead the technology organization building safe generative conversational AI for healthcare. As an AI and technology executive, I spearhead next-generation multimodal AI models that align to human reasoning, emotion, and safety. My executive scope spans building AI platforms and leading organizations across speech, large language models, and efficient inference for frontier-scale systems. Hippocratic AI has received $404 million in funding at a $3.5B valuation.

Prior to this, I led large-scale foundation-model initiatives as a Principal Researcher at Microsoft Research, with a decade of AI work across Microsoft, Amazon, IBM, and Google and 100+ publications. I earned my PhD summa cum laude from the Max Planck Institute for Informatics and was the 2018 SIGKDD dissertation runner-up.

Subho Mukherjee, Co-Founder and Chief Scientific Officer at Hippocratic AI
Safe AI · Human touch
250M+real patient - AI interactions
5TPolaris 5.0 · parameter constellation
7,200+citations · h-index 37
7US patents

Building the organization
behind safe AI at scale.

Every deployment creates signals that improve the system; every research result has a path to improving real human - AI interaction.

Polaris

A safety-first system—not a single model.

A stateful primary model holds natural, long-form voice conversations while specialist models for medication safety, labs, compliance, escalation, and memory work in parallel. A shared orchestration layer coordinates the system in real time.

In-house built, state-of-the-art ASR and TTS make conversations accurate, natural, and clinically fluent.

An in-house inference stack serves a 700B+ primary LLM at real-time voice latency.

99.89% clinical accuracy · Polaris 55.5% in-house ASR word error rate92% drug pronunciation accuracy · in-house TTS230 ms TTFT · in-house inference · DeepSeek 671B
Benchmarks against frontier cascaded and speech-to-speech systems
One technologyOrganizationresearch → production
01

Speech

In-house ASR · in-house TTS · real-time voice

02

Frontier models

Post-training · large-scale reinforcement learning · recursive self-improvement

03

Alignment

Reasoning · emotion · safety

04

Inference

Quantization · kernels · decoding · network & communication

05

Clinical platform

Orchestration · evaluation

06

Applied research

Paper → product → real patient

$3.5Bcompany valuation
$404Mtotal funding

Backed by leading investors including Andreessen Horowitz, General Catalyst, Kleiner Perkins, NVIDIA Ventures, Alphabet CapitalG, Avenir, and Premji Invest.

The public voice
of the technology.

Keynotes, features, and invited talks on safe agentic AI, alignment, and building systems that scale.

Feature interview

The Agentic AI Advantage

On infusing generative conversational AI with genuine human touch—and the engineering behind agents that understand, reason, and connect safely with patients.

Research
that ships.

100+ publications and patents across a decade at Microsoft Research, Amazon, IBM, and Google—now compounding inside a production system.

7,200+citations37h-index
JMLR 2024 · Foundational

Orca

Progressive learning from complex explanation traces helped define a new era of open-model post-training.

650+ citationsRead paper ↗
arXiv 2024—26 · The system

Polaris

A safety-focused LLM constellation built for real-time voice conversations in healthcare.

250M+ interactionsRead paper ↗
arXiv 2026 · Production signals

Human–AI interaction

Beyond what real patient conversations teach, this work shows the system building blocks: in-house ASR, TTS, inference, and natural human–AI interaction alignment—and how research is translated into a real-world production system.

NeurIPS keynoteRead paper ↗
Emotional support benchmark · 2026

HEART

A unified benchmark for assessing humans and language models in emotional-support dialogue—connecting rigorous evaluation with the empathy, reasoning, and safety required in real conversations.