~/writing ttheme  ·  gtop  ·  bbottom

Writing

I write at Donuts with Chopsticks on Substack.

the gap series — benchmarks & evaluation

A Benchmark Score Is an Engine on a Stand — F1 wireframe diagram

A Benchmark Score Is an Engine on a Standdraft

visual essay · standalone · 2026

The power unit alone, and the same engine in a package — a wireframe F1 car where every part is a member of the clinical team. Engines don’t win championships; packages do.

One Chart, Five Questions — scroll-driven chart

One Chart, Five Questionsdraft

visual essay · 4 articles · 2026

Healthcare AI benchmarks are saturated, and the evaluation that answers “does this help?” looks like product analytics — in private.

State of Healthcare AI Benchmarks — title slide

State of Healthcare AI Benchmarksdraft

visual essay deck · 14 slides, one per page · 2026

Five things true at once: the record, the ceiling, the wall, and the instrument — the full landscape of healthcare AI evaluation in four parts.

Question Stream — the precharting funnel as live particles

Question Streamdraft

visual essay · interactive · WebGPU

The precharting funnel rendered as live particles — swarm → the doctor → extraction and reasoning lanes → the usage log.

Model Evaluation as Product Analytics — long-form essay

Model Evaluation as Product Analyticsdraft

long-form essay

Capability benchmarks are saturated; the evaluation that answers “does this help?” looks like product analytics, and it happens in private.

tutorials

Intro to Biomedical Ontologies — Owlready2 tutorial

Data Poetry: Intro to Biomedical Ontologies (Owlready2)draft

tutorial · data-poetry

Ontologies as data poetry — reasoning over biomedical knowledge graphs in Python, landing on a rare-disease example.