Clinician Dictation

Dictate the note, get the note. Built for clinicians speaking findings, prescriptions and case summaries, with every drug name, dose and term intact.

Dictations up to 2 min · English

Upload a dictation · or hold Space and dictate

0:00
Where general-purpose recognisers break on dictation

Indian clinicians dictating drug names, findings and case notes. Same audio to both models; listen, then read. Underlined: the term that differs.

Spokenrosiglitazone
Med Veloxrosiglitazone
General-purpose modelrosy glitter zone
Spokenthiamine
Med Veloxthiamine
General-purpose modelTire menu
Spokenmometasone
Med Veloxmometasone
General-purpose modelMomento soon
Spokenloperamide
Med Veloxloperamide
General-purpose modelLow Poramide
Spokendorzolamide
Med Veloxdorzolamide
General-purpose modelDorzola Mind
Spokencalf pain
Med Veloxcalf pain
General-purpose modelCaffeine
Spokenolmesartan intolerance
Med Veloxolmesartan intolerance
General-purpose modelOnly certain intolerance
Spokenquery malar rash
Med Veloxquery malar rash
General-purpose modelKriyamala Resh
Spokenchest pain at rest
Med Veloxchest pain at rest
General-purpose modelJust being at rest.
Spokenpain left great toe
Med Veloxpain left great toe
General-purpose modelPain left grain toe.
SpokenOME otitis media with effusion
Med VeloxOME otitis media with effusion
General-purpose modelOME, OTT's Media with Effusion.
Spoken… chest x-ray in the posteroanterior PA view showing bilateral infiltrates …
Med Velox… in the posteroanterior PA view showing bilateral infiltrates …
General-purpose model… in the posterior anterior PA view showing bilateral infiltrates …
SpokenHowever, dehiscence of the repair had occurred and the patient continued to leak urine …
Med VeloxHowever, dehiscence of the repair had occurred …
General-purpose modelHowever, dyescence of the repair had occurred …

Audio from the Eka Medical ASR Evaluation Dataset (ekacare, MIT licence), English test split. General-purpose model: Whisper large-v3, default English transcription. Outputs shown verbatim.

Dictated Note

00:00:00
Idle
Performance Overview
Latency
-
End-to-end
Recorded
0.00s
Clip length
Language
English
Dictation
Server RTF
-
Processing ÷ audio
Status
Ready
Connected
Why Conscious Engines

Built for the dictating clinician: the note comes out the way it was spoken.

01

Post-trained on clinician speech

A speech model post-trained on clinicians speaking medicine, so anatomy, pathology, procedures and pharmacology come out as the words that were dictated, not their nearest everyday sound-alike.

Terminology survives the transcript
02

Drug names, doses and units

Generic and brand names, milligram and microgram doses, frequencies and routes are the parts of a dictated note that matter most and that general-purpose recognisers break first. They are treated as first-class vocabulary here.

“metformin 500 mg twice daily” stays exactly that
03

Measured on unseen clinical audio

Accuracy is reported on held-out recordings of Indian clinicians the model never saw in training, scored the same way as the clips above. No demo-only tuning.

Word error rate down by a quarter after post-training
04

The note is ready before the patient leaves

A compact model on a single GPU turns a 30-second dictation into text in well under a second of processing.

Server RTF below 0.1 on live clips
05

Readable notes, not raw tokens

Sentence casing, punctuation and section structure are restored on the server so the dictation drops straight into an EMR field, prescription or discharge summary without a clean-up pass.

Copy, paste, sign
06

Private by design

Audio is processed on infrastructure we run, never sent to a third-party speech API. Deployable inside a hospital network or a dedicated cloud tenancy, with no model or vendor names exposed to end users.

Your clinicians’ voices stay with you