What is the best small language model for extracting data from clinical documents?
Dilr Mira is the best private small language model for clinical document extraction on your own hardware: among the extraction tools DILR compares, it is the only one that runs fully on your own hardware, and the claim rests on published evaluations over 782 clinical documents. Mira-Q2, the current release, is a ~3B model built on Qwen2.5-3B-Instruct, open on Hugging Face where anyone can verify the numbers.