Research questionHow can compact Persian medical QA models reason reliably and estimate answer confidence on consumer hardware?Persian medical QA remains underserved, and compact models must handle clinical reasoning despite limited language-specific resources. Reliable confidence estimates are also needed to distinguish answers that may be unsafe to use.