Performance of Clinical AI System and Physicians and Frontier Language Models in primary care diagnostics
Clinical AI evaluation should encompass diagnosis and management after adaptive information gathering. We compared Doctorina, eight physicians and four standalone frontier language models in 150 synthetic Polish-language primary-care consultations. Doctorina achieved 82.0% Top-1 concordance versus 57.0% for physicians (difference, 25.0 percentage points; 95% confidence interval, 17.7-32.7) and 97.3% versus 85.0% primary-or-reference-differential concordance. Across 149 case pairs, normalized workup and treatment scores were 89.4 versus 66.9 and 83.7 versus 61.2. Doctorina had the highest diagn
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 58%The Parallel Consultation: A Literature Review of Patient Use of Large Language Models and Its Implications for Psychiatric Clinical Decision-Making - Cureus →
- PossiblePossibly related (embedding) · 64%Clinician Use of a General-Purpose Large Language Model in Hospital Medicine: A Mixed-Methods Pilot Study - Cureus →
- PossiblePossibly related (embedding) · 61%Comparative Performance of Large Language Models in the Polish State Specialization Examination in Anesthesiology and Intensive Care Medicine - Cureus →
- PossiblePossibly related (embedding) · 60%Embracing Large Language Models for Medical Applications, Part II: Building a Framework for Clinical Stewardship - Cureus →
- PossiblePossibly related (embedding) · 59%Co-pilot, Not Autopilot: A Practical Method for Using Large Language Models in Interventional Cardiology - EMJ →
- LinkedLinked via arxiv author · 85%Andy Nkansah →
“Performance of Clinical AI System and Physicians and Frontier Language Models in primary care diagnostics”
- LinkedLinked via arxiv author · 85%Hanna Plotnitskaya →
“Performance of Clinical AI System and Physicians and Frontier Language Models in primary care diagnostics”
- LinkedLinked via arxiv author · 85%Stanislau Salavei →
“Performance of Clinical AI System and Physicians and Frontier Language Models in primary care diagnostics”
