newsGoogle News — LLMTrust 62 · AggregatorPublished 7d agoLive · 6d ago
Comparative Performance of Large Language Models in the Polish State Specialization Examination in Anesthesiology and Intensive Care Medicine - Cureus
Comparative Performance of Large Language Models in the Polish Stat
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 55%EuroEval/EuroEval →
- PossiblePossibly related (embedding) · 53%Furyton/awesome-language-model-analysis →
- PossiblePossibly related (embedding) · 51%The strength of clinical evidence is recoverable from language model representations but not from their stated grades →
- PossiblePossibly related (embedding) · 50%tyang816/Awesome-TCM-LLM →
- PossiblePossibly related (embedding) · 48%How Surprising Is Historical Italian to Language Models? Tokenization Tax, Comprehension Tax, and a Simple Mitigation →
- PossiblePossibly related (embedding) · 51%Safety That Does Not Transfer: Cross-Lingual Clinical Correctness Drift in Deployable Medical Language Models →
Covers
repoEuroEval/EuroEvalrepoFuryton/awesome-language-model-analysispaperThe strength of clinical evidence is recoverable from language model representations but not from their stated gradesrepotyang816/Awesome-TCM-LLMpaperHow Surprising Is Historical Italian to Language Models? Tokenization Tax, Comprehension Tax, and a Simple Mitigation
Covers (incoming)
Related across the graph
repotyang816/Awesome-TCM-LLMpaperHow Surprising Is Historical Italian to Language Models? Tokenization Tax, Comprehension Tax, and a Simple MitigationpaperThe strength of clinical evidence is recoverable from language model representations but not from their stated gradesrepoEuroEval/EuroEvalrepoFuryton/awesome-language-model-analysispaperSafety That Does Not Transfer: Cross-Lingual Clinical Correctness Drift in Deployable Medical Language Models
