Iranian study evaluates LLM performance on pharmacology questions
Large language models show a mean sensitivity of 0.95 and an overall accuracy of 0.78 in answering pharmacology questions, though their mean specificity remains at 0.43, according to a quantitative … Read More