<?xml version="1.0" encoding="utf-8"?>
<journal>
<title>Medical Journal of the Islamic Republic Of Iran</title>
<title_fa>مجله پزشکی جمهوری اسلامی ایران</title_fa>
<short_title>Med J Islam Repub Iran</short_title>
<subject>Medical Sciences</subject>
<web_url>http://mjiri.iums.ac.ir</web_url>
<journal_hbi_system_id>2</journal_hbi_system_id>
<journal_hbi_system_user>journal2</journal_hbi_system_user>
<journal_id_issn>1016-1430</journal_id_issn>
<journal_id_issn_online>2251-6840</journal_id_issn_online>
<journal_id_pii>8</journal_id_pii>
<journal_id_doi>10.18869/mjiri</journal_id_doi>
<journal_id_iranmedex></journal_id_iranmedex>
<journal_id_magiran></journal_id_magiran>
<journal_id_sid>14</journal_id_sid>
<journal_id_nlai>8888</journal_id_nlai>
<journal_id_science>13</journal_id_science>
<language>en</language>
<pubdate>
	<type>jalali</type>
	<year>1404</year>
	<month>10</month>
	<day>1</day>
</pubdate>
<pubdate>
	<type>gregorian</type>
	<year>2026</year>
	<month>1</month>
	<day>1</day>
</pubdate>
<volume>40</volume>
<number>1</number>
<publish_type>online</publish_type>
<publish_edition>1</publish_edition>
<article_type>fulltext</article_type>
<articleset>
	<article>


	<language>en</language>
	<article_id_doi></article_id_doi>
	<title_fa></title_fa>
	<title>Interdisciplinary Evaluation of Urogenital Radiological Findings: Human Expert vs Multimodal Generative AI Classification Performance</title>
	<subject_fa>Urology and Nephrology</subject_fa>
	<subject>Urology and Nephrology</subject>
	<content_type_fa>Original Research</content_type_fa>
	<content_type>Original Research</content_type>
	<abstract_fa></abstract_fa>
	<abstract>&lt;span style=&quot;font-size:13pt&quot;&gt;&lt;span style=&quot;text-justify:kashida&quot;&gt;&lt;span style=&quot;text-kashida:0%&quot;&gt;&lt;span style=&quot;text-autospace:none&quot;&gt;&lt;span style=&quot;font-family:&amp;quot;Times New Roman&amp;quot;,serif&quot;&gt;&lt;span style=&quot;font-style:italic&quot;&gt;&lt;b&gt;&amp;nbsp;&amp;nbsp;&lt;/b&gt;&lt;b&gt;&lt;span style=&quot;font-size:9.0pt&quot;&gt;&lt;span style=&quot;font-style:normal&quot;&gt;&amp;nbsp; Background: &lt;/span&gt;&lt;/span&gt;&lt;/b&gt;&lt;span style=&quot;font-size:9.0pt&quot;&gt;&lt;span style=&quot;font-style:normal&quot;&gt;Multimodal generative artificial intelligence (AI) systems are increasingly evaluated for medical image classification, but their performance in interdisciplinary urogenital radiological assessment remains uncertain. This study aimed to compare the overall and domain-specific classification performance and uncertainty patterns of two domain-specific human specialists and three multimodal generative AI systems in a standardized single-image urogenital radiology task.&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;br&gt;
&lt;span style=&quot;font-size:13pt&quot;&gt;&lt;span style=&quot;text-justify:kashida&quot;&gt;&lt;span style=&quot;text-kashida:0%&quot;&gt;&lt;span style=&quot;text-autospace:none&quot;&gt;&lt;span style=&quot;font-family:&amp;quot;Times New Roman&amp;quot;,serif&quot;&gt;&lt;span style=&quot;font-style:italic&quot;&gt;&lt;span style=&quot;font-size:9.0pt&quot;&gt;&lt;span style=&quot;font-style:normal&quot;&gt;&amp;nbsp;&amp;nbsp; &lt;b&gt;Methods:&lt;/b&gt; This exploratory pilot comparative study evaluated 60 publicly available anonymized radiological cases selected from Radiopaedia, including 20 urological anomaly/variation cases, 20 gynecological anomaly/variation cases, and 20 normal anatomy cases. Two human specialists and three multimodal AI systems (GPT-4o, Gemini 1.5 Pro, and Microsoft Copilot) independently classified randomized single images into predefined categories under blinded conditions. Overall accuracy was summarized with 95% confidence intervals, while domain-specific accuracy and uncertain/unable-to-classify responses were analyzed descriptively.&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;br&gt;
&lt;span style=&quot;font-size:13pt&quot;&gt;&lt;span style=&quot;text-justify:kashida&quot;&gt;&lt;span style=&quot;text-kashida:0%&quot;&gt;&lt;span style=&quot;text-autospace:none&quot;&gt;&lt;span style=&quot;font-family:&amp;quot;Times New Roman&amp;quot;,serif&quot;&gt;&lt;span style=&quot;font-style:italic&quot;&gt;&lt;span style=&quot;font-size:9.0pt&quot;&gt;&lt;span style=&quot;font-style:normal&quot;&gt;&amp;nbsp;&amp;nbsp; &lt;b&gt;Results:&lt;/b&gt; Overall accuracy was highest for Gemini 1.5 Pro (36/60, 60.0%; 95% CI, 47.4-71.4), followed by the urologist (34/60, 56.7%; 95% CI, 44.1-68.4), GPT-4o (31/60, 51.7%; 95% CI, 39.3-63.8), the gynecologic oncologist (29/60, 48.3%; 95% CI, 36.2-60.7), and Copilot (19/60, 31.7%; 95% CI, 21.3-44.2). Human specialists performed best within their own domains. Gemini performed strongly in urological and gynecological categories but had low normal-category accuracy (2/20, 10.0%), suggesting overclassification.&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;br&gt;
&lt;span style=&quot;font-size:13pt&quot;&gt;&lt;span style=&quot;text-justify:kashida&quot;&gt;&lt;span style=&quot;text-kashida:0%&quot;&gt;&lt;span style=&quot;text-autospace:none&quot;&gt;&lt;span style=&quot;font-family:&amp;quot;Times New Roman&amp;quot;,serif&quot;&gt;&lt;span style=&quot;font-style:italic&quot;&gt;&lt;span style=&quot;font-size:9.0pt&quot;&gt;&lt;span style=&quot;font-style:normal&quot;&gt;&amp;nbsp;&amp;nbsp; &lt;b&gt;Conclusion:&lt;/b&gt; In this standardized single-image pilot task, multimodal AI systems showed potential as supportive classification tools but also demonstrated clinically relevant limitations, particularly false-positive interpretation of normal anatomy. These findings should be interpreted as hypothesis-generating and not as evidence of autonomous diagnostic capability.&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;/span&gt;&lt;br&gt;
&amp;nbsp;</abstract>
	<keyword_fa></keyword_fa>
	<keyword>Artificial Intelligence, Decision Support Systems, Clinical, Human-AI Interaction, Medical Image Interpretation, Multimodal Large Language Models</keyword>
	<start_page>639</start_page>
	<end_page>644</end_page>
	<web_url>http://mjiri.iums.ac.ir/browse.php?a_code=A-10-10256-1&amp;slc_lang=en&amp;sid=1</web_url>


<author_list>
	<author>
	<first_name>Büşra Asena </first_name>
	<middle_name></middle_name>
	<last_name>Torun</last_name>
	<suffix></suffix>
	<first_name_fa></first_name_fa>
	<middle_name_fa></middle_name_fa>
	<last_name_fa></last_name_fa>
	<suffix_fa></suffix_fa>
	<email>b.asena.torun@gmail.com</email>
	<code>200319475328460098866</code>
	<orcid>200319475328460098866</orcid>
	<coreauthor>Yes
</coreauthor>
	<affiliation>Department of Gynecologic Oncology, Adana City Training and Research Hospital, University of Health Sciences, Adana, Turkey</affiliation>
	<affiliation_fa></affiliation_fa>
	 </author>


	<author>
	<first_name>Alpkon</first_name>
	<middle_name></middle_name>
	<last_name> Torun</last_name>
	<suffix></suffix>
	<first_name_fa></first_name_fa>
	<middle_name_fa></middle_name_fa>
	<last_name_fa></last_name_fa>
	<suffix_fa></suffix_fa>
	<email>alpkontorun@gmail.com</email>
	<code>200319475328460098867</code>
	<orcid>0009-0006-7487-7456</orcid>
	<coreauthor>No</coreauthor>
	<affiliation>Department of Urology, Faculty of Medicine, Hitit University, Çorum, Turkey</affiliation>
	<affiliation_fa></affiliation_fa>
	 </author>


	<author>
	<first_name>Mustafa Serdar</first_name>
	<middle_name></middle_name>
	<last_name> Çağlayan</last_name>
	<suffix></suffix>
	<first_name_fa></first_name_fa>
	<middle_name_fa></middle_name_fa>
	<last_name_fa></last_name_fa>
	<suffix_fa></suffix_fa>
	<email>serdar.09.09@hotmail.com</email>
	<code>200319475328460098868</code>
	<orcid>200319475328460098868</orcid>
	<coreauthor>No</coreauthor>
	<affiliation>Department of Urology, Faculty of Medicine, Hitit University, Çorum, Turkey</affiliation>
	<affiliation_fa></affiliation_fa>
	 </author>


</author_list>


	</article>
</articleset>
</journal>
