<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE ArticleSet PUBLIC "-//NLM//DTD PubMed 2.7//EN" "https://dtd.nlm.nih.gov/ncbi/pubmed/in/PubMed.dtd">
<ArticleSet>
<Article>
<Journal>
				<PublisherName></PublisherName>
				<JournalTitle>Transactions on Machine Intelligence</JournalTitle>
				<Issn>2821-1693</Issn>
				<Volume>9</Volume>
				<Issue>2</Issue>
				<PubDate PubStatus="epublish">
					<Year>2026</Year>
					<Month>06</Month>
					<Day>01</Day>
				</PubDate>
			</Journal>
<ArticleTitle>Evidence-Gated Autonomy in Healthcare AI Agents: A Systematic Review of Action Scope, Human-Oversight Effectiveness, and Multi-Step Safety Failures</ArticleTitle>
<VernacularTitle></VernacularTitle>
			<FirstPage>84</FirstPage>
			<LastPage>111</LastPage>
			<ELocationID EIdType="pii">248202</ELocationID>
			
<ELocationID EIdType="doi">10.22034/tmi.2026.248202</ELocationID>
			
			<Language>EN</Language>
<AuthorList>
<Author>
					<FirstName>H.</FirstName>
					<LastName>Ghayoumizadeh</LastName>
<Affiliation>Associate Professor of Biomedical Engineering, Vali-e-Asr University of Rafsanjan</Affiliation>
<Identifier Source="ORCID">0000-0002-5390-3938</Identifier>

</Author>
<Author>
					<FirstName>Kh.</FirstName>
					<LastName>Rezaee</LastName>
<Affiliation>Department of Biomedical Engineering, Meybod University, Meybod, Iran</Affiliation>

</Author>
<Author>
					<FirstName>A.</FirstName>
					<LastName>Fayazi</LastName>
<Affiliation>Department of Engineering.Vali-e-Asr University of Rafsanjan,Iran</Affiliation>

</Author>
<Author>
					<FirstName>O.</FirstName>
					<LastName>Rahmani Seryasat</LastName>
<Affiliation>Assistant Professor,Faculty of Electrical and Computer Engineering, Shams Gonbad Higher Education Institute, Gorgan, Iran</Affiliation>
<Identifier Source="ORCID">0000-0002-9289-6128</Identifier>

</Author>
</AuthorList>
				<PublicationType>Journal Article</PublicationType>
			<History>
				<PubDate PubStatus="received">
					<Year>2025</Year>
					<Month>11</Month>
					<Day>23</Day>
				</PubDate>
			</History>
		<Abstract>&lt;strong style=&quot;mso-bidi-font-weight: normal;&quot;&gt;Background:&lt;/strong&gt; Healthcare AI agents increasingly perform multi-step workflows and tool use, but evidence on their operational autonomy, safety, and human oversight remains limited. This review examines whether greater autonomy is supported by empirical safety and oversight evidence.&lt;br&gt;&lt;strong style=&quot;mso-bidi-font-weight: normal;&quot;&gt;Methods:&lt;/strong&gt; PubMed/MEDLINE, Scopus, and Web of Science were searched from January 2022 to July 2026. Of 4,096 records, 2,557 remained after deduplication. Fifty priority reports were selected for full-text review; 34 were assessed and 28 met the inclusion criteria. Data on workflow, architecture, autonomy, safety, oversight, and evaluation were narratively synthesized.&lt;br&gt;&lt;strong style=&quot;mso-bidi-font-weight: normal;&quot;&gt;Results:&lt;/strong&gt; Among 28 studies, 18 were published in 2026. Thirteen addressed diagnostic or clinical decision support, nine EHR or administrative workflows, four treatment planning or prescribing, and two patient-facing or cross-domain applications. Autonomy levels were L0 in five studies, L1 in 13, L2 in three, L3 in one, and L4 in six; none reached L5. Most L4 systems were evaluated only in sandbox or retrospective settings. Only three studies empirically assessed oversight effectiveness.&lt;br&gt;&lt;strong style=&quot;mso-bidi-font-weight: normal;&quot;&gt;Conclusions:&lt;/strong&gt; Current evidence supports tool-using and workflow-capable healthcare agents, but not unrestricted clinical autonomy. Greater autonomy should require stronger evidence of safety, failure containment, and effective human oversight.</Abstract>
		<ObjectList>
			<Object Type="keyword">
			<Param Name="value">Artificial intelligence</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">AI Agents</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Agentic AI</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Healthcare Workflow</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">patient safety</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Operational Autonomy</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Action Scope</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Human Oversight</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Oversight Effectiveness</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Multi-Step Failure</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Evidence-Gated Autonomy</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Systematic review</Param>
			</Object>
		</ObjectList>
<ArchiveCopySource DocType="pdf">https://www.tmachineintelligence.ir/article_248202_3b490b92ec85b3e3cb8258a886ca61e9.pdf</ArchiveCopySource>
</Article>
</ArticleSet>
