<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE ArticleSet PUBLIC "-//NLM//DTD PubMed 2.7//EN" "https://dtd.nlm.nih.gov/ncbi/pubmed/in/PubMed.dtd">
<ArticleSet>
<Article>
<Journal>
				<PublisherName>Shahrood University of Technology</PublisherName>
				<JournalTitle>Journal of AI and Data Mining</JournalTitle>
				<Issn>2322-5211</Issn>
				<Volume>14</Volume>
				<Issue>4</Issue>
				<PubDate PubStatus="epublish">
					<Year>2026</Year>
					<Month>10</Month>
					<Day>01</Day>
				</PubDate>
			</Journal>
<ArticleTitle>Graph-Transformer Reinforcement Learning for Scalable Multi-Agent Coordination</ArticleTitle>
<VernacularTitle></VernacularTitle>
			<FirstPage>587</FirstPage>
			<LastPage>607</LastPage>
			<ELocationID EIdType="pii">3865</ELocationID>
			
<ELocationID EIdType="doi">10.22044/jadm.2026.17636.2918</ELocationID>
			
			<Language>EN</Language>
<AuthorList>
<Author>
					<FirstName>Sajad</FirstName>
					<LastName>Bastami</LastName>
<Affiliation>Computer Engineering Department, Kurdistan University, Sanandaj, Iran.</Affiliation>

</Author>
<Author>
					<FirstName>Mohammad Bagher</FirstName>
					<LastName>Dowlatshahi</LastName>
<Affiliation>Computer Engineering Department, Lorestan University, Khorramabad, Iran.</Affiliation>

</Author>
<Author>
					<FirstName>Rojiar</FirstName>
					<LastName>Pir Mohammadiani</LastName>
<Affiliation>Computer Engineering Department, Kurdistan University, Sanandaj, Iran.</Affiliation>

</Author>
<Author>
					<FirstName>Seyedeh Zahra</FirstName>
					<LastName>Mousavi</LastName>
<Affiliation>Computer Engineering Department, Lorestan University, Khorramabad, Iran.</Affiliation>

</Author>
</AuthorList>
				<PublicationType>Journal Article</PublicationType>
			<History>
				<PubDate PubStatus="received">
					<Year>2026</Year>
					<Month>04</Month>
					<Day>24</Day>
				</PubDate>
			</History>
		<Abstract>Multi-agent reinforcement learning (MARL) is a key paradigm for coordination in robotics, autonomous systems, and distributed control. However, existing MARL methods face fundamental limitations in scalability, adaptability to dynamic environments, and stability under evolving interactions. To address these challenges, we propose Adaptive Graph-Transformer Reinforcement Learning (AGTRL), a framework integrating graph-based relational modelling with transformer attention for adaptive coordination in large-scale multi-agent systems. AGTRL unifies graph-based perception and attention-based coordination in an end-to-end pipeline, encoding role information and adaptively weighting interactions by context. This combination, missing in prior MARL methods, bridges scalability and robustness in dynamic environments. AGTRL constructs a dynamic graph of agent relationships and uses multi-head self-attention to prioritize relevant interactions in real-time, ensuring robust performance under perturbations. We evaluate robustness under communication dropout (up to 40% link removal) and dynamic edge removal, measuring performance via episode reward and win rate. The framework incorporates an adaptive stability-performance trade-off mechanism that maintains learning efficacy in the presence of communication constraints and environmental uncertainty. We introduce a graph-enhanced policy architecture that jointly optimizes individual agent policies and inter-agent coordination through attention-weighted message passing. Comprehensive evaluations on benchmark environments—including StarCraft II micromanagement scenarios, cooperative navigation (Spread), and adversarial tasks (Predator-Prey)—demonstrate that AGTRL achieves superior sample efficiency, scalability, and robustness compared to state-of-the-art MARL baselines. Experimental results show AGTRL improves convergence speed by 32% on average and maintains stable performance with up to 40% communication dropout, establishing its viability for real-world deployment in dynamic multi-agent domains.</Abstract>
		<ObjectList>
			<Object Type="keyword">
			<Param Name="value">Keywords: Multi-Agent Reinforcement Learning</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Graph Neural Networks</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Transformer Attention Mechanisms</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Adaptive Coordination</Param>
			</Object>
			<Object Type="keyword">
			<Param Name="value">Scalable Multi-Agent Systems</Param>
			</Object>
		</ObjectList>
<ArchiveCopySource DocType="pdf">https://jad.shahroodut.ac.ir/article_3865_55a767557212ff3135a185d626b76d70.pdf</ArchiveCopySource>
</Article>
</ArticleSet>
