<?xml version="1.0" encoding="utf-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.0 20120330//EN" "http://jats.nlm.nih.gov/publishing/1.0/JATS-journalpublishing1.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" dtd-version="1.0" article-type="review-article">
  <front>
    <journal-meta>
      <journal-id journal-id-type="nlm-ta">Intell. Netw. Comput.</journal-id>
      <journal-id journal-id-type="publisher-id">inect</journal-id>
      <journal-title-group>
        <journal-title>Intelligent Networking and Computing</journal-title>
      </journal-title-group>
      <issn pub-type="epub"/>
      <publisher>
        <publisher-name>OAE Publishing Inc.</publisher-name>
      </publisher>
    </journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.20517/inect.2026.04</article-id>
      <article-id pub-id-type="publisher-id">INECT-2026-4</article-id>
      <article-categories>
        <subj-group>
          <subject>Review</subject>
        </subj-group>
      </article-categories>
      <title-group>
        <article-title>Resource-efficient federated large language models: challenges and mechanisms</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <name>
            <surname>Chen</surname>
            <given-names>Xiaojing</given-names>
          </name>
          <xref ref-type="aff" rid="I1">
            <sup>1</sup>
          </xref>
        </contrib>
        <contrib contrib-type="author">
          <name>
            <surname>Wang</surname>
            <given-names>Chenchen</given-names>
          </name>
          <xref ref-type="aff" rid="I1">
            <sup>1</sup>
          </xref>
        </contrib>
        <contrib contrib-type="author">
          <name>
            <surname>Li</surname>
            <given-names>Bingcong</given-names>
          </name>
          <xref ref-type="aff" rid="I2">
            <sup>2</sup>
          </xref>
        </contrib>
        <contrib contrib-type="author" corresp="yes">
          <contrib-id contrib-id-type="orcid">https://orcid.org/0000-0003-2292-3845</contrib-id>
          <name>
            <surname>Wang</surname>
            <given-names>Xin</given-names>
          </name>
          <xref ref-type="aff" rid="I3">
            <sup>3</sup>
          </xref>
          <xref ref-type="corresp" rid="cor1">*</xref>
        </contrib>
      </contrib-group>
      <aff id="I1"><sup>1</sup>Key Laboratory of Specialty Fiber Optics and Optical Access Networks, Shanghai University, Shanghai 200444, China.</aff>
      <aff id="I2"><sup>2</sup>Dept of CS, ETH Zurich, Zurich 8092, Switzerland.</aff>
      <aff id="I3"><sup>3</sup>College of Future Information Technology, Fudan University, Shanghai 200433, China.</aff>
      <author-notes>
        <corresp id="cor1">Correspondence to: Prof. Xin Wang, College of Future Information Technology, Fudan University, Shanghai 200433, China. E-mail: <email>xwang11@fudan.edu.cn</email> </corresp>
        <fn fn-type="other">
          <p><bold>Received:</bold> 17 Jul 2026 | <bold>First Decision:</bold> 6 Aug 2026 | <bold>Revised:</bold> 20 Aug 2026 | <bold>Accepted:</bold> 21 Aug 2026 | <bold>Published:</bold> 23 Sep 2026</p>
        </fn>
        <fn fn-type="other">
          <p><bold>Academic Editor:</bold> Hongliang Zhang | <bold>Copy Editor:</bold> Shu-Yuan Duan | <bold>Production Editor:</bold> Shu-Yuan Duan </p>
        </fn>
      </author-notes>
      <pub-date pub-type="ppub">
        <year>2026</year>
      </pub-date>
      <pub-date pub-type="epub">
        <day>23</day>
        <month>9</month>
        <year>2026</year>
      </pub-date>
      <volume>1</volume>
	  <issue>1</issue>
      <elocation-id>3</elocation-id>
      <permissions>
        <copyright-statement>© The Author(s) 2026.</copyright-statement>
        <license xlink:href="https://creativecommons.org/licenses/by/4.0/">
          <license-p>© The Author(s) 2026.<bold>Open Access</bold>This article is licensed under a Creative Commons Attribution 4.0 International License (<uri xlink:href="https://creativecommons.org/licenses/by/4.0/">https://creativecommons.org/licenses/by/4.0/</uri>), which permits unrestricted use, sharing, adaptation, distribution and reproduction in any medium or format, for any purpose, even commercially, as long as you give appropriate credit to the original author(s) and the source, provide a link to the Creative Commons license, and indicate if changes were made.</license-p>
        </license>
      </permissions>
      <abstract>
        <p>Adapting large language models (LLMs) over distributed private data is increasingly important for domain specialization, personalization, and alignment. Federated learning (FL) provides a natural paradigm for this goal, but federated LLMs are not simply conventional FL systems with larger models. The pretrained backbone, Transformer execution, activation memory, structured parameter-efficient fine-tuning (PEFT) updates, heterogeneous clients, and protection mechanisms create coupled resource costs across the full adaptation workflow. This review examines federated LLM adaptation from a resource-efficiency perspective. We first identify five recurring challenges: local adaptation feasibility, communication overhead, aggregation compatibility, system orchestration, and trustworthy operation. We then organize existing mechanisms into a taxonomy covering local adaptation and training, structured update and knowledge exchange, structure-aware aggregation, resource-aware orchestration, and trustworthy resource efficiency. The review shows that resource savings should be evaluated across the complete workflow, because reductions in trainable parameters, transmitted objects, or client memory may be accompanied by additional communication, server processing, coordination, or protection costs. Finally, we summarize evaluation requirements and research priorities for deployable federated LLMs, emphasizing end-to-end accounting, comparable utility targets, transparent resource evidence, realistic heterogeneity, and lifecycle-aware workloads.</p>
      </abstract>
      <kwd-group>
        <kwd>Federated large language models</kwd>
        <kwd>parameter-efficient fine-tuning</kwd>
        <kwd>resource efficiency</kwd>
        <kwd>heterogeneous aggregation</kwd>
        <kwd>edge-cloud collaboration</kwd>
        <kwd>trustworthy federated learning</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec1">
      <title>INTRODUCTION</title>
      <p>Large language models (LLMs) support a wide range of knowledge-intensive applications, but their practical value increasingly depends on post-training adaptation<sup>[<xref ref-type="bibr" rid="B1">1</xref>-<xref ref-type="bibr" rid="B3">3</xref>]</sup> for domain specialization, personalization, and alignment. Such adaptation often relies on private, domain-specific, or user-generated data that remain distributed because of regulation, ownership, or confidentiality, or because centralization would impose excessive data-transfer overhead or violate application latency requirements. Federated learning (FL) addresses this access problem by keeping data local while clients optimize and exchange model-related information for collaborative updating. It preserves local control over distributed data and draws on computing resources across organizations and edge devices, making it a natural route to private and domain-specific LLM adaptation<sup>[<xref ref-type="bibr" rid="B1">1</xref>,<xref ref-type="bibr" rid="B4">4</xref>,<xref ref-type="bibr" rid="B5">5</xref>]</sup>. More broadly, the evolution of edge intelligence toward generative AI places increasing demands on distributed computation, communication, memory, and service orchestration<sup>[<xref ref-type="bibr" rid="B6">6</xref>]</sup>, further motivating resource-aware LLM adaptation across edge and cloud infrastructure.</p>
      <p>In this context, federated LLMs refer to FL systems that collaboratively adapt a pretrained LLM, or its parameter-efficient adaptation modules, across distributed clients without sharing raw data. The adaptation process may involve full-model fine-tuning, low-rank adaptation (LoRA) or adapter optimization, prompt tuning, intermediate-representation exchange, or personalized component learning. This review focuses on federated post-training adaptation, including fine-tuning, alignment, and personalization, rather than federated pretraining from scratch.</p>
      <p>Federated LLMs retain the basic collaborative-learning principle of conventional FL, but substantially change its resource profile. Typical FL systems often assume a common model architecture and directly aggregatable parameter updates, whereas federated LLM adaptation starts from a large pretrained backbone and may exchange LoRA factors, adapters, prompts, intermediate representations, or other structured adaptation objects. Even with parameter-efficient fine-tuning (PEFT), clients may still incur substantial backbone-residency, Transformer-execution, and activation-memory costs. Moreover, clients can differ in adaptation rank, trainable layers, precision, or model-access mode, making communication and aggregation more heterogeneous and potentially requiring alignment, reconstruction, distillation, or personalized fusion.</p>
      <p>These characteristics amplify conventional FL challenges, including communication overhead, partial participation, stragglers, non-independent and non-identically distributed (non-IID) data, and privacy risks<sup>[<xref ref-type="bibr" rid="B4">4</xref>,<xref ref-type="bibr" rid="B5">5</xref>,<xref ref-type="bibr" rid="B7">7</xref>,<xref ref-type="bibr" rid="B8">8</xref>]</sup>, while creating stronger coupling among computation, communication, aggregation, and system resources. In particular, reducing trainable or transmitted parameters does not necessarily reduce end-to-end cost: Client-side savings may reappear as repeated activation/gradient exchange in split or offloaded adaptation<sup>[<xref ref-type="bibr" rid="B9">9</xref>]</sup>, while privacy, verification, and lifecycle-management mechanisms introduce additional computation, communication, or storage overhead<sup>[<xref ref-type="bibr" rid="B10">10</xref>-<xref ref-type="bibr" rid="B12">12</xref>]</sup>. More broadly, large-scale AI deployment in communication networks also couples model capabilities with communication, computing, resource allocation, and scalability constraints<sup>[<xref ref-type="bibr" rid="B13">13</xref>]</sup>. Resource efficiency in federated LLMs should therefore be evaluated across the complete adaptation workflow rather than inferred from trainable parameter count or per-round communication payload alone.</p>
      <p>To clarify the distinction from existing surveys on federated LLMs<sup>[<xref ref-type="bibr" rid="B14">14</xref>-<xref ref-type="bibr" rid="B17">17</xref>]</sup>, <xref ref-type="table" rid="t1">Table 1</xref> compares representative reviews in terms of their primary emphasis, resource aspects that receive less attention, and the additional perspective provided by this review. Existing surveys provide complementary perspectives on Federated LLM frameworks, PEFT, federated foundation models, privacy and robustness, and edge learning<sup>[<xref ref-type="bibr" rid="B18">18</xref>-<xref ref-type="bibr" rid="B20">20</xref>]</sup>, whereas this review takes end-to-end resource efficiency in federated LLM post-training adaptation as its organizing perspective. Specifically, it connects local execution, communication, aggregation, orchestration, and trust/lifecycle costs, while emphasizing cross-stage cost analysis, quantitative resource evidence, and utility-normalized Pareto benchmarking.</p>
      <table-wrap id="t1">
          <label>Table 1</label>
          <caption>
            <p>Comparison of representative surveys and the distinctive scope of this review</p>
          </caption>
          <table frame="hsides" rules="groups">
            <thead>
              <tr>
                <td style="border-bottom:1;">
                  <bold>Review</bold>
                </td>
                <td style="border-bottom:1;">
                  <bold>Primary emphasis</bold>
                </td>
                <td style="border-bottom:1;">
                  <bold>Coverage not emphasized</bold>
                </td>
                <td style="border-bottom:1;">
                  <bold>Added by this review</bold>
                </td>
              </tr>
            </thead>
            <tbody>
              <tr>
                <td>Hu <italic>et al</italic>.<sup>[<xref ref-type="bibr" rid="B14">14</xref>]</sup></td>
                <td>Federated LLM solutions, challenges, and future directions</td>
                <td>Systematic resource accounting across different stages of federated adaptation</td>
                <td>Workflow-level resource taxonomy and cross-stage cost analysis</td>
              </tr>
              <tr>
                <td>Wen <italic>et al</italic>.<sup>[<xref ref-type="bibr" rid="B15">15</xref>]</sup></td>
                <td>Federated PEFT mechanisms for LLMs</td>
                <td>Resource costs beyond parameter-efficient adaptation, including aggregation,<break />orchestration, and trust</td>
                <td>End-to-end analysis beyond trainable-parameter and update efficiency</td>
              </tr>
              <tr>
                <td>Ren <italic>et al</italic>.<sup>[<xref ref-type="bibr" rid="B16">16</xref>]</sup></td>
                <td>Federated foundation models, architectures, and open challenges</td>
                <td>LLM-specific post-training resource accounting and quantitative cross-stage<break />evaluation</td>
                <td>Focused analysis of resource-efficient federated LLM adaptation with<break />quantitative resource evidence</td>
              </tr>
              <tr>
                <td>Yan <italic>et al</italic>.<sup>[<xref ref-type="bibr" rid="B17">17</xref>]</sup></td>
                <td>Comparison of FedLLM, KD-FedLLM, and Split-FedLLM frameworks</td>
                <td>Aggregation heterogeneity, system orchestration, trust/lifecycle costs, and<break />full-workflow resource accounting</td>
                <td>Unified analysis of local execution, communication, aggregation,<break />orchestration, and trust costs</td>
              </tr>
              <tr>
                <td>Adhikari <italic>et al</italic>.<sup>[<xref ref-type="bibr" rid="B18">18</xref>]</sup></td>
                <td>Robustness, privacy, trustworthiness, and edge deployment</td>
                <td>Resource interactions among adaptation, aggregation, orchestration, and<break />protection mechanisms</td>
                <td>Trust and protection incorporated into end-to-end resource accounting</td>
              </tr>
              <tr>
                <td>Piccialli <italic>et al</italic>. <sup>[<xref ref-type="bibr" rid="B21">21</xref>]</sup></td>
                <td>Federated and edge learning, efficient LLM training and deployment</td>
                <td>Federated post-training-specific aggregation, orchestration, and cross-stage<break />resource shifting</td>
                <td>Resource analysis specialized to federated LLM post-training adaptation</td>
              </tr>
              <tr>
                <td>This review</td>
                <td>Resource-efficient federated LLM post-training adaptation</td>
                <td>-</td>
                <td>Workflow-level resource taxonomy, cross-stage cost-shift analysis,<break />representative quantitative resource evidence, and utility-normalized<break />Pareto benchmarking</td>
              </tr>
            </tbody>
          </table>
          <table-wrap-foot>
            <fn>
              <p>LLM: Large language model; PEFT: parameter-efficient fine-tuning.</p>
            </fn>
          </table-wrap-foot>
        </table-wrap>
      <p>Related surveys on efficient FL examine communication reduction, client selection, resource allocation, and system optimization, whereas surveys on efficient and edge LLMs mainly discuss compression, memory reduction, inference acceleration, and hardware-aware deployment<sup>[<xref ref-type="bibr" rid="B7">7</xref>,<xref ref-type="bibr" rid="B21">21</xref>]</sup>. Low-rank and other parameter-efficient adaptation methods reduce trainable state<sup>[<xref ref-type="bibr" rid="B2">2</xref>,<xref ref-type="bibr" rid="B3">3</xref>]</sup> and complement these broader deployment-oriented efficiency techniques. Together, these studies provide important foundations, but they do not fully explain how LLM-scale adaptation costs interact with FL-specific communication, aggregation, orchestration, and trust overheads across the complete federated adaptation workflow.</p>
      <p>Accordingly, we make three contributions:</p>
      <p>• We develop a workflow-level resource analysis of federated LLM adaptation, identifying five key challenges: local adaptation infeasibility, communication overhead, aggregation incompatibility, orchestration complexity, and privacy, security, and lifecycle-trust overheads. </p>
      <p>• We propose a matching resource-efficiency taxonomy covering resource-efficient local adaptation and training, communication-efficient exchange of structured updates and knowledge, structure-aware aggregation and heterogeneous adaptation, resource-aware system orchestration, and trustworthy resource efficiency. </p>
      <p>• We summarize evaluation requirements and research priorities for deployable federated LLMs, emphasizing end-to-end resource accounting, comparable utility targets, evidence transparency, realistic heterogeneity, and lifecycle-aware workloads.</p>
      <p><xref ref-type="fig" rid="fig1">Figure 1</xref> summarizes the organization of this review. Section <bold>BACKGROUND AND RESOURCE ACCOUNTING FOR FEDERATED LLMS</bold> presents the background and resource-accounting framework for federated LLM adaptation. Section <bold>RESOURCE CHALLENGES IN FEDERATED LLMS</bold> analyzes the main resource challenges, and Section <bold>RESOURCE-EFFICIENT MECHANISMS FOR FEDERATED LLMS</bold> reviews the corresponding resource-efficient mechanisms. Section <bold>EVALUATION REQUIREMENTS AND OPEN RESEARCH DIRECTIONS</bold> discusses evaluation requirements and research priorities for deployable federated LLMs. Section <bold>CONCLUSION AND OUTLOOK</bold> concludes the review.</p>
      <fig id="fig1" position="float">
        <label>Figure 1</label>
        <caption>
          <p>Organization and logical progression of this review. LLM: Large language model.</p>
        </caption>
        <graphic xlink:href="inect1004.fig.1.jpg"/>
      </fig>
    </sec>
    <sec id="sec2">
      <title>BACKGROUND AND RESOURCE ACCOUNTING FOR FEDERATED LLMS</title>
      <sec id="sec2-1">
        <title>Federated learning workflow</title>
        <p>FL coordinates model optimization across distributed data holders without requiring raw data to be centralized. In a typical server-client workflow, the server initializes and distributes a global model, selected clients perform local optimization on private data, and the server aggregates the returned model-related information to update the global model<sup>[<xref ref-type="bibr" rid="B4">4</xref>,<xref ref-type="bibr" rid="B5">5</xref>]</sup>. Federated Averaging (FedAvg) is the canonical instance, where client updates are commonly averaged according to local data volume<sup>[<xref ref-type="bibr" rid="B4">4</xref>]</sup>. This repeated distribution-local update-aggregation cycle makes FL suitable for privacy-preserving collaborative learning, but also introduces resource costs from communication, synchronization, client heterogeneity, and server-side aggregation.</p>
        <p>The workflow can be described along three dimensions that are especially relevant to federated LLM adaptation:</p>
        <p>• Data organization: Horizontal FL considers different samples in a shared feature space; vertical FL considers overlapping entities with different feature sets; and federated transfer learning addresses limited overlap in both samples and features.</p>
        <p>• Coordination architecture: Server-based FL relies on a central aggregator; hierarchical FL introduces intermediate edge servers; and decentralized FL uses peer-to-peer communication and local consensus.</p>
        <p>• Synchronization mode: Synchronous FL aggregates updates from a round-specific client set, asynchronous FL incorporates updates as they arrive, and bounded-asynchronous FL limits staleness while retaining partial flexibility<sup>[<xref ref-type="bibr" rid="B5">5</xref>,<xref ref-type="bibr" rid="B8">8</xref>]</sup>.</p>
        <p>These dimensions determine where data remain, how model-related information moves, and when updates are fused. They therefore provide the basic workflow assumptions for analyzing communication, computation, aggregation, and synchronization costs in federated LLM systems.</p>
      </sec>
      <sec id="sec2-2">
        <title>Large language models and adaptation</title>
        <p>LLMs are foundation models trained on large-scale text or multimodal corpora to acquire transferable capabilities for language-centered tasks. Most contemporary LLMs use the Transformer architecture, whose self-attention captures long-range token dependencies and permits parallel processing of token positions within each training layer, although autoregressive generation remains sequential. Through pretraining, typically with next-token prediction, LLMs learn general linguistic, semantic, and task-relevant patterns that support generation, reasoning, coding, and domain-specific text processing<sup>[<xref ref-type="bibr" rid="B22">22</xref>,<xref ref-type="bibr" rid="B23">23</xref>]</sup>.</p>
        <p>Post-training adapts this pretrained foundation to specific domains, users, or preferences through supervised fine-tuning, instruction tuning, domain adaptation, or preference alignment. Full-model fine-tuning updates most or all parameters, whereas PEFT trains only low-rank matrices, adapters, prompts, prefixes, or other small components while freezing most of the backbone<sup>[<xref ref-type="bibr" rid="B2">2</xref>,<xref ref-type="bibr" rid="B3">3</xref>,<xref ref-type="bibr" rid="B24">24</xref>]</sup>. It aims to preserve adaptation utility while reducing trainable state, optimizer storage, and, in federated settings, update size.</p>
        <p>However, PEFT does not remove the main execution burden of LLM adaptation. Clients may still need to host the pretrained backbone, perform forward and backward propagation through Transformer layers, store activations and intermediate states, and process long input sequences. Long-context adaptation further increases attention computation and activation memory<sup>[<xref ref-type="bibr" rid="B22">22</xref>,<xref ref-type="bibr" rid="B23">23</xref>]</sup>. As a result, LLM adaptation is not only a model-update problem, but also an execution problem shaped by memory, computation, bandwidth, latency, and energy. This distinction is central to resource accounting in federated LLMs, where fewer trainable parameters do not necessarily imply lower end-to-end adaptation cost.</p>
      </sec>
      <sec id="sec2-3">
        <title>Federated LLM framework and resource accounting</title>
        <p>Although FL can support language-model pretraining, this review focuses on the more common and practically relevant setting of federated post-training adaptation. In this setting, clients collaboratively adapt a pretrained LLM, or an adaptation interface built around it, over distributed private data rather than training a language model from random initialization. Depending on the system design, clients may update the full model, optimize PEFT modules<sup>[<xref ref-type="bibr" rid="B1">1</xref>,<xref ref-type="bibr" rid="B14">14</xref>,<xref ref-type="bibr" rid="B17">17</xref>]</sup>, tune prompts, maintain personalized components, or exchange split-model representations<sup>[<xref ref-type="bibr" rid="B9">9</xref>]</sup>, while raw data remain local<sup>[<xref ref-type="bibr" rid="B25">25</xref>,<xref ref-type="bibr" rid="B26">26</xref>]</sup>.</p>
        <p>This setting changes three key objects in the federated loop:</p>
        <p>• Adaptation object: the full model, a PEFT module, a prompt representation, a split-model component, or a personalized adapter.</p>
        <p>• Communication object: compressed parameter deltas, low-rank factors, prompts, logits, activations, gradients, or other task-related signals.</p>
        <p>• Federated output: a global LLM, a shared adaptation module, personalized components, or a deployment configuration.</p>
        <p>These objects may vary across clients. In server-coordinated, hierarchical, and emerging service-assisted federated LLM systems<sup>[<xref ref-type="bibr" rid="B27">27</xref>]</sup>, clients can differ in adaptation configuration, model access, objectives, and available resources. Such differences create two coupled constraints: local feasibility, which determines whether a client can produce an update, and aggregation compatibility, which determines whether heterogeneous updates can be fused into a useful shared or personalized model.</p>
        <p>These choices complicate resource accounting in federated LLMs. Unlike conventional FL, model size, computation, and communication can become decoupled: A small PEFT module may still require executing a large pretrained backbone<sup>[<xref ref-type="bibr" rid="B1">1</xref>,<xref ref-type="bibr" rid="B24">24</xref>,<xref ref-type="bibr" rid="B28">28</xref>]</sup>, while compact updates may introduce frequent synchronization, activation transfer, reconstruction, or alignment costs<sup>[<xref ref-type="bibr" rid="B9">9</xref>]</sup>. Trust mechanisms, including privacy protection, verification, robustness, provenance, and unlearning, further add computation, communication, storage, or utility costs<sup>[<xref ref-type="bibr" rid="B10">10</xref>-<xref ref-type="bibr" rid="B12">12</xref>]</sup>. </p>
        <p>Resource efficiency should therefore be assessed across the complete adaptation workflow at a comparable utility target, covering client execution, bidirectional communication, server-side processing, orchestration, and trust-related costs. As summarized in <xref ref-type="fig" rid="fig2">Figure 2</xref>, trainable parameter count or per-round payload alone cannot represent the end-to-end cost of federated LLM adaptation. </p>
        <fig id="fig2" position="float">
          <label>Figure 2</label>
          <caption>
            <p>Federated LLM framework and resource-accounting view. Heterogeneous clients adapt LLM-specific objects through communication and aggregation, while total cost is accounted across client-side execution, network exchange, server-side processing, orchestration, and trust/lifecycle operations at a comparable utility target. LLM: Large language model; PEFT: parameter-efficient fine-tuning.</p>
          </caption>
          <graphic xlink:href="inect1004.fig.2.jpg"/>
        </fig>
        <p><xref ref-type="table" rid="t2">Table 2</xref> summarizes the main differences between federated LLM adaptation and related model collaboration paradigms. Federated LLM adaptation occupies a specific design point among these collaboration paradigms. It combines local data control with collaborative adaptation of large pretrained models. Its resource challenge extends beyond conventional FL communication and local computation because LLM-scale execution, structured and potentially heterogeneous adaptation objects, and their aggregation and orchestration must be considered jointly. This distinction motivates the workflow-level resource analysis developed in Sections <bold>RESOURCE CHALLENGES IN FEDERATED LLMS</bold> and <bold>RESOURCE-EFFICIENT MECHANISMS FOR FEDERATED LLMS</bold>.</p>
        <table-wrap id="t2">
          <label>Table 2</label>
          <caption>
            <p>Comparison of federated LLM adaptation with adjacent model-collaboration paradigms</p>
          </caption>
          <table frame="hsides" rules="groups">
            <thead>
            <tr>
              <td style="border-bottom:1;">
                  <bold>Paradigm</bold>
                </td>
                 <td style="border-bottom:1;">
                  <bold>Data/model placement</bold>
                </td>
                 <td style="border-bottom:1;">
                  <bold>Collaboration object</bold>
                </td>
                 <td style="border-bottom:1;">
                  <bold>Primary resource burden</bold>
                </td>
                 <td style="border-bottom:1;">
                  <bold>Appropriate setting/relation to FLLM</bold>
                </td>
              </tr>
          </thead>
          <tbody>
            <tr>
                <td>Centralized LLM adaptation<sup>[<xref ref-type="bibr" rid="B2">2</xref>-<xref ref-type="bibr" rid="B3">3</xref>,<xref ref-type="bibr" rid="B23">23</xref>]</sup></td>
                <td>Training data and computation are centralized</td>
                <td>Raw data are transferred to a central trainer; checkpoints are later deployed</td>
                <td>Centralized data transfer, governance, and compute cost</td>
                <td>Preferable when data movement and centralized governance are acceptable; it does not preserve local data control during adaptation</td>
              </tr>
              <tr>
                <td>Conventional FL<sup>[<xref ref-type="bibr" rid="B4">4</xref>,<xref ref-type="bibr" rid="B5">5</xref>]</sup></td>
                <td>Raw data remain local; clients commonly share a task model</td>
                <td>Parameters or gradients in a compatible structure</td>
                <td>Communication, local computation, stragglers, and statistical/system heterogeneity</td>
                <td>Suitable for manageable, structurally compatible task models; it does not capture the backbone and structured-adaptation costs of LLMs</td>
              </tr>
              <tr>
                <td>Split/split-federated learning<sup>[<xref ref-type="bibr" rid="B9">9</xref>,<xref ref-type="bibr" rid="B29">29</xref>,<xref ref-type="bibr" rid="B30">30</xref>]</sup></td>
                <td>Model execution is partitioned between client and server</td>
                <td>Intermediate activations and gradients</td>
                <td>Reduced device-side execution, but repeated bidirectional traffic and server-side load</td>
                <td>Complementary to FLLM when edge assistance is available; it changes, rather than removes, the resource bottleneck</td>
              </tr>
              <tr>
                <td>Large-small/edge-cloud model collaboration<sup>[<xref ref-type="bibr" rid="B31">31</xref>]</sup></td>
                <td>Small models run locally; large models/services reside at edge/cloud</td>
                <td>Inputs, intermediate information, knowledge, or service outputs</td>
                <td>Local computation versus communication, service latency, and edge/cloud resource use</td>
                <td>Primarily deployment-time collaboration; it can complement FLLM after adaptation but does not itself perform collaborative training on local data</td>
              </tr>
              <tr>
                <td>Decentralized peer collaboration<sup>[<xref ref-type="bibr" rid="B32">32</xref>,<xref ref-type="bibr" rid="B33">33</xref>]</sup></td>
                <td>Raw data remain local; no conventional aggregation server is required</td>
                <td>Peer updates, PEFT modules, or knowledge</td>
                <td>Network-wide peer exchange, topology/peer selection, consensus, and trust cost</td>
                <td>A coordination variant of FLLM when central aggregation is undesirable; it trades server dependence for distributed coordination cost<break /></td>
              </tr>
              <tr>
                <td>Federated LLM adaptation<sup>[<xref ref-type="bibr" rid="B1">1</xref>,<xref ref-type="bibr" rid="B14">14</xref>-<xref ref-type="bibr" rid="B17">17</xref>,<xref ref-type="bibr" rid="B25">25</xref>,<xref ref-type="bibr" rid="B26">26</xref>]</sup></td>
                <td>Raw data remain local; a pretrained LLM or PEFT interface is adapted across participants</td>
                <td>Full deltas, LoRA/adapters, prompts, logits, activations, gradients, or other structured signals</td>
                <td>Backbone execution, structured communication, heterogeneous aggregation, orchestration, and trust/lifecycle cost</td>
                <td>Distinct when local data control and collaborative shared/personalized adaptation of a common backbone are both required; this is the scope of the review.</td>
              </tr>
            </tbody>
          </table>
          <table-wrap-foot>
            <fn>
              <p>FLLM denotes federated LLM. LLM: Large language model; PEFT: parameter-efficient fine-tuning.</p>
            </fn>
          </table-wrap-foot>
        </table-wrap>
        <p>The review is organized around five resource questions that follow the federated adaptation workflow: Can a client <italic>produce</italic> a useful update within its memory, computation, time, and energy budget? Can the required information be <italic>exchanged</italic> efficiently? Can heterogeneous client contributions be <italic>aggregated</italic> into a meaningful shared or personalized model? Can participation, workload, synchronization, computation placement, and communication resources be <italic>orchestrated</italic> efficiently? Finally, can <italic>privacy, security, and lifecycle trust</italic> be maintained without excessive additional resource cost? These questions cover the principal resource-bearing stages of federated post-training adaptation and explain the one-to-one organization of the challenges in Section <bold>RESOURCE CHALLENGES IN FEDERATED LLMS</bold> and the corresponding mechanisms in Section <bold>RESOURCE CHALLENGES IN FEDERATED LLMS</bold>.</p>
      </sec>
    </sec>
    <sec id="sec3">
      <title>RESOURCE CHALLENGES IN FEDERATED LLMS</title>
      <p>Federated LLMs inherit classical FL challenges, including communication overhead, partial participation, stragglers, non-IID data, and privacy risks, while intensifying them through large-scale model execution and adaptation. Structured updates, heterogeneous modules, multi-tier orchestration, and LLM-specific trust risks add further complexity. Resource efficiency therefore requires adaptation, communication, aggregation, orchestration, and protection to remain jointly feasible under practical resource constraints.</p>
      <sec id="sec3-1">
        <title>Client-side adaptation feasibility</title>
        <p>The reported link rates in <xref ref-type="table" rid="t3">Table 3</xref> are experimental configurations adopted by the corresponding studies and should not be interpreted as constant-rate assumptions for practical wireless deployments.</p>
        <p>A basic assumption in FL is that each selected client can complete its assigned local update. This assumption becomes fragile in federated LLMs. Representative edge platforms differ in accelerator support, memory capacity, power envelope, network conditions, and measured local-training latency, as summarized in <xref ref-type="table" rid="t3">Table 3</xref><sup>[<xref ref-type="bibr" rid="B34">34</xref>-<xref ref-type="bibr" rid="B36">36</xref>]</sup> The reported link rates in <xref ref-type="table" rid="t3">Table 3</xref> are experimental configurations adopted by the corresponding studies and should not be interpreted as constant-rate assumptions for practical wireless deployments. Even when adaptation starts from a pretrained model, clients may still need to host the backbone, store activations and intermediate states, propagate gradients through Transformer layers, and maintain optimizer states. These requirements can make some local adaptation configurations infeasible, rather than merely slower, on resource-constrained clients<sup>[<xref ref-type="bibr" rid="B1">1</xref>,<xref ref-type="bibr" rid="B37">37</xref>-<xref ref-type="bibr" rid="B39">39</xref>]</sup>. Local feasibility is also affected by the data used for adaptation. Redundant, noisy, or low-value instructions consume expensive Transformer updates without proportional utility, whereas data-scarce settings may require additional augmentation, privacy-preserving adaptation, or sample-selection mechanisms<sup>[<xref ref-type="bibr" rid="B40">40</xref>-<xref ref-type="bibr" rid="B42">42</xref>]</sup>.</p>
		<table-wrap id="t3">
          <label>Table 3</label>
          <caption>
            <p>Representative edge devices and resource heterogeneity in federated language-model adaptation studies</p>
          </caption>
          <table frame="hsides" rules="groups">
  <thead>
            <tr>
              <td style="border-bottom:1;">
        <bold>Device class</bold>
      </td>
      <td style="border-bottom:1;">
        <bold>Compute and memory profile</bold>
      </td>
      <td style="border-bottom:1;">
        <bold>Power and network budget</bold>
      </td>
      <td style="border-bottom:1;">
        <bold>Observed heterogeneity and implication for federated adaptation</bold>
      </td>
    </tr>
          </thead>
          <tbody>
            <tr>
      <td>Jetson AGX Orin 64 GB<sup>[<xref ref-type="bibr" rid="B34">34</xref>]</sup></td>
      <td>2048-core Ampere GPU; 12-core Arm CPU; 64 GB unified LPDDR5</td>
      <td>15-60 W configurable power; 1 Gbit/s link in the edge-LLM study</td>
      <td>High-end edge client; supports PEFT of FLAN-T5 models up to 3B parameters, but memory bandwidth and model-update communication remain bottlenecks</td>
    </tr>
    <tr>
      <td>Jetson Orin Nano 8 GB<sup>[<xref ref-type="bibr" rid="B35">35</xref>]</sup></td>
      <td>1024-core Ampere GPU; 6-core Arm CPU; 8 GB LPDDR5</td>
      <td>7-15 W power modes; evaluated at 15 W and 1 MB/s in FedARA</td>
      <td>Mid-range accelerator; lower latency than CPU-only clients, but memory, thermal, and energy budgets favor adaptive rank, precision, and participation</td>
    </tr>
    <tr>
      <td>Jetson TX2 8 GB<sup>[<xref ref-type="bibr" rid="B36">36</xref>]</sup></td>
      <td>256-core Pascal GPU; 8 GB LPDDR4</td>
      <td>7.5-15 W device class; 1 MB/s default experimental link</td>
      <td>0.88 s per BERT batch; communication dominates on the stronger client, making compact adapters and fewer rounds important</td>
    </tr>
    <tr>
      <td>Jetson Nano 4 GB<sup>[<xref ref-type="bibr" rid="B36">36</xref>]</sup></td>
      <td>128-core Maxwell GPU; 4 GB LPDDR4</td>
      <td>5/10 W power modes; 1 MB/s default experimental link</td>
      <td>1.89 s per BERT batch; tighter memory and compute than TX2 increase sensitivity to adapter depth and width</td>
    </tr>
    <tr>
      <td>Raspberry Pi 4B<sup>[<xref ref-type="bibr" rid="B36">36</xref>]</sup></td>
      <td>Quad-core Cortex-A72 CPU; 1-8 GB LPDDR4; no CUDA-class GPU</td>
      <td>15 W recommended supply; 1 MB/s default experimental link</td>
      <td>18.27 s per BERT batch; CPU-bound local training motivates shallow adapters, caching, and selective participation</td>
    </tr>
    <tr>
      <td>Raspberry Pi 5<sup>[<xref ref-type="bibr" rid="B35">35</xref>]</sup></td>
      <td>Quad-core Cortex-A76 CPU; 4/8 GB LPDDR4X variants; exact evaluated RAM not reported</td>
      <td>27 W recommended supply; 1 MB/s experimental link in FedARA</td>
      <td>1.00/2.01 s per DistilBERT/BERT batch; faster than Pi 4B but still markedly slower than Orin devices</td>
    </tr>
  </tbody>
</table>
          <table-wrap-foot>
            <fn id="t3FN1">
              <p>GPU: Graphics processing unit; LLM: Large language model; RAM: random access memory; CPU: central processing unit; BERT: Bidirectional Encoder Representations from Transformers.</p>
            </fn>
          </table-wrap-foot>
        </table-wrap>
        <p>The resource evidence in existing federated language-model studies spans substantially different model scales. Results obtained with DistilBERT, BERT, and BART are useful for studying device heterogeneity and resource-control mechanisms, but they should not be interpreted as direct feasibility evidence for multi-billion-parameter generative LLMs. The latter introduce a much larger backbone-residency and execution burden, in addition to activation and optimizer-state memory during training. For example, AssyLLM<sup>[<xref ref-type="bibr" rid="B43">43</xref>]</sup> reports that full fine-tuning of LLaMA-7B with batch size 16 requires more than 40 GB of memory, compared with the 4-16 GB memory available to many of the edge devices considered in its evaluation. Its block-assembly design reduces memory consumption by up to 92%, illustrating both the severity of the memory gap and the need for LLM-specific execution mechanisms.</p>
        <p>This creates a gap between parameter efficiency and system feasibility. PEFT methods such as LoRA and adapters reduce the number of trainable parameters<sup>[<xref ref-type="bibr" rid="B2">2</xref>-<xref ref-type="bibr" rid="B3">3</xref>,<xref ref-type="bibr" rid="B24">24</xref>,<xref ref-type="bibr" rid="B44">44</xref>-<xref ref-type="bibr" rid="B46">46</xref>]</sup>, but producing even a lightweight update may still require substantial memory, computation, and energy because the frozen backbone must be executed. Local feasibility is best assessed through the full adaptation process, where uploaded update size is considered together with model residency, memory footprint, training time, energy use, and data utility.</p>
      </sec>
      <sec id="sec3-2">
        <title>Communication overhead from structured exchanges</title>
        <p>Communication has long been a core bottleneck in FL, but federated LLMs change both the size and the form of exchanged information. Since full-model transmission is often impractical, clients may instead communicate PEFT modules<sup>[<xref ref-type="bibr" rid="B47">47</xref>,<xref ref-type="bibr" rid="B48">48</xref>]</sup>, dense or compressed parameter deltas<sup>[<xref ref-type="bibr" rid="B49">49</xref>,<xref ref-type="bibr" rid="B50">50</xref>]</sup>, prompts, logits, activations, gradients, or other intermediate representations<sup>[<xref ref-type="bibr" rid="B9">9</xref>,<xref ref-type="bibr" rid="B51">51</xref>]</sup>. These objects are usually much smaller than a complete LLM, but they may be exchanged repeatedly across local batches, communication rounds, or split-model interactions, creating substantial uplink, downlink, and synchronization costs.</p>
        <p>Thus, smaller messages do not necessarily imply lower end-to-end communication cost. Compression and low-rank transmission may reduce each payload but introduce extra processing, metadata, or reconstruction costs, and lower-fidelity updates may require more rounds to reach the same utility.</p>
        <p>In practical wireless deployments, network conditions may vary substantially over time. Channel fading, interference, link errors and retransmissions, MAC-layer contention, bandwidth fluctuations, and uplink/downlink asymmetry can reduce effective goodput and increase communication-latency variability. This variability directly affects federated synchronization: A temporarily weak client link may become a communication straggler in synchronous FL, whereas asynchronous or partial aggregation can reduce waiting at the cost of increased update staleness. The effect can be particularly pronounced for split-federated LLMs, where activations and gradients may be exchanged repeatedly during local adaptation. Communication efficiency should therefore account not only for nominal link rates, but also for link variability, reliability, tail latency, and their effects on synchronization.</p>
      </sec>
      <sec id="sec3-3">
        <title>Aggregation compatibility under heterogeneous adaptation</title>
        <p>FedAvg assumes that client updates share a common parameter structure and can be directly averaged<sup>[<xref ref-type="bibr" rid="B4">4</xref>,<xref ref-type="bibr" rid="B5">5</xref>]</sup>. This assumption weakens in federated LLMs because resource constraints often push clients toward heterogeneous adaptation configurations. A resource-rich client may train a higher-rank LoRA module or more Transformer layers, whereas a constrained client may use lower ranks, fewer trainable layers<sup>[<xref ref-type="bibr" rid="B52">52</xref>-<xref ref-type="bibr" rid="B54">54</xref>]</sup>, lower-bit precision<sup>[<xref ref-type="bibr" rid="B37">37</xref>]</sup>, or a prompt-based interface<sup>[<xref ref-type="bibr" rid="B55">55</xref>]</sup>. These choices improve local feasibility, but they also change the structure and functional meaning of the updates received by the server.</p>
        <p>Aggregation therefore becomes a compatibility problem, not merely an averaging operation. Client updates may differ in structure, precision, or training objective. While mild structural differences can be handled through simple alignment<sup>[<xref ref-type="bibr" rid="B52">52</xref>,<xref ref-type="bibr" rid="B53">53</xref>]</sup>, stronger heterogeneity often requires reconstruction<sup>[<xref ref-type="bibr" rid="B54">54</xref>]</sup>, distillation<sup>[<xref ref-type="bibr" rid="B56">56</xref>]</sup>, or personalized fusion<sup>[<xref ref-type="bibr" rid="B55">55</xref>]</sup>. Uniform configurations simplify aggregation but may exclude weaker clients, whereas heterogeneous configurations improve participation at the cost of additional fusion complexity. The key challenge is therefore to preserve the functional or semantic content of local adaptations while keeping aggregation efficient. </p>
      </sec>
      <sec id="sec3-4">
        <title>System heterogeneity and orchestration</title>
        <p>Aggregation compatibility concerns the structure of updates, whereas system heterogeneity concerns the clients and infrastructure that produce them. In federated LLMs, system heterogeneity arises jointly from computing and communication capacities. Computational heterogeneity includes memory and accelerator availability, training speed, numerical-precision support, and energy or thermal constraints, while communication heterogeneity includes bandwidth, uplink/downlink asymmetry, time-varying link quality, and intermittent connectivity. Existing studies adapt quantization, LoRA configuration, or local workload to heterogeneous client capabilities<sup>[<xref ref-type="bibr" rid="B35">35</xref>,<xref ref-type="bibr" rid="B37">37</xref>]</sup>, while hierarchical and split-federated designs coordinate workload, synchronization, and communication resources across heterogeneous devices and links<sup>[<xref ref-type="bibr" rid="B29">29</xref>,<xref ref-type="bibr" rid="B57">57</xref>,<xref ref-type="bibr" rid="B58">58</xref>]</sup>. A uniform workload may exclude resource-constrained clients or create stragglers<sup>[<xref ref-type="bibr" rid="B59">59</xref>,<xref ref-type="bibr" rid="B60">60</xref>]</sup>, whereas edge assistance<sup>[<xref ref-type="bibr" rid="B61">61</xref>]</sup> and asynchronous participation<sup>[<xref ref-type="bibr" rid="B8">8</xref>]</sup> can shift resource burdens toward communication, edge servers, or coordination overhead<sup>[<xref ref-type="bibr" rid="B57">57</xref>,<xref ref-type="bibr" rid="B62">62</xref>]</sup>. These effects are tightly coupled: offloading can reduce client computation while increasing communication demand, whereas communication-limited clients may require lighter adaptation workloads or less frequent participation.</p>
       <p>System orchestration must therefore assign feasible and useful roles to heterogeneous participants. This involves deciding which clients participate, what adaptation configuration they use, where computation is placed, when updates are synchronized, and how communication and computing resources are allocated. These decisions are coupled: Lowering local computation through edge assistance may increase repeated transfers of activations and gradients<sup>[<xref ref-type="bibr" rid="B61">61</xref>,<xref ref-type="bibr" rid="B62">62</xref>]</sup>; relaxing synchronization may reduce waiting time but introduce staleness<sup>[<xref ref-type="bibr" rid="B8">8</xref>,<xref ref-type="bibr" rid="B57">57</xref>]</sup>; and favoring capable clients may improve efficiency but reduce data coverage or fairness<sup>[<xref ref-type="bibr" rid="B55">55</xref>,<xref ref-type="bibr" rid="B59">59</xref>]</sup>. Effective orchestration should therefore balance local feasibility, communication and computation cost, system latency, and data representativeness.</p>
        </sec>
      <sec id="sec3-5">
        <title>Privacy, security, and trust overheads</title>
        <p>Keeping raw data local provides an important basis for privacy-preserving LLM adaptation, but the exchanged information still requires protection. Privacy mainly concerns leakage from shared updates or intermediate representations, which may reveal sensitive training content or user-specific patterns<sup>[<xref ref-type="bibr" rid="B63">63</xref>-<xref ref-type="bibr" rid="B65">65</xref>]</sup>. Security concerns the robustness and integrity of adaptation, where malicious, biased, or backdoored updates may distort the shared model or propagate harmful behavior<sup>[<xref ref-type="bibr" rid="B66">66</xref>-<xref ref-type="bibr" rid="B68">68</xref>]</sup>. Trust further concerns whether the adaptation process is verifiable, accountable, and auditable across the model lifecycle.</p>
        <p>These protections introduce resource overheads. Privacy mechanisms may reduce utility or increase communication and cryptographic costs<sup>[<xref ref-type="bibr" rid="B10">10</xref>,<xref ref-type="bibr" rid="B18">18</xref>,<xref ref-type="bibr" rid="B65">65</xref>]</sup>; security mechanisms such as robust aggregation and attack detection require additional computation<sup>[<xref ref-type="bibr" rid="B66">66</xref>]</sup>; and trust mechanisms such as verifiable aggregation, provenance, and unlearning introduce storage, auditing, rollback, or retraining costs<sup>[<xref ref-type="bibr" rid="B11">11</xref>,<xref ref-type="bibr" rid="B12">12</xref>]</sup>. Privacy, security, and trust mechanisms should therefore be co-designed with resource optimization, because they jointly shape the attainable balance among utility, efficiency, and trustworthiness.</p>
      </sec>
    </sec>
    <sec id="sec4">
      <title>RESOURCE-EFFICIENT MECHANISMS FOR FEDERATED LLMS</title>
      <p>Following the challenges in Section <bold>RESOURCE CHALLENGES IN FEDERATED LLMS</bold>, <xref ref-type="fig" rid="fig3">Figure 3</xref> organizes resource-efficient mechanisms according to the main stages of the federated adaptation workflow: local adaptation and training, structured exchange, heterogeneous aggregation, system orchestration, and trust-constrained operation. This section examines how these mechanisms reduce resource costs and where their trade-offs arise.</p>
      <fig id="fig3" position="float">
        <label>Figure 3</label>
        <caption>
          <p>Taxonomy of resource-efficient mechanisms for federated LLM adaptation. LLM: Large language model; PEFT: parameter-efficient fine-tuning.</p>
        </caption>
        <graphic xlink:href="inect1004.fig.3.jpg"/>
      </fig>
      <sec id="sec4-1">
        <title>Resource-efficient local adaptation and training</title>
        <p>Local adaptation is the first step at which resource feasibility is tested. A selected client must be able to produce a useful update within its memory, computation, time, and energy budget. Existing methods improve local feasibility by reducing trainable state, lowering execution cost, or selecting the data that justify expensive Transformer updates.</p>
        <p><bold>Parameter-efficient adaptation. </bold>PEFT adapts a pretrained LLM by freezing most backbone parameters and updating only a small set of trainable components. Many federated LLM systems therefore replace full fine-tuning with PEFT methods such as LoRA, adapters, prefix tuning, and prompt tuning<sup>[<xref ref-type="bibr" rid="B1">1</xref>,<xref ref-type="bibr" rid="B44">44</xref>-<xref ref-type="bibr" rid="B46">46</xref>]</sup>. For LLaMA-7B, FederatedScope-LLM<sup>[<xref ref-type="bibr" rid="B1">1</xref>]</sup> reports per-round messages of approximately 12,852 MB for full-model exchange, 21.40 MB for LoRA, and 0.17 MB for prompt tuning, while model-only memory remains about 13.4 GB under the study’s accounting. PEFT therefore provides its most direct savings in trainable and transmitted state, while backbone hosting and execution remain important parts of local resource accounting.</p>
        <p><bold>Resource-aware local execution.</bold> Beyond the choice of trainable modules, resource-aware execution controls the cost of producing each local update. Existing methods adjust numerical precision<sup>[<xref ref-type="bibr" rid="B37">37</xref>]</sup>, LoRA rank, trainable layers, or active modules<sup>[<xref ref-type="bibr" rid="B39">39</xref>]</sup> according to client capability and module sensitivity<sup>[<xref ref-type="bibr" rid="B35">35</xref>,<xref ref-type="bibr" rid="B52">52</xref>,<xref ref-type="bibr" rid="B69">69</xref>,<xref ref-type="bibr" rid="B70">70</xref>]</sup>. FedARA<sup>[<xref ref-type="bibr" rid="B35">35</xref>]</sup>, for example, dynamically allocates and prunes low-rank modules and reports an average 2.40-fold improvement in communication efficiency, up to 48.90% shorter total training time, and 46.95% lower energy consumption on Orin Nano. It also reports a 31.67% reduction in average peak GPU memory usage per round relative to FedLoRA. These results are obtained mainly with DistilBERT, BERT, and BART and therefore provide PLM-scale evidence for adaptive execution rather than direct feasibility evidence for multi-billion-parameter LLMs.</p>
        <p>Recent generative-LLM studies have addressed the larger memory constraint more directly. AssyLLM<sup>[<xref ref-type="bibr" rid="B43">43</xref>]</sup> assembles and adapts selected pretrained blocks to reduce the memory required for federated LLaMA-7B fine-tuning, while FedBiOT<sup>[<xref ref-type="bibr" rid="B71">71</xref>]</sup> enables federated adaptation of LLaMA-2 without requiring clients to hold the complete model. Together with the LLaMA-7B evidence reported above for FederatedScope-LLM<sup>[<xref ref-type="bibr" rid="B1">1</xref>]</sup>, these studies show that reducing trainable parameters alone is insufficient; backbone residency, activation memory, and Transformer execution remain first-order constraints for edge-side generative-LLM adaptation.</p>
        <p><bold>Data-efficient local training. </bold>Local cost also depends on how many examples undergo expensive Transformer updates. Data-efficient methods reduce this cost by selecting higher-value training samples or filtering low-quality instructions. FedDQC<sup>[<xref ref-type="bibr" rid="B40">40</xref>]</sup> filters low-quality instructions through a scoring stage that consumes roughly 1% of training time, whereas FedHDS<sup>[<xref ref-type="bibr" rid="B41">41</xref>]</sup> retains less than 1.5% of the available samples and reports 6.66- to 48.8-fold wall-clock speedups across different settings. These results extend local efficiency from choosing trainable parameters to choosing training data.</p>
        <p>In summary, PEFT reduces what is updated, resource-aware execution reduces the cost of each update, and data-efficient training reduces the number of updates performed on low-value samples. End-to-end deployability depends on combining these mechanisms with communication, aggregation, orchestration, and trust-aware design across the full federated adaptation workflow.</p>
      </sec>
      <sec id="sec4-2">
        <title>Communication-efficient exchange of structured updates and knowledge</title>
        <p>Communication efficiency in federated LLMs depends on what is exchanged, how it is encoded, and how often information is exchanged. Existing methods improve communication efficiency by exchanging compact adaptation parameters, selectively transmitting structured updates, or replacing parameter exchange with knowledge or intermediate representations.</p>
        <p><bold>Compact structured parameter exchange. </bold>The most direct approach is to exchange trainable adaptation parameters instead of full model weights. LoRA modules, adapters, and prompts preserve the conventional server-client training loop while substantially reducing the communicated object<sup>[<xref ref-type="bibr" rid="B1">1</xref>,<xref ref-type="bibr" rid="B44">44</xref>]</sup>. More compact parameterizations further redesign the adaptation object itself. FedTT and FedTT+<sup>[<xref ref-type="bibr" rid="B47">47</xref>]</sup> represent tensor-train-based federated adaptation, with FedTT+ further freezing selected factors. These approaches are most effective when clients share a pretrained backbone and compatible adaptation structures. Their end-to-end benefit should be assessed together with model initialization, downlink broadcasts, metadata, and convergence rounds.</p>
        <p><bold>Selective and compressed transmission. </bold>These methods reduce the amount of structured information sent in each exchange while keeping the adaptation representation largely unchanged. Typical strategies include low-precision encoding, sparsity, component selection, and partial transmission. EcoLoRA<sup>[<xref ref-type="bibr" rid="B48">48</xref>]</sup>, for example, retains the LoRA structure while rotating transmitted segments, exploiting matrix-specific sparsity, and compactly encoding indices. Other methods transmit selected low-rank components or active update portions<sup>[<xref ref-type="bibr" rid="B69">69</xref>,<xref ref-type="bibr" rid="B72">72</xref>]</sup>. These methods mainly optimize the scheduling and encoding of structured updates rather than redesigning the adaptation representation. Their gains depend on the balance among payload reduction, update fidelity, and the number of rounds required to reach a target utility.</p>
        <p><bold>Knowledge and intermediate-representation exchange. </bold>Instead of transmitting model parameters, these methods use task-level signals or internal representations as the communication object. Black-box language and vision-language settings may exchange discrete prompts or compact task signals<sup>[<xref ref-type="bibr" rid="B73">73</xref>-<xref ref-type="bibr" rid="B75">75</xref>]</sup>, while split systems transmit activations and gradients across model partitions<sup>[<xref ref-type="bibr" rid="B9">9</xref>,<xref ref-type="bibr" rid="B51">51</xref>]</sup>. These designs can support heterogeneous access modes and reduce parameter transmission, but they may introduce repeated queries, proxy-data dependence, distillation overhead, or frequent bidirectional communication. TITANIC<sup>[<xref ref-type="bibr" rid="B9">9</xref>]</sup> provides a useful example of this cost shift. Under its reported representation size and exchange pattern, cumulative activation and gradient traffic can exceed a compact LoRA update after only a few batches.</p>
        <p>A simple break-even condition helps clarify when split adaptation is preferable to direct PEFT exchange. For one local adaptation round, let <italic>D<sub>P</sub></italic> denote the total bidirectional PEFT-update traffic, <italic>D<sub>S</sub></italic> the bidirectional activation/gradient traffic per split interaction, and <italic>K</italic> the number of such interactions. With effective network throughput <italic>B<sub>eff</sub></italic>, the first-order completion times of PEFT and split adaptation are</p>
          <p><disp-formula><label>(1)</label> <tex-math id="E1"> $$ T_{\mathrm{P}}=T_{\text {local }}+\frac{D_{\mathrm{P}}}{B_{\text {eff }}}, \quad T_{\mathrm{S}}=T_{\mathrm{c}}+T_{\mathrm{e}}+\frac{K D_{\mathrm{S}}}{B_{\text {eff }}}, \\ $$ </tex-math></disp-formula></p>
          <p>where <italic>T<sub>local</sub></italic> is the local PEFT execution time, and <italic>T<sub>c</sub></italic> and <italic>T<sub>e</sub></italic> are the client-side and edge-side execution times under splitting, respectively. Defining the relative split-computation ratio as</p>
          <p><disp-formula><label>(2)</label> <tex-math id="E2"> $$ \rho=\frac{T_{\mathrm{c}}+T_{\mathrm{e}}}{T_{\text {local }}}, \\ $$ </tex-math></disp-formula></p>
          <p>when <italic>ρ</italic> &lt; 1 and <italic>KD<sub>S</sub></italic> > <italic>D<sub>P</sub></italic>, the break-even throughput is</p>
          <p><disp-formula><label>(3)</label> <tex-math id="E3"> $$ B^{*}=\frac{K D_{\mathrm{S}}-D_{\mathrm{P}}}{(1-\rho) T_{\text {local }}}, \\ $$ </tex-math></disp-formula></p>
          <p>and split adaptation reduces latency when <italic>B<sub>eff</sub></italic> > <italic>B</italic><sup>*</sup>. Thus, greater computation offloading lowers the required bandwidth, whereas larger or more frequent activation/gradient exchanges raise it. The threshold varies with the split point, sequence length, batch size, compression, hardware capability, and network conditions, motivating joint model-partitioning and communication-resource optimization<sup>[<xref ref-type="bibr" rid="B29">29</xref>,<xref ref-type="bibr" rid="B58">58</xref>]</sup>.</p>
          <p>Overall, communication efficiency should therefore be evaluated at the level of the complete exchange process and its computation-communication trade-off. In addition to the size of each transmitted object, assessment should include total bidirectional traffic, the number and pattern of interactions, encoding and reconstruction overhead, and the time or rounds needed to reach a comparable utility target.</p>
        </sec>
      <sec id="sec4-3">
        <title>Structure-aware aggregation and heterogeneous adaptation</title>
        <p>Aggregation becomes a resource problem when clients adopt different adaptation configurations. FedAvg is straightforward when updates share the same structure, but uniform configurations may overburden weak clients or underuse capable ones. Heterogeneous adaptation improves local feasibility by allowing clients to use different ranks, trainable layers, precisions, or adaptation modules, while introducing additional aggregation complexity. Existing methods address this problem at three levels: shape compatibility, functional equivalence, and semantic compatibility.</p>
        <p><bold>Shape compatibility. </bold>Shape-compatible methods handle heterogeneous updates through layer allocation<sup>[<xref ref-type="bibr" rid="B52">52</xref>]</sup>, zero-padding and sparsity-weighted aggregation<sup>[<xref ref-type="bibr" rid="B53">53</xref>]</sup>, padding- or knowledge-distillation-based alignment<sup>[<xref ref-type="bibr" rid="B76">76</xref>]</sup>, masking, or partial aggregation<sup>[<xref ref-type="bibr" rid="B72">72</xref>]</sup>. HETLORA<sup>[<xref ref-type="bibr" rid="B53">53</xref>]</sup>, for example, combines local rank self-pruning with server-side zero-padding, sparsity-weighted aggregation, and rank-specific truncation. These operations make updates structurally aggregatable, while functional consistency still depends on how the adapted parameters affect the model.</p>
        <p><bold>Functional equivalence. </bold>For LoRA-based adaptation, the update produced by client <italic>i </italic>is not represented by a single matrix, but by the product of two low-rank factors, i.e., ΔW<sub>i</sub>=B<sub>i</sub>A<sub>i</sub>. Therefore, the desired global update is the weighted average of the actual updates, <inline-formula><tex-math id="M1">$$ \Delta W_{\text {global }}=\sum_{i=1}^{K} p_{i} B_{i} A_{i} \\ $$</tex-math></inline-formula>, where <italic>p<sub>i</sub></italic> denotes the normalized aggregation weight. A naïve parameter-wise extension of FedAvg would instead average the factors separately, but this does not generally recover the desired global update<sup>[<xref ref-type="bibr" rid="B54">54</xref>]</sup>. The product of the averaged factors introduces cross-client combinations that were not optimized by any client<sup>[<xref ref-type="bibr" rid="B46">46</xref>,<xref ref-type="bibr" rid="B54">54</xref>]</sup>. This creates a functional mismatch: The aggregated LoRA factors may have the right shape but fail to represent the intended average model update. Methods such as FLoRA<sup>[<xref ref-type="bibr" rid="B54">54</xref>]</sup>, FedOTAB<sup>[<xref ref-type="bibr" rid="B77">77</xref>]</sup>, and FedPipe<sup>[<xref ref-type="bibr" rid="B70">70</xref>]</sup> address this issue by preserving or reconstructing LoRA updates in a functionally meaningful space rather than relying on naïve factor averaging: FLoRA<sup>[<xref ref-type="bibr" rid="B54">54</xref>]</sup> stacks factors to preserve the weighted update, FedOTAB<sup>[<xref ref-type="bibr" rid="B77">77</xref>]</sup> alternates the optimized factor to reduce communication, and FedPipe<sup>[<xref ref-type="bibr" rid="B70">70</xref>]</sup> aggregates in the ΔW space before refactorization.</p>
        <p><bold>Semantic compatibility and personalization. </bold>When clients differ in backbone, adapter type, task, or output space, parameter-level alignment may be insufficient. Related federated foundation-model methods use distillation, representation alignment, or common-subspace mapping to connect heterogeneous client knowledge<sup>[<xref ref-type="bibr" rid="B76">76</xref>,<xref ref-type="bibr" rid="B78">78</xref>,<xref ref-type="bibr" rid="B79">79</xref>]</sup>. Structural-bias-aware partial-layer tuning addresses a related layer-coverage heterogeneity problem<sup>[<xref ref-type="bibr" rid="B56">56</xref>]</sup>. Personalized designs, including dual adapters<sup>[<xref ref-type="bibr" rid="B80">80</xref>]</sup>, clustered aggregation<sup>[<xref ref-type="bibr" rid="B79">79</xref>]</sup>, and Rest-of-World LoRA<sup>[<xref ref-type="bibr" rid="B81">81</xref>]</sup>, further separate globally transferable knowledge from client-specific behavior<sup>[<xref ref-type="bibr" rid="B82">82</xref>,<xref ref-type="bibr" rid="B83">83</xref>]</sup>. These methods improve flexibility but may introduce proxy data, server inference, or multi-module management costs.</p>
        <p>Structure-aware aggregation therefore determines whether heterogeneous local adaptations can be made shape-compatible, functionally meaningful, and semantically useful. Effective aggregation should not only align update dimensions, but also preserve the model behavior contributed by different clients and support either a coherent global model or personalized local components.</p>
      </sec>
      <sec id="sec4-4">
        <title>Resource-aware system orchestration and end-edge-cloud collaboration</title>
        <p>System orchestration determines how federated LLM adaptation is mapped onto heterogeneous clients, networks, and edge/cloud infrastructure. It controls which clients participate, what adaptation workload they receive, where computation is executed, when updates are incorporated, and how communication and computing resources are allocated.</p>
        <p><bold>Participation and role assignment. </bold>Client participation can be guided by data value, resource availability, energy status, and expected completion time. Resource-aware selection should assign training roles to clients that can provide useful updates with manageable delay, energy use, or carbon emissions, while maintaining sufficient data coverage and participation fairness<sup>[<xref ref-type="bibr" rid="B59">59</xref>]</sup>. FedSustain<sup>[<xref ref-type="bibr" rid="B59">59</xref>]</sup> shows how energy-aware participation can affect both learning and system cost. Its renewable-energy-aware configuration achieved 73.8% accuracy with estimated 13 kWh energy use and 9 kg CO₂ emissions, compared with 71.2%, 21 kWh, and 14 kg CO₂ under random selection. This result shows that sustainability-aware participation can jointly improve learning utility and resource efficiency when client selection is aligned with energy and deployment conditions.</p>
        <p><bold>Workload and synchronization control.</bold> Orchestration determines how much work each selected client performs and how this workload is coordinated with available communication resources. LoRA rank, trainable layers, numerical precision, sparsity<sup>[<xref ref-type="bibr" rid="B37">37</xref>,<xref ref-type="bibr" rid="B52">52</xref>,<xref ref-type="bibr" rid="B69">69</xref>]</sup>, local training time, and participation<sup>[<xref ref-type="bibr" rid="B62">62</xref>]</sup> can be adapted jointly with communication and computing resources<sup>[<xref ref-type="bibr" rid="B60">60</xref>,<xref ref-type="bibr" rid="B84">84</xref>-<xref ref-type="bibr" rid="B86">86</xref>]</sup>. HierFedLoRA<sup>[<xref ref-type="bibr" rid="B57">57</xref>]</sup> provides a representative example by combining near-IID grouping, group-specific aggregation frequency, and fine-tuning depth across 80 Jetson devices, optimizing time-to-accuracy rather than single-round efficiency. Dynamic network conditions further couple these decisions with synchronization. Channel fading, interference, and bandwidth variation can change client communication latency over time, causing communication stragglers in synchronous FL or increased update staleness in asynchronous FL. Accordingly, client participation, bandwidth allocation, adaptation workload, and synchronization frequency can be adjusted according to both device and channel states. Wang <italic>et al</italic>.<sup>[<xref ref-type="bibr" rid="B58">58</xref>]</sup> consider device scheduling and bandwidth allocation under time-varying wireless conditions; Pang <italic>et al</italic>.<sup>[<xref ref-type="bibr" rid="B84">84</xref>]</sup> jointly optimize client-specific pruning and bandwidth allocation for low-latency federated LLM fine-tuning, and Zhang <italic>et al</italic>.<sup>[<xref ref-type="bibr" rid="B87">87</xref>]</sup> adapt the communicated knowledge according to available channel resources. More broadly, hierarchical and asynchronous coordination can reduce cloud traffic or waiting time, but their benefits depend on how network variability, staleness, grouping overhead, and controller complexity are managed.</p>
        <p><bold>Computation placement and end-edge-cloud collaboration. </bold>Split and offloaded training move part of the adaptation workload from devices to edge or cloud servers. By placing selected Transformer blocks or backward computation outside the client, these methods can reduce client-side memory and computation, while shifting part of the cost to activation exchange, edge/cloud load, and coordination overhead<sup>[<xref ref-type="bibr" rid="B61">61</xref>,<xref ref-type="bibr" rid="B88">88</xref>-<xref ref-type="bibr" rid="B90">90</xref>]</sup>.</p>
        <p>Related wireless split-learning and cloud-edge FL studies provide broader evidence that computation placement and representation design can reshape communication and computing costs across heterogeneous networks<sup>[<xref ref-type="bibr" rid="B30">30</xref>,<xref ref-type="bibr" rid="B91">91</xref>]</sup>. In wireless split-federated LLM systems, these trade-offs become directly coupled with LLM adaptation and communication-resource control. Existing studies jointly optimize split points, LoRA ranks, bandwidth, transmit power, subchannel allocation, or multiple-access schemes<sup>[<xref ref-type="bibr" rid="B61">61</xref>,<xref ref-type="bibr" rid="B88">88</xref>,<xref ref-type="bibr" rid="B89">89</xref>]</sup>, and some designs further incorporate anti-jamming, sensing assistance, or privacy constraints<sup>[<xref ref-type="bibr" rid="B92">92</xref>-<xref ref-type="bibr" rid="B94">94</xref>]</sup>.</p>
        <p><bold>Decentralized and multi-participant collaboration.</bold> Federated LLM adaptation can extend beyond conventional server-coordinated aggregation to direct collaboration among multiple clients or model instances. Dec-LoRA<sup>[<xref ref-type="bibr" rid="B32">32</xref>]</sup> coordinates decentralized LoRA updates through topology-based peer interaction, while Balija <italic>et al</italic>.<sup>[<xref ref-type="bibr" rid="B33">33</xref>]</sup> consider asynchronous peer-to-peer model exchange. Such decentralized designs can reduce dependence on a central aggregator and alleviate communication congestion at the central node, but shift resource costs toward repeated peer exchange, heterogeneous Device-to-Device (D2D) links, topology management, synchronization, and distributed coordination. Compact PEFT exchange can reduce peer-transfer payloads, while topology-aware peer selection and clustered communication can limit unnecessary transfers. Asynchronous coordination can further reduce waiting for slow or intermittently connected peers, although update staleness must be controlled.</p>
        <p>Beyond client-level decentralized federation, broader multi-participant training can involve collaboration among multiple model agents. MAPoRL<sup>[<xref ref-type="bibr" rid="B95">95</xref>]</sup>, for example, studies multi-agent post-training in which multiple LLMs generate, exchange, and discuss responses during joint optimization. Such collaboration introduces additional interaction rounds, token processing, peer communication, and coordination overhead.</p>
        <p>Resource-aware orchestration therefore goes beyond client selection by jointly designing participation, workload configuration, synchronization, computation placement, and resource allocation across the end-edge-cloud continuum.</p>
      </sec>
      <sec id="sec4-5">
        <title>Trustworthy resource efficiency</title>
        <p>Privacy<sup>[<xref ref-type="bibr" rid="B96">96</xref>,<xref ref-type="bibr" rid="B97">97</xref>]</sup>, security, and lifecycle trust constraints reshape the resource-efficiency problem in federated LLM adaptation. Protection mechanisms<sup>[<xref ref-type="bibr" rid="B98">98</xref>,<xref ref-type="bibr" rid="B99">99</xref>]</sup> improve the reliability of distributed adaptation<sup>[<xref ref-type="bibr" rid="B100">100</xref>-<xref ref-type="bibr" rid="B102">102</xref>]</sup>, but they also consume the same computation, communication, storage, and utility budgets required for training and aggregation<sup>[<xref ref-type="bibr" rid="B103">103</xref>]</sup>.</p>
        <p><bold>Privacy-aware resource efficiency. </bold>Privacy-aware mechanisms reduce leakage from shared updates and intermediate representations. Differential privacy protects updates through clipping and perturbation<sup>[<xref ref-type="bibr" rid="B10">10</xref>]</sup>, while encrypted or homomorphic aggregation strengthens confidentiality at the cost of additional communication and cryptographic processing<sup>[<xref ref-type="bibr" rid="B65">65</xref>,<xref ref-type="bibr" rid="B98">98</xref>]</sup>. Selective perturbation can reduce utility loss by protecting sensitive components<sup>[<xref ref-type="bibr" rid="B96">96</xref>,<xref ref-type="bibr" rid="B97">97</xref>,<xref ref-type="bibr" rid="B99">99</xref>]</sup>, and reconstruction risk depends on what information the protocol exposes, such as gradients, adapters, activations, or split representations<sup>[<xref ref-type="bibr" rid="B63">63</xref>,<xref ref-type="bibr" rid="B64">64</xref>]</sup>. The resource challenge is to provide sufficient privacy protection while controlling communication, computation, and convergence overhead.</p>
        <p><bold>Security-aware resource efficiency. </bold>Security-aware mechanisms address malicious or biased contributions that may distort the shared model, introduce backdoors, or weaken robustness<sup>[<xref ref-type="bibr" rid="B67">67</xref>,<xref ref-type="bibr" rid="B68">68</xref>,<xref ref-type="bibr" rid="B104">104</xref>]</sup>. Robust aggregation, poisoning detection, and adversarially robust prompt-tuning can improve resilience, while adding server-side computation, screening latency, or extra optimization stages<sup>[<xref ref-type="bibr" rid="B66">66</xref>,<xref ref-type="bibr" rid="B105">105</xref>,<xref ref-type="bibr" rid="B106">106</xref>]</sup>. Related security-aware resource-allocation studies in broader cloud-edge-terminal systems also show that security requirements can be coupled with communication and computing-resource allocation<sup>[<xref ref-type="bibr" rid="B107">107</xref>]</sup>. This provides a complementary system-level perspective on security-resource co-design for federated LLMs. In the Shakespeare next-token-prediction experiment, GPT-2 Medium was evaluated with 100 clients and 10 selected per round; aggregation required approximately 3.9 s for a perturbation-based defense, 27.03 s for Multi-Krum, and 63.20 s for FLAME. The sentence-trigger Neurotoxin variant could nevertheless evade both screening methods<sup>[<xref ref-type="bibr" rid="B66">66</xref>]</sup>. Security-aware resource efficiency requires robustness improvements to be evaluated alongside aggregation latency, server-side computation, and learning utility.</p>
        <p><bold>Trust- and lifecycle-aware resource efficiency. </bold>Trust- and lifecycle-aware mechanisms focus on whether the federated adaptation process is verifiable, accountable, and auditable over time. Verifiable aggregation strengthens confidence that structured updates are processed correctly, but introduces additional computation and verification overhead. FLAGuard<sup>[<xref ref-type="bibr" rid="B11">11</xref>]</sup>, for example, improves the efficiency of LoRA verification by more than two orders of magnitude relative to prior schemes, with approximately 8% overhead over unverified FLoRA. Provenance tracking and federated unlearning extend trust beyond the training round by requiring historical states, audit trails, rollback, retraining, or deletion verification<sup>[<xref ref-type="bibr" rid="B12">12</xref>]</sup>. The resource challenge is to integrate verification, provenance, and unlearning with compact federated adaptation, so that accountability can be supported without excessive storage or retraining cost.</p>
        <p>Privacy, security, and lifecycle trust protect different aspects of federated LLM adaptation, but they draw on the same resource budgets as training and communication. Resource-efficient design should therefore optimize protection strength, learning utility, and system cost jointly.</p>
      </sec>
      <sec id="sec4-6">
        <title>Lessons learned</title>
        <p>Across these mechanism classes, local adaptation reduces the cost of producing client updates; structured exchange reduces communication payload; structure-aware aggregation preserves the meaning of heterogeneous updates; orchestration matches workloads to available resources; and privacy, security, and trust mechanisms define the protection cost required for reliable adaptation. Resource savings should therefore be assessed over the full workflow. A reduction in trainable parameters, message size, or client memory is valuable only when it also improves end-to-end efficiency at a comparable utility target. <xref ref-type="table" rid="t4">Table 4</xref> summarizes representative methods according to their mechanism, primary efficiency contribution, and main limitation.</p>
        <table-wrap id="t4">
        <label>Table 4</label>
        <caption>
          <p>Comparison of representative mechanisms for resource-efficient federated LLM adaptation</p>
        </caption>
        <table frame="hsides" rules="groups">
          <thead>
            <tr>
              <td style="border-bottom:1;">
                <bold>Mechanism class</bold>
              </td>
              <td style="border-bottom:1;">
                <bold>Representative method</bold>
              </td>
              <td style="border-bottom:1;">
                <bold>Mechanism</bold>
              </td>
              <td style="border-bottom:1;">
                <bold>Reported method-specific efficiency evidence</bold>
              </td>
              <td style="border-bottom:1;">
                <bold>Remaining limitation</bold>
              </td>
            </tr>
          </thead>
          <tbody>
            <tr>
              <td rowspan="3">Local adaptation</td>
              <td>FederatedScope-LLM<sup>[<xref ref-type="bibr" rid="B1">1</xref>]</sup></td>
              <td>LoRA/prompt update exchange</td>
              <td>Serialized-adapter message size for one server-client communication: 21.40 MB with LoRA and 0.17 MB with prompt tuning (LLaMA-7B)</td>
              <td>Approximately 13.4 GB model-only memory; backbone execution remains</td>
            </tr>
            <tr>
              <td>FedARA<sup>[<xref ref-type="bibr" rid="B35">35</xref>]</sup></td>
              <td>Dynamic rank allocation and rank-based module pruning</td>
              <td>2.40× average communication-efficiency improvement; up to 48.90% shorter total training time relative to FedLoRA; on Orin Nano (15 W), 46.95% lower energy consumption relative to FedLoRA over 100 rounds with 10 clients per round</td>
              <td>Rank-based module pruning reduces average peak GPU memory usage per round by 31.67% relative to FedLoRA; based on DistilBERT, BERT, and BART</td>
            </tr>
            <tr>
              <td>FedDQC/FedHDS<sup>[<xref ref-type="bibr" rid="B40">40</xref>-<xref ref-type="bibr" rid="B41">41</xref>]</sup></td>
              <td>Quality and representativeness selection</td>
              <td>FedDQC scoring consumes approximately 1% of training time; FedHDS variants use less than 1.5% of the data with up to 48.8× speedup</td>
              <td>Coverage and robustness depend on the quality-scoring and subset-selection criteria</td>
            </tr>
            <tr>
              <td rowspan="3">Structured exchange</td>
              <td>FedTT/FedTT+<sup>[<xref ref-type="bibr" rid="B47">47</xref>]</sup></td>
              <td>Tensor-train factorization and adaptive factor freezing</td>
              <td>Approximately 10× lower communication overhead for FedTT and 30× for FedTT+ in the reported LLaMA2-13B cross-silo setting</td>
              <td>Compatible backbone and adapter structures are required; evidence remains specific to tensor-train parameterization</td>
            </tr>
            <tr>
              <td>EcoLoRA<sup>[<xref ref-type="bibr" rid="B48">48</xref>]</sup></td>
              <td>Rotating sparse segments and indices</td>
              <td>Up to 79% less communication time and 65% less total training time under the reported 1/5 Mbps uplink/downlink setting</td>
              <td>Potential staleness and fidelity-round trade-off; limited real-device evidence</td>
            </tr>
            <tr>
              <td>TITANIC<sup>[<xref ref-type="bibr" rid="B9">9</xref>]</sup></td>
              <td>Activation/gradient partitioning</td>
              <td>Lower client residency and execution burden</td>
              <td>Frequent bidirectional traffic; batch- and placement-sensitive cost</td>
            </tr>
            <tr>
              <td rowspan="3">Structure-aware aggregation</td>
              <td>HETLORA<sup>[<xref ref-type="bibr" rid="B53">53</xref>]</sup></td>
              <td>Rank self-pruning, zero-padding, and sparsity-weighted aggregation</td>
              <td>Faster convergence with reduced computation and communication</td>
              <td>Shape compatibility does not ensure functional equivalence</td>
            </tr>
            <tr>
              <td>FLoRA<sup>[<xref ref-type="bibr" rid="B54">54</xref>]</sup></td>
              <td>Function-preserving factor stacking</td>
              <td>Preserves the weighted Δ<italic>W</italic> update</td>
              <td>Global rank grows with participating ranks</td>
            </tr>
            <tr>
              <td>FedPipe/FedOTAB<sup>[<xref ref-type="bibr" rid="B70">70</xref>,<xref ref-type="bibr" rid="B77">77</xref>]</sup></td>
              <td>ΔW-space refactorization or alternating one-factor optimization/transmission</td>
              <td>Function-preserving aggregation; half-LoRA transmission with FedOTAB</td>
              <td>Additional server reconstruction or optimization</td>
            </tr>
            <tr>
              <td rowspan="3">System orchestration</td>
              <td>FedSustain<sup>[<xref ref-type="bibr" rid="B59">59</xref>]</sup></td>
              <td>Utility- and renewable-aware scheduling</td>
              <td>Up to 38% lower energy use and 46% lower CO<sub>2</sub> emissions</td>
              <td>Deployment-, grid-, and renewable-dependent gains</td>
            </tr>
            <tr>
              <td>HierFedLoRA<sup>[<xref ref-type="bibr" rid="B57">57</xref>]</sup></td>
              <td>Adaptive grouping, depth, and frequency</td>
              <td>At least 2.1× faster fine-tuning and 1.6%-4.2% higher final accuracy</td>
              <td>Grouping, staleness, and controller overhead</td>
            </tr>
            <tr>
              <td>Split-federated LLMs<sup>[<xref ref-type="bibr" rid="B61">61</xref>,<xref ref-type="bibr" rid="B88">88</xref>,<xref ref-type="bibr" rid="B89">89</xref>]</sup></td>
              <td>Split-point, rank, and/or radio-resource co-design</td>
              <td>For FedsLLM, approximately 47.63% lower training delay on average than the BA strategy (which optimizes neither <italic>η</italic> nor bandwidth) in the reported MATLAB simulation</td>
              <td>Repeated transfers of activations and gradients, server load, synchronization, and idealized channels</td>
            </tr>
            <tr>
              <td rowspan="3">Trustworthy resource efficiency</td>
              <td>DP and privacy-aware adaptation<sup>[<xref ref-type="bibr" rid="B10">10</xref>,<xref ref-type="bibr" rid="B96">96</xref>-<xref ref-type="bibr" rid="B97">97</xref>]</sup></td>
              <td>Calibrated perturbation, selective protection, and privacy-aware alignment</td>
              <td>Formal DP guarantees or targeted leakage mitigation</td>
              <td>Utility loss, extra rounds, or narrower formal coverage</td>
            </tr>
            <tr>
              <td>FLAGuard/FedHE<sup>[<xref ref-type="bibr" rid="B11">11</xref>,<xref ref-type="bibr" rid="B65">65</xref>]</sup></td>
              <td>LoRA verification or CKKS encryption</td>
              <td>More than 100× faster verification with FLAGuard; encrypted aggregation with FedHE</td>
              <td>Cryptographic overhead; FedHE is limited to compact BERT</td>
            </tr>
            <tr>
              <td>Federated TrustChain<sup>[<xref ref-type="bibr" rid="B12">12</xref>]</sup></td>
              <td>Provenance and unlearning</td>
              <td>Auditable training and unlearning workflows</td>
              <td>Ledger, rollback, and retraining costs; formal deletion guarantees remain open</td>
            </tr>
          </tbody>
        </table>
        <table-wrap-foot>
          <fn>
            <p>BERT: Bidirectional Encoder Representations from Transformers; CKKS: Cheon-Kim-Kim-Song homomorphic encryption scheme; CO<sub>2</sub>: carbon dioxide; DP: differential privacy; GPU: graphics processing unit; Δ<italic>W</italic>: model-weight update; LLM: large language model.</p>
          </fn>
        </table-wrap-foot>
      </table-wrap>
      </sec>
    </sec>
    <sec id="sec5">
      <title>EVALUATION REQUIREMENTS AND OPEN RESEARCH DIRECTIONS</title>
      <p>Current evidence on resource-efficient federated LLMs remains fragmented across model scales, hardware platforms, network assumptions, and reporting conventions. Progress requires end-to-end evaluation at comparable utility targets, together with research that addresses adaptive system control, realistic heterogeneity, and emerging LLM workloads.</p>
      <sec id="sec5-1">
        <title>End-to-end evaluation and resource trade-offs</title>
        <p>Resource-efficient federated LLM methods should be evaluated by the total cost required to reach a comparable utility target. A complete accounting should include backbone distribution, client-side memory and computation<sup>[<xref ref-type="bibr" rid="B1">1</xref>]</sup>, bidirectional traffic, server-side processing, orchestration and synchronization overhead<sup>[<xref ref-type="bibr" rid="B57">57</xref>,<xref ref-type="bibr" rid="B59">59</xref>]</sup>, and privacy/security/trust costs<sup>[<xref ref-type="bibr" rid="B11">11</xref>,<xref ref-type="bibr" rid="B65">65</xref>]</sup>. Studies should report time, energy, and communication cost at comparable utility levels, together with task performance, personalization quality, tail-client performance, fairness, privacy, and robustness.</p>
        <p>Transparent reporting is also needed for fair comparison. Key information includes the backbone model, adaptation object, client participation pattern, heterogeneity setting<sup>[<xref ref-type="bibr" rid="B35">35</xref>,<xref ref-type="bibr" rid="B37">37</xref>,<xref ref-type="bibr" rid="B52">52</xref>]</sup>, hardware platform, network assumptions, peak memory, transferred bytes, wall-clock time, and energy measurement or estimation method<sup>[<xref ref-type="bibr" rid="B57">57</xref>,<xref ref-type="bibr" rid="B70">70</xref>]</sup>. Privacy and security studies should further specify the threat model, server-visible information, and protected object. Resource claims should be labeled as directly measured, modeled or estimated, proxy-based, or unevaluated, while formal guarantees should be reported separately from empirical efficiency results.</p>
        <p>Wireless evaluations should additionally report the adopted channel model or network trace, uplink and downlink variability, reliability and retransmission assumptions, synchronization protocol, and mean and tail communication latency. For asynchronous or partial aggregation, update staleness and acceptance or timeout policies should also be specified. These factors are necessary to distinguish performance under nominal link configurations from resource efficiency under realistic time-varying communication conditions.</p>
        <p><bold>Utility-normalized Pareto benchmarking.</bold> Motivated by multi-objective resource evaluation in FL<sup>[<xref ref-type="bibr" rid="B108">108</xref>]</sup>, we propose a utility-normalized Pareto framework. Let <italic>U<sup>*</sup></italic> denote a task-specific target utility and <inline-formula><tex-math id="M2">$$ T_{i}^{*} $$</tex-math></inline-formula> the wall-clock time required by method <italic>i</italic> to first reach this target. For each method that reaches <italic>U<sup>*</sup></italic> within the evaluation budget, we define its resource vector at the target utility as</p>
          <p><disp-formula><label>(4)</label> <tex-math id="E4"> $$ R_{i}\left(U^{*}\right)=\left(T_{i}^{*}, E_{i, \text { client }}, E_{i, \text { server }}, C_{i, \text { up }}, C_{\mathrm{i}, \text { down }}, M_{\mathrm{i}, \text { peak }}\right), $$ </tex-math></disp-formula></p>
          <p>where <italic>E<sub>i,client</sub></italic> and <italic>E<sub>i,server</sub> </italic>are the accumulated client- and server/edge-side energy, respectively; <italic>C<sub>i,up</sub></italic> and <italic>C<sub>i,down</sub></italic> are the cumulative uplink and downlink traffic, respectively; and <italic>M<sub>i,peak</sub></italic> is the peak memory, all measured up to <inline-formula><tex-math id="M3">$$ T_{i}^{*} $$</tex-math></inline-formula>. At the same <italic>U<sup>*</sup></italic>, method <italic>A </italic>Pareto-dominates method <italic>B</italic> if it is no worse in every reported resource dimension and strictly better in at least one. Methods that do not reach the target should instead report their terminal utility and budget-end resource vector. Fair comparisons should use the same task, backbone, data partition, client-participation setting, hardware/network scenario, and resource-accounting boundary. This framework avoids labeling a method resource-efficient when savings in one dimension are mainly obtained by shifting cost to another.</p>
          <p>Several trade-offs recur across the literature. Lower ranks, lower bit widths, or smaller selected training subsets reduce local cost but may affect utility or convergence. Heterogeneous configurations improve participation flexibility but complicate functionally correct aggregation. Privacy, security, and lifecycle-trust mechanisms improve protection while consuming bandwidth, computation, storage, or utility. Favoring fast or well-provisioned clients can improve average efficiency, but may reduce data coverage and participation fairness<sup>[<xref ref-type="bibr" rid="B55">55</xref>,<xref ref-type="bibr" rid="B59">59</xref>]</sup>. Methods should therefore be compared through utility-normalized and Pareto-aware evaluation rather than a single favorable metric.</p>
        </sec>
      <sec id="sec5-2">
        <title>Research priorities for deployable federated LLMs</title>
        <p><bold>Adaptive cross-layer control. </bold>Federated LLM systems require joint control across model, communication, system, and trust layers. Future methods should adapt participation<sup>[<xref ref-type="bibr" rid="B60">60</xref>,<xref ref-type="bibr" rid="B85">85</xref>]</sup>, PEFT configuration, exchange strategy, aggregation rule<sup>[<xref ref-type="bibr" rid="B35">35</xref>,<xref ref-type="bibr" rid="B37">37</xref>,<xref ref-type="bibr" rid="B69">69</xref>]</sup>, computation placement<sup>[<xref ref-type="bibr" rid="B60">60</xref>,<xref ref-type="bibr" rid="B85">85</xref>]</sup>, and protection level<sup>[<xref ref-type="bibr" rid="B94">94</xref>]</sup> according to changing memory, bandwidth, energy, data value, and risk conditions. Such cross-layer control is important because savings in one stage may otherwise reappear as additional rounds, server reconstruction, communication traffic, or protection overhead.</p>
        <p><bold>Realistic-scale validation and joint heterogeneity. </bold>Current evidence remains concentrated on BERT-family models, small client populations, homogeneous backbones, and simulated network settings. Future validation should extend to larger generative LLMs, heterogeneous devices<sup>[<xref ref-type="bibr" rid="B35">35</xref>]</sup>, dynamic client availability, variable bandwidth<sup>[<xref ref-type="bibr" rid="B61">61</xref>,<xref ref-type="bibr" rid="B88">88</xref>,<xref ref-type="bibr" rid="B89">89</xref>]</sup>, asynchronous participation, and explicit client- and server-side energy accounting. Aggregation also remains challenging when clients differ simultaneously in backbone, LoRA rank, layer coverage, adapter type, precision, or model-access mode<sup>[<xref ref-type="bibr" rid="B53">53</xref>-<xref ref-type="bibr" rid="B54">54</xref>,<xref ref-type="bibr" rid="B56">56</xref>,<xref ref-type="bibr" rid="B78">78</xref>]</sup>. Addressing such joint heterogeneity is essential for moving from controlled experiments to deployable federated LLM systems.</p>
        <p><bold>Lifecycle-aware and emerging post-training settings.</bold> Future federated LLM systems will increasingly face diverse post-training workloads and lifecycle constraints. Multimodal and agentic workloads add visual representations<sup>[<xref ref-type="bibr" rid="B109">109</xref>-<xref ref-type="bibr" rid="B111">111</xref>]</sup>, persistent memory, planning, and repeated interaction. In multi-model or multi-agent deployment, additional collaboration costs arise from agent participation, repeated model interaction, token processing, dialogue-history management, and service latency. Wang <italic>et al</italic>.<sup>[<xref ref-type="bibr" rid="B112">112</xref>]</sup>, for example, study agent-level collaboration strategies and their accuracy-token-cost trade-offs, while large-small model collaboration<sup>[<xref ref-type="bibr" rid="B31">31</xref>]</sup> illustrates cooperation between models with different capabilities across edge resources. Personalization<sup>[<xref ref-type="bibr" rid="B81">81</xref>]</sup>, provenance and unlearning<sup>[<xref ref-type="bibr" rid="B12">12</xref>]</sup>, and renewable-energy-aware operation<sup>[<xref ref-type="bibr" rid="B59">59</xref>,<xref ref-type="bibr" rid="B113">113</xref>]</sup> further extend resource costs beyond a single adaptation round. Evaluation should therefore consider not only federated adaptation but also the downstream collaboration and operational costs over the lifetime of the adapted model.</p>
      </sec>
    </sec>
    <sec id="sec6">
      <title>CONCLUSION AND OUTLOOK</title>
      <p>Federated LLM adaptation changes the resource profile of FL by altering what is trained, what is exchanged, how updates are aggregated, where computation is placed, and which protection mechanisms are required. This review shows that resource efficiency cannot be inferred from trainable parameter count or per-round payload alone. PEFT, structured exchange, heterogeneous aggregation, system orchestration, and privacy/security/trust mechanisms reduce different parts of the federated adaptation cost, but their gains may also shift costs to additional communication rounds, repeated transfer of intermediate activations and gradients, server-side processing, coordination, or protection.</p>
      <p>Current evidence remains uneven across model scales, client populations, hardware platforms, network settings, and reporting conventions. Future research should therefore emphasize utility-normalized end-to-end accounting, function-preserving aggregation under joint model and resource heterogeneity, adaptive cross-layer control, and realistic evaluation of energy, privacy, security, and lifecycle costs. As federated LLMs move toward larger generative models, multimodal and agentic workloads, and continuously personalized services, resource efficiency should be assessed over the full operational lifetime of the system rather than within a single training round.</p>
    </sec>
  </body>
  <back>
    <sec>
      <title>DECLARATIONS</title>
      <sec>
        <title>Authors’ contributions</title>
        <p>Made substantial contributions to the conception and design of the review, literature search and screening, analysis and synthesis of the literature, and manuscript drafting: Chen, X.; Wang, C.; Li, B.</p>
        <p>Provided supervision, critical revision of the manuscript, and administrative, technical, and material support: Wang, X.</p>
      </sec>
      <sec>
        <title>Availability of data and materials</title>
        <p>Not applicable.</p>
      </sec>
      <sec>
        <title>AI and AI-assisted tools statement</title>
        <p>During the preparation of this manuscript, the AI tool ChatGPT (version GPT-5.4, released 2026-03-05) was used for language editing and to assist in generating illustrative icons for <xref ref-type="fig" rid="fig1">Figures 1</xref>-<xref ref-type="fig" rid="fig3">3</xref> based on textual descriptions provided by the authors. In the Graphical Abstract, the AI-assisted icons include those representing distributed private data, LLM-scale adaptation workloads, heterogeneous resources, local adaptation, structured exchange, heterogeneous aggregation, system orchestration, and trustworthy resource efficiency. The scientific concepts, taxonomy, text, logical relationships, and overall layout were developed by the authors, who subsequently selected, edited, and arranged the illustrative icons. The tool did not influence the study design, data collection, analysis, interpretation, or the scientific content of the work. All authors take full responsibility for the accuracy, integrity, and final content of the manuscript.</p>
      </sec>
      <sec>
        <title>Financial support and sponsorship</title>
        <p>This work was supported by the National Key R&amp;D Program of China (No. 2022YFB2902303) and Shanghai Municipal Science and Technology Commission Foundation (No. 25DP1500300).</p>
      </sec>
      <sec>
        <title>Conflicts of interest</title>
        <p>All authors declared that there are no conflicts of interest.</p>
      </sec>
      <sec>
        <title>Ethical approval and consent to participate</title>
        <p>Not applicable.</p>
      </sec>
      <sec>
        <title>Consent for publication</title>
        <p>Not applicable.</p>
      </sec>
      <sec>
        <title>Copyright</title>
        <p>© The Author(s) 2026.</p>
      </sec>
    </sec>
    <ref-list>
      <ref id="B1">
        <label>1</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Kuang</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Qian</surname>
              <given-names>B</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>Z</given-names>
            </name>
            <etal />
          </person-group>
          <comment>FederatedScope-LLM: A comprehensive package for fine-tuning large language models in federated learning. KDD '24: The 30th ACM SIGKDD Conference on Knowledge Discovery and Data Mining; Barcelona Spain. New York, NY, USA: ACM; 2024. pp. 5260-71.</comment>
          <pub-id pub-id-type="doi">10.1145/3637528.3671573</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B2">
        <label>2</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Hu</surname>
              <given-names>EJ</given-names>
            </name>
            <name>
              <surname>Shen</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Wallis</surname>
              <given-names>P</given-names>
            </name>
            <etal />
          </person-group>
          <comment>LoRA: Low-rank adaptation of large language models. The International Conference on Learning Representations; 2022 April 25-29; <uri xlink:href="https://openreview.net/forum?id=nZeVKeeFYf9">https://openreview.net/forum?id=nZeVKeeFYf9</uri> (accessed 2026-09-14)</comment>
        </nlm-citation>
      </ref>
      <ref id="B3">
        <label>3</label>
        <nlm-citation publication-type="web">
          <person-group person-group-type="author">
            <name>
              <surname>Lialin</surname>
              <given-names>V</given-names>
            </name>
            <name>
              <surname>Deshpande</surname>
              <given-names>V</given-names>
            </name>
            <name>
              <surname>Yao</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Rumshisky</surname>
              <given-names>A</given-names>
            </name>
          </person-group>
          <comment>Scaling down to scale up: a guide to parameter-efficient fine-tuning. arXiv 2023, arXiv:2303.15647. Available online: <uri xlink:href="https://arxiv.org/abs/2303.15647">https://arxiv.org/abs/2303.15647</uri></comment>
        </nlm-citation>
      </ref>
      <ref id="B4">
        <label>4</label>
        <nlm-citation publication-type="confproc">
          <comment>McMahan, B.; Moore, E.; Ramage, D.; Hampson, S.; Agüera y Arcas, B. Communication-efficient learning of deep networks from decentralized data. In Proceedings of the 20th International Conference on Artificial Intelligence and Statistics; 2017 April 20-22; Ft. Lauderdale, FL USA. ACM ICPS; 2017. pp 1273-82. <uri xlink:href="https://proceedings.mlr.press/v54/mcmahan17a.html">https://proceedings.mlr.press/v54/mcmahan17a.html</uri> (accessed 2026-09-14)</comment>
        </nlm-citation>
      </ref>
      <ref id="B5">
        <label>5</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Kairouz</surname>
              <given-names>P</given-names>
            </name>
            <name>
              <surname>Mcmahan</surname>
              <given-names>HB</given-names>
            </name>
          </person-group>
          <article-title>Advances and open problems in federated learning</article-title>
          <source>Found Trends Mach Learn.</source>
          <year>2021</year>
          <volume>14</volume>
          <fpage>1</fpage>
          <lpage>210</lpage>
          <pub-id pub-id-type="doi">10.1561/2200000083</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B6">
        <label>6</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Zhang</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Xie</surname>
              <given-names>G</given-names>
            </name>
            <name>
              <surname>Huang</surname>
              <given-names>Y</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Edge intelligence in the generative artificial intelligence era</article-title>
          <source>IEEE Wireless Commun.</source>
          <year>2025</year>
          <volume>32</volume>
          <fpage>60</fpage>
          <lpage>8</lpage>
          <pub-id pub-id-type="doi">10.1109/mwc.2025.3599652</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B7">
        <label>7</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Lim</surname>
              <given-names>WYB</given-names>
            </name>
            <name>
              <surname>Luong</surname>
              <given-names>NC</given-names>
            </name>
            <name>
              <surname>Hoang</surname>
              <given-names>DT</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Federated learning in mobile edge networks: a comprehensive survey</article-title>
          <source>IEEE Commun. Surv. Tutorials.</source>
          <year>2020</year>
          <volume>22</volume>
          <fpage>2031</fpage>
          <lpage>63</lpage>
          <pub-id pub-id-type="doi">10.1109/comst.2020.2986024</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B8">
        <label>8</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Nguyen</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Malik</surname>
              <given-names>K</given-names>
            </name>
            <name>
              <surname>Zhan</surname>
              <given-names>H</given-names>
            </name>
            <etal />
          </person-group>
          <comment>Federated learning with buffered asynchronous aggregation. 25th International Conference on Artificial Intelligence and Statistics; 2022 Mar 28-30; Valencia, Spain. 2022, pp 3581-607. <uri xlink:href="https://proceedings.mlr.press/v151/nguyen22b.html">https://proceedings.mlr.press/v151/nguyen22b.html</uri> (accessed 2026-09-14)</comment>
        </nlm-citation>
      </ref>
      <ref id="B9">
        <label>9</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Su</surname>
              <given-names>N</given-names>
            </name>
            <name>
              <surname>Hu</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>B</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>B</given-names>
            </name>
          </person-group>
          <comment>Titanic: towards production federated learning with large language models. IEEE INFOCOM 2024 - IEEE Conference on Computer Communications; 2024 May 20-23; Vancouver, BC, Canada. IEEE; 2024. pp. 611-20.</comment>
          <pub-id pub-id-type="doi">10.1109/infocom52122.2024.10621164</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B10">
        <label>10</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Liu</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Zhu</surname>
              <given-names>R</given-names>
            </name>
            <name>
              <surname>Zha</surname>
              <given-names>D</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Differentially private low-rank adaptation of large language model using federated learning</article-title>
          <source>ACM Trans Manage Inf Syst.</source>
          <year>2025</year>
          <volume>16</volume>
          <fpage>1</fpage>
          <lpage>24</lpage>
          <pub-id pub-id-type="doi">10.1145/3682068</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B11">
        <label>11</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Zhang</surname>
              <given-names>T</given-names>
            </name>
            <name>
              <surname>Yu</surname>
              <given-names>H</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Yang</surname>
              <given-names>Z</given-names>
            </name>
          </person-group>
          <article-title>FLAGuard: efficient verifiable federated LoRA of large language models</article-title>
          <source>IEEE Trans Mobile Comput.</source>
          <year>2026</year>
          <volume>25</volume>
          <fpage>7182</fpage>
          <lpage>95</lpage>
          <pub-id pub-id-type="doi">10.1109/tmc.2025.3641570</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B12">
        <label>12</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Zuo</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Zhu</surname>
              <given-names>T</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Federated TrustChain: blockchain-enhanced LLM training and unlearning</article-title>
          <source>IEEE Trans. Dependable and Secure Comput.</source>
          <year>2026</year>
          <volume>23</volume>
          <fpage>6457</fpage>
          <lpage>73</lpage>
          <pub-id pub-id-type="doi">10.1109/tdsc.2026.3665277</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B13">
        <label>13</label>
        <nlm-citation publication-type="web">
          <person-group person-group-type="author">
            <name>
              <surname>Shahid</surname>
              <given-names />
            </name>
            <name>
              <surname>A</surname>
              <given-names />
            </name>
          </person-group>
          <comment>; Kliks, A.; Al-Tahmeesschi, A.; et al. Large-scale AI in telecom: charting the roadmap for innovation, scalability, and enhanced digital experiences. arXiv 2025, arXiv:2503.04184. Available online: <uri xlink:href="https://doi.org/10.48550/arXiv.2503.04184">https://doi.org/10.48550/arXiv.2503.04184</uri></comment>
        </nlm-citation>
      </ref>
      <ref id="B14">
        <label>14</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Hu</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>D</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>Z</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Federated large language model: solutions, challenges and future directions</article-title>
          <source>IEEE Wireless Commun.</source>
          <year>2025</year>
          <volume>32</volume>
          <fpage>82</fpage>
          <lpage>9</lpage>
          <pub-id pub-id-type="doi">10.1109/mwc.009.2400244</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B15">
        <label>15</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Wen</surname>
              <given-names>Q</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Xiang</surname>
              <given-names>N</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>J</given-names>
            </name>
          </person-group>
          <comment>A survey on federated parameter-efficient fine-tuning for large language models. 2025 11th International Conference on Big Data and Information Analytics (BigDIA); 2025 Nov 8-11; Nha Trang, Vietnam. IEEE; 2025. pp. 637-42.</comment>
          <pub-id pub-id-type="doi">10.1109/bigdia68682.2025.11383022</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B16">
        <label>16</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Ren</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Yu</surname>
              <given-names>H</given-names>
            </name>
            <name>
              <surname>Peng</surname>
              <given-names>H</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Advances and open challenges in federated foundation models</article-title>
          <source>IEEE Commun. Surv. Tutorials.</source>
          <year>2026</year>
          <volume>28</volume>
          <fpage>2087</fpage>
          <lpage>126</lpage>
          <pub-id pub-id-type="doi">10.1109/comst.2025.3552524</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B17">
        <label>17</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Yan</surname>
              <given-names>N</given-names>
            </name>
            <name>
              <surname>Su</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Deng</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Schober</surname>
              <given-names>R</given-names>
            </name>
          </person-group>
          <article-title>Federated fine-tuning of LLMs: framework comparison and research directions</article-title>
          <source>IEEE Commun. Mag.</source>
          <year>2025</year>
          <volume>63</volume>
          <fpage>52</fpage>
          <lpage>8</lpage>
          <pub-id pub-id-type="doi">10.1109/mcom.001.2400770</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B18">
        <label>18</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Adhikari</surname>
              <given-names>D</given-names>
            </name>
            <name>
              <surname>Ullah</surname>
              <given-names>I</given-names>
            </name>
            <name>
              <surname>Khadim</surname>
              <given-names>M</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>A comprehensive survey on robustness and privacy in federated learning meets large language model at edge</article-title>
          <source>J. Reliab. Secur. Comput.</source>
          <year>2026</year>
          <volume>2</volume>
          <fpage>111</fpage>
          <lpage>55</lpage>
          <pub-id pub-id-type="doi">10.62762/jrsc.2026.942513</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B19">
        <label>19</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Akhmetov</surname>
              <given-names>A</given-names>
            </name>
            <name>
              <surname>Ala’Anzy</surname>
              <given-names>MA</given-names>
            </name>
            <name>
              <surname>Ibraheem</surname>
              <given-names>A</given-names>
            </name>
          </person-group>
          <comment>Federated learning strategies for fine-tuning large language models: a systematic literature review. 2026 11th International Conference on Information and Network Technologies (ICINT); 2026 Mar 6-8; Sydney, Australia. IEEE; 2026. pp. 27-32.</comment>
          <pub-id pub-id-type="doi">10.1109/icint69379.2026.11549926</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B20">
        <label>20</label>
        <nlm-citation publication-type="book">
          <person-group person-group-type="author">
            <name>
              <surname>Wu</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Lin</surname>
              <given-names>W</given-names>
            </name>
          </person-group>
          <comment>Federated learning with large language models. 2025 IEEE Cyber Science and Technology Congress (CyberSciTech); 2025 Oct 21-24; Hakodate, Japan. IEEE; 2025. pp. 620-5.</comment>
          <pub-id pub-id-type="doi">10.1109/cyberscitech68397.2025.00129</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B21">
        <label>21</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Piccialli</surname>
              <given-names>F</given-names>
            </name>
            <name>
              <surname>Chiaro</surname>
              <given-names>D</given-names>
            </name>
            <name>
              <surname>Qi</surname>
              <given-names>P</given-names>
            </name>
            <name>
              <surname>Bellandi</surname>
              <given-names>V</given-names>
            </name>
            <name>
              <surname>Damiani</surname>
              <given-names>E</given-names>
            </name>
          </person-group>
          <article-title>Federated and edge learning for large language models</article-title>
          <source>Information Fusion.</source>
          <year>2025</year>
          <volume>117</volume>
          <fpage>102840</fpage>
          <pub-id pub-id-type="doi">10.1016/j.inffus.2024.102840</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B22">
        <label>22</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Vaswani</surname>
              <given-names>A</given-names>
            </name>
            <name>
              <surname>Shazeer</surname>
              <given-names>N</given-names>
            </name>
            <name>
              <surname>Parmar</surname>
              <given-names>N</given-names>
            </name>
            <etal />
          </person-group>
          <comment>Attention is all you need. 31st International Conference on Neural Information Processing Systems; 2017 Dec 4-9; Long Beach, California, USA; Curran Associates Inc.; 2017. pp 5998-6008. <uri xlink:href="https://papers.nips.cc/paper/7181-attention-is-all-you-need">https://papers.nips.cc/paper/7181-attention-is-all-you-need</uri> (accessed 2026-09-14)</comment>
          <pub-id pub-id-type="doi">10.65215/2q58a426</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B23">
        <label>23</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Zhao</surname>
              <given-names>WX</given-names>
            </name>
            <name>
              <surname>Zhou</surname>
              <given-names>K</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>J</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>A survey of large language models</article-title>
          <source>Front. Comput. Sci.</source>
          <year>2026</year>
          <volume>20</volume>
          <fpage>2012627</fpage>
          <pub-id pub-id-type="doi">10.1007/s11704-025-50472-3</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B24">
        <label>24</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Sun</surname>
              <given-names>G</given-names>
            </name>
            <name>
              <surname>Khalid</surname>
              <given-names>U</given-names>
            </name>
            <name>
              <surname>Mendieta</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>P</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>C</given-names>
            </name>
          </person-group>
          <comment>Exploring parameter-efficient fine-tuning to enable foundation models in federated learning. 2024 IEEE International Conference on Big Data (BigData); 2024 Dec 15-18; Washington, DC, USA. IEEE; 2024. pp. 8015-24.</comment>
          <pub-id pub-id-type="doi">10.1109/bigdata62323.2024.10825712</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B25">
        <label>25</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Chen</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Cai</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Zheng</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Yu</surname>
              <given-names>PS</given-names>
            </name>
          </person-group>
          <article-title>A federated adaptive large language model fine-tuning framework for software development</article-title>
          <source>IEEE Trans. Serv. Comput.</source>
          <year>2026</year>
          <volume>19</volume>
          <fpage>32</fpage>
          <lpage>43</lpage>
          <pub-id pub-id-type="doi">10.1109/tsc.2025.3623626</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B26">
        <label>26</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Baali</surname>
              <given-names>FA</given-names>
            </name>
            <name>
              <surname>Ait-Mlouk</surname>
              <given-names>A</given-names>
            </name>
            <name>
              <surname>Agouti</surname>
              <given-names>T</given-names>
            </name>
          </person-group>
          <comment>Federated instruction tuning with DeepSeek: towards scalable and private LLM adaptation. 2025 3rd International Conference on Federated Learning Technologies and Applications (FLTA); 2025 Oct 14-17; Dubrovnik, Croatia. IEEE; 2025. pp. 373-9.</comment>
          <pub-id pub-id-type="doi">10.1109/flta67013.2025.11336718</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B27">
        <label>27</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Yuan</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Yang</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Ye</surname>
              <given-names>G</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>T</given-names>
            </name>
            <name>
              <surname>Nguyen</surname>
              <given-names>QVH</given-names>
            </name>
            <name>
              <surname>Yin</surname>
              <given-names>H</given-names>
            </name>
          </person-group>
          <article-title>FELLAS: Enhancing federated sequential recommendation with LLM as external services</article-title>
          <source>ACM Trans. Inf. Syst.</source>
          <year>2025</year>
          <volume>43</volume>
          <fpage>1</fpage>
          <lpage>24</lpage>
          <pub-id pub-id-type="doi">10.1145/3709138</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B28">
        <label>28</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Wang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Guo</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Tang</surname>
              <given-names>J</given-names>
            </name>
          </person-group>
          <comment>An investigation of parameter efficient federated learning with foundation model. 2025 10th International Conference on Machine Learning Technologies (ICMLT); 2025 May 23-25; Helsinki, Finland. IEEE; 2025. pp. 233-9.</comment>
          <pub-id pub-id-type="doi">10.1109/icmlt65785.2025.11193423</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B29">
        <label>29</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Zhang</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Cheng</surname>
              <given-names>G</given-names>
            </name>
            <name>
              <surname>Wu</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Huang</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Song</surname>
              <given-names>L</given-names>
            </name>
            <name>
              <surname>Shen</surname>
              <given-names>X</given-names>
            </name>
          </person-group>
          <article-title>Split fine-tuning for large language models in wireless networks</article-title>
          <source>IEEE J. Sel. Top. Signal Process.</source>
          <year>2025</year>
          <volume>19</volume>
          <fpage>1376</fpage>
          <lpage>91</lpage>
          <pub-id pub-id-type="doi">10.1109/jstsp.2025.3581484</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B30">
        <label>30</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Wu</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Qu</surname>
              <given-names>K</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Split learning over wireless networks: parallel design and resource management</article-title>
          <source>IEEE J. Select. Areas Commun.</source>
          <year>2023</year>
          <volume>41</volume>
          <fpage>1051</fpage>
          <lpage>66</lpage>
          <pub-id pub-id-type="doi">10.1109/jsac.2023.3242704</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B31">
        <label>31</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Cheng</surname>
              <given-names>L</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>H</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Large-small model collaboration in mobile edge networks with heterogeneous computational resources</article-title>
          <source>IEEE J. Sel. Areas Commun.</source>
          <year>2026</year>
          <volume>44</volume>
          <fpage>2733</fpage>
          <lpage>49</lpage>
          <pub-id pub-id-type="doi">10.1109/jsac.2025.3644295</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B32">
        <label>32</label>
        <nlm-citation publication-type="book">
          <person-group person-group-type="author">
            <name>
              <surname>Ghiasvand</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Alizadeh</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Pedarsani</surname>
              <given-names>R</given-names>
            </name>
          </person-group>
          <comment>Decentralized low-rank fine-tuning of large language models. 1st Workshop for Research on Agent Language Models (REALM 2025); 2025 Jun; Vienna, Austria. Stroudsburg, PA, USA: Association for Computational Linguistics; 2025. pp. 334-45.</comment>
          <pub-id pub-id-type="doi">10.18653/v1/2025.realm-1.24</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B33">
        <label>33</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Balija</surname>
              <given-names>SB</given-names>
            </name>
            <name>
              <surname>Nanda</surname>
              <given-names>A</given-names>
            </name>
            <name>
              <surname>Sahoo</surname>
              <given-names>D</given-names>
            </name>
          </person-group>
          <article-title>Building communication efficient asynchronous peer-to-peer federated LLMs with blockchain</article-title>
          <source>AAAI-SS.</source>
          <year>2024</year>
          <volume>3</volume>
          <fpage>288</fpage>
          <lpage>92</lpage>
          <pub-id pub-id-type="doi">10.1609/aaaiss.v3i1.31212</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B34">
        <label>34</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Woisetschläger</surname>
              <given-names>H</given-names>
            </name>
            <name>
              <surname>Erben</surname>
              <given-names>A</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Mayer</surname>
              <given-names>R</given-names>
            </name>
            <name>
              <surname>Jacobsen</surname>
              <given-names>H</given-names>
            </name>
          </person-group>
          <comment>Federated fine-tuning of LLMs on the very edge: the good, the bad, the ugly. SIGMOD/PODS '24: International Conference on Management of Data; Santiago AA Chile. New York, NY, USA: ACM; 2024. pp. 39-50.</comment>
          <pub-id pub-id-type="doi">10.1145/3650203.3663331</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B35">
        <label>35</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Wu</surname>
              <given-names>F</given-names>
            </name>
            <name>
              <surname>Hu</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Min</surname>
              <given-names>G</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>S</given-names>
            </name>
          </person-group>
          <article-title>Adaptive rank allocation for federated parameter-efficient fine-tuning of language models</article-title>
          <source>IEEE Trans. Comput.</source>
          <year>2026</year>
          <volume>75</volume>
          <fpage>1650</fpage>
          <lpage>63</lpage>
          <pub-id pub-id-type="doi">10.1109/tc.2026.3655161</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B36">
        <label>36</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Cai</surname>
              <given-names>D</given-names>
            </name>
            <name>
              <surname>Wu</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Lin</surname>
              <given-names>FX</given-names>
            </name>
            <name>
              <surname>Xu</surname>
              <given-names>M</given-names>
            </name>
          </person-group>
          <comment>Efficient federated learning for modern NLP. ACM MobiCom '23: 29th Annual International Conference on Mobile Computing and Networking; Madrid Spain. New York, NY, USA: ACM; 2023. pp. 1-16.</comment>
          <pub-id pub-id-type="doi">10.1145/3570361.3592505</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B37">
        <label>37</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Gao</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Guo</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Gong</surname>
              <given-names>Y</given-names>
            </name>
          </person-group>
          <comment>Federated adaptive fine-tuning of large language models with heterogeneous quantization and LoRA. IEEE INFOCOM 2025 - IEEE Conference on Computer Communications; 2025 May 19-22; London, United Kingdom. IEEE; 2025. pp. 1-10.</comment>
          <pub-id pub-id-type="doi">10.1109/infocom55648.2025.11044641</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B38">
        <label>38</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Cai</surname>
              <given-names>D</given-names>
            </name>
          </person-group>
          <comment>Federated LLM pre-training on mobile phones. MobiSys '25: 23rd Annual International Conference on Mobile Systems, Applications and Services; Hilton Anaheim Anaheim CA USA. New York, NY, USA: ACM; 2025. pp. 657-8.</comment>
          <pub-id pub-id-type="doi">10.1145/3711875.3736664</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B39">
        <label>39</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Bai</surname>
              <given-names>G</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Zhao</surname>
              <given-names>L</given-names>
            </name>
            <name>
              <surname>Kim</surname>
              <given-names>K</given-names>
            </name>
          </person-group>
          <comment>FedSpaLLM: federated pruning of large language models. Proceedings of the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers); 2025 Mar; Albuquerque, New Mexico. Stroudsburg, PA, USA: Association for Computational Linguistics; 2025. pp. 8361-73.</comment>
          <pub-id pub-id-type="doi">10.18653/v1/2025.naacl-long.424</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B40">
        <label>40</label>
        <nlm-citation publication-type="book">
          <person-group person-group-type="author">
            <name>
              <surname>Du</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Ye</surname>
              <given-names>R</given-names>
            </name>
            <name>
              <surname>Yuchi</surname>
              <given-names>F</given-names>
            </name>
            <etal />
          </person-group>
          <comment>FedDQC: Data quality control in federated instruction-tuning of large language models. Findings of the Association for Computational Linguistics: ACL 2025; 2025 Jun; Vienna, Austria. Stroudsburg, PA, USA: Association for Computational Linguistics; 2025. pp. 15267-91.</comment>
          <pub-id pub-id-type="doi">10.18653/v1/2025.findings-acl.791</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B41">
        <label>41</label>
        <nlm-citation publication-type="book">
          <person-group person-group-type="author">
            <name>
              <surname>Qin</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Wu</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>He</surname>
              <given-names>B</given-names>
            </name>
            <name>
              <surname>Deng</surname>
              <given-names>S</given-names>
            </name>
          </person-group>
          <comment>Federated data-efficient instruction tuning for large language models. Findings of the Association for Computational Linguistics: ACL 2025; 2025 Jun; Vienna, Austria. Stroudsburg, PA, USA: Association for Computational Linguistics; 2025. pp. 15550-68.</comment>
          <pub-id pub-id-type="doi">10.18653/v1/2025.findings-acl.803</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B42">
        <label>42</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Zhang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Huang</surname>
              <given-names>J</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>PPFedIT: towards privacy-preserving federated instruction tuning with few-shot local examples</article-title>
          <source>ACM Trans. Intell. Syst. Technol.</source>
          <year>2026</year>
          <volume>17</volume>
          <fpage>1</fpage>
          <lpage>23</lpage>
          <pub-id pub-id-type="doi">10.1145/3806196</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B43">
        <label>43</label>
        <nlm-citation publication-type="confproc">
          <comment>Zhan, S.; Li, L.; Xu, C. AssyLLM: Efficient federated fine-tuning of LLMs via assembling pre-trained blocks. 2025 USENIX Annual Technical Conference; 2025; pp 1677-91. Available online: <uri xlink:href="https://www.usenix.org/conference/atc25/presentation/zhan">https://www.usenix.org/conference/atc25/presentation/zhan</uri> (accessed 2026-09-14)</comment>
        </nlm-citation>
      </ref>
      <ref id="B44">
        <label>44</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Che</surname>
              <given-names>T</given-names>
            </name>
            <name>
              <surname>Liu</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Zhou</surname>
              <given-names>Y</given-names>
            </name>
            <etal />
          </person-group>
          <comment>Federated learning of large language models with parameter-efficient prompt tuning and adaptive optimization. 2023 Conference on Empirical Methods in Natural Language Processing; 2023 Nov; Singapore. Stroudsburg, PA, USA: Association for Computational Linguistics; 2023. pp. 7871-88.</comment>
          <pub-id pub-id-type="doi">10.18653/v1/2023.emnlp-main.488</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B45">
        <label>45</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Kim</surname>
              <given-names>G</given-names>
            </name>
            <name>
              <surname>Yoo</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Kang</surname>
              <given-names>S</given-names>
            </name>
          </person-group>
          <article-title>Efficient federated learning with pre-trained large language model using several adapter mechanisms</article-title>
          <source>Mathematics.</source>
          <year>2023</year>
          <volume>11</volume>
          <fpage>4479</fpage>
          <pub-id pub-id-type="doi">10.3390/math11214479</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B46">
        <label>46</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Saadati</surname>
              <given-names>N</given-names>
            </name>
            <name>
              <surname>Jiang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Balu</surname>
              <given-names>A</given-names>
            </name>
            <name>
              <surname>Liu</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Hegde</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Sarkar</surname>
              <given-names>S</given-names>
            </name>
          </person-group>
          <comment>Foundation model efficient fine-tuning in centralized and federated settings. 2025 IEEE International Conference on Big Data (BigData); 2025 Dec 8-11; Macau, China. IEEE; 2025. pp. 1857-66.</comment>
          <pub-id pub-id-type="doi">10.1109/bigdata66926.2025.11400875</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B47">
        <label>47</label>
        <nlm-citation publication-type="book">
          <person-group person-group-type="author">
            <name>
              <surname>Ghiasvand</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Yang</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Xue</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Alizadeh</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Pedarsani</surname>
              <given-names>R</given-names>
            </name>
          </person-group>
          <comment>Communication-efficient and tensorized federated fine-tuning of large language models. Findings of the Association for Computational Linguistics: ACL 2025; 2025 Jun; Vienna, Austria. Stroudsburg, PA, USA: Association for Computational Linguistics; 2025. pp. 24192-207.</comment>
          <pub-id pub-id-type="doi">10.18653/v1/2025.findings-acl.1241</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B48">
        <label>48</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Liu</surname>
              <given-names>H</given-names>
            </name>
            <name>
              <surname>Wen</surname>
              <given-names>R</given-names>
            </name>
            <name>
              <surname>Nair</surname>
              <given-names>S</given-names>
            </name>
            <etal />
          </person-group>
          <comment>EcoLoRA: communication-efficient federated fine-tuning of large language models. 2025 Conference on Empirical Methods in Natural Language Processing; 2025 Oct; Suzhou, China. Stroudsburg, PA, USA: Association for Computational Linguistics; 2025. pp. 20743-57.</comment>
          <pub-id pub-id-type="doi">10.18653/v1/2025.emnlp-main.1046</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B49">
        <label>49</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Liang</surname>
              <given-names>P</given-names>
            </name>
            <name>
              <surname>Guo</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Zhao</surname>
              <given-names>M</given-names>
            </name>
          </person-group>
          <comment>Federated fine-tuning large language models with LoRA and non-orthogonal transmission. GLOBECOM 2025 - 2025 IEEE Global Communications Conference; 2025 Dec 8-12; Taipei, Taiwan. IEEE; 2025. pp. 405-10.</comment>
          <pub-id pub-id-type="doi">10.1109/globecom59602.2025.11432233</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B50">
        <label>50</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Faiyaz</surname>
              <given-names>A</given-names>
            </name>
            <name>
              <surname>Salman</surname>
              <given-names>T</given-names>
            </name>
          </person-group>
          <comment>GradualDiff-Fed: A federated learning specialized framework for large language model. 2025 IEEE 4th International Conference on Computing and Machine Intelligence (ICMI); 2025 Apr 5-6; MI, USA. IEEE; 2025. pp. 1-5.</comment>
          <pub-id pub-id-type="doi">10.1109/icmi65310.2025.11141097</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B51">
        <label>51</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Rahman</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Rahman</surname>
              <given-names>R</given-names>
            </name>
          </person-group>
          <article-title>Semantic communication-aware federated fine-tuning of large language models</article-title>
          <source>IEEE Commun. Lett.</source>
          <year>2025</year>
          <volume>29</volume>
          <fpage>2974</fpage>
          <lpage>7</lpage>
          <pub-id pub-id-type="doi">10.1109/lcomm.2025.3621354</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B52">
        <label>52</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Zhang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Liu</surname>
              <given-names>P</given-names>
            </name>
            <name>
              <surname>Xu</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Hu</surname>
              <given-names>R</given-names>
            </name>
          </person-group>
          <article-title>Fed-HeLLo: Efficient federated foundation model fine-tuning with heterogeneous LoRA allocation</article-title>
          <source>IEEE Trans. Neural Netw. Learning Syst.</source>
          <year>2025</year>
          <volume>36</volume>
          <fpage>17556</fpage>
          <lpage>69</lpage>
          <pub-id pub-id-type="doi">10.1109/tnnls.2025.3580495</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B53">
        <label>53</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Cho</surname>
              <given-names>YJ</given-names>
            </name>
            <name>
              <surname>Liu</surname>
              <given-names>L</given-names>
            </name>
            <name>
              <surname>Xu</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Fahrezi</surname>
              <given-names>A</given-names>
            </name>
            <name>
              <surname>Joshi</surname>
              <given-names>G</given-names>
            </name>
          </person-group>
          <comment>Heterogeneous LoRA for federated fine-tuning of on-device foundation models. 2024 Conference on Empirical Methods in Natural Language Processing; 2024 Oct; Miami, Florida, USA. Stroudsburg, PA, USA: Association for Computational Linguistics; 2024. pp. 12903-13.</comment>
          <pub-id pub-id-type="doi">10.18653/v1/2024.emnlp-main.717</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B54">
        <label>54</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Wang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Shen</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>He</surname>
              <given-names>Y</given-names>
            </name>
            <etal />
          </person-group>
          <comment>FLoRA: federated fine-tuning large language models with heterogeneous low-rank adaptations. Advances in Neural Information Processing Systems 37; 2024 Dec 10-15; Vancouver, BC, Canada. San Diego, California, USA: Neural Information Processing Systems Foundation, Inc. (NeurIPS); 2024. pp. 22513-33.</comment>
          <pub-id pub-id-type="doi">10.52202/079017-0708</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B55">
        <label>55</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Jiang</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Song</surname>
              <given-names>B</given-names>
            </name>
          </person-group>
          <article-title>Fine-tuning large language models in federated learning with fairness-aware prompt selection</article-title>
          <source>Neural Netw.</source>
          <year>2026</year>
          <volume>194</volume>
          <fpage>108160</fpage>
          <pub-id pub-id-type="doi">10.1016/j.neunet.2025.108160</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B56">
        <label>56</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Zhang</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Sun</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>F</given-names>
            </name>
          </person-group>
          <article-title>FedBRICK: structural bias aware heterogeneous foundation model federated tuning</article-title>
          <source>AAAI.</source>
          <year>2026</year>
          <volume>40</volume>
          <fpage>28528</fpage>
          <lpage>36</lpage>
          <pub-id pub-id-type="doi">10.1609/aaai.v40i34.40083</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B57">
        <label>57</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Liu</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Liao</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Xu</surname>
              <given-names>H</given-names>
            </name>
            <name>
              <surname>Xu</surname>
              <given-names>Y</given-names>
            </name>
          </person-group>
          <comment>Tackling data heterogeneity in parameter-efficient federated fine-tuning of large language models. ICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP); 2026 May 3-8; Barcelona, Spain. IEEE; 2026. pp. 17202-6.</comment>
          <pub-id pub-id-type="doi">10.1109/icassp55912.2026.11460924</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B58">
        <label>58</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Wang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Zhou</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Shi</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Letaief</surname>
              <given-names>KB</given-names>
            </name>
          </person-group>
          <article-title>Federated fine-tuning for pre-trained foundation models over wireless networks</article-title>
          <source>IEEE Trans. Wireless Commun.</source>
          <year>2025</year>
          <volume>24</volume>
          <fpage>3450</fpage>
          <lpage>64</lpage>
          <pub-id pub-id-type="doi">10.1109/twc.2025.3531128</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B59">
        <label>59</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Iftikhar</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Alsamhi</surname>
              <given-names>SH</given-names>
            </name>
            <name>
              <surname>Davy</surname>
              <given-names>S</given-names>
            </name>
          </person-group>
          <article-title>Enhancing sustainability in LLM training: leveraging federated learning and parameter-efficient fine-tuning</article-title>
          <source>IEEE Trans. Sustain. Comput.</source>
          <year>2025</year>
          <volume>10</volume>
          <fpage>1158</fpage>
          <lpage>72</lpage>
          <pub-id pub-id-type="doi">10.1109/tsusc.2025.3592043</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B60">
        <label>60</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Wang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Zhou</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Shi</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Letaief</surname>
              <given-names>KB</given-names>
            </name>
          </person-group>
          <comment>Federated low-rank adaptation for large language model fine-tuning over wireless networks. GLOBECOM 2024 - 2024 IEEE Global Communications Conference; 2024 Dec 8-12; Cape Town, South Africa. IEEE; 2024. pp. 3063-8.</comment>
          <pub-id pub-id-type="doi">10.1109/globecom52923.2024.10901572</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B61">
        <label>61</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Zhao</surname>
              <given-names>K</given-names>
            </name>
            <name>
              <surname>Zhu</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Huang</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Yang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>Z</given-names>
            </name>
          </person-group>
          <comment>SflLLM: efficient split federated learning for large language model over wireless networks. GLOBECOM 2025 - 2025 IEEE Global Communications Conference; 2025 Dec 8-12; Taipei, Taiwan. IEEE; 2025. pp. 1835-40.</comment>
          <pub-id pub-id-type="doi">10.1109/globecom59602.2025.11432069</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B62">
        <label>62</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Otoum</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Danish</surname>
              <given-names>SM</given-names>
            </name>
            <name>
              <surname>Ahmad</surname>
              <given-names>I</given-names>
            </name>
            <name>
              <surname>Alkhrijah</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Asad</surname>
              <given-names>A</given-names>
            </name>
          </person-group>
          <article-title>Efficient federated LLM framework for intelligent IoMT management</article-title>
          <source>IEEE Trans. Consumer Electron.</source>
          <year>2026</year>
          <volume>72</volume>
          <fpage>5832</fpage>
          <lpage>47</lpage>
          <pub-id pub-id-type="doi">10.1109/tce.2026.3689383</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B63">
        <label>63</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Wang</surname>
              <given-names>F</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>B</given-names>
            </name>
          </person-group>
          <article-title>Data reconstruction and protection in federated learning for fine-tuning large language models</article-title>
          <source>IEEE Trans. Big Data.</source>
          <year>2024</year>
          <fpage>1</fpage>
          <lpage>13</lpage>
          <pub-id pub-id-type="doi">10.1109/tbdata.2024.3524105</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B64">
        <label>64</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Zheng</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>H</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>L</given-names>
            </name>
            <name>
              <surname>Qiu</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Zheng</surname>
              <given-names>H</given-names>
            </name>
            <name>
              <surname>Zheng</surname>
              <given-names>Z</given-names>
            </name>
          </person-group>
          <comment>Safely learning with private data: a federated learning framework for large language model. 2024 Conference on Empirical Methods in Natural Language Processing; 2024 Oct; Miami, Florida, USA. Stroudsburg, PA, USA: Association for Computational Linguistics; 2024. pp. 5293-306.</comment>
          <pub-id pub-id-type="doi">10.18653/v1/2024.emnlp-main.303</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B65">
        <label>65</label>
        <nlm-citation publication-type="book">
          <person-group person-group-type="author">
            <name>
              <surname>Pontes</surname>
              <given-names>MF</given-names>
            </name>
            <name>
              <surname>Pedrosa</surname>
              <given-names>RC</given-names>
            </name>
            <name>
              <surname>Lopes</surname>
              <given-names>PH</given-names>
            </name>
            <name>
              <surname>Luz</surname>
              <given-names>EJ</given-names>
            </name>
          </person-group>
          <comment>Evaluating federated learning with homomorphic encryption for medical named entity recognition using compact BERT models. Simpósio Brasileiro de Tecnologia da Informação e da Linguagem Humana; Brasil. Sociedade Brasileira de Computação; 2024. pp. 48-56.</comment>
          <pub-id pub-id-type="doi">10.5753/stil.2024.245381</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B66">
        <label>66</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Kim</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Lim</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Ryu</surname>
              <given-names>G</given-names>
            </name>
            <name>
              <surname>Kim</surname>
              <given-names>H</given-names>
            </name>
          </person-group>
          <article-title>How robust are language models against backdoors in federated learning?</article-title>
          <source>CMES.</source>
          <year>2025</year>
          <volume>145</volume>
          <fpage>2617</fpage>
          <lpage>30</lpage>
          <pub-id pub-id-type="doi">10.32604/cmes.2025.071190</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B67">
        <label>67</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Zhao</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Fang</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Zhong</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Zheng</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>L</given-names>
            </name>
            <name>
              <surname>Pechenizkiy</surname>
              <given-names>M</given-names>
            </name>
          </person-group>
          <article-title>Investigating social bias propagation in federated fine-tuning of large language models</article-title>
          <source>AAAI.</source>
          <year>2026</year>
          <volume>40</volume>
          <fpage>39637</fpage>
          <lpage>45</lpage>
          <pub-id pub-id-type="doi">10.1609/aaai.v40i46.41316</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B68">
        <label>68</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Li</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Wu</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>J</given-names>
            </name>
          </person-group>
          <comment>Foundation models in federated learning: assessing backdoor vulnerabilities. 2025 International Joint Conference on Neural Networks (IJCNN); 2025 Jun 30-Jul 5; Rome, Italy. IEEE; 2025. pp. 1-8.</comment>
          <pub-id pub-id-type="doi">10.1109/ijcnn64981.2025.11229350</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B69">
        <label>69</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Xie</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Wen</surname>
              <given-names>D</given-names>
            </name>
            <name>
              <surname>You</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>Q</given-names>
            </name>
            <name>
              <surname>Bennis</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Huang</surname>
              <given-names>K</given-names>
            </name>
          </person-group>
          <article-title>FedLoDrop: Federated LoRA with dropout for generalized LLM fine-tuning</article-title>
          <source>IEEE J. Sel. Areas Commun.</source>
          <year>2026</year>
          <volume>44</volume>
          <fpage>3541</fpage>
          <lpage>56</lpage>
          <pub-id pub-id-type="doi">10.1109/jsac.2026.3660935</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B70">
        <label>70</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Fang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Lin</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Gao</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Fang</surname>
              <given-names>Y</given-names>
            </name>
          </person-group>
          <article-title>Automated federated pipeline for parameter-efficient fine-tuning of large language models</article-title>
          <source>IEEE Trans. on Mobile Comput.</source>
          <year>2026</year>
          <volume>25</volume>
          <fpage>8782</fpage>
          <lpage>97</lpage>
          <pub-id pub-id-type="doi">10.1109/tmc.2025.3649881</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B71">
        <label>71</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Wu</surname>
              <given-names>F</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Ding</surname>
              <given-names>B</given-names>
            </name>
            <name>
              <surname>Gao</surname>
              <given-names>J</given-names>
            </name>
          </person-group>
          <comment>FedBiOT: LLM local fine-tuning in federated learning without full model. KDD '24: The 30th ACM SIGKDD Conference on Knowledge Discovery and Data Mining; Barcelona Spain. New York, NY, USA: ACM; 2024. pp. 3345-55.</comment>
          <pub-id pub-id-type="doi">10.1145/3637528.3671897</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B72">
        <label>72</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Su</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Yan</surname>
              <given-names>N</given-names>
            </name>
            <name>
              <surname>Deng</surname>
              <given-names>Y</given-names>
            </name>
          </person-group>
          <comment>Federated LLMs fine-tuned with adaptive importance-aware LoRA. ICC 2025 - IEEE International Conference on Communications; 2025 Jun 8-12; Montreal, QC, Canada. IEEE; 2025. pp. 6112-7.</comment>
          <pub-id pub-id-type="doi">10.1109/icc52391.2025.11161447</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B73">
        <label>73</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Li</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Sun</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Liu</surname>
              <given-names>Y</given-names>
            </name>
            <etal />
          </person-group>
          <comment>Federated black-box prompt tuning system for large language models on the edge. ACM MobiCom '24: 30th Annual International Conference on Mobile Computing and Networking; Washington D.C. DC USA. New York, NY, USA: ACM; 2024. pp. 1775-7.</comment>
          <pub-id pub-id-type="doi">10.1145/3636534.3698856</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B74">
        <label>74</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Tanimura</surname>
              <given-names>T</given-names>
            </name>
            <name>
              <surname>Nakano</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Kitagawa</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Takase</surname>
              <given-names>M</given-names>
            </name>
          </person-group>
          <comment>Federated discrete prompt tuning for language models using synthetic examples. 2025 IEEE 22nd Consumer Communications &amp; Networking Conference (CCNC); 2025 Jan 10-13; Las Vegas, NV, USA. IEEE; 2025. pp. 1-4.</comment>
          <pub-id pub-id-type="doi">10.1109/ccnc54725.2025.10975936</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B75">
        <label>75</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Wu</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Tang</surname>
              <given-names>J</given-names>
            </name>
            <etal />
          </person-group>
          <comment>FDPT: Federated discrete prompt tuning for black-box visual-language models. 2025 IEEE/CVF International Conference on Computer Vision (ICCV); 2025 Oct 19-25; Honolulu, HI, USA. IEEE; 2025. pp. 1-10.</comment>
          <pub-id pub-id-type="doi">10.1109/iccv51701.2025.00237</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B76">
        <label>76</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Fan</surname>
              <given-names>B</given-names>
            </name>
            <name>
              <surname>Su</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Tarkoma</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Hui</surname>
              <given-names>P</given-names>
            </name>
          </person-group>
          <article-title>HeLoRA: LoRA-heterogeneous federated fine-tuning for foundation Models</article-title>
          <source>ACM Trans. Internet Technol.</source>
          <year>2025</year>
          <volume>25</volume>
          <fpage>1</fpage>
          <lpage>22</lpage>
          <pub-id pub-id-type="doi">10.1145/3723877</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B77">
        <label>77</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Chen</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>R</given-names>
            </name>
            <name>
              <surname>Shao</surname>
              <given-names>J</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Federated LoRA fine-tuning of LLMs with only transmitting matrix A or B</article-title>
          <source>IEEE Trans. Mobile Comput.</source>
          <year>2026</year>
          <volume>25</volume>
          <fpage>17503</fpage>
          <lpage>19</lpage>
          <pub-id pub-id-type="doi">10.1109/tmc.2026.3696859</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B78">
        <label>78</label>
        <nlm-citation publication-type="book">
          <person-group person-group-type="author">
            <name>
              <surname>Guo</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Lu</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Tong</surname>
              <given-names>Y</given-names>
            </name>
            <etal />
          </person-group>
          <comment>H2Tune: federated foundation model fine-tuning with hybrid heterogeneity. In: Lynce I, Murano N, Vallati M, Villata S, Chesani F, Milano M, Omicini A, Dastani M, Editors. ECAI 2025. IOS Press; 2025.</comment>
          <pub-id pub-id-type="doi">10.3233/faia413</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B79">
        <label>79</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Wang</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Qiao</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Wu</surname>
              <given-names>D</given-names>
            </name>
            <name>
              <surname>Wu</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>F</given-names>
            </name>
          </person-group>
          <article-title>Cluster based heterogeneous federated foundation model adaptation and fine-tuning</article-title>
          <source>AAAI.</source>
          <year>2025</year>
          <volume>39</volume>
          <fpage>21269</fpage>
          <lpage>77</lpage>
          <pub-id pub-id-type="doi">10.1609/aaai.v39i20.35426</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B80">
        <label>80</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Yang</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Long</surname>
              <given-names>G</given-names>
            </name>
            <name>
              <surname>Shen</surname>
              <given-names>T</given-names>
            </name>
            <name>
              <surname>Jiang</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Blumenstein</surname>
              <given-names>M</given-names>
            </name>
          </person-group>
          <comment>Dual-personalizing adapter for federated foundation models. Advances in Neural Information Processing Systems 37; 2024 Dec 10-15; Vancouver, BC, Canada. San Diego, California, USA: Neural Information Processing Systems Foundation, Inc. (NeurIPS); 2024. pp. 39409-33.</comment>
          <pub-id pub-id-type="doi">10.52202/079017-1245</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B81">
        <label>81</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Bian</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>L</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>L</given-names>
            </name>
            <name>
              <surname>Xu</surname>
              <given-names>J</given-names>
            </name>
          </person-group>
          <article-title>FedALT: federated fine-tuning through adaptive local training with rest-of-world LoRA</article-title>
          <source>AAAI.</source>
          <year>2026</year>
          <volume>40</volume>
          <fpage>19728</fpage>
          <lpage>36</lpage>
          <pub-id pub-id-type="doi">10.1609/aaai.v40i24.39054</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B82">
        <label>82</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Chang</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Shi</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Zhao</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Ma</surname>
              <given-names>D</given-names>
            </name>
          </person-group>
          <article-title>Dual prompt personalized federated learning in foundation models</article-title>
          <source>Sci Rep.</source>
          <year>2025</year>
          <volume>15</volume>
          <fpage>28026</fpage>
          <pub-id pub-id-type="doi">10.1038/s41598-025-11864-4</pub-id>
          <pub-id pub-id-type="pmid">40745444</pub-id>
          <pub-id pub-id-type="pmcid">PMC12313890</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B83">
        <label>83</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Su</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Yang</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>B</given-names>
            </name>
            <name>
              <surname>Xue</surname>
              <given-names>X</given-names>
            </name>
          </person-group>
          <article-title>Federated adaptive prompt tuning for multi-domain collaborative learning</article-title>
          <source>AAAI.</source>
          <year>2024</year>
          <volume>38</volume>
          <fpage>15117</fpage>
          <lpage>25</lpage>
          <pub-id pub-id-type="doi">10.1609/aaai.v38i13.29434</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B84">
        <label>84</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Pang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Wei</surname>
              <given-names>K</given-names>
            </name>
            <name>
              <surname>Shi</surname>
              <given-names>L</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Shu</surname>
              <given-names>F</given-names>
            </name>
          </person-group>
          <article-title>Low-latency federated fine-tuning for large language models over wireless networks</article-title>
          <source>IEEE Wireless Commun. Lett.</source>
          <year>2026</year>
          <volume>15</volume>
          <fpage>2179</fpage>
          <lpage>83</lpage>
          <pub-id pub-id-type="doi">10.1109/lwc.2026.3672798</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B85">
        <label>85</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Chen</surname>
              <given-names>H</given-names>
            </name>
            <name>
              <surname>Yuan</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>H</given-names>
            </name>
          </person-group>
          <article-title>Edge-assisted federated learning for large language models in IoT sensor systems</article-title>
          <source>IEEE J. Sel. Areas Sensors.</source>
          <year>2026</year>
          <volume>3</volume>
          <fpage>125</fpage>
          <lpage>38</lpage>
          <pub-id pub-id-type="doi">10.1109/jsas.2026.3663337</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B86">
        <label>86</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Chen</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Ni</surname>
              <given-names>W</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Toward dynamic resource allocation and client scheduling in hierarchical federated learning: a two-phase deep reinforcement learning approach</article-title>
          <source>IEEE Trans. Commun.</source>
          <year>2024</year>
          <volume>72</volume>
          <fpage>7798</fpage>
          <lpage>813</lpage>
          <pub-id pub-id-type="doi">10.1109/tcomm.2024.3420733</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B87">
        <label>87</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Zhang</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Yan</surname>
              <given-names>N</given-names>
            </name>
            <name>
              <surname>Su</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Deng</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Mahmoodi</surname>
              <given-names>T</given-names>
            </name>
          </person-group>
          <comment>Communication-aware knowledge distillation for federated LLM fine-tuning over wireless networks. GLOBECOM 2025 - 2025 IEEE Global Communications Conference; 2025 Dec 8-12; Taipei, Taiwan. IEEE; 2025. pp. 1956-61.</comment>
          <pub-id pub-id-type="doi">10.1109/globecom59602.2025.11432269</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B88">
        <label>88</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Zhao</surname>
              <given-names>K</given-names>
            </name>
            <name>
              <surname>Yang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Huang</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>Z</given-names>
            </name>
          </person-group>
          <comment>FedsLLM: federated split learning for large language models over communication networks. 2024 International Conference on Ubiquitous Communication (Ucom); 2024 Jul 5-7; Xi'an, China. IEEE; 2024. pp. 438-43.</comment>
          <pub-id pub-id-type="doi">10.1109/ucom62433.2024.10695888</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B89">
        <label>89</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Dai</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Tang</surname>
              <given-names>R</given-names>
            </name>
            <name>
              <surname>Jiang</surname>
              <given-names>F</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Federated split learning for large language models with RSMA</article-title>
          <source>IET Communications.</source>
          <year>2026</year>
          <volume>20</volume>
          <fpage>e70152</fpage>
          <pub-id pub-id-type="doi">10.1049/cmu2.70152</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B90">
        <label>90</label>
        <nlm-citation publication-type="book">
          <person-group person-group-type="author">
            <name>
              <surname>Zhao</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Xu</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Ng</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Chua</surname>
              <given-names>T</given-names>
            </name>
          </person-group>
          <comment>A federated framework for LLM-based recommendation. Findings of the Association for Computational Linguistics: NAACL 2025; 2025 Mar; Albuquerque, New Mexico. Stroudsburg, PA, USA: Association for Computational Linguistics; 2025. pp. 2852-65.</comment>
          <pub-id pub-id-type="doi">10.18653/v1/2025.findings-naacl.155</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B91">
        <label>91</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Gao</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Hou</surname>
              <given-names>L</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>B</given-names>
            </name>
            <name>
              <surname>Yao</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Suo</surname>
              <given-names>Z</given-names>
            </name>
          </person-group>
          <article-title>Compressive-learning-based federated learning for intelligent IoT with cloud-edge collaboration</article-title>
          <source>IEEE Internet Things J.</source>
          <year>2025</year>
          <volume>12</volume>
          <fpage>2291</fpage>
          <lpage>4</lpage>
          <pub-id pub-id-type="doi">10.1109/jiot.2024.3505838</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B92">
        <label>92</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Andrei</surname>
              <given-names>VC</given-names>
            </name>
            <name>
              <surname>Djuhera</surname>
              <given-names>A</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Mönich</surname>
              <given-names>UJ</given-names>
            </name>
            <name>
              <surname>Saad</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Boche</surname>
              <given-names>H</given-names>
            </name>
          </person-group>
          <comment>Resilient, Federated large language models over wireless networks: why the PHY matters. GLOBECOM 2024 - 2024 IEEE Global Communications Conference; 2024 Dec 8-12; Cape Town, South Africa. IEEE; 2024. pp. 5211-6.</comment>
          <pub-id pub-id-type="doi">10.1109/globecom52923.2024.10901454</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B93">
        <label>93</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Djuhera</surname>
              <given-names>A</given-names>
            </name>
            <name>
              <surname>Andrei</surname>
              <given-names>VC</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Mönich</surname>
              <given-names>UJ</given-names>
            </name>
            <name>
              <surname>Boche</surname>
              <given-names>H</given-names>
            </name>
            <name>
              <surname>Saad</surname>
              <given-names>W</given-names>
            </name>
          </person-group>
          <article-title>R-SFLLM: jamming resilient framework for split federated learning with large language models</article-title>
          <source>IEEE Trans. Inform. Forensic Secur.</source>
          <year>2025</year>
          <volume>20</volume>
          <fpage>8296</fpage>
          <lpage>311</lpage>
          <pub-id pub-id-type="doi">10.1109/tifs.2025.3594107</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B94">
        <label>94</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Wang</surname>
              <given-names>H</given-names>
            </name>
            <name>
              <surname>Yin</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>B</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>ROFED-LLM: robust federated learning for large language models in adversarial wireless environments</article-title>
          <source>IEEE Trans. Netw. Sci. Eng.</source>
          <year>2026</year>
          <volume>13</volume>
          <fpage>1084</fpage>
          <lpage>96</lpage>
          <pub-id pub-id-type="doi">10.1109/tnse.2025.3590975</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B95">
        <label>95</label>
        <nlm-citation publication-type="book">
          <person-group person-group-type="author">
            <name>
              <surname>Park</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Han</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Guo</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Ozdaglar</surname>
              <given-names>AE</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>K</given-names>
            </name>
            <name>
              <surname>Kim</surname>
              <given-names>J</given-names>
            </name>
          </person-group>
          <comment>MAPoRL: Multi-agent post-co-training for collaborative large language models with reinforcement learning. 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers); 2025 Jun; Vienna, Austria. Stroudsburg, PA, USA: Association for Computational Linguistics; 2025. pp. 30215-48.</comment>
          <pub-id pub-id-type="doi">10.18653/v1/2025.acl-long.1459</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B96">
        <label>96</label>
        <nlm-citation publication-type="book">
          <person-group person-group-type="author">
            <name>
              <surname>Pan</surname>
              <given-names>Q</given-names>
            </name>
            <name>
              <surname>Wu</surname>
              <given-names>J</given-names>
            </name>
          </person-group>
          <comment>Selective privacy-preserving federated learning for large language model fine-tuning. 2025 International Wireless Communications and Mobile Computing (IWCMC); 2025 May 12-16; Abu Dhabi, United Arab Emirates. IEEE; 2025. pp. 1626-31.</comment>
          <pub-id pub-id-type="doi">10.1109/iwcmc65282.2025.11059634</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B97">
        <label>97</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Wang</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Han</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Cheng</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>Y</given-names>
            </name>
          </person-group>
          <comment>Federated fine-tuning of large language models with privacy preservation and cross-domain semantic alignment. 2025 6th International Conference on Computer Vision and Data Mining (ICCVDM); 2025 Sep 12-14; London, United Kingdom. IEEE; 2025. pp. 494-8.</comment>
          <pub-id pub-id-type="doi">10.1109/iccvdm66874.2025.11290023</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B98">
        <label>98</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Kumar</surname>
              <given-names>GS</given-names>
            </name>
            <name>
              <surname>Vankudothu</surname>
              <given-names>B</given-names>
            </name>
            <name>
              <surname>Tiwari</surname>
              <given-names>AK</given-names>
            </name>
            <name>
              <surname>Pushparathi</surname>
              <given-names>VG</given-names>
            </name>
            <name>
              <surname>Sathiya</surname>
              <given-names>B</given-names>
            </name>
            <name>
              <surname>Prasan</surname>
              <given-names>UD</given-names>
            </name>
          </person-group>
          <comment>Privacy-aware federated large language model adaptation using encrypted gradient aggregation. 2026 International Conference on Electronic Systems and Intelligent Computing (ICESIC); 2026 Mar 13-14; Chennai, India. IEEE; 2026. pp. 420-5.</comment>
          <pub-id pub-id-type="doi">10.1109/icesic67389.2026.11496607</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B99">
        <label>99</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Shanmugam</surname>
              <given-names>K</given-names>
            </name>
            <name>
              <surname>B</surname>
              <given-names>N</given-names>
            </name>
            <name>
              <surname>K</surname>
              <given-names>KK</given-names>
            </name>
          </person-group>
          <comment>Federated instruction-tuning of large language models with privacy and communication efficiency. 2026 Contemporary Computing Innovations Conference (CCIC); 2026 Feb 6-7; Tirupati, India. IEEE; 2026. pp. 1-7.</comment>
          <pub-id pub-id-type="doi">10.1109/ccic68129.2026.11486066</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B100">
        <label>100</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Li</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Fu</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Li</surname>
              <given-names>J</given-names>
            </name>
          </person-group>
          <comment>Privacy-preserving and efficient aggregation for federated large language models. 2025 International Conference on Artificial Intelligence Security and Governance (ICAISG); 2025 Dec 12-14; Hangzhou, China. IEEE; 2025. pp. 65-9.</comment>
          <pub-id pub-id-type="doi">10.1109/icaisg68699.2025.11452122</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B101">
        <label>101</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Zhang</surname>
              <given-names>T</given-names>
            </name>
            <name>
              <surname>Yu</surname>
              <given-names>H</given-names>
            </name>
            <name>
              <surname>Yang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Yu</surname>
              <given-names>S</given-names>
            </name>
          </person-group>
          <article-title>LaVFL: efficient verifiable federated learning for large language models</article-title>
          <source>IEEE Trans. Dependable and Secure Comput.</source>
          <year>2025</year>
          <volume>22</volume>
          <fpage>6214</fpage>
          <lpage>29</lpage>
          <pub-id pub-id-type="doi">10.1109/tdsc.2025.3581728</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B102">
        <label>102</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Wu</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Ren</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Guo</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Huang</surname>
              <given-names>M</given-names>
            </name>
          </person-group>
          <article-title>Privacy-preserving personalized federated prompt learning for vision-language models</article-title>
          <source>Neural Netw.</source>
          <year>2026</year>
          <volume>195</volume>
          <fpage>108220</fpage>
          <pub-id pub-id-type="doi">10.1016/j.neunet.2025.108220</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B103">
        <label>103</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Ni</surname>
              <given-names>F</given-names>
            </name>
            <name>
              <surname>Zhou</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Ni</surname>
              <given-names>W</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Scheduling and securing asynchronous federated learning through cooperative jamming</article-title>
          <source>IEEE Trans. Cogn. Commun. Netw.</source>
          <year>2026</year>
          <volume>12</volume>
          <fpage>3209</fpage>
          <lpage>22</lpage>
          <pub-id pub-id-type="doi">10.1109/tccn.2025.3623377</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B104">
        <label>104</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Faiyaz</surname>
              <given-names>A</given-names>
            </name>
            <name>
              <surname>Olapojoye</surname>
              <given-names>R</given-names>
            </name>
            <name>
              <surname>Karwa</surname>
              <given-names>G</given-names>
            </name>
            <name>
              <surname>Salman</surname>
              <given-names>T</given-names>
            </name>
          </person-group>
          <comment>A study of model poisoning attacks on federated large language models. 2026 IEEE 5th International Conference on Computing and Machine Intelligence (ICMI); 2026 Apr 9-10; Al-Ahsa, Saudi Arabia. IEEE; 2026. pp. 1-6.</comment>
          <pub-id pub-id-type="doi">10.1109/icmi68585.2026.11539939</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B105">
        <label>105</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Zhai</surname>
              <given-names>K</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Ma</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Jiang</surname>
              <given-names>Y</given-names>
            </name>
          </person-group>
          <comment>FedAPT: federated adversarial prompt tuning for vision-language models. MM '25: The 33rd ACM International Conference on Multimedia; Dublin Ireland. New York, NY, USA: ACM; 2025. pp. 4310-8.</comment>
          <pub-id pub-id-type="doi">10.1145/3746027.3755387</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B106">
        <label>106</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Nehara</surname>
              <given-names>T</given-names>
            </name>
            <name>
              <surname>Samaraweera</surname>
              <given-names>CK</given-names>
            </name>
            <name>
              <surname>Nettasinghe</surname>
              <given-names>O</given-names>
            </name>
            <etal />
          </person-group>
          <comment>DistilGuard - large language models for poisoning detection in federated learning. 2025 IEEE Conference on Communications and Network Security (CNS); 2025 Sep 8-11; Avignon, France. IEEE; 2025. pp. 1-9.</comment>
          <pub-id pub-id-type="doi">10.1109/cns66487.2025.11195054</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B107">
        <label>107</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Zhang</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Jiang</surname>
              <given-names>C</given-names>
            </name>
            <name>
              <surname>Zhang</surname>
              <given-names>P</given-names>
            </name>
          </person-group>
          <article-title>Security-aware resource allocation scheme based on DRL in cloud-edge-terminal cooperative vehicular network</article-title>
          <source>IEEE Internet Things J.</source>
          <year>2024</year>
          <volume>11</volume>
          <fpage>95</fpage>
          <lpage>104</lpage>
          <pub-id pub-id-type="doi">10.1109/jiot.2023.3293497</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B108">
        <label>108</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Lu</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Pan</surname>
              <given-names>B</given-names>
            </name>
            <name>
              <surname>Yu</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Jiang</surname>
              <given-names>W</given-names>
            </name>
            <name>
              <surname>Han</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Ye</surname>
              <given-names>Z</given-names>
            </name>
          </person-group>
          <article-title>Towards energy-efficient and time-sensitive task assignment in cross-silo federated learning</article-title>
          <source>J King Saud Univ-Com.</source>
          <year>2023</year>
          <volume>35</volume>
          <fpage>63</fpage>
          <lpage>74</lpage>
          <pub-id pub-id-type="doi">10.1016/j.jksuci.2023.03.003</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B109">
        <label>109</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Xiong</surname>
              <given-names>B</given-names>
            </name>
            <name>
              <surname>Yang</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Song</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>Y</given-names>
            </name>
            <name>
              <surname>Xu</surname>
              <given-names>C</given-names>
            </name>
          </person-group>
          <article-title>Pilot: building the federated multimodal instruction tuning framework</article-title>
          <source>AAAI.</source>
          <year>2025</year>
          <volume>39</volume>
          <fpage>21716</fpage>
          <lpage>24</lpage>
          <pub-id pub-id-type="doi">10.1609/aaai.v39i20.35476</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B110">
        <label>110</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Singha</surname>
              <given-names>M</given-names>
            </name>
            <name>
              <surname>Roy</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Mehrotra</surname>
              <given-names>S</given-names>
            </name>
            <etal />
          </person-group>
          <comment>FedMVP: federated multimodal visual prompt tuning for vision-language models. 2025 IEEE/CVF International Conference on Computer Vision (ICCV); 2025 Oct 19-25; Honolulu, HI, USA. IEEE; 2025. pp. 1-10.</comment>
          <pub-id pub-id-type="doi">10.1109/iccv51701.2025.01660</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B111">
        <label>111</label>
        <nlm-citation publication-type="confproc">
          <person-group person-group-type="author">
            <name>
              <surname>Bala</surname>
              <given-names>A</given-names>
            </name>
            <name>
              <surname>Vereshchaka</surname>
              <given-names>A</given-names>
            </name>
          </person-group>
          <comment>Multimodal LLM using federated visual instruction tuning for visually impaired. ICMI '25: International Conference on Multimodal Interaction; Canberra Australia. New York, NY, USA: ACM; 2025. pp. 191-9.</comment>
          <pub-id pub-id-type="doi">10.1145/3716553.3750763</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B112">
        <label>112</label>
        <nlm-citation publication-type="book">
          <person-group person-group-type="author">
            <name>
              <surname>Wang</surname>
              <given-names>H</given-names>
            </name>
            <name>
              <surname>Zhao</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Wang</surname>
              <given-names>J</given-names>
            </name>
            <name>
              <surname>Qiang</surname>
              <given-names>Z</given-names>
            </name>
            <name>
              <surname>Qin</surname>
              <given-names>B</given-names>
            </name>
            <name>
              <surname>Liu</surname>
              <given-names>T</given-names>
            </name>
          </person-group>
          <comment>Beyond frameworks: unpacking collaboration strategies in multi-agent systems. 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers); 2025 Jun; Vienna, Austria. Stroudsburg, PA, USA: Association for Computational Linguistics; 2025. pp. 21361-75.</comment>
          <pub-id pub-id-type="doi">10.18653/v1/2025.acl-long.1037</pub-id>
        </nlm-citation>
      </ref>
      <ref id="B113">
        <label>113</label>
        <nlm-citation publication-type="journal">
          <person-group person-group-type="author">
            <name>
              <surname>Chen</surname>
              <given-names>X</given-names>
            </name>
            <name>
              <surname>Chen</surname>
              <given-names>S</given-names>
            </name>
            <name>
              <surname>Ni</surname>
              <given-names>W</given-names>
            </name>
            <etal />
          </person-group>
          <article-title>Optimal two-timescale configuration of mobile edge computing with mixed energy supply</article-title>
          <source>IEEE Trans. Smart Grid.</source>
          <year>2024</year>
          <volume>15</volume>
          <fpage>4765</fpage>
          <lpage>78</lpage>
          <pub-id pub-id-type="doi">10.1109/tsg.2024.3390772</pub-id>
        </nlm-citation>
      </ref>
    </ref-list>
  </back>
</article>