Skip to main navigation Skip to search Skip to main content

A Comprehensive Analysis of Adapter Efficiency.

  • Nandini Mundra
  • , Sumanth Doddapaneni
  • , Raj Dabre
  • , Anoop Kunchukuttan
  • , Ratish Puduppully
  • , Mitesh M. Khapra
  • Indian Institute of Technology Madras
  • National Institute Of Information And Communications Technology, Japan
  • Microsoft India
  • Agency for Science, Technology and Research (A*Star)

Research output: Conference Article in Proceeding or Book/Report chapterArticle in proceedingsResearchpeer-review

Abstract

Adapters have been positioned as a parameter-efficient fine-tuning (PEFT) approach, whereby a minimal number of parameters are added to the model and fine-tuned. However, adapters have not been sufficiently analyzed to understand if PEFT translates to benefits in training/deployment efficiency and maintainability/extensibility. Through extensive experiments on many adapters, tasks, and languages in supervised and cross-lingual zero-shot settings, we clearly show that for Natural Language Understanding (NLU) tasks, the parameter efficiency in adapters does not translate to efficiency gains compared to full fine-tuning of models. More precisely, adapters are relatively expensive to train and have slightly higher deployment latency. Furthermore, the maintainability/extensibility benefits of adapters can be achieved with simpler approaches like multi-task training via full fine-tuning, which also provide relatively faster training times. We, therefore, recommend that for moderately sized models for NLU tasks, practitioners should rely on full fine-tuning or multi-task training rather than using adapters. Our code is available at https://github.com/AI4Bharat/adapter-efficiency.
Original languageEnglish
Title of host publicationCODS-COMAD '24: Proceedings of the 7th Joint International Conference on Data Science & Management of Data (11th ACM IKDD CODS and 29th COMAD)
Number of pages19
PublisherAssociation for Computing Machinery
Publication date2024
Pages136-154
DOIs
Publication statusPublished - 2024
Externally publishedYes
EventInternational Conference on Data Science & Management of Data - Bangalore, India
Duration: 4 Jan 20247 Jan 2024
Conference number: 7
https://dl.acm.org/doi/proceedings/10.1145/3632410

Conference

ConferenceInternational Conference on Data Science & Management of Data
Number7
Country/TerritoryIndia
CityBangalore
Period04/01/202407/01/2024
Internet address

Keywords

  • computational efficiency
  • Deep Learning
  • Adapter
  • Parameter-efficient fine-tuning
  • Cross-lingual zero-shot learning
  • Multi-task training
  • NLU tasks
  • Training efficiency

Fingerprint

Dive into the research topics of 'A Comprehensive Analysis of Adapter Efficiency.'. Together they form a unique fingerprint.

Cite this