Zero-Shot Cross-Lingual Transfer using Prefix-Based Adaptation

2510.24619v1 cs.CL, cs.AI, cs.LG, I.2.7 2025-10-30

Авторы:

Snegha A, Sayambhu Sen, Piyush Singh Pasi, Abhishek Singhania, Preethi Jyothi

Abstract

With the release of new large language models (LLMs) like Llama and Mistral, zero-shot cross-lingual transfer has become increasingly feasible due to their multilingual pretraining and strong generalization capabilities. However, adapting these decoder-only LLMs to new tasks across languages remains challenging. While parameter-efficient fine-tuning (PeFT) techniques like Low-Rank Adaptation (LoRA) are widely used, prefix-based techniques such as soft prompt tuning, prefix tuning, and Llama Adapter are less explored, especially for zero-shot transfer in decoder-only models. We present a comprehensive study of three prefix-based methods for zero-shot cross-lingual transfer from English to 35+ high- and low-resource languages. Our analysis further explores transfer across linguistic families and scripts, as well as the impact of scaling model sizes from 1B to 24B. With Llama 3.1 8B, prefix methods outperform LoRA-baselines by up to 6% on the Belebele benchmark. Similar improvements were observed with Mistral v0.3 7B as well. Despite using only 1.23M learning parameters with prefix tuning, we achieve consistent improvements across diverse benchmarks. These findings highlight the potential of prefix-based techniques as an effective and scalable alternative to LoRA, particularly in low-resource multilingual settings.

Ссылки и действия

Читать на arXiv Скачать PDF

Дополнительные ресурсы:

Zero-Shot Cross-Lingual Transfer using Prefix-Based Adaptation

Авторы:

Abstract

Ссылки и действия

Связанные статьи

Slim-SC: Thought Pruning for Efficient Scaling with Self-Consistency

ObfusQAte: A Proposed Framework to Evaluate LLM Robustness on Obfuscated Factual...

Навигация