Monolingual or Multilingual Instruction Tuning: Which Makes a Better Alpaca

Pinzhen Chen,Shaoxiong Ji,Nikolay Bogoychev, Andrey Kutuzov,Barry Haddow,Kenneth Heafield

arXiv (Cornell University)（2023）

引用 0|浏览24

暂无评分

摘要

Foundational large language models (LLMs) can be instruction-tuned to perform open-domain question answering, facilitating applications like chat assistants. While such efforts are often carried out in a single language, we empirically analyze cost-efficient strategies for multilingual scenarios. Our study employs the Alpaca dataset and machine translations of it to form multilingual data, which is then used to tune LLMs through either low-rank adaptation or full-parameter training. Under a controlled computation budget, comparisons show that multilingual tuning is on par or better than tuning a model for each language. Furthermore, multilingual tuning with downsampled data can be as powerful and more robust. Our findings serve as a guide for expanding language support through instruction tuning.

查看译文

关键词

multilingual instruction tuning,alpaca

AI 理解论文

溯源树

样例

生成溯源树，研究论文发展脉络

Chat Paper

正在生成论文摘要