Adapting Large Language Models for Document-Level Machine Translation

Adapting Large Language Models for Document-Level Machine Translation [49.7]
大規模言語モデル(LLM)は、様々な自然言語処理(NLP)タスクにおいて大きな進歩を遂げている。近年の研究では、中程度のLLMはタスク固有の微調整の後、より大きなLLMよりも優れていることが示されている。
論文参考訳（メタデータ） (Fri, 12 Jan 2024 09:29:13 GMT)
LLMの機械翻訳への応用。fine tuningの効果など実験結果が多く参考になる。
「We find that the PEFT approach yields superior overall performance compared to the FFT approach」（ただしFFTのほうがデータ効率は高いとのこと）がとても興味深い

コメントを残す

コメントを残す コメントをキャンセル