haoranxu
/

ALMA-13B-R

Text Generation

text-generation-inference

Inference Endpoints

Model card Files Files and versions Community

haoranxu commited on Jan 18, 2024

Commit

8e68401

·

verified ·

1 Parent(s): 13e39b0

Update README.md

Files changed (1) hide show

README.md +10 -1

README.md CHANGED Viewed

@@ -3,7 +3,16 @@ license: mit
 ---
 **[ALMA-R](https://arxiv.org/abs/2401.08417)** builds upon [ALMA models](https://arxiv.org/abs/2309.11674), with further LoRA fine-tuning with our proposed **Contrastive Preference Optimization (CPO)** as opposed to the Supervised Fine-tuning used in ALMA. CPO fine-tuning requires our [triplet preference data](https://huggingface.co/datasets/haoranxu/ALMA-R-Preference) for preference learning. ALMA-R now can matches or even exceeds GPT-4 or WMT winners!
 # Download ALMA(-R) Models and Dataset 🚀
 We release six translation models presented in the paper:

 ---
 **[ALMA-R](https://arxiv.org/abs/2401.08417)** builds upon [ALMA models](https://arxiv.org/abs/2309.11674), with further LoRA fine-tuning with our proposed **Contrastive Preference Optimization (CPO)** as opposed to the Supervised Fine-tuning used in ALMA. CPO fine-tuning requires our [triplet preference data](https://huggingface.co/datasets/haoranxu/ALMA-R-Preference) for preference learning. ALMA-R now can matches or even exceeds GPT-4 or WMT winners!
+```
+@misc{xu2024contrastive,
+      title={Contrastive Preference Optimization: Pushing the Boundaries of LLM Performance in Machine Translation},
+      author={Haoran Xu and Amr Sharaf and Yunmo Chen and Weiting Tan and Lingfeng Shen and Benjamin Van Durme and Kenton Murray and Young Jin Kim},
+      year={2024},
+      eprint={2401.08417},
+      archivePrefix={arXiv},
+      primaryClass={cs.CL}
+}
+```
 # Download ALMA(-R) Models and Dataset 🚀
 We release six translation models presented in the paper: