Orion-14B: Open-source Multilingual Large Language Models
Chen, Du, Huang, Yi, Li, Xiaopu, Li, Yongqiang, Liu, Yongqiang, Pan, Haihui, Xu, Leichao, Zhang, Dacheng, Zhang, Zhipeng, Han, Kun
–arXiv.org Artificial Intelligence
In this study, we introduce Orion-14B, a collection of multilingual large language models with 14 billion parameters. We utilize a data scheduling approach to train a foundational model on a diverse corpus of 2.5 trillion tokens, sourced from texts in English, Chinese, Japanese, Korean, and other languages. Additionally, we fine-tuned a series of models tailored for conversational applications and other specific use cases. Our evaluation results demonstrate that Orion-14B achieves state-of-the-art performance across a broad spectrum of tasks.
arXiv.org Artificial Intelligence
Jan-20-2024
- Country:
- Europe
- Asia
- Middle East > Jordan (0.04)
- China (0.04)
- Afghanistan > Parwan Province
- Charikar (0.04)
- Genre:
- Research Report > New Finding (0.68)
- Industry:
- Education (0.68)
- Technology: