Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models
Neha Sengupta, Sunil Kumar Sahu, Bokang Jia, Satheesh Katipomu, Haonan, Li, Fajri Koto, William Marshall, Gurpreet Gosal, Cynthia Liu, Zhiming Chen,, Osama Mohammed Afzal, Samta Kamboj, Onkar Pandit, Rahul Pal, Lalit Pradhan,, Zain Muhammad Mujahid, Massa Baali, Xudong Han

TL;DR
Jais and Jais-chat are Arabic-centric large language models with 13 billion parameters, outperforming existing models in Arabic and showing competitive English performance, aimed at advancing Arabic NLP research.
Contribution
Introduction of the first open Arabic-centric foundation and instruction-tuned LLMs based on GPT-3 architecture, with extensive evaluation and open release for research.
Findings
Superior Arabic language understanding and reasoning
Competitive English performance despite less English data
Open release to foster Arabic NLP research
Abstract
We introduce Jais and Jais-chat, new state-of-the-art Arabic-centric foundation and instruction-tuned open generative large language models (LLMs). The models are based on the GPT-3 decoder-only architecture and are pretrained on a mixture of Arabic and English texts, including source code in various programming languages. With 13 billion parameters, they demonstrate better knowledge and reasoning capabilities in Arabic than any existing open Arabic and multilingual models by a sizable margin, based on extensive evaluation. Moreover, the models are competitive in English compared to English-centric open models of similar size, despite being trained on much less English data. We provide a detailed description of the training, the tuning, the safety alignment, and the evaluation of the models. We release two open versions of the model -- the foundation Jais model, and an instruction-tuned…
Peer Reviews
No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.
Code & Models
- 🤗inceptionai/jais-13bmodel· 489 dl· ♡ 177489 dl♡ 177
- 🤗inceptionai/jais-13b-chatmodel· 9.5k dl· ♡ 1659.5k dl♡ 165
- 🤗asas-ai/jais_13B_8bitmodel· 65 dl· ♡ 965 dl♡ 9
- 🤗asas-ai/jais-13b-chat-8bitmodel· 9 dl· ♡ 29 dl♡ 2
- 🤗poiccard/jais-13b-chat-adnmodel· 9 dl9 dl
- 🤗hussain2030/jais13bchat2model· 7 dl· ♡ 17 dl♡ 1
- 🤗inceptionai/jais-30b-v1model· 5 dl· ♡ 255 dl♡ 25
- 🤗derek-thomas/jais-13b-chat-hfmodel· 7 dl· ♡ 57 dl♡ 5
- 🤗inceptionai/jais-30b-chat-v1model· 4 dl· ♡ 244 dl♡ 24
- 🤗brainiac-origin/jais-chat-30b-8bitmodel· 4 dl· ♡ 14 dl♡ 1
Videos
No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.
Taxonomy
TopicsTopic Modeling · Natural Language Processing Techniques · Machine Learning in Healthcare
MethodsMulti-Head Attention · Attention Is All You Need · Residual Connection · 15 Ways to Contact How can i speak to someone at Delta Airlines · Layer Normalization · Softmax · {Dispute@FaQ-s}How to file a dispute with Expedia? · Refunds@Expedia|||How do I get a full refund from Expedia? · Dense Connections · Cosine Annealing
