English
Related papers

Related papers: MaLLaM -- Malaysia Large Language Model

200 papers

Large Language Models (LLMs) have shown remarkable capabilities, not only in generating human-like text, but also in acquiring knowledge. This highlights the need to go beyond the typical Natural Language Processing downstream benchmarks…

Computation and Language · Computer Science 2025-01-03 Ahmad Mustapha , Hadi Al-Khansa , Hadi Al-Mubasher , Aya Mourad , Ranam Hamoud , Hasan El-Husseini , Marwah Al-Sakkaf , Mariette Awad

The advancements in the Large Language Model (LLM) have helped in solving several problems related to language processing. Most of the researches have focused on the English language only, because of its popularity and abundance on the…

Computation and Language · Computer Science 2024-12-31 Sanjay Chouhan , Shubha Brata Nath , Aparajita Dutta

Various large language models (LLMs) have been proposed in recent years, including closed- and open-source ones, continually setting new records on multiple benchmarks. However, the development of LLMs still faces several issues, such as…

Computation and Language · Computer Science 2024-04-05 Ruyi Gan , Ziwei Wu , Renliang Sun , Junyu Lu , Xiaojun Wu , Dixiang Zhang , Kunhao Pan , Junqing He , Yuanhe Tian , Ping Yang , Qi Yang , Hao Wang , Jiaxing Zhang , Yan Song

Large language models (LLMs) and their variants have shown extraordinary efficacy across numerous downstream natural language processing (NLP) tasks, which has presented a new vision for the development of NLP. Despite their remarkable…

Computation and Language · Computer Science 2024-01-18 Yazhou Zhang , Mengyao Wang , Youxi Wu , Prayag Tiwari , Qiuchi Li , Benyou Wang , Jing Qin

Large language models (LLMs) have revolutionized natural language processing (NLP), yet open-source multilingual LLMs remain scarce, with existing models often limited in language coverage. Such models typically prioritize well-resourced…

Computation and Language · Computer Science 2025-03-04 Yiran Zhao , Chaoqun Liu , Yue Deng , Jiahao Ying , Mahani Aljunied , Zhaodonghui Li , Lidong Bing , Hou Pong Chan , Yu Rong , Deli Zhao , Wenxuan Zhang

In this study, we introduce CT-LLM, a 2B large language model (LLM) that illustrates a pivotal shift towards prioritizing the Chinese language in developing LLMs. Uniquely initiated from scratch, CT-LLM diverges from the conventional…

This paper embarks on an exploration into the Large Language Model (LLM) datasets, which play a crucial role in the remarkable advancements of LLMs. The datasets serve as the foundational infrastructure analogous to a root system that…

Computation and Language · Computer Science 2024-02-29 Yang Liu , Jiahuan Cao , Chongyu Liu , Kai Ding , Lianwen Jin

Large Language Models (LLMs) are the cornerstones of modern artificial intelligence systems. This paper introduces Juhaina, a Arabic-English bilingual LLM specifically designed to align with the values and preferences of Arabic speakers.…

Computation and Language · Computer Science 2024-09-25 Zhaozhi Qian , Faroq Altam , Muhammad Alqurishi , Riad Souissi

As retrieval-augmented generation prevails in large language models, embedding models are becoming increasingly crucial. Despite the growing number of general embedding models, prior work often overlooks the critical role of training data…

Computation and Language · Computer Science 2025-01-16 Xinshuo Hu , Zifei Shan , Xinping Zhao , Zetian Sun , Zhenyu Liu , Dongfang Li , Shaolin Ye , Xinyuan Wei , Qian Chen , Baotian Hu , Haofen Wang , Jun Yu , Min Zhang

We introduce TableLLM, a robust large language model (LLM) with 8 billion parameters, purpose-built for proficiently handling tabular data manipulation tasks, whether they are embedded within documents or spreadsheets, catering to…

Computation and Language · Computer Science 2025-02-18 Xiaokang Zhang , Sijia Luo , Bohan Zhang , Zeyao Ma , Jing Zhang , Yang Li , Guanlin Li , Zijun Yao , Kangli Xu , Jinchang Zhou , Daniel Zhang-Li , Jifan Yu , Shu Zhao , Juanzi Li , Jie Tang

Multimodal language analysis is a rapidly evolving field that leverages multiple modalities to enhance the understanding of high-level semantics underlying human conversational utterances. Despite its significance, little research has…

Computation and Language · Computer Science 2025-04-25 Hanlei Zhang , Zhuohang Li , Yeshuang Zhu , Hua Xu , Peiwu Wang , Haige Zhu , Jie Zhou , Jinchao Zhang

We explore optimally training protein language models, an area of significant interest in biological research where guidance on best practices is limited. Most models are trained with extensive compute resources until performance gains…

Machine Learning · Computer Science 2024-11-05 Xingyi Cheng , Bo Chen , Pan Li , Jing Gong , Jie Tang , Le Song

Tokenization is a central component of natural language processing in current large language models (LLMs), enabling models to convert raw text into processable units. Although learned tokenizers are widely adopted, they exhibit notable…

Large Language Models (LLMs) play a central role in modern artificial intelligence, yet their development has been primarily focused on English, resulting in limited support for other languages. We present PLLuM (Polish Large Language…

Computation and Language · Computer Science 2025-11-07 Jan Kocoń , Maciej Piasecki , Arkadiusz Janz , Teddy Ferdinan , Łukasz Radliński , Bartłomiej Koptyra , Marcin Oleksy , Stanisław Woźniak , Paweł Walkowiak , Konrad Wojtasik , Julia Moska , Tomasz Naskręt , Bartosz Walkowiak , Mateusz Gniewkowski , Kamil Szyc , Dawid Motyka , Dawid Banach , Jonatan Dalasiński , Ewa Rudnicka , Bartłomiej Alberski , Tomasz Walkowiak , Aleksander Szczęsny , Maciej Markiewicz , Tomasz Bernaś , Hubert Mazur , Kamil Żyta , Mateusz Tykierko , Grzegorz Chodak , Tomasz Kajdanowicz , Przemysław Kazienko , Agnieszka Karlińska , Karolina Seweryn , Anna Kołos , Maciej Chrabąszcz , Katarzyna Lorenc , Aleksandra Krasnodębska , Artur Wilczek , Katarzyna Dziewulska , Paula Betscher , Zofia Cieślińska , Katarzyna Kowol , Daria Mikoś , Maciej Trzciński , Dawid Krutul , Marek Kozłowski , Sławomir Dadas , Rafał Poświata , Michał Perełkiewicz , Małgorzata Grębowiec , Maciej Kazuła , Marcin Białas , Roman Roszko , Danuta Roszko , Jurgita Vaičenonienė , Andrius Utka , Paweł Levchuk , Paweł Kowalski , Irena Prawdzic-Jankowska , Maciej Ogrodniczuk , Monika Borys , Anna Bulińska , Wiktoria Gumienna , Witold Kieraś , Dorota Komosińska , Katarzyna Krasnowska-Kieraś , Łukasz Kobyliński , Martyna Lewandowska , Marek Łaziński , Mikołaj Łątkowski , Dawid Mastalerz , Beata Milewicz , Agnieszka Anna Mykowiecka , Angelika Peljak-Łapińska , Sandra Penno , Zuzanna Przybysz , Michał Rudolf , Piotr Rybak , Karolina Saputa , Aleksandra Tomaszewska , Aleksander Wawer , Marcin Woliński , Joanna Wołoszyn , Alina Wróblewska , Bartosz Żuk , Filip Żarnecki , Konrad Kaczyński , Anna Cichosz , Zuzanna Deckert , Monika Garnys , Izabela Grabarczyk , Wojciech Janowski , Sylwia Karasińska , Aleksandra Kujawiak , Piotr Misztela , Maria Szymańska , Karolina Walkusz , Igor Siek , Jakub Kwiatkowski , Piotr Pęzik

In recent years, Large Language Models (LLMs) have demonstrated exceptional proficiency across a broad spectrum of Natural Language Processing (NLP) tasks, including Machine Translation. However, previous methods predominantly relied on…

Large language models (LLMs) have achieved remarkable success across various natural language processing tasks. However, most LLM models use traditional tokenizers like BPE and SentencePiece, which fail to capture the finer nuances of a…

Computation and Language · Computer Science 2025-05-26 Pramit Bhattacharyya , Arnab Bhattacharya

Multilingual Large Language Models (LLMs) often provide suboptimal performance on low-resource languages like Urdu. This paper introduces UrduLLaMA 1.0, a model derived from the open-source Llama-3.1-8B-Instruct architecture and continually…

Computation and Language · Computer Science 2025-02-25 Layba Fiaz , Munief Hassan Tahir , Sana Shams , Sarmad Hussain

Although instruction-tuned large language models (LLMs) have exhibited remarkable capabilities across various NLP tasks, their effectiveness on other data modalities beyond text has not been fully studied. In this work, we propose…

Computation and Language · Computer Science 2023-06-16 Chenyang Lyu , Minghao Wu , Longyue Wang , Xinting Huang , Bingshuai Liu , Zefeng Du , Shuming Shi , Zhaopeng Tu