English
Related papers

Related papers: Command R7B Arabic: A Small, Enterprise Focused, M…

200 papers

While recent Arabic NLP benchmarks focus on scale, they often rely on synthetic or translated data which may benefit from deeper linguistic verification. We introduce ALPS (Arabic Linguistic & Pragmatic Suite), a native, expert-curated…

Computation and Language · Computer Science 2026-02-20 Hussein S. Al-Olimat , Ahmad Alshareef

Large language models (LLMs) exhibit remarkable capabilities across diverse tasks, yet aligning them efficiently and effectively with human expectations remains a critical challenge. This thesis advances LLM alignment by introducing novel…

Computation and Language · Computer Science 2025-06-12 Yuxin Jiang

The rapid advancements in Large Language Models (LLMs) have led to significant improvements in various natural language processing tasks. However, the evaluation of LLMs' legal knowledge, particularly in non-English languages such as…

The performance of large language models (LLMs) and large multimodal models (LMMs) depends heavily on the quality and scale of their pre-training datasets. Recent research shows that large multimodal models trained on natural documents…

Computation and Language · Computer Science 2025-11-11 Khalil Hennara , Ahmad Bastati , Muhammad Hreden , Mohamed Motasim Hamed , Zeina Aldallal , Sara Chrouf , Safwan AlModhayan

The rapid advancement of Large Language Models (LLMs) in the realm of mathematical reasoning necessitates comprehensive evaluations to gauge progress and inspire future directions. Existing assessments predominantly focus on problem-solving…

Computation and Language · Computer Science 2024-06-05 Xiaoyuan Li , Wenjie Wang , Moxin Li , Junrong Guo , Yang Zhang , Fuli Feng

Large language models (LLMs) deployed as agents solve user-specified tasks over multiple steps while keeping the required manual engagement to a minimum. Crucially, such LLMs need to ground their generations in any feedback obtained to…

Computation and Language · Computer Science 2025-02-19 Jonas Gehring , Kunhao Zheng , Jade Copet , Vegard Mella , Quentin Carbonneaux , Taco Cohen , Gabriel Synnaeve

This research's primary motivation of this study is to address the high hardware and computational demands typically associated with LLMs.Therefore,our goal is to find a balance between model lightness and performance,striving to maximize…

Computation and Language · Computer Science 2024-03-27 Chih-Wei Song , Yin-Te Tsai

In this report we describe the development of Command A, a powerful large language model purpose-built to excel at real-world enterprise use cases. Command A is an agent-optimised and multilingual-capable model, with support for 23…

Computation and Language · Computer Science 2025-04-15 Team Cohere , : , Aakanksha , Arash Ahmadian , Marwan Ahmed , Jay Alammar , Milad Alizadeh , Yazeed Alnumay , Sophia Althammer , Arkady Arkhangorodsky , Viraat Aryabumi , Dennis Aumiller , Raphaël Avalos , Zahara Aviv , Sammie Bae , Saurabh Baji , Alexandre Barbet , Max Bartolo , Björn Bebensee , Neeral Beladia , Walter Beller-Morales , Alexandre Bérard , Andrew Berneshawi , Anna Bialas , Phil Blunsom , Matt Bobkin , Adi Bongale , Sam Braun , Maxime Brunet , Samuel Cahyawijaya , David Cairuz , Jon Ander Campos , Cassie Cao , Kris Cao , Roman Castagné , Julián Cendrero , Leila Chan Currie , Yash Chandak , Diane Chang , Giannis Chatziveroglou , Hongyu Chen , Claire Cheng , Alexis Chevalier , Justin T. Chiu , Eugene Cho , Eugene Choi , Eujeong Choi , Tim Chung , Volkan Cirik , Ana Cismaru , Pierre Clavier , Henry Conklin , Lucas Crawhall-Stein , Devon Crouse , Andres Felipe Cruz-Salinas , Ben Cyrus , Daniel D'souza , Hugo Dalla-Torre , John Dang , William Darling , Omar Darwiche Domingues , Saurabh Dash , Antoine Debugne , Théo Dehaze , Shaan Desai , Joan Devassy , Rishit Dholakia , Kyle Duffy , Ali Edalati , Ace Eldeib , Abdullah Elkady , Sarah Elsharkawy , Irem Ergün , Beyza Ermis , Marzieh Fadaee , Boyu Fan , Lucas Fayoux , Yannis Flet-Berliac , Nick Frosst , Matthias Gallé , Wojciech Galuba , Utsav Garg , Matthieu Geist , Mohammad Gheshlaghi Azar , Ellen Gilsenan-McMahon , Seraphina Goldfarb-Tarrant , Tomas Goldsack , Aidan Gomez , Victor Machado Gonzaga , Nithya Govindarajan , Manoj Govindassamy , Nathan Grinsztajn , Nikolas Gritsch , Patrick Gu , Shangmin Guo , Kilian Haefeli , Rod Hajjar , Tim Hawes , Jingyi He , Sebastian Hofstätter , Sungjin Hong , Sara Hooker , Tom Hosking , Stephanie Howe , Eric Hu , Renjie Huang , Hemant Jain , Ritika Jain , Nick Jakobi , Madeline Jenkins , JJ Jordan , Dhruti Joshi , Jason Jung , Trushant Kalyanpur , Siddhartha Rao Kamalakara , Julia Kedrzycki , Gokce Keskin , Edward Kim , Joon Kim , Wei-Yin Ko , Tom Kocmi , Michael Kozakov , Wojciech Kryściński , Arnav Kumar Jain , Komal Kumar Teru , Sander Land , Michael Lasby , Olivia Lasche , Justin Lee , Patrick Lewis , Jeffrey Li , Jonathan Li , Hangyu Lin , Acyr Locatelli , Kevin Luong , Raymond Ma , Lukáš Mach , Marina Machado , Joanne Magbitang , Brenda Malacara Lopez , Aryan Mann , Kelly Marchisio , Olivia Markham , Alexandre Matton , Alex McKinney , Dominic McLoughlin , Jozef Mokry , Adrien Morisot , Autumn Moulder , Harry Moynehan , Maximilian Mozes , Vivek Muppalla , Lidiya Murakhovska , Hemangani Nagarajan , Alekhya Nandula , Hisham Nasir , Shauna Nehra , Josh Netto-Rosen , Daniel Ohashi , James Owers-Bardsley , Jason Ozuzu , Dennis Padilla , Gloria Park , Sam Passaglia , Jeremy Pekmez , Laura Penstone , Aleksandra Piktus , Case Ploeg , Andrew Poulton , Youran Qi , Shubha Raghvendra , Miguel Ramos , Ekagra Ranjan , Pierre Richemond , Cécile Robert-Michon , Aurélien Rodriguez , Sudip Roy , Sebastian Ruder , Laura Ruis , Louise Rust , Anubhav Sachan , Alejandro Salamanca , Kailash Karthik Saravanakumar , Isha Satyakam , Alice Schoenauer Sebag , Priyanka Sen , Sholeh Sepehri , Preethi Seshadri , Ye Shen , Tom Sherborne , Sylvie Shang Shi , Sanal Shivaprasad , Vladyslav Shmyhlo , Anirudh Shrinivason , Inna Shteinbuk , Amir Shukayev , Mathieu Simard , Ella Snyder , Ava Spataru , Victoria Spooner , Trisha Starostina , Florian Strub , Yixuan Su , Jimin Sun , Dwarak Talupuru , Eugene Tarassov , Elena Tommasone , Jennifer Tracey , Billy Trend , Evren Tumer , Ahmet Üstün , Bharat Venkitesh , David Venuto , Pat Verga , Maxime Voisin , Alex Wang , Donglu Wang , Shijian Wang , Edmond Wen , Naomi White , Jesse Willman , Marysia Winkels , Chen Xia , Jessica Xie , Minjie Xu , Bowen Yang , Tan Yi-Chern , Ivan Zhang , Zhenyu Zhao , Zhoujie Zhao

In recent years, Large Language Models (LLMs) have achieved almost human-like performance on various tasks. While some LLMs have been trained on multilingual data, most of the training data is in English; hence, their performance in English…

HTR models development has become a conventional step for digital humanities projects. The performance of these models, often quite high, relies on manual transcription and numerous handwritten documents. Although the method has proven…

Computer Vision and Pattern Recognition · Computer Science 2022-11-30 Lucas Noëmie , Clément Salah , Chahan Vidal-Gorène

The emergence of ChatGPT marked a transformative milestone for Artificial Intelligence (AI), showcasing the remarkable potential of Large Language Models (LLMs) to generate human-like text. This wave of innovation has revolutionized how we…

Computation and Language · Computer Science 2025-10-16 Shahad Al-Khalifa , Nadir Durrani , Hend Al-Khalifa , Firoj Alam

Large Language Models (LLMs) have the potential to accelerate small molecule drug design due to their ability to reason about information from diverse sources and formats. However, their practical utility remains unclear due to the lack of…

Despite the success of large language models (LLMs) in Text-to-SQL tasks, open-source LLMs encounter challenges in contextual understanding and response coherence. To tackle these issues, we present \ours, a systematic methodology tailored…

Computation and Language · Computer Science 2024-05-14 Xiaojun Chen , Tianle Wang , Tianhao Qiu , Jianbin Qin , Min Yang

Multimodal Machine Learning (MML) aims to integrate and analyze information from diverse modalities, such as text, audio, and visuals, enabling machines to address complex tasks like sentiment analysis, emotion recognition, and multimedia…

Computation and Language · Computer Science 2025-08-22 Abdelhamid Haouhat , Slimane Bellaouar , Attia Nehar , Hadda Cherroun , Ahmed Abdelali

This paper addresses the classification of Arabic text data in the field of Natural Language Processing (NLP), with a particular focus on Natural Language Inference (NLI) and Contradiction Detection (CD). Arabic is considered a…

Computation and Language · Computer Science 2023-07-28 Mohammad Majd Saad Al Deen , Maren Pielka , Jörn Hees , Bouthaina Soulef Abdou , Rafet Sifa

The deployment of Large Language Models (LLMs) in real-world applications presents both opportunities and challenges, particularly in multilingual and code-mixed communication settings. This research evaluates the performance of seven…

Computation and Language · Computer Science 2024-06-14 Millicent Ochieng , Varun Gumma , Sunayana Sitaram , Jindong Wang , Vishrav Chaudhary , Keshet Ronen , Kalika Bali , Jacki O'Neill

Large Language Models (LLMs) have demonstrated remarkable capabilities on various tasks, while the further evolvement is limited to the lack of high-quality training data. In addition, traditional training approaches rely too much on…

Computation and Language · Computer Science 2025-02-14 Peidong Wang , Ming Wang , Zhiming Ma , Xiaocui Yang , Shi Feng , Daling Wang , Yifei Zhang , Kaisong Song

In this report, we introduce Qwen2.5, a comprehensive series of large language models (LLMs) designed to meet diverse needs. Compared to previous iterations, Qwen 2.5 has been significantly improved during both the pre-training and…

We present DialectalArabicMMLU, a new benchmark for evaluating the performance of large language models (LLMs) across Arabic dialects. While recently developed Arabic and multilingual benchmarks have advanced LLM evaluation for Modern…