English
Related papers

Related papers: Nyonic Technical Report

200 papers

Despite their remarkable abilities in various tasks, large language models (LLMs) still struggle with real-time information (e.g., new facts and terms) due to the knowledge cutoff in their development process. However, existing benchmarks…

Computation and Language · Computer Science 2024-10-29 Hexuan Deng , Wenxiang Jiao , Xuebo Liu , Min Zhang , Zhaopeng Tu

Large Language Models (LLMs) have become ubiquitous across various domains, transforming the way we interact with information and conduct research. However, most high-performing LLMs remain confined behind proprietary walls, hindering…

Using representations provided by a large pre-trained model has become the primary strategy for achieving state-of-the-art results in a wide range of tasks. A recently proposed large pre-trained model, wav2vec 2.0, was seminal for several…

Computation and Language · Computer Science 2025-12-01 Jonatas Grosman , Cassio Almeida , Guilherme Schardong , Hélio Lopes

We introduce two multilingual, multimodal foundation language models that power Apple Intelligence features across Apple devices and services: i a 3B-parameter on-device model optimized for Apple silicon through architectural innovations…

Machine Learning · Computer Science 2025-08-28 Ethan Li , Anders Boesen Lindbo Larsen , Chen Zhang , Xiyou Zhou , Jun Qin , Dian Ang Yap , Narendran Raghavan , Xuankai Chang , Margit Bowler , Eray Yildiz , John Peebles , Hannah Gillis Coleman , Matteo Ronchi , Peter Gray , Keen You , Anthony Spalvieri-Kruse , Ruoming Pang , Reed Li , Yuli Yang , Emad Soroush , Zhiyun Lu , Crystal Xiao , Rong Situ , Jordan Huffaker , David Griffiths , Zaid Ahmed , Peng Zhang , Daniel Parilla , Asaf Liberman , Jennifer Mallalieu , Parsa Mazaheri , Qibin Chen , Manjot Bilkhu , Aonan Zhang , Eric Wang , Dave Nelson , Michael FitzMaurice , Thomas Voice , Jeremy Liu , Josh Shaffer , Shiwen Zhao , Prasanth Yadla , Farzin Rasteh , Pengsheng Guo , Arsalan Farooq , Jeremy Snow , Stephen Murphy , Tao Lei , Minsik Cho , George Horrell , Sam Dodge , Lindsay Hislop , Sumeet Singh , Alex Dombrowski , Aiswarya Raghavan , Sasha Sirovica , Mandana Saebi , Faye Lao , Max Lam , TJ Lu , Zhaoyang Xu , Karanjeet Singh , Marc Kirchner , David Mizrahi , Rajat Arora , Haotian Zhang , Henry Mason , Lawrence Zhou , Yi Hua , Ankur Jain , Felix Bai , Joseph Astrauskas , Floris Weers , Josh Gardner , Mira Chiang , Yi Zhang , Pulkit Agrawal , Tony Sun , Quentin Keunebroek , Matthew Hopkins , Bugu Wu , Tao Jia , Chen Chen , Xingyu Zhou , Nanzhu Wang , Peng Liu , Ruixuan Hou , Rene Rauch , Yuan Gao , Afshin Dehghan , Jonathan Janke , Zirui Wang , Cha Chen , Xiaoyi Ren , Feng Nan , Josh Elman , Dong Yin , Yusuf Goren , Jeff Lai , Yiran Fei , Syd Evans , Muyang Yu , Guoli Yin , Yi Qin , Erin Feldman , Isha Garg , Aparna Rajamani , Karla Vega , Walker Cheng , TJ Collins , Hans Han , Raul Rea Menacho , Simon Yeung , Sophy Lee , Phani Mutyala , Ying-Chang Cheng , Zhe Gan , Sprite Chu , Justin Lazarow , Alessandro Pappalardo , Federico Scozzafava , Jing Lu , Erik Daxberger , Laurent Duchesne , Jen Liu , David Güera , Stefano Ligas , Mary Beth Kery , Brent Ramerth , Ciro Sannino , Marcin Eichner , Haoshuo Huang , Rui Qian , Moritz Schwarzer-Becker , David Riazati , Mingfei Gao , Bailin Wang , Jack Cackler , Yang Lu , Ransen Niu , John Dennison , Guillaume Klein , Jeffrey Bigham , Deepak Gopinath , Navid Shiee , Darren Botten , Guillaume Tartavel , Alex Guillen Garcia , Sam Xu , Victoria MönchJuan Haladjian , Zi-Yi Dou , Matthias Paulik , Adolfo Lopez Mendez , Zhen Li , Hong-You Chen , Chao Jia , Dhaval Doshi , Zhengdong Zhang , Raunak Manjani , Aaron Franklin , Zhile Ren , David Chen , Artsiom Peshko , Nandhitha Raghuram , Hans Hao , Jiulong Shan , Kavya Nerella , Ramsey Tantawi , Vivek Kumar , Saiwen Wang , Brycen Wershing , Bhuwan Dhingra , Dhruti Shah , Ob Adaranijo , Xin Zheng , Tait Madsen , Hadas Kotek , Chang Liu , Yin Xia , Hanli Li , Suma Jayaram , Yanchao Sun , Ahmed Fakhry , Vasileios Saveris , Dustin Withers , Yanghao Li , Alp Aygar , Andres Romero Mier Y Teran , Kaiwei Huang , Mark Lee , Xiujun Li , Yuhong Li , Tyler Johnson , Jay Tang , Joseph Yitan Cheng , Futang Peng , Andrew Walkingshaw , Lucas Guibert , Abhishek Sharma , Cheng Shen , Piotr Maj , Yasutaka Tanaka , You-Cyuan Jhang , Vivian Ma , Tommi Vehvilainen , Kelvin Zou , Jeff Nichols , Matthew Lei , David Qiu , Yihao Qian , Gokul Santhanam , Wentao Wu , Yena Han , Dominik Moritz , Haijing Fu , Mingze Xu , Vivek Rathod , Jian Liu , Louis D'hauwe , Qin Ba , Haitian Sun , Haoran Yan , Philipp Dufter , Anh Nguyen , Yihao Feng , Emma Wang , Keyu He , Rahul Nair , Sanskruti Shah , Jiarui Lu , Patrick Sonnenberg , Jeremy Warner , Yuanzhi Li , Bowen Pan , Ziyi Zhong , Joe Zhou , Sam Davarnia , Olli Saarikivi , Irina Belousova , Rachel Burger , Shang-Chen Wu , Di Feng , Bas Straathof , James Chou , Yuanyang Zhang , Marco Zuliani , Eduardo Jimenez , Abhishek Sundararajan , Xianzhi Du , Chang Lan , Nilesh Shahdadpuri , Peter Grasch , Sergiu Sima , Josh Newnham , Varsha Paidi , Jianyu Wang , Kaelen Haag , Alex Braunstein , Daniele Molinari , Richard Wei , Brenda Yang , Nicholas Lusskin , Joanna Arreaza-Taylor , Meng Cao , Nicholas Seidl , Simon Wang , Jiaming Hu , Yiping Ma , Mengyu Li , Kieran Liu , Hang Su , Sachin Ravi , Chong Wang , Xin Wang , Kevin Smith , Haoxuan You , Binazir Karimzadeh , Rui Li , Jinhao Lei , Wei Fang , Alec Doane , Sam Wiseman , Ismael Fernandez , Jane Li , Andrew Hansen , Javier Movellan , Christopher Neubauer , Hanzhi Zhou , Chris Chaney , Nazir Kamaldin , Valentin Wolf , Fernando Bermúdez-Medina , Joris Pelemans , Peter Fu , Howard Xing , Xiang Kong , Wayne Shan , Gabriel Jacoby-Cooper , Dongcai Shen , Tom Gunter , Guillaume Seguin , Fangping Shi , Shiyu Li , Yang Xu , Areeba Kamal , Dan Masi , Saptarshi Guha , Qi Zhu , Jenna Thibodeau , Changyuan Zhang , Rebecca Callahan , Charles Maalouf , Wilson Tsao , Boyue Li , Qingqing Cao , Naomy Sabo , Cheng Leong , Yi Wang , Anupama Mann Anupama , Colorado Reed , Kenneth Jung , Zhifeng Chen , Mohana Prasad Sathya Moorthy , Yifei He , Erik Hornberger , Devi Krishna , Senyu Tong , Michael , Lee , David Haldimann , Yang Zhao , Bowen Zhang , Chang Gao , Chris Bartels , Sushma Rao , Nathalie Tran , Simon Lehnerer , Co Giang , Patrick Dong , Junting Pan , Biyao Wang , Dongxu Li , Mehrdad Farajtabar , Dongseong Hwang , Grace Duanmu , Eshan Verma , Sujeeth Reddy , Qi Shan , Hongbin Gao , Nan Du , Pragnya Sridhar , Forrest Huang , Yingbo Wang , Nikhil Bhendawade , Diane Zhu , Sai Aitharaju , Fred Hohman , Lauren Gardiner , Chung-Cheng Chiu , Yinfei Yang , Alper Kokmen , Frank Chu , Ke Ye , Kaan Elgin , Oron Levy , John Park , Donald Zhang , Eldon Schoop , Nina Wenzel , Michael Booker , Hyunjik Kim , Chinguun Erdenebileg , Nan Dun , Eric Liang Yang , Priyal Chhatrapati , Vishaal Mahtani , Haiming Gang , Kohen Chia , Deepa Seshadri , Donghan Yu , Yan Meng , Kelsey Peterson , Zhen Yang , Yongqiang Wang , Carina Peng , Doug Kang , Anuva Agarwal , Albert Antony , Juan Lao Tebar , Albin Madappally Jose , Regan Poston , Andy De Wang , Gerard Casamayor , Elmira Amirloo , Violet Yao , Wojciech Kryscinski , Kun Duan , Lezhi L

Large Language Models (LLMs) have demonstrated significant potential in transforming clinical applications. In this study, we investigate the efficacy of four techniques in adapting LLMs for clinical use-cases: continuous pretraining,…

Modern language models rely on static vocabularies, fixed before pretraining, in contrast to the adaptive vocabulary acquisition observed in human language learning. To bridge this gap, we introduce vocabulary curriculum learning, an…

Computation and Language · Computer Science 2025-02-26 Fangyuan Yu

Large language models (LLMs) have demonstrated prowess in a wide range of tasks. However, many LLMs exhibit significant performance discrepancies between high- and low-resource languages. To mitigate this challenge, we present FuxiTranyu,…

Computation and Language · Computer Science 2024-10-29 Haoran Sun , Renren Jin , Shaoyang Xu , Leiyu Pan , Supryadi , Menglong Cui , Jiangcun Du , Yikun Lei , Lei Yang , Ling Shi , Juesi Xiao , Shaolin Zhu , Deyi Xiong

We introduce RakutenAI-7B, a suite of Japanese-oriented large language models that achieve the best performance on the Japanese LM Harness benchmarks among the open 7B models. Along with the foundation model, we release instruction- and…

State-of-the-art multilingual models depend on vocabularies that cover all of the languages the model will expect to see at inference time, but the standard methods for generating those vocabularies are not ideal for massively multilingual…

Computation and Language · Computer Science 2020-10-27 Hyung Won Chung , Dan Garrette , Kiat Chuan Tan , Jason Riesa

Large transformer-based language models, e.g. BERT and GPT-3, outperform previous architectures on most natural language processing tasks. Such language models are first pre-trained on gigantic corpora of text and later used as base-model…

Computation and Language · Computer Science 2022-11-16 Pieter Delobelle , Thomas Winters , Bettina Berendt

High-stakes decision making involves reasoning under uncertainty about the future. In this work, we train language models to make predictions on open-ended forecasting questions. To scale up training data, we synthesize novel forecasting…

Machine Learning · Computer Science 2026-01-06 Nikhil Chandak , Shashwat Goel , Ameya Prabhu , Moritz Hardt , Jonas Geiping

Alignment is a crucial step to enhance the instruction-following and conversational abilities of language models. Despite many recent work proposing new algorithms, datasets, and training pipelines, there is a lack of comprehensive studies…

Computation and Language · Computer Science 2024-10-04 Xiao Yu , Qingyang Wu , Yu Li , Zhou Yu

Large language models (LLMs) have revolutionized various domains but still struggle with non-Latin scripts and low-resource languages. This paper addresses the critical challenge of improving multilingual performance without extensive…

Computation and Language · Computer Science 2025-01-08 Somnath Kumar , Vaibhav Balloli , Mercy Ranjit , Kabir Ahuja , Sunayana Sitaram , Kalika Bali , Tanuja Ganu , Akshay Nambi

As retrieval-augmented generation prevails in large language models, embedding models are becoming increasingly crucial. Despite the growing number of general embedding models, prior work often overlooks the critical role of training data…

Computation and Language · Computer Science 2025-01-16 Xinshuo Hu , Zifei Shan , Xinping Zhao , Zetian Sun , Zhenyu Liu , Dongfang Li , Shaolin Ye , Xinyuan Wei , Qian Chen , Baotian Hu , Haofen Wang , Jun Yu , Min Zhang

Multimodal foundation models, such as Gemini and ChatGPT, have revolutionized human-machine interactions by seamlessly integrating various forms of data. Developing a universal spoken language model that comprehends a wide range of natural…

Machine learning has brought striking advances in multilingual natural language processing capabilities over the past year. For example, the latest techniques have improved the state-of-the-art performance on the XTREME multilingual…

Computation and Language · Computer Science 2021-10-08 Sebastian Ruder , Noah Constant , Jan Botha , Aditya Siddhant , Orhan Firat , Jinlan Fu , Pengfei Liu , Junjie Hu , Dan Garrette , Graham Neubig , Melvin Johnson

The recent breakthroughs in Large Language Models (LLMs) have mostly focused on languages with easily available and sufficient resources, such as English. However, there remains a significant gap for languages that lack sufficient…

Computation and Language · Computer Science 2024-03-20 Louis Owen , Vishesh Tripathi , Abhay Kumar , Biddwan Ahmed

We introduce llama-embed-nemotron-8b, an open-weights text embedding model that achieves state-of-the-art performance on the Multilingual Massive Text Embedding Benchmark (MMTEB) leaderboard as of October 21, 2025. While recent models show…

Computation and Language · Computer Science 2025-11-11 Yauhen Babakhin , Radek Osmulski , Ronay Ak , Gabriel Moreira , Mengyao Xu , Benedikt Schifferer , Bo Liu , Even Oldridge

Most large language models are fine-tuned using either expensive human-annotated data or GPT-4 generated data which cannot guarantee performance in certain domains. We argue that although the web-crawled data often has formatting errors…

Computation and Language · Computer Science 2024-08-16 Jing Zhou , Chenglin Jiang , Wei Shen , Xiao Zhou , Xiaonan He

Large language models (LLMs) have exhibited remarkable capabilities and achieved significant breakthroughs across various domains, leading to their widespread adoption in recent years. Building on this progress, we investigate their…

Artificial Intelligence · Computer Science 2025-10-27 Xiaochong Lan , Jie Feng , Jiahuan Lei , Xinlei Shi , Yong Li