English
Related papers

Related papers: SAGE Celer 2.6 Technical Card

200 papers

Multilingual end-to-end automatic speech recognition models are attractive due to its simplicity in training and deployment. Recent work on large-scale training of such models has shown promising results compared to monolingual models.…

Computation and Language · Computer Science 2022-10-13 Ke Hu , Bo Li , Tara N. Sainath

Evaluating relevance in large-scale search systems is fundamentally constrained by the governance gap between nuanced, resource-constrained human oversight and the high-throughput requirements of production systems. While traditional…

Large Language Models (LLMs) have demonstrated remarkable capabilities at solving complex reasoning tasks with Chain-of-Thought (CoT) prompting, but their decision-making processes remain somewhat blackbox. We introduce textbfinverse…

Artificial Intelligence · Computer Science 2025-07-02 Basab Jha , Firoj Paudel , Ujjwal Puri , Zhang Yuting , Choi Donghyuk , Wang Junhao

In this report, we introduce Qwen3-ASR family, which includes two powerful all-in-one speech recognition models and a novel non-autoregressive speech forced alignment model. Qwen3-ASR-1.7B and Qwen3-ASR-0.6B are ASR models that support…

Computation and Language · Computer Science 2026-02-02 Xian Shi , Xiong Wang , Zhifang Guo , Yongqi Wang , Pei Zhang , Xinyu Zhang , Zishan Guo , Hongkun Hao , Yu Xi , Baosong Yang , Jin Xu , Jingren Zhou , Junyang Lin

Recent innovations in architecture, pre-training, and fine-tuning have led to the remarkable in-context learning and reasoning abilities of large auto-regressive language models such as LLaMA and DeepSeek. In contrast, encoders like BERT…

Computation and Language · Computer Science 2025-06-10 Lola Le Breton , Quentin Fournier , Mariam El Mezouar , John X. Morris , Sarath Chandar

Large language models often fail on multi-step reasoning due to fixed reasoning strategies that ignore problem specific difficulty. We introduce CARD (Complexity Agnostic Recursive Decomposition), a framework that predicts problem…

Computation and Language · Computer Science 2026-01-09 Kaleem Ullah Qasim , Jiashu Zhang , Hafiz Saif Ur Rehman

Retrieval-augmented generation (RAG) extends large language models (LLMs) with external data sources to enhance factual correctness and domain coverage. Modern RAG pipelines rely on large datastores, creating a significant system challenge:…

Distributed, Parallel, and Cluster Computing · Computer Science 2026-05-19 Chien-Yu Lin , Keisuke Kamahori , Yiyu Liu , Xiaoxiang Shi , Madhav Kashyap , Yile Gu , Rulin Shao , Zihao Ye , Kan Zhu , Rohan Kadekodi , Stephanie Wang , Arvind Krishnamurthy , Luis Ceze , Baris Kasikci

Sentiment Analysis (SA) is a crucial aspect of Natural Language Processing (NLP), focusing on identifying and interpreting subjective assessments in textual content. Syntactic parsing is useful in SA as it improves accuracy and provides…

Computation and Language · Computer Science 2026-02-04 Muhammad Imran , Olga Kellert , Carlos Gómez-Rodríguez

Semantically guided conditional Generative Adversarial Networks (cGANs) have become a popular approach for face editing in recent years. However, most existing methods introduce semantic masks as direct conditional inputs to the generator…

Computer Vision and Pattern Recognition · Computer Science 2022-06-17 Jiaze Sun , Binod Bhattarai , Zhixiang Chen , Tae-Kyun Kim

This technical report introduces the EXAONE 3.5 instruction-tuned language models, developed and released by LG AI Research. The EXAONE 3.5 language models are offered in three configurations: 32B, 7.8B, and 2.4B. These models feature…

This technical report introduces EXAONE 4.0, which integrates a Non-reasoning mode and a Reasoning mode to achieve both the excellent usability of EXAONE 3.5 and the advanced reasoning abilities of EXAONE Deep. To pave the way for the…

Large language models (LLMs) have proven to work well in question-answering scenarios, but real-world applications often require access to tools for live information or actuation. For this, LLMs can be extended with tools, which are often…

Software Engineering · Computer Science 2026-01-16 Robert K. Strehlow , Tobias Küster , Oskar F. Kupke , Brandon Llanque Kurps , Fikret Sivrikaya , Sahin Albayrak

Large language models (LLMs) often face a bottleneck in inference speed due to their reliance on auto-regressive decoding. Recently, parallel decoding has shown significant promise in enhancing inference efficiency. However, we have…

Computation and Language · Computer Science 2024-10-18 Yuxuan Liu , Wenyuan Li , Laizhong Cui , Hailiang Yang

While frontier large language models (LLMs) continue to push capability boundaries, their deployment remains confined to GPU-powered cloud infrastructure. We challenge this paradigm with SmallThinker, a family of LLMs natively designed -…

Large language models (LLMs) with explicit reasoning capabilities excel at mathematical reasoning yet still commit process errors, such as incorrect calculations, brittle logic, and superficially plausible but invalid steps. In this paper,…

Artificial Intelligence · Computer Science 2026-03-26 Qihao Liu , Luoxin Ye , Wufei Ma , Yu-Cheng Chou , Alan Yuille

Context-augmented generation (CAG) techniques, including RAG and ICL, require the efficient combination of multiple contexts to generate responses to user queries. Directly inputting these contexts as a sequence introduces a considerable…

Machine Learning · Computer Science 2025-02-13 Xinyu Yang , Tianqi Chen , Beidi Chen

Real-world image super-resolution (Real-ISR) must handle complex degradations and inherent reconstruction ambiguities. While generative models have improved perceptual quality, a key trade-off remains with computational cost. One-step…

Computer Vision and Pattern Recognition · Computer Science 2025-10-23 Yun Kai Zhuang

In this report we describe the development of Command A, a powerful large language model purpose-built to excel at real-world enterprise use cases. Command A is an agent-optimised and multilingual-capable model, with support for 23…

Computation and Language · Computer Science 2025-04-15 Team Cohere , : , Aakanksha , Arash Ahmadian , Marwan Ahmed , Jay Alammar , Milad Alizadeh , Yazeed Alnumay , Sophia Althammer , Arkady Arkhangorodsky , Viraat Aryabumi , Dennis Aumiller , Raphaël Avalos , Zahara Aviv , Sammie Bae , Saurabh Baji , Alexandre Barbet , Max Bartolo , Björn Bebensee , Neeral Beladia , Walter Beller-Morales , Alexandre Bérard , Andrew Berneshawi , Anna Bialas , Phil Blunsom , Matt Bobkin , Adi Bongale , Sam Braun , Maxime Brunet , Samuel Cahyawijaya , David Cairuz , Jon Ander Campos , Cassie Cao , Kris Cao , Roman Castagné , Julián Cendrero , Leila Chan Currie , Yash Chandak , Diane Chang , Giannis Chatziveroglou , Hongyu Chen , Claire Cheng , Alexis Chevalier , Justin T. Chiu , Eugene Cho , Eugene Choi , Eujeong Choi , Tim Chung , Volkan Cirik , Ana Cismaru , Pierre Clavier , Henry Conklin , Lucas Crawhall-Stein , Devon Crouse , Andres Felipe Cruz-Salinas , Ben Cyrus , Daniel D'souza , Hugo Dalla-Torre , John Dang , William Darling , Omar Darwiche Domingues , Saurabh Dash , Antoine Debugne , Théo Dehaze , Shaan Desai , Joan Devassy , Rishit Dholakia , Kyle Duffy , Ali Edalati , Ace Eldeib , Abdullah Elkady , Sarah Elsharkawy , Irem Ergün , Beyza Ermis , Marzieh Fadaee , Boyu Fan , Lucas Fayoux , Yannis Flet-Berliac , Nick Frosst , Matthias Gallé , Wojciech Galuba , Utsav Garg , Matthieu Geist , Mohammad Gheshlaghi Azar , Ellen Gilsenan-McMahon , Seraphina Goldfarb-Tarrant , Tomas Goldsack , Aidan Gomez , Victor Machado Gonzaga , Nithya Govindarajan , Manoj Govindassamy , Nathan Grinsztajn , Nikolas Gritsch , Patrick Gu , Shangmin Guo , Kilian Haefeli , Rod Hajjar , Tim Hawes , Jingyi He , Sebastian Hofstätter , Sungjin Hong , Sara Hooker , Tom Hosking , Stephanie Howe , Eric Hu , Renjie Huang , Hemant Jain , Ritika Jain , Nick Jakobi , Madeline Jenkins , JJ Jordan , Dhruti Joshi , Jason Jung , Trushant Kalyanpur , Siddhartha Rao Kamalakara , Julia Kedrzycki , Gokce Keskin , Edward Kim , Joon Kim , Wei-Yin Ko , Tom Kocmi , Michael Kozakov , Wojciech Kryściński , Arnav Kumar Jain , Komal Kumar Teru , Sander Land , Michael Lasby , Olivia Lasche , Justin Lee , Patrick Lewis , Jeffrey Li , Jonathan Li , Hangyu Lin , Acyr Locatelli , Kevin Luong , Raymond Ma , Lukáš Mach , Marina Machado , Joanne Magbitang , Brenda Malacara Lopez , Aryan Mann , Kelly Marchisio , Olivia Markham , Alexandre Matton , Alex McKinney , Dominic McLoughlin , Jozef Mokry , Adrien Morisot , Autumn Moulder , Harry Moynehan , Maximilian Mozes , Vivek Muppalla , Lidiya Murakhovska , Hemangani Nagarajan , Alekhya Nandula , Hisham Nasir , Shauna Nehra , Josh Netto-Rosen , Daniel Ohashi , James Owers-Bardsley , Jason Ozuzu , Dennis Padilla , Gloria Park , Sam Passaglia , Jeremy Pekmez , Laura Penstone , Aleksandra Piktus , Case Ploeg , Andrew Poulton , Youran Qi , Shubha Raghvendra , Miguel Ramos , Ekagra Ranjan , Pierre Richemond , Cécile Robert-Michon , Aurélien Rodriguez , Sudip Roy , Sebastian Ruder , Laura Ruis , Louise Rust , Anubhav Sachan , Alejandro Salamanca , Kailash Karthik Saravanakumar , Isha Satyakam , Alice Schoenauer Sebag , Priyanka Sen , Sholeh Sepehri , Preethi Seshadri , Ye Shen , Tom Sherborne , Sylvie Shang Shi , Sanal Shivaprasad , Vladyslav Shmyhlo , Anirudh Shrinivason , Inna Shteinbuk , Amir Shukayev , Mathieu Simard , Ella Snyder , Ava Spataru , Victoria Spooner , Trisha Starostina , Florian Strub , Yixuan Su , Jimin Sun , Dwarak Talupuru , Eugene Tarassov , Elena Tommasone , Jennifer Tracey , Billy Trend , Evren Tumer , Ahmet Üstün , Bharat Venkitesh , David Venuto , Pat Verga , Maxime Voisin , Alex Wang , Donglu Wang , Shijian Wang , Edmond Wen , Naomi White , Jesse Willman , Marysia Winkels , Chen Xia , Jessica Xie , Minjie Xu , Bowen Yang , Tan Yi-Chern , Ivan Zhang , Zhenyu Zhao , Zhoujie Zhao

Pre-trained text encoders have drawn sustaining attention in natural language processing (NLP) and shown their capability in obtaining promising results in different tasks. Recent studies illustrated that external self-supervised signals…

Computation and Language · Computer Science 2021-05-05 Yan Song , Tong Zhang , Yonggang Wang , Kai-Fu Lee

Current high-performing intracortical speech neuroprostheses achieve low word error rates but typically rely on external language models during inference, increasing memory, computation, and latency. In this work, we investigate whether…

Computation and Language · Computer Science 2026-05-26 Owais Mujtaba Khanday , Jose A. Gonzalez-Lopez , Marc Ouellet , Alberto Galdon , Gonzalo Olivares Granados
‹ Prev 1 4 5 6 7 8 10 Next ›