中文
相关论文

相关论文: TULIP: Adapting Open-Source Large Language Models …

200 篇论文

Large language models (LLMs) have demonstrated remarkable open-domain capabilities. LLMs tailored for a domain are typically trained entirely on domain corpus to excel at handling domain-specific tasks. In this work, we explore an…

计算与语言 · 计算机科学 2026-01-13 Yong Xie , Karan Aggarwal , Aitzaz Ahmad

This paper presents a novel approach to fine-tuning the Qwen2-1.5B model for Arabic language processing using Quantized Low-Rank Adaptation (QLoRA) on a system with only 4GB VRAM. We detail the process of adapting this large language model…

计算与语言 · 计算机科学 2024-12-24 Prakash Aryan

Configuring computational fluid dynamics (CFD) simulations typically demands extensive domain expertise, limiting broader access. Although large language models (LLMs) have advanced scientific computing, their use in automating CFD…

流体动力学 · 物理学 2025-12-30 Zhehao Dong , Zhen Lu , Yue Yang

Large Language Models (LLMs) play a central role in modern artificial intelligence, yet their development has been primarily focused on English, resulting in limited support for other languages. We present PLLuM (Polish Large Language…

计算与语言 · 计算机科学 2025-11-07 Jan Kocoń , Maciej Piasecki , Arkadiusz Janz , Teddy Ferdinan , Łukasz Radliński , Bartłomiej Koptyra , Marcin Oleksy , Stanisław Woźniak , Paweł Walkowiak , Konrad Wojtasik , Julia Moska , Tomasz Naskręt , Bartosz Walkowiak , Mateusz Gniewkowski , Kamil Szyc , Dawid Motyka , Dawid Banach , Jonatan Dalasiński , Ewa Rudnicka , Bartłomiej Alberski , Tomasz Walkowiak , Aleksander Szczęsny , Maciej Markiewicz , Tomasz Bernaś , Hubert Mazur , Kamil Żyta , Mateusz Tykierko , Grzegorz Chodak , Tomasz Kajdanowicz , Przemysław Kazienko , Agnieszka Karlińska , Karolina Seweryn , Anna Kołos , Maciej Chrabąszcz , Katarzyna Lorenc , Aleksandra Krasnodębska , Artur Wilczek , Katarzyna Dziewulska , Paula Betscher , Zofia Cieślińska , Katarzyna Kowol , Daria Mikoś , Maciej Trzciński , Dawid Krutul , Marek Kozłowski , Sławomir Dadas , Rafał Poświata , Michał Perełkiewicz , Małgorzata Grębowiec , Maciej Kazuła , Marcin Białas , Roman Roszko , Danuta Roszko , Jurgita Vaičenonienė , Andrius Utka , Paweł Levchuk , Paweł Kowalski , Irena Prawdzic-Jankowska , Maciej Ogrodniczuk , Monika Borys , Anna Bulińska , Wiktoria Gumienna , Witold Kieraś , Dorota Komosińska , Katarzyna Krasnowska-Kieraś , Łukasz Kobyliński , Martyna Lewandowska , Marek Łaziński , Mikołaj Łątkowski , Dawid Mastalerz , Beata Milewicz , Agnieszka Anna Mykowiecka , Angelika Peljak-Łapińska , Sandra Penno , Zuzanna Przybysz , Michał Rudolf , Piotr Rybak , Karolina Saputa , Aleksandra Tomaszewska , Aleksander Wawer , Marcin Woliński , Joanna Wołoszyn , Alina Wróblewska , Bartosz Żuk , Filip Żarnecki , Konrad Kaczyński , Anna Cichosz , Zuzanna Deckert , Monika Garnys , Izabela Grabarczyk , Wojciech Janowski , Sylwia Karasińska , Aleksandra Kujawiak , Piotr Misztela , Maria Szymańska , Karolina Walkusz , Igor Siek , Jakub Kwiatkowski , Piotr Pęzik

Recent research has shown that smaller language models can acquire substantial reasoning abilities when fine-tuned with reasoning exemplars crafted by a significantly larger teacher model. We explore this paradigm for the financial domain,…

Large language models (LLMs) have significantly advanced the field of natural language processing, with GPT models at the forefront. While their remarkable performance spans a range of tasks, adapting LLMs for real-world business scenarios…

计算与语言 · 计算机科学 2023-07-11 Minh-Tien Nguyen , Duy-Hung Nguyen , Shahab Sabahi , Hung Le , Jeff Yang , Hajime Hotta

High-quality textual training data is essential for the success of multimodal data processing tasks, yet outputs from image captioning models like BLIP and GIT often contain errors and anomalies that are difficult to rectify using…

计算与语言 · 计算机科学 2025-02-25 Elyas Meguellati , Nardiena Pratama , Shazia Sadiq , Gianluca Demartini

Large Language Models (LLMs) have shown impressive versatility as general purpose models. However, their broad applicability comes at a high-cost computational overhead, particularly in auto-regressive decoding where each step requires a…

计算与语言 · 计算机科学 2025-08-04 Itay Nakash , Nitay Calderon , Eyal Ben David , Elad Hoffer , Roi Reichart

Large language models (LLMs) have showcased profound capabilities in language understanding and generation, facilitating a wide array of applications. However, there is a notable paucity of detailed, open-sourced methodologies on…

While Large Language Models (LLMs) demonstrate exceptional natural language capabilities, general-purpose models lack specialized domain knowledge for effective cybersecurity analysis. In this work, we investigate Domain-Adaptive Continuous…

计算与语言 · 计算机科学 2025-07-08 Salahuddin Salahuddin , Ahmed Hussain , Jussi Löppönen , Toni Jutila , Panos Papadimitratos

Large language models are powerful but often limited by high computational cost, privacy concerns, and English-centric training. Recent progress demonstrates that small, efficient models with around one billion parameters can deliver strong…

计算与语言 · 计算机科学 2025-12-16 Anna Aksenova , Boris Zverkov , Nicola Dainese , Alexander Nikitin , Pekka Marttinen

High-resource language models often fall short in the African context, where there is a critical need for models that are efficient, accessible, and locally relevant, even amidst significant computing and data constraints. This paper…

This paper addresses the critical need for democratizing large language models (LLM) in the Arab world, a region that has seen slower progress in developing models comparable to state-of-the-art offerings like GPT-4 or ChatGPT 3.5, due to a…

Recent advancements in Large Language Models (LLMs) have revealed new capabilities and opportunities across the technological landscape. However, the practicality of very large LLMs is challenged by their high compute cost, which does not…

In recent years, large language models (e.g., Open AI's GPT-4, Meta's LLaMa, Google's PaLM) have become the dominant approach for building AI systems to analyze and generate language online. However, the automated systems that increasingly…

计算与语言 · 计算机科学 2023-06-14 Gabriel Nicholas , Aliya Bhatia

Scientific workflow systems are increasingly popular for expressing and executing complex data analysis pipelines over large datasets, as they offer reproducibility, dependability, and scalability of analyses by automatic parallelization on…

分布式、并行与集群计算 · 计算机科学 2024-07-09 Mario Sänger , Ninon De Mecquenem , Katarzyna Ewa Lewińska , Vasilis Bountris , Fabian Lehmann , Ulf Leser , Thomas Kosch

Recent releases of pre-trained Large Language Models (LLMs) have gained considerable traction, yet research on fine-tuning and employing domain-specific LLMs remains scarce. This study investigates approaches for fine-tuning and leveraging…

计算与语言 · 计算机科学 2024-05-29 Cheonsu Jeong

While Large Language Models (LLMs) have exhibited remarkable emergent capabilities through extensive pre-training, they still face critical limitations in generalizing to specialized domains and handling diverse linguistic variations, known…

计算与语言 · 计算机科学 2025-05-28 Jinwu Hu , Zhitian Zhang , Guohao Chen , Xutao Wen , Chao Shuai , Wei Luo , Bin Xiao , Yuanqing Li , Mingkui Tan

Large language models (LLMs) are increasingly used to support the analysis of complex financial disclosures, yet their reliability, behavioral consistency, and transparency remain insufficiently understood in high-stakes settings. This…

计算与语言 · 计算机科学 2026-01-21 Md Talha Mohsin

This study introduces a benchmark framework for evaluating the financial decision-making capabilities of large language models (LLMs) through portfolio optimization problems with mathematically explicit solutions. Unlike existing financial…

投资组合管理 · 定量金融 2026-05-28 Hanyong Cho , Jang Ho Kim