English
Related papers

Related papers: RadegastXDB - Prototype of Native XML Database Man…

200 papers

Large Language Model-based (LLM-based) Text-to-SQL methods have achieved important progress in generating SQL queries for real-world applications. When confronted with table content-aware questions in real-world scenarios, ambiguous data…

Databases · Computer Science 2025-11-07 Wenbo Xu , Liang Yan , Chuanyi Liu , Peiyi Han , Haifeng Zhu , Yong Xu , Yingwei Liang , Bob Zhang

The integration of tabular data from diverse sources is often hindered by inconsistencies in formatting and representation, posing significant challenges for data analysts and personal digital assistants. Existing methods for automating…

Databases · Computer Science 2025-08-20 Arash Dargahi Nobari , Davood Rafiei

Text-to-SQL bridges the gap between natural language and structured database language, thus allowing non-technical users to easily query databases. Traditional approaches model text-to-SQL as a direct translation task, where a given Natural…

Machine Learning · Computer Science 2025-08-12 Anurag Tripathi , Vaibhav Patle , Abhinav Jain , Ayush Pundir , Sairam Menon , Ajeet Kumar Singh , Dorien Herremans

Business Intelligence (BI) analysis is evolving towards Exploratory BI, an iterative, multi-round exploration paradigm where analysts progressively refine their understanding. However, traditional BI systems impose critical limits for…

Databases · Computer Science 2026-03-27 Yunkai Lou , Shunyang Li , Longbin Lai , Jianke Yu , Wenyuan Yu , Ying Zhang

Large-scale low-background detectors are increasingly used in rare-event searches as experimental collaborations push for enhanced sensitivity. However, building such detectors, in practice, creates an abundance of radioassay data…

Instrumentation and Detectors · Physics 2025-08-08 R. H. M. Tsang , A. Piepke , S. Al Kharusi , E. Angelico , I. J. Arnquist , A. Atencio , I. Badhrees , J. Bane , V. Belov , E. P. Bernard , A. Bhat , T. Bhatta , A. Bolotnikov , P. A. Breur , J. P. Brodsky , E. Brown , T. Brunner , E. Caden , G. F. Cao , L. Q. Cao , D. Cesmecioglu , C. Chambers , E. Chambers , B. Chana , S. A. Charlebois , D. Chernyak , M. Chiu , B. Cleveland , J. R. Cohen , R. Collister , M. Cvitan , J. Dalmasson , L. Darroch , K. Deslandes , R. DeVoe , M. L. di Vacri , Y. Y. Ding , M. J. Dolinski , J. Echevers , B. Eckert , M. Elbeltagi , R. Elmansali , L. Fabris , W. Fairbank , J. Farine , Y. S. Fu , D. Gallacher , G. Gallina , P. Gautam , G. Giacomini , W. Gillis , C. Gingras , D. Goeldi , R. Gornea , G. Gratta , Y. D. Guan , C. A. Hardy , S. Hedges , M. Heffner , E. Hein , J. Holt , E. W. Hoppe , A. House , W. Hunt , A. Iverson , A. Jamil , X. S. Jiang , A. Karelin , L. J. Kaufman , I. Kotov , R. Krücken , A. Kuchenkov , K. S. Kumar , A. Larson , K. G. Leach , B. G. Lenardo , D. S. Leonard , G. Li , S. Li , Z. Li , C. Licciardi , R. Lindsay , R. MacLellan , M. Mahtab , S. Majidi , C. Malbrunot , P. Martel-Dion , J. Masbou , N. Massacret , K. McMichael , B. Mong , D. C. Moore , K. Murray , J. Nattress , C. R. Natzke , X. E. Ngwadla , K. Ni , A. Nolan , S. C. Nowicki , J. C. Nzobadila Ondze , J. L. Orrell , G. S. Ortega , C. T. Overman , H. Peltz-Smalley , A. Perna , T. Pinto Franco , A. Pocar , J. -F. Pratte , V. Radeka , E. Raguzin , H. Rasiwala , D. Ray , B. Rebeiro , S. Rescia , F. Retière , G. Richardson , J. Ringuette , V. Riot , P. C. Rowson , N. Roy , L. Rudolph , R. Saldanha , S. Sangiorgio , S. Schwartz , J. Soderstrom , A. K. Soma , F. Spadoni , V. Stekhanov , X. L. Sun , E. Teimoori Barakoohi , S. Thibado , A. Tidball , T. Totev , S. Triambak , T. Tsang , O. A. Tyuka , R. Underwood , E. van Bruggen , V. Veeraraghavan , M. Vidal , S. Viel , M. Walent , K. Wamba , Q. D. Wang , W. Wang , Y. G. Wang , M. Watts , W. Wei , L. J. Wen , U. Wichoski , S. Wilde , M. Worcester , S. Wu , X. M. Wu , H. Yang , L. Yang , M. Yvaine , O. Zeldovich , J. Zhao , T. Ziegler

Various automated testing approaches have been proposed for Database Management Systems (DBMSs). Many such approaches generate pairs of equivalent queries to identify bugs that cause DBMSs to compute incorrect results, and have found…

Software Engineering · Computer Science 2025-05-06 Suyang Zhong , Manuel Rigger

Trajectory mining has attracted significant attention. This paper addresses the Top-k Representative Similar Subtrajectory Query (TRSSQ) problem, which aims to find the k most representative subtrajectories similar to a query. Existing…

Databases · Computer Science 2025-07-09 Mingchang Ge , Liping Wang , Xuemin Lin , Yuang Zhang , Kunming Wang

With the rise of XML as a standard for representing business data, XML data warehouses appear as suitable solutions for Web-based decision-support applications. In this context, it is necessary to allow OLAP analyses over XML data cubes…

Databases · Computer Science 2008-09-17 Marouane Hachicha , Hadj Mahboubi , Jérôme Darmont

Facts change over time, making it essential for Large Language Models (LLMs) to handle time-sensitive factual knowledge accurately and reliably. Although factual Time-Sensitive Question-Answering (TSQA) tasks have been widely developed,…

Computation and Language · Computer Science 2026-03-03 Soyeon Kim , Jindong Wang , Xing Xie , Steven Euijong Whang

RDF has seen increased adoption in recent years, prompting the standardization of the SPARQL query language for RDF, and the development of local and distributed engines for processing SPARQL queries. This survey paper provides a…

Databases · Computer Science 2021-10-14 Waqas Ali , Muhammad Saleem , Bin Yao , Aidan Hogan , Axel-Cyrille Ngonga Ngomo

With the recent proliferation of sensor data, there is an increasing need for the efficient evaluation of analytical queries over multiple sensor datasets. The magnitude of such datasets makes exact query answering infeasible, leading…

Schema Matching, i.e. the process of discovering semantic correspondences between concepts adopted in different data source schemas, has been a key topic in Database and Artificial Intelligence research areas for many years. In the past, it…

Databases · Computer Science 2014-07-11 Santa Agreste , Pasquale De Meo , Emilio Ferrara , Domenico Ursino

Extreme Multi-label classification (XML) is an important yet challenging machine learning task, that assigns to each instance its most relevant candidate labels from an extremely large label collection, where the numbers of labels, features…

Machine Learning · Computer Science 2019-04-15 Bingyu Wang , Li Chen , Wei Sun , Kechen Qin , Kefeng Li , Hui Zhou

Many database columns contain string or numerical data that conforms to a pattern, such as phone numbers, dates, addresses, product identifiers, and employee ids. These patterns are useful in a number of data processing applications,…

Databases · Computer Science 2017-12-07 Andrew Ilyas , Joana M. F. da Trindade , Raul Castro Fernandez , Samuel Madden

HRDBMS is a novel distributed relational database that uses a hybrid model combining the best of traditional distributed relational databases and Big Data analytics platforms such as Hive. This allows HRDBMS to leverage years worth of…

Databases · Computer Science 2019-01-28 Jason Arnold , Boris Glavic , Ioan Raicu

Neural networks have proved to be very robust at processing unstructured data like images, text, videos, and audio. However, it has been observed that their performance is not up to the mark in tabular data; hence tree-based models are…

Machine Learning · Computer Science 2022-04-25 Tushar Sarkar

This paper introduces RG (Relational Genetic) model, a revised relational model to represent graph-structured data in RDBMS while preserving its topology, for efficiently and effectively extracting data in different formats from disparate…

Databases · Computer Science 2024-02-01 Wenzhi Fu

GPUs are uniquely suited to accelerate (SQL) analytics workloads thanks to their massive compute parallelism and High Bandwidth Memory (HBM) -- when datasets fit in the GPU HBM, performance is unparalleled. Unfortunately, GPU HBMs remain…

We present DrJAX, a JAX-based library designed to support large-scale distributed and parallel machine learning algorithms that use MapReduce-style operations. DrJAX leverages JAX's sharding mechanisms to enable native targeting of TPUs and…

Distributed, Parallel, and Cluster Computing · Computer Science 2024-07-19 Keith Rush , Zachary Charles , Zachary Garrett , Sean Augenstein , Nicole Mitchell

In this systems paper, we present MillenniumDB: a novel graph database engine that is modular, persistent, and open source. MillenniumDB is based on a graph data model, which we call domain graphs, that provides a simple abstraction upon…