中文
相关论文

相关论文: Introducing Morphit, a new type of spreadsheet tec…

200 篇论文

Numerous indexing databases keep track of the number of publications, citations, etc. in order to maintain the progress of science and individual. However, the choice of journals and articles varies among these indexing databases, hence the…

数字图书馆 · 计算机科学 2021-06-03 Parul Khurana , Geetha Ganesan , Gulshan Kumar , Kiran Sharma

We intend to demonstrate the innate problems with existing spreadsheet products and to show how to tackle these issues using a new type of spreadsheet program called Resolver. It addresses the issues head-on and thereby moves the 1980's…

软件工程 · 计算机科学 2008-03-10 Patrick Kemmis , Giles Thomas

Forms are a widespread type of template-based document used in a great variety of fields including, among others, administration, medicine, finance, or insurance. The automatic extraction of the information included in these documents is…

计算与语言 · 计算机科学 2021-12-15 María Villota , César Domínguez , Jónathan Heras , Eloy Mata , Vico Pascual

In this paper, we discuss the problem of the software engineering of a class of business spreadsheet models. A methodology for structured software development is proposed, which is based on structured analysis of data, represented as…

软件工程 · 计算机科学 2008-05-29 Brian Knight , David Chadwick , Kamalesen Rajalingham

We present a framework for creating small, informative sub-tables of large data tables to facilitate the first step of data science: data exploration. Given a large data table table T, the goal is to create a sub-table of small, fixed…

数据库 · 计算机科学 2022-03-08 Kathy Razmadze , Yael Amsterdamer , Amit Somech , Susan B. Davidson , Tova Milo

This paper presents a batch classifier that has been improved from the earlier version and fixed a mistake in the earlier paper. Two important changes have been made. Each category is represented by a classifier, where each classifier…

机器学习 · 计算机科学 2021-12-03 Kieran Greer

With today's public data sets containing billions of data items, more and more companies are looking to integrate external data with their traditional enterprise data to improve business intelligence analysis. These distributed data sources…

数据库 · 计算机科学 2012-05-16 Ahmad Assaf , Eldad Louw , Aline Senart , Corentin Follenfant , Raphaël Troncy , David Trastour

Most organizations use large and complex spreadsheets that are embedded in their mission-critical processes and are used for decision-making purposes. Identification of the various types of errors that can be present in these spreadsheets…

软件工程 · 计算机科学 2011-11-30 Zbigniew Przasnyski , Linda Leon , Kala Chand Seal

We have developed the Model Master (MM) language for describing spreadsheets, and tools for converting MM programs to and from spreadsheets. The MM decompiler translates a spreadsheet into an MM program which gives a concise summary of its…

编程语言 · 计算机科学 2024-12-31 Jocelyn Paine

Data-driven applications rely on the correctness of their data to function properly and effectively. Errors in data can be incredibly costly and disruptive, leading to loss of revenue, incorrect conclusions, and misguided policy decisions.…

数据库 · 计算机科学 2016-02-15 Xiaolan Wang , Alexandra Meliou , Eugene Wu

A number of automated techniques and tools were proposed in the research literature over the years which aim to support the spreadsheet developer in the process of testing and debugging a faulty spreadsheet. One underlying assumption of…

软件工程 · 计算机科学 2015-03-12 Dietmar Jannach , Thomas Schmitz

Tables have been an ever-existing structure to store data. There exist now different approaches to store tabular data physically. PDFs, images, spreadsheets, and CSVs are leading examples. Being able to parse table structures and extract…

计算机视觉与模式识别 · 计算机科学 2022-01-06 Susie Xi Rao , Johannes Rausch , Peter Egger , Ce Zhang

Numerical reasoning over hybrid data containing both textual and tabular content (e.g., financial reports) has recently attracted much attention in the NLP community. However, existing question answering (QA) benchmarks over hybrid data…

人工智能 · 计算机科学 2022-06-06 Yilun Zhao , Yunxiang Li , Chenying Li , Rui Zhang

Classifying journals or publications into research areas is an essential element of many bibliometric analyses. Classification usually takes place at the level of journals, where the Web of Science subject categories are the most popular…

数字图书馆 · 计算机科学 2012-03-05 Ludo Waltman , Nees Jan van Eck

Cell types are at the root of modern biology, and describing them is a core task of the Human Cell Atlas project. Surprisingly, there are no standards for reporting new cell types, leading to a gap between classes mentioned in biomedical…

Table summarization is a crucial task aimed at condensing information from tabular data into concise and comprehensible textual summaries. However, existing approaches often fall short of adequately meeting users' information and quality…

计算与语言 · 计算机科学 2024-08-27 Weijia Zhang , Vaishali Pal , Jia-Hong Huang , Evangelos Kanoulas , Maarten de Rijke

These are a set of lecture notes on generalized global symmetries in quantum field theory. The focus is on invertible symmetries with a few comments regarding non-invertible symmetries. The main topics covered are the basics of higher-form…

In \cite{Spi}, we developed a category of databases in which the schema of a database is represented as a simplicial set. Each simplex corresponds to a table in the database. There, our main concern was to find a categorical formulation of…

数据库 · 计算机科学 2010-03-16 David I. Spivak

Many earth science applications require data at both high spatial and temporal resolution for effective monitoring of various ecosystem resources. Due to practical limitations in sensor design, there is often a trade-off in different…

机器学习 · 计算机科学 2017-11-17 Ankush Khandelwal , Anuj Karpatne , Vipin Kumar

In this paper, we report some on-going focused research, but are further keen to set it in the context of a proposed bigger picture, as follows. There is a certain depressing pattern about the attitude of industry to spreadsheet error…

计算机与社会 · 计算机科学 2008-03-10 V. R. Vemula , David Ball , Simon Thorne