English
Related papers

Related papers: Accessing United States Bulk Patent Data with pate…

200 papers

We report on the development of an interface to the US Patent and Trademark Office (USPTO) that allows for the mapping of patent portfolios as overlays to basemaps constructed from citation relations among all patents contained in this…

Digital Libraries · Computer Science 2012-11-20 Loet Leydesdorff , Duncan Kushnir , Ismael Rafols

The USPTO disseminates one of the largest publicly accessible repositories of scientific, technical, and commercial data worldwide. USPTO data has historically seen frequent use in fields such as patent analytics, economics, and prosecution…

Machine Learning · Computer Science 2022-07-13 Scott Beliveau , Jerry Ma

Innovation is a major driver of economic and social development, and information about many kinds of innovation is embedded in semi-structured data from patents and patent applications. Although the impact and novelty of innovations…

Computation and Language · Computer Science 2022-07-11 Mirac Suzgun , Luke Melas-Kyriazi , Suproteem K. Sarkar , Scott Duke Kominers , Stuart M. Shieber

A technique is developed using patent information available online (at the US Patent and Trademark Office) for the generation of Google Maps. The overlays indicate both the quantity and quality of patents at the city level. This information…

Computers and Society · Computer Science 2012-02-10 Loet Leydesdorff , Lutz Bornmann

The effective and ethical use of data to inform decision-making offers huge value to the public sector, especially when delivered by transparent, reproducible, and robust data processing workflows. One way that governments are unlocking…

Computers and Society · Computer Science 2023-05-26 Federico Botta , Robin Lovelace , Laura Gilbert , Arthur Turrell

Drafting patent claims is time-intensive, costly, and requires professional skill. Therefore, researchers have investigated large language models (LLMs) to assist inventors in writing claims. However, existing work has largely relied on…

Computation and Language · Computer Science 2025-05-20 Lekang Jiang , Chengzu Li , Stephan Goetz

We introduce FDApy, an open-source Python package for the analysis of functional data. The package provides tools for the representation of (multivariate) functional data defined on different dimensional domains and for functional data that…

Mathematical Software · Computer Science 2025-03-10 Steven Golovkine

We report the findings of a month-long online competition in which participants developed algorithms for augmenting the digital version of patent documents published by the United States Patent and Trademark Office (USPTO). The goal was to…

Computer Vision and Pattern Recognition · Computer Science 2016-03-11 Christoph Riedl , Richard Zanibbi , Marti A. Hearst , Siyu Zhu , Michael Menietti , Jason Crusan , Ivan Metelsky , Karim R. Lakhani

Three periods can be distinguished in university patenting at the U.S. Patent and Trade Office (USPTO) since the Bayh-Dole Act of 1980: (1) a first period of exponential increase in university patenting till 1995 (filing date) or 1999…

Digital Libraries · Computer Science 2019-08-19 Loet Leydesdorff , Martin Meyer

Via the Internet, information scientists can obtain cost-free access to large databases in the hidden or deep web. These databases are often structured far more than the Internet domains themselves. The patent database of the U.S. Patent…

Digital Libraries · Computer Science 2009-12-08 Loet Leydesdorff

For scientific knowledge to be findable, accessible, interoperable, and reusable, it needs to be machine-readable. Moving forward from post-publication extraction of knowledge, we adopted a pre-publication approach to write research…

Digital Libraries · Computer Science 2025-12-12 Olga Lezhnina , Manuel Prinz , Markus Stocker

With the ever increasing number of filed patent applications every year, the need for effective and efficient systems for managing such tremendous amounts of data becomes inevitably important. Patent Retrieval (PR) is considered the pillar…

Information Retrieval · Computer Science 2018-12-21 Walid Shalaby , Wlodek Zadrozny

Writing generic data processing pipelines requires that the algorithmic code does not ever have to know about data formats of files, or the locations of those files. At LSST we have a software system known as "the Data Butler," that…

Instrumentation and Methods for Astrophysics · Physics 2018-12-20 Tim Jenness , James F. Bosch , Pim Schellart , Kian-Ta Lim , Andrei Salnikov , Michelle Gower

In this paper, we extend some usual techniques of classification resulting from a large-scale data-mining and network approach. This new technology, which in particular is designed to be suitable to big data, is used to construct an open…

Physics and Society · Physics 2017-07-05 Antonin Bergeaud , Yoann Potiron , Juste Raimbault

Robust estimation provides essential tools for analyzing data that contain outliers, ensuring that statistical models remain reliable even in the presence of some anomalous data. While robust methods have long been available in R, users of…

Computation · Statistics 2024-11-05 Sarah Leyder , Jakob Raymaekers , Peter J. Rousseeuw , Thomas Servotte , Tim Verdonck

We present fplyr, a new package for the R language to deal with big files. It allows users to easily implement the split-apply-combine strategy for files that are too big to fit into the available memory, without relying on data bases nor…

Computation · Statistics 2020-06-23 Federico Marotta

Most existing text summarization datasets are compiled from the news domain, where summaries have a flattened discourse structure. In such datasets, summary-worthy content often appears in the beginning of input articles. Moreover, large…

Computation and Language · Computer Science 2019-06-11 Eva Sharma , Chen Li , Lu Wang

Various stakeholders, such as researchers, government agencies, businesses, and research laboratories require a large volume of reliable scientific research outcomes including research articles and patent data to support their work. These…

Databases · Computer Science 2024-10-01 Xinran Wu , Hui Zou , Yidan Xing , Jingjing Qu , Qiongxiu Li , Renxia Xue , Xiaoming Fu

The treatment of missing data can be difficult in multilevel research because state-of-the-art procedures such as multiple imputation (MI) may require advanced statistical knowledge or a high degree of familiarity with certain statistical…

Computation · Statistics 2016-11-11 Simon Grund , Oliver Lüdtke , Alexander Robitzsch

Power laws are theoretically interesting probability distributions that are also frequently used to describe empirical data. In recent years effective statistical methods for fitting power laws have been developed, but appropriate use of…

Data Analysis, Statistics and Probability · Physics 2014-02-03 Jeff Alstott , Ed Bullmore , Dietmar Plenz
‹ Prev 1 2 3 10 Next ›