English
Related papers

Related papers: Strategic Exploration for Innovation

200 papers

In deep reinforcement learning (RL) research, there has been a concerted effort to design more efficient and productive exploration methods while solving sparse-reward problems. These exploration methods often share common principles (e.g.,…

Machine Learning · Computer Science 2024-04-04 Jonathan C. Balloch , Rishav Bhagat , Geigh Zollicoffer , Ruoran Jia , Julia Kim , Mark O. Riedl

The efficient use of available resources is a key factor in achieving success on both personal and organizational levels. One of the crucial resources in knowledge economy is time. The ability to force others to adapt to our schedule even…

Multiagent Systems · Computer Science 2017-06-06 Michal Kakol , Radoslaw Nielek , Adam Wierzbicki

Strategic classification regards the problem of learning in settings where users can strategically modify their features to improve outcomes. This setting applies broadly and has received much recent attention. But despite its practical…

Machine Learning · Computer Science 2021-06-15 Sagi Levanon , Nir Rosenfeld

Human creativity is the ultimate driving force behind scientific progress. While the building blocks of innovations are often embodied in existing knowledge, it is creativity that blends seemingly disparate ideas. Existing studies have made…

Social and Information Networks · Computer Science 2016-12-06 Xinyang Zhang , Dashun Wang , Ting Wang

Designing protocols enhancing cooperation for multi-agent systems remains a grand challenge. Cheap talk, defined as costless, non-binding communication before formal action, serves as a pivotal solution. However, existing theoretical…

Multiagent Systems · Computer Science 2026-03-03 Zhao Song , Chen Shen , Zhen Wang , The Anh Han

New ideas are often thought to arise from recombining existing knowledge. Yet despite rapid publication growth - and expanding opportunities for recombination - scientific breakthroughs remain rare. This gap between productivity and…

Digital Libraries · Computer Science 2025-12-04 Linzhuo Li , Yiling Lin , Lingfei Wu

Controllable Markov chains describe the dynamics of sequential decision making tasks and are the central component in optimal control and reinforcement learning. In this work, we give the general form of an optimal policy for learning…

Machine Learning · Computer Science 2025-12-24 Peter N. Loxley

Information foraging connects optimal foraging theory in ecology with how humans search for information. The theory suggests that, following an information scent, the information seeker must optimize the tradeoff between exploration by…

Information Retrieval · Computer Science 2016-11-18 Peter Wittek , Ying-Hsang Liu , Sándor Darányi , Tom Gedeon , Ik Soo Lim

The ability to learn from others (social learning) is often deemed a cause of human species success. But if social learning is indeed more efficient (whether less costly or more accurate) than individual learning, it raises the question of…

Physics and Society · Physics 2021-01-01 Benoît de Courson , Léo Fitouchi , Jean-Philippe Bouchaud , Michael Benzaquen

The notion of the "adjacent possible" has been advanced to theorize the generation of novelty across many different research domains. This study is an attempt to examine in what way the notion can be made empirically useful for innovation…

General Economics · Economics 2025-08-28 Josef Taalbi

If scientific discovery is one of the main driving forces of human progress, insight is the fuel for the engine, which has long attracted behavior-level research to understand and model its underlying cognitive process. However, current…

Artificial Intelligence · Computer Science 2022-12-06 Yu-Zhe Shi , Manjie Xu , Wenjuan Han , Yixin Zhu

We identify a distinct motive for search, termed catalytic exploration, where agents rationally explore alternatives they expect to reject to resolve uncertainty about the status quo. By decomposing option value into switching and catalytic…

Theoretical Economics · Economics 2025-11-25 Zeyu He

How do you incentivize self-interested agents to $\textit{explore}$ when they prefer to $\textit{exploit}$? We consider complex exploration problems, where each agent faces the same (but unknown) MDP. In contrast with traditional…

Machine Learning · Computer Science 2023-02-21 Max Simchowitz , Aleksandrs Slivkins

We solve for the equilibrium dynamics of information sharing in a large population. Each agent is endowed with signals regarding the likely outcome of a random variable of common concern. Individuals choose the effort with which they search…

Probability · Mathematics 2008-11-20 Darrell Duffie , Semyon Malamud , Gustavo Manso

Agents exert hidden effort to produce randomly-sized innovations in a technology they share. Flow payoffs grow as the technology develops, but so does the marginal cost of effort. I characterise the unique symmetric MPE with the quality of…

Theoretical Economics · Economics 2025-11-11 Gregorio Curello

This paper develops a theory of scientific and technological peer effects to study how individuals' productivity responds to the behavior and network positions of their collaborators across both scientific and inventive activities. Building…

General Economics · Economics 2026-02-03 Michael Balzer , Adhen Benlahlou

We propose and design recommendation systems that incentivize efficient exploration. Agents arrive sequentially, choose actions and receive rewards, drawn from fixed but unknown action-specific distributions. The recommendation system…

Computer Science and Game Theory · Computer Science 2026-04-02 Nicole Immorlica , Jieming Mao , Aleksandrs Slivkins , Zhiwei Steven Wu

Efficient exploration is a long-standing problem in sensorimotor learning. Major advances have been demonstrated in noise-free, non-stochastic domains such as video games and simulation. However, most of these formulations either get stuck…

Machine Learning · Computer Science 2019-06-11 Deepak Pathak , Dhiraj Gandhi , Abhinav Gupta

Balancing exploration and exploitation is a central goal in reinforcement learning (RL). Despite recent advances in enhancing large language model (LLM) reasoning, most methods lean toward exploitation, and increasingly encounter…

Computation and Language · Computer Science 2025-11-11 Daixuan Cheng , Shaohan Huang , Xuekai Zhu , Bo Dai , Wayne Xin Zhao , Zhenliang Zhang , Furu Wei

An online labor platform faces an online learning problem in matching workers with jobs and using the performance on these jobs to create better future matches. This learning problem is complicated by the rise of complex tasks on these…

Machine Learning · Computer Science 2018-10-16 Ramesh Johari , Vijay Kamble , Anilesh K. Krishnaswamy , Hannah Li