







This review summarizes results of field experiments examining individual behaviors across several market settings—from open-air markets to rideshare markets to tax-compliance markets—where people sort themselves into market roles wherein they make consequential decisions. Using three distinct examples from my own research on the endowment effect, left-digit bias, and omission bias, I showcase how field experiments can help researchers understand mediators, heterogeneity, and causal moderation involved in judgment biases in the field. In this manner, the review highlights that economic field experiments can serve an invaluable intellectual role alongside traditional laboratory research.
The Effects of Financial Incentives in Experiments: A Review and Capital-Labor-Production Framework
We review 74 experiments with no, low, or high performance-based financial incentives. The modal result has no effect on mean performance (though variance is usually reduced by higher payment). Higher incentive does improve performance often, typically judgment tasks that are responsive to better effort. Incentives also reduce “presentation” effects (e.g., generosity and risk-seeking). Incentive effects are comparable to effects of other variables, particularly “cognitive capital” and task “production” demands, and interact with those variables, so a narrow-minded focus on incentives alone is misguided. We also note that no replicated study has made rationality violations disappear purely by raising incentives.
Cheating in the Lab Predicts Fraud in the Field: An Experiment in Public Transportation
We conduct an artefactual field experiment using a diversified sample of passengers of public transportation to study attitudes toward dishonesty. We find that the diversity of behavior in terms of (dis)honesty in laboratory tasks and in the field correlate. Moreover, individuals who have just been fined in the field behave more honestly in the lab than the other fare dodgers, except when context is introduced. Overall, we show that simple tests of dishonesty in the lab can predict moral firmness in life, although fraudsters who care about social image cheat less when behavior can be verified ex post by the experimenter. Data and the online appendix are available at https://doi.org/10.1287/mnsc.2016.2616 . This paper was accepted by Uri Gneezy, behavioral economics.

Beware of samples! A cognitive-ecological sampling approach to judgment biases.
Integrative experiments identify how punishment affects welfare in public goods games
Despite decades of research, the conditions under which punishment promotes cooperation remain unclear. Through an integrative experiment varying 14 design parameters of public goods games across 360 experimental conditions (147,618 decisions from 7100 participants), we reveal substantial heterogeneity in punishment effectiveness: Its impact on welfare ranges from 43% improvement to 44% reduction depending on the game parameters. To characterize these patterns, we developed models that outperformed human forecasters in predicting punishment effectiveness in new experiments. Communication emerges as the most important factor, followed by contribution framing (opt out versus opt in), contribution type (variable versus all-or-nothing), game length, and outcome visibility, though these factors often interact. The results reframe the debate from whether punishment works to when it does, demonstrating how integrative experiments enable discovery of generalizable patterns in social phenomena. , Editor’s summary People face conflicts between maximizing personal gain versus supporting collective interests. If we cooperatively recycle or donate to charities, it benefits society, but it also costs us time and resources that could be selfishly preserved for ourselves. We impose penalties to deter those undesirable or selfish behaviors, but under what conditions do punishments or penalties effectively modify behavior to benefit group welfare? Alsobay et al . systematically and simultaneously varied 14 factors together instead of in isolation. Punishment was unequivocally most effective when paired with consistent communication, particularly over time. Another effective factor was “opting out” or withdrawing some, but not all, endowments already in the public fund. These methodological advances revealed when, rather than whether, punishment works. —Ekeoma Uzogara , INTRODUCTION Human societies face many situations where individual and collective interests conflict, often referred to as social dilemmas. Costly peer punishment has been studied for more than 25 years in public goods games (stylized behavioral experiments in which individuals decide how much to contribute to a shared pool that benefits everyone) as a mechanism to promote cooperation. Prior research has identified many contextual factors that moderate punishment’s effectiveness, including game length, communication, group size, punishment cost, and so on. However, the specific conditions under which punishment improves group welfare remain unclear. RATIONALE We argue that this lack of clarity derives from the dominant experimental paradigm, in which any given study manipulates only one or a few theoretically informed factors. Because such studies differ in many ways (different experimental procedures, populations), their results are often difficult to compare or integrate. Consequently, one can list many factors that have some effect, but cannot say how much each matters relative to the others, or how they work together, and as a result, cannot predict when punishment will help or harm welfare in new settings. To address this fundamental knowledge gap, we use an integrative experimental design and systematically vary 14 parameters across 360 conditions (147,618 decisions from 7100 participants) to elucidate when punishment improves versus undermines welfare in public goods games, which factors matter most, and how they interact. RESULTS The effect of punishment on welfare ranged from 43% improvement to 44% reduction depending on the specific combination of game parameters. To characterize this heterogeneity, we trained a model that outperformed all 553 human forecasters (laypeople and experts) in predicting whether punishment would help or harm welfare in new experiments. Communication emerged as roughly three times more important than any other factor, followed by contribution framing (opt in versus opt out), contribution type (variable versus all-or-nothing), game length, and peer outcome visibility (whether participants can see others’ earnings). These factors often interact. For example, longer games enhance punishment’s effectiveness only when communication is available, and contribution framing effects depend on both contribution type and outcome visibility. CONCLUSION Many phenomena in social science are shaped by many factors whose interactions are consequential, yet the dominant experimental paradigm often limits its inquiry to “does a given effect exist?” and examines hypothesized factors in isolation. As a result, research programs can accumulate many partial explanations without a clear picture of how they combine to determine outcomes across settings. Knowing that factors matter individually is fundamentally different from knowing how much each matters and how they interact. The integrative approach implemented here offers one way forward. It varies many factors simultaneously within a shared design space, evaluates models by their predictive accuracy on new experiments, and probes those models to constrain and develop theory. Our hope is that integrative experiment designs, combined with models that integrate prediction and explanation, represent a path toward more cumulative social science. Integrative experiment reveals when punishment helps versus harms. We systematically varied 14 design parameters across 360 experimental conditions. The effect of punishment on cooperation efficiency ranged from −44% to +43% depending on the specific game parameters. Communication emerged as three times more important than any other factor, followed by contribution framing, contribution type, and game length.

Mechanism Experiments and Policy Evaluations
Randomized controlled trials are increasingly used to evaluate policies. How can we make these experiments as useful as possible for policy purposes? We argue greater use should be made of experiments that identify the behavioral mechanisms that are central to clearly specified policy questions, what we call "mechanism experiments." These types of experiments can be of great policy value even if the intervention that is tested (or its setting) does not correspond exactly to any realistic policy option.
A sampling model of social judgment.
Behavioural economics, consumer behaviour and consumer policy: state of the art
Counter to the traditional assumption of neoclassical economics that individuals are rational Homo oeconomici that always seek to maximize their utility and follow their ‘true’ preferences, research in behavioural economics has demonstrated that people's judgements and decisions are often subject to systematic biases and heuristics, and are strongly dependent on the context of the decision. In this article, we briefly review the transition of research from neoclassical economics to behavioural economics, and discuss how the latter has influenced research in consumer behaviour and consumer policy. In particular, we discuss the impacts of key principles such as status quo bias, the endowment effect, mental accounting and the sunk-cost effect, other heuristics and biases related to availability, salience, the anchoring effect and simplicity rules, as well as the effects of other supposedly irrelevant factors such as music, temperature and physical markers on consumers’ decisions. These principles not only add significantly to research on consumer behaviour – they also offer readily available practical implications for consumer policy to nudge behaviour in beneficial directions in consumption domains including financial decision making, product choice, healthy eating and sustainable consumption.

Unwillingness to pay for privacy: A field experiment
We measure willingness to pay for privacy in a field experiment. Participants bought at most one DVD from one of two competing online stores. One store consistently required more sensitive personal data than the other, but otherwise the stores were identical. In one treatment, DVDs were one Euro cheaper at the store requesting more personal information, and almost all buyers chose the cheaper store. Surprisingly, in the second treatment when prices were identical, participants bought from both shops equally often.
Choice Bracketing
When making many choices, a person can broadly bracket them by assessing the consequences of all of them taken together, or narrowly bracket them by making each choice in isolation. We integrate research conducted in a wide range of decision contexts which shows that choice bracketing is an important determinant of behavior. Because broad bracketing allows people to take into account all the consequences of their actions, it generally leads to choices that yield higher utility. The evidence that we review, however, shows that people often fail to bracket broadly when it would be feasible for them to do so. In addition to documenting the diverse effects of bracketing, we also discuss factors that determine whether people bracket narrowly or broadly. We conclude with a discussion of normative aspects of bracketing and argue that there are some situations in which narrower bracketing results in superior decision making.

Echo Chambers and Their Effects on Economic and Political Outcomes
In this review, we survey the economics literature on echo chambers. We identify echo chambers as arising from a combination of two phenomena: ( a) the choice of individuals to segregate with like-minded ones, i.e., the creation of chambers, and ( b) behavioral biases that induce polarization when individuals exchange beliefs in these chambers, i.e., the echo. We summarize the literatures on these two phenomena and suggest how to combine the two literatures to gain insights about the effects of echo chambers on economic and political outcomes. We end by suggesting pathways for future research and discussing policy interventions to alleviate echo chambers.

Field Experimentation in Marketing Research
Despite increasing efforts to encourage the adoption of field experiments in marketing research (e.g., Campbell 1969 ; Cialdini 1980 ; Li et al. 2015 ), the majority of scholars continue to rely primarily on laboratory studies ( Cialdini 2009 ). For example, of the 50 articles published in Journal of Marketing Research in 2013, only three (6%) were based on field experiments. The goal of this article is to motivate a methodological shift in marketing research and increase the proportion of empirical findings obtained using field experiments. The author begins by making a case for field experiments and offers a description of their defining features. She then demonstrates the unique value that field experiments can offer and concludes with a discussion of key considerations that researchers should be mindful of when designing, planning, and running field experiments.

Monetary incentives, what are they good for?
This paper is a critical reflection on the use of monetary incentives in economic experiments. The argument is that incentives have their effect through their influence on one or more of three fact...

You’ve Got Mail: A Randomized Field Experiment on Tax Evasion
We report from a large-scale randomized field experiment conducted on a unique sample of more than 15,000 taxpayers in Norway who were likely to have misreported their foreign income. By randomly manipulating a letter from the tax authorities, we cleanly identify that moral suasion and the perceived detection probability play a crucial role in shaping taxpayer behavior. The moral letter mainly works on the intensive margin, while the detection letter has a strong effect on the extensive margin. We further show that only the detection letter has long-term effects on tax compliance. This paper was accepted by Yan Chen, behavioral economics.

Adult age differences in monetary decisions with real and hypothetical reward
Abstract Age differences in monetary decisions may emerge because younger and older adults perceive the value of outcomes differently. Yet, age‐differential effects of monetary rewards on decisions are not well understood. Most laboratory studies on aging and decision making have used scenarios in which rewards were merely hypothetical (decisions did not have any real consequences) or in which only small amounts of money were at stake. In the current study, we compared younger adults' (20–29 years) and older adults' (61–82 years) decisions in probabilistic choice problems with real or hypothetical rewards. Decision‐contingent rewards were in a typical range of previous studies (gains of up to ~4.25 USD) or substantially scaled up (gains of up to ~85 USD per participant). Reward type (real vs. hypothetical) affected decision quality, including value maximization, switching between options, and dominance violations (choices of an option that was inferior to another option in all respects). Decision quality was markedly better with real than hypothetical rewards in older adults and correlated with numeracy in both age groups. However, we found no evidence that reward type affected people's risk preferences. Overall, the findings portray a fairly positive picture regarding the use of hypothetical scenarios to assess preferences: With carefully prepared instructions, people from different age groups indicate preferences in hypothetical scenarios that match their decisions with real and much higher rewards. One advantage of using real rewards is that they help to reduce decision noise.

Cognitive Bias Lab | Learn to Make Better Decisions
Explore cognitive biases with interactive tests, simulations, and real-world examples. Free platform to sharpen decision-making and critical thinking — no sign-up needed.
