Abstract
Building on principles of evolutionary psychology and sociometer theory, we propose that people feel worse about the extent to which they have forgiven when their forgiveness level increases their risk of exploitation or their risk of spoiling a valuable relationship. We predicted that people would feel worse about their forgiveness level when they grant a high level of forgiveness to a perpetrator who has made weak (vs. strong) amends, thereby heightening their risk of exploitation (H1). We also predicted that people would feel worse about their forgiveness level when they grant a low (vs. high) level of forgiveness to a perpetrator who has made strong amends, thereby putting the value of their relationship with the perpetrator at risk (H2). We conducted a longitudinal study of transgressions occurring in romantic relationships and two experiments to test these ideas. H1 was supported in two of the three studies; H2 was supported in all three. A mini meta-analysis indicated that both effects were reliable across the program of research. These results suggest that feelings about one’s forgiveness level serve a functional purpose: Feeling bad about one’s forgiveness level signals that the current combination of amends and forgiveness levels may be causing an adaptation risk.
Throughout human history, our ancestors have faced adaptation problems caused by other people. Lying, cheating, stealing, breaking promises, withholding resources or information, infidelity, and violence are but a few ways in which one person’s actions can negatively affect another person’s fitness (Buss & Duntley, 2008). Individuals can respond to these transgressions in a variety of ways, some of which more effectively minimize the fitness costs of being the victim of a transgression than others. Those who respond most effectively to interpersonal transgressions would be more likely to survive and reproduce, causing their adaptive response mechanisms to become widespread through natural selection (Petersen, Sell, Tooby, & Cosmides, 2010).
At first blush, one might suppose that revenge is the most adaptive way to respond to transgressions. Revenge can lead perpetrators to make up for the harm they caused and discourage them from repeating their transgressions (Burnette, McCullough, Van Tongeren, & Davis, 2012; McCullough, Kurzban, & Tabak, 2010; Petersen et al., 2010; Sell, Tooby, & Cosmides, 2009). In other words, seeking revenge would minimize one’s exploitation risk, whereas forgiving—especially in the absence of evidence that the perpetrator is unlikely to repeat their transgression—would heighten one’s exploitation risk. However, seeking revenge creates its own adaptation problems in that it estranges one from another person, harming a relationship from which one might gain fitness benefits such as cooperation or shared resources (Burnette et al., 2012; Petersen et al., 2010). In other words, seeking revenge—especially when the perpetrator has signaled that they are unlikely to repeat their transgression—would heighten one’s risk of spoiling a valuable relationship. Using this analysis, Burnette, McCullough, Van Tongeren, and Davis (2012) proposed that a forgiveness system has evolved in humans (see also McCullough, 2008), inclining them to forgive when their risk of exploitation is low and the potential value of a continued relationship with their perpetrator is high. Across correlational and experimental studies examining hypothetical and actual offenses, they found that people were most likely to forgive when they perceived low exploitation risk combined with high relationship value (Burnette et al., 2012). 1
Although Burnette et al. (2012) examined perceptions of exploitation risk and relationship value, they did not identify specific cues that indicate the level of one’s exploitation risk or potential value of a continued relationship with one’s perpetrator. We posit that amends serves as a cue for both. Amends consists of accepting responsibility, offering sincere apology, and making genuine atonement for a transgression (Hannon, Rusbult, Finkel, & Kumashiro, 2010). Amends relates to exploitation risk in that forgiving when one has received weak (vs. strong) amends heightens one’s risk of future exploitation. Amends relates to relationship value in that withholding (vs. granting) forgiveness when one has received strong amends heightens one’s risk of spoiling a valuable relationship. It follows that the forgiveness system would incline people to forgive when they have received strong amends. It would also incline people to withhold forgiveness when they have received weak amends. Indeed, past research has shown that stronger amends predicts greater forgiveness (e.g., Hannon et al., 2010), particularly when the victim perceives the amends as sincere (Pansera & La Guardia, 2012).
Over evolutionary time, natural selection favors adaptive behavior (Petersen et al., 2010); Burnette and colleagues’ (2012) analysis suggests that forgiving when exploitation risk is low and relationship value is high may be one such set of adaptive behaviors. But this time course is far too slow to change the behavior of an individual who has forgiven in the absence of amends or has withheld forgiveness despite having receiving strong amends. Might there be a more immediate signal that one has put oneself at risk by not forgiving in accordance with one’s forgiveness system? Sociometer theory suggests that certain affective experiences can function as an alert for social adaptation risks (Leary & Baumeister, 2000; Leary, Tambor, Terdal, & Downs, 1995). Specifically, low self-esteem alerts an individual to potential social exclusion. Because having low self-esteem is unpleasant, it prompts people to build and strengthen relationships with others. In that way, self-esteem serves a functional purpose: helping people avoid the fitness costs of social exclusion.
In a parallel manner, we propose that feelings about the extent to which one has forgiven are an indicator of the degree to which one has forgiven in accordance with one’s forgiveness system. Feeling regret or unhappy about one’s forgiveness alerts an individual to the social adaptation risk caused by a mismatch between the extent to which one should forgive according to one’s forgiveness system and the extent to which one actually has forgiven. We hypothesize that people feel the happiest about their forgiveness level when their risk of exploitation is minimized at the same time as their relational value is maximized. This would occur when both parties act in a prosocial manner such that one’s perpetrator has made strong amends and one matches those amends by granting a high level of forgiveness.
In some cases, only one party acts in a prosocial manner. Two types of mismatches between amends and forgiveness would put one at risk and decrease one’s happiness about one’s forgiveness level. First, if one’s perpetrator has made only weak amends but one grants a high level of forgiveness anyway, one is at heighted risk of exploitation. Forgiving an unapologetic or insincere perpetrator may lead to a continued, harmful relationship, rife with future hurt and exploitation. Therefore, we hypothesize that, when one has granted a high level of forgiveness, receiving weaker amends will have a negative effect on happiness about forgiveness. That is, we predict a significant simple effect of amends on happiness about forgiveness when forgiveness is high (H1). Second, if one’s perpetrator has made strong amends but one withholds forgiveness, one is at risk of spoiling the value of one’s relationship with the perpetrator. Failing to reengage with a person of relational value who has sufficiently taken responsibility and sincerely apologized is a significant relational and social loss. Therefore, we hypothesize that, when one has received strong amends, granting less forgiveness will have a negative effect on happiness about forgiveness level. That is, we predict a significant simple effect of forgiveness on happiness about forgiveness when amends are high (H2). Just as self-esteem serves a functional purpose by motivating people to avoid social exclusion, we suggest that the extent to which one feels happy about one’s forgiveness level can serve a functional purpose, motivating people to forgive in alignment with their forgiveness system. When people regret their forgiveness or lack thereof, they may adjust their level of forgiveness in the future to lower their adaptation risks.
The preceding analysis makes several novel contributions to the existing literature. First, it combines principles of evolutionary psychology (Buss & Duntley, 2008; Petersen et al., 2010; Sell et al., 2009), the forgiveness system analysis of Burnette et al. (2012), and ideas from sociometer theory (Leary & Baumeister, 2000; Leary et al., 1995). Second, it identifies amends as a cue for both exploitation risk and relationship value, thereby extending prior analysis of the forgiveness system and connecting it to the literature on amends. Third, it builds on this theoretical framework to make specific predictions about the joint influence of forgiveness and amends on happiness about forgiveness. Fourth, it suggests that happiness about forgiveness is an important construct because it can serve a functional purpose, alerting one to an adaptation risk caused by a mismatch between the extent to which one should forgive according to one’s forgiveness system and the extent to which one has actually forgiven.
To summarize, we predict that people would feel happiest about their forgiveness when they and their perpetrators act in a prosocial manner by forgiving and making amends, respectively. We derived two hypotheses about the circumstances under which people feel worse about their forgiveness level. H1 predicts a simple effect of amends on happiness about forgiveness when forgiveness is high. That is, it predicts that people would feel worse about their forgiveness level when they grant a high level of forgiveness to a perpetrator who has made weak (vs. strong) amends, thereby putting themselves at heightened risk of exploitation. H2 predicts a simple effect of forgiveness on happiness about forgiveness when amends are high. That is, it predicts that people would feel worse about their forgiveness level when they grant a low (vs. high) level of forgiveness to a perpetrator who has made strong amends, thereby putting the value of their relationship with the perpetrator at risk. It is helpful to note that these hypotheses are both simple effects, but they are not parallel simple effects. Instead, our hypotheses represent the two effects derived from the preceding theoretical analysis.
We conducted three studies to test these hypotheses. Study 1 was an intensive, 5-month longitudinal study of established romantic couples in which participants reported on transgressions their partner had committed every 2 weeks. This study provided a preliminary test of our ideas in the context of transgressions occurring in ongoing relationships. Study 2 was a laboratory experiment in which participants’ perceptions of their own forgiveness and a close relationship partner’s amends for a past transgression were manipulated through false feedback; Study 3 was a preregistered, higher powered online replication of Study 2. These experiments extended Study 1 by examining causal effects among model variables. Finally, we conducted a mini meta-analysis of our hypotheses to assess how reliably they were supported across the three studies.
Study 1
As an initial exploration of our hypotheses, we conducted a longitudinal study in which participants reported transgressions committed by their romantic partner at biweekly intervals over a 5-month period. Participants rated each transgression’s severity, their partner’s amends, their own forgiveness, and their happiness about their forgiveness level. By assessing these variables, we were able to examine whether (a) increasing one’s risk of exploitation by forgiving when one has received only weak amends or (b) potentially spoiling one’s valuable relationships by withholding forgiveness when one has received strong amends has a negative association with happiness about one’s forgiveness level. If substantiated, these effects would suggest that feelings about one’s forgiveness level serve a functional purpose by signaling that one is at risk of exploitation or at risk of spoiling a valuable relationship. Study 1’s longitudinal research design, in which participants report actual transgressions shortly after they occur, provides an ecologically valid initial test of our hypotheses.
Method
Participants
Members of 104 heterosexual couples (N = 208) from the U.S. who were married or had been dating for at least 6 months were recruited through advertisements on community bulletin boards, a graduate student listserv, and Craigslist. Sample size was determined by the amount of grant funding available and the ability to recruit couples in a timely manner. Fourteen participants did not report any transgressions and were not included in this study. The final sample included 194 participants (101 women, 93 men) who were 26.87 years of age on average (SD = 7.60); 83% were White/Caucasian, 9% were Black/African American, 4% were Asian/Asian American, and 4% indicated another racial identity. Thirty-four percent of participants were married and had been married for an average of 5.55 years (SD = 8.14); 48% were dating; and 17% were engaged and had been in their relationships for an average of 2.27 years (SD = 1.85). Participants were paid US$126 if they completed all parts of the study and a prorated amount if they did not.
Procedure
This study was part of a larger investigation of relationships that included 10 online questionnaires: 1 every 2 weeks for 5 months, each lasting 10–15 min. On each questionnaire, participants reported whether their partner committed each of 20 specific transgressions (see Table 1). On average, participants reported 8.94 (SD = 8.59) transgressions over the course of the study. For each transgression, participants completed single-item measures assessing transgression severity (“How hurtful do you feel this behavior was?”; 1 = not at all hurtful, 5 = very hurtful), their partner’s amends (“To what extent did your partner try to make up for this hurtful behavior [e.g., apologize]?”; 1 = not at all, 5 = very much), their own forgiveness (“To what extent have you forgiven your partner for this hurtful behavior?”; 1 = strong unforgiveness, 6 = strong forgiveness), and their happiness about their forgiveness level (“How happy do you feel about your amount of forgiveness?”; 1 = very sad, 7 = very happy). 2
Number of participants reporting and total number of occurrences of each transgression (Study 1).
Analysis strategy
Data had a four-level structure in which reports about transgressions were nested within transgression type within person within couple. We used multilevel data analytic strategies (Raudenbush & Bryk, 2002) for analyzing nested diary data (Bolger, Davis, & Rafaeli, 2003; Nezlek, 2001), which provide unbiased hypothesis tests by simultaneously examining variance associated with every level of nesting. All variables were grand-mean standardized (M = 0, SD = 1). To create Figure 1, in which happiness about one’s forgiveness level is presented in its raw metric, we ran a second set of analyses in which the dependent variable was not standardized.

Happiness about one’s forgiveness level as a function of forgiveness, amends, and transgression severity (Study 1).
Results
We conducted a multilevel regression analysis predicting happiness about one’s forgiveness level from forgiveness, amends, and transgression severity.
Preliminary results
All three main effects were significant. At the mean levels of the other two variables, both forgiveness and amends were positively associated with happiness about forgiveness level, whereas transgression severity was negatively associated with happiness about forgiveness level, β = .58, t(770) = 23.30, p < .001, 95% CI [0.53, 0.63]; β = .15, t(770) = 4.78, p < .001, 95% CI [0.09, 0.21]; and β = −.25, t(770) = −7.88, p < .001, 95% CI [−0.31, −0.18], respectively. The three-way interaction was also significant, β = .06, t(770) = 3.20, p = .001, 95% CI [0.02, 0.10]. Therefore, we examined mild (−1 SD) and severe (+1 SD) transgressions separately with simple effects tests (Aiken, West, & Reno, 1991). As shown in the left panel of Figure 1, the forgiveness × amends interaction was not significant for mild transgressions, β = .02, t(770) = 0.50, p = .62, 95% CI [−0.05, 0.09]. For mild transgressions, stronger forgiveness predicted greater happiness about forgiveness, and amends did not moderate this association. However, as shown in the right panel, this forgiveness × amends interaction was significant for severe transgressions, β = .14, t(770) = 6.20, p < .001, 95% CI [0.10, 0.19]. For severe transgressions, stronger forgiveness predicted greater happiness about forgiveness, and this association was more robust when amends were strong than when amends were weak. We decompose the forgiveness × amends interactions into their relevant simple effects below.
H1 results
H1 predicted that people would feel worse about their forgiveness level when they grant a high level of forgiveness to a perpetrator who has made weak (vs. strong) amends, thereby heightening their risk of exploitation. Simple effects tests of amends supported this hypothesis, both for mild and severe transgressions. Among those granting relatively high (+1 SD) forgiveness for mild transgressions, receiving weaker amends was associated with feeling less happy about one’s forgiveness level, β = .13, t(770) = 3.63, p < .001, 95% CI [0.06, 0.21]. This was also the case among those granting relatively high forgiveness for severe transgressions, β = .39, t(770) = 7.38, p < .001, 95% CI [0.28, 0.49]. Furthermore, the H1 effect was significantly stronger for severe transgressions than it was for mild transgressions, β = .13, t(770) = 4.36, p < .001, 95% CI [0.07, 0.18]. These associations are illustrated by comparing the two points on the right side of each panel of Figure 1.
H2 results
H2 predicted that people would feel worse about their forgiveness level when they grant a low (vs. high) level of forgiveness to a perpetrator who has made strong amends, thereby putting the value of their relationship with the perpetrator at risk. Simple effects tests of forgiveness supported this hypothesis, both for mild and severe transgressions. Among those receiving relatively strong (+1 SD) amends for mild transgressions, granting less forgiveness was associated with feeling less happy about one’s forgiveness level, β = .79 t(770) = 8.97, p < .001, 95% CI [0.62, 0.96]. This was also the case among those receiving relatively strong amends for severe transgressions, β = 1.04, t(770) = 17.84, p < .001, 95% CI [0.93, 1.16]. Furthermore, the H2 effect was significantly stronger for mild transgressions than it was for severe transgressions, β = .13, t(770) = 2.54, p = .011, 95% CI [0.03, 0.23]. These associations are illustrated by the solid line in each panel of Figure 1. 3
Discussion
Study 1 examined the associations of forgiveness, amends, and transgression severity with happiness about one’s forgiveness level following transgressions occurring in ongoing romantic relationships. H1 was supported in that receiving only weak amends when one has granted high forgiveness, which could increase one’s risk of exploitation, was associated with less happiness about one’s forgiveness level. H2 was supported in that withholding forgiveness when one has received strong amends, which could jeopardize the value of one’s relationship, was also associated with less happiness about one’s forgiveness level. Furthermore, both hypotheses were supported both when examining relatively mild offenses and when examining relatively severe offenses.
Strengths of Study 1 include its strong external validity, intensive longitudinal procedures examining processes in ongoing romantic relationships, and high statistical power. Together, these strengths provide evidence for the hypothesized associations in established, ongoing relationships. However, Study 1 has modest internal validity and cannot demonstrate causal relationships. We conducted Study 2 to address these limitations.
Study 2
To test our hypotheses further, we conducted a laboratory experiment in which we presented participants with false feedback regarding (a) the extent to which they forgave a perpetrator for a transgression that they had actually experienced and (b) the extent to which their perpetrator made amends for that transgression. By experimentally manipulating participants’ perceptions of forgiveness and amends, we were able to test whether increasing participants’ perception of their exploitation risk by leading them to believe they have forgiven when they received weak (vs. strong) amends has a negative effect on their happiness about forgiveness. We were also able to test whether increasing participants’ perception of their risk of spoiling a valuable relationship by withholding (vs. granting) forgiveness when they have received strong amends has a negative effect on their happiness about forgiveness. These effects would provide further evidence for the idea that feelings about one’s forgiveness are important because they can signal that one has behaved in a way that heightens an adaptation risk. Study 2 compliments the strengths of Study 1 by providing an experimental design to examine the causal effects among model variables.
Method
Participants
Ninety-four undergraduates (59 women, 35 men) participated in Study 2. Sample size was determined by the number of participants from the participant pool that were allocated to the first author during the Spring 2011 academic term (before the psychology replication crisis began). Eight participants’ data were excluded because they did not correctly recall the false feedback, did not report a transgression committed by a specific person, or reported suspicion. The remaining 86 participants (53 women, 33 men) were 19.05 years of age on average (SD = 1.52); 67% were White/Caucasian, 20% were Asian/Asian American, 7% were Hispanic/Latino/a, 4% were Black/African American, and 2% indicated another racial identity.
Procedure
Participants received instructions via computers in individual cubicles. Participants recalled and described a time a close other did something that hurt, angered, or upset them. They were instructed to choose a severe incident that has not been completely resolved. Participants then provided the perpetrator’s first name, indicated their relationship to the perpetrator, and reported how long ago the incident had occurred. Then participants answered questions about their perpetrator’s amends (e.g., “[Perpetrator] made amends for his or her behavior”; 1 = strongly disagree, 9 = strongly agree).
Manipulation of forgiveness
Participants read about the “forgiveness test,” which would assess the extent to which they had forgiven their perpetrator. In reality, the forgiveness test was used to provide participants with false feedback about their forgiveness. We used the procedure Luchies, Finkel, McNulty, and Kumashiro (2010) adapted from Karremans, Van Lange, Ouwerkerk, and Kluwer (2003). The forgiveness test was a version of the implicit association test (Greenwald, McGhee, & Schwartz, 1998), which assesses implicit associations between target categories by comparing their reaction times in blocks of trials. The target categories were (a) the perpetrator’s first name and filler first names and (b) words with positive valence and words with negative valence. For each block of trials, participants identified each word that appeared on the computer screen by pressing keys corresponding to the categories presented at the top of the screen. There were two critical blocks of trials. In one block, participants responded with the same key to positive words and the perpetrator’s name. In the other block, participants responded with the same key to negative words and the perpetrator’s name.
Next, participants read about the rationale of the forgiveness test, which was that when a person has forgiven, mental associations (measured through reaction times) between positive words and the name of the perpetrator would be stronger, but, when a person has not forgiven, associations between negative words and the name of the perpetrator would be stronger. Participants were randomly assigned to one of the two forgiveness feedback conditions, ostensibly based on their reaction times. Participants in the high forgiveness condition read that they had responded faster when the perpetrator’s name was paired with positive words. They read, “As your reaction times reveal, the associations between [Perpetrator] and positive words are stronger…you have largely forgiven [Perpetrator].” Participants in the low forgiveness condition read that they had responded faster when the perpetrator’s name was paired with negative words. They read, “As your reaction times reveal, the associations between [Perpetrator] and negative words are stronger…you have not completely forgiven [Perpetrator].”
Manipulation of amends
Next, using the Luchies et al. (2010) procedure, participants were randomly assigned to one of two amends feedback conditions, ostensibly based on their responses to the questions concerning amends. Participants were told that their responses to the amends questions had been compared to the responses of other participants. Participants in the weak amends condition read that their perpetrator’s amends was in the 17th percentile, meaning that “83% of other offenders made more amends than [Perpetrator]. According to these results, [Perpetrator] has made only weak amends.” Participants in the strong amends condition read that their perpetrator’s amends was in the 83rd percentile, meaning that “83% of other offenders made less amends than [Perpetrator]. According to these results, [Perpetrator] has made quite strong amends.”
Assessment of feelings about one’s forgiveness level
Participants reported how happy, sad, good, bad, satisfied, and dissatisfied they felt about the extent to which they had forgiven their perpetrator (e.g., “How happy are you about the extent to which you forgave [Perpetrator]?”; 1 = not at all happy, 9 = extremely happy; α = .92). The negatively valenced items were reverse-scored before averaging all six items to create a composite measure.
Manipulation and suspicion checks
Participants reported whether the forgiveness test had indicated that they had forgiven, that they had not forgiven, or that the test was inconclusive. They also reported the extent to which they felt that they had forgiven their perpetrator (1 = not at all forgiven, 9 = completely forgiven). Then, participants recalled their perpetrator’s amends by reporting the percentile the computer had calculated. They also reported the extent to which they felt that the perpetrator had made amends (1 = very weak amends, 9 = very strong amends). Participants reported what hypothesis they thought the study was testing and how they thought it was tested. Finally, participants were debriefed to ensure that they understood that the feedback was determined by random assignment.
Results
Descriptive analyses
Participants reported transgressions committed by friends (47%), current or former romantic partners (28%), family members (23%), and others (2%). Participants reported a variety of transgressions, including negative communication/lies (29%), lack of communication/secrets (26%), disrespectful/rude behavior (18%), breakups (10%), unfaithfulness/flirtation (7%), irresponsible behavior (6%), and lack of support/emotional distance (4%). Transgressions occurred an average of 7.70 months before the study (SD = 14.69).
Manipulation checks
Between-subjects t-tests showed that the manipulations were successful. Participants in the high forgiveness condition reported that they had forgiven to a greater extent (M = 6.89, SD = 1.64) than those in the low forgiveness condition (M = 4.76, SD = 1.91), t(84) = 5.57, p < .001, d = 1.20, 95% CI [0.74, 1.66]. Participants in the strong amends condition reported that the perpetrator had made stronger amends (M = 4.58, SD = 2.28) than those in the weak amends condition (M = 3.19, SD = 1.79), t(84) = 3.16, p = .002, d = 0.68, 95% CI [0.24, 1.11].
Primary analysis of the experimental effects
We conducted a 2 (forgiveness: low vs. high) × 2 (amends: weak vs. strong) between-subjects analysis of variance (ANOVA) on feelings about forgiveness level, followed by simple effects tests to examine our hypotheses. Results are presented in Figure 2.

Feelings about one’s forgiveness level as a function of forgiveness and amends feedback conditions (Study 2 primary analysis of the experimental effects).
Preliminary results
The main effect of forgiveness feedback condition was significant, indicating that, when collapsing across amends feedback conditions, participants who were led to believe that they had forgiven reported feeling better about their forgiveness level (M = 6.34, SD = 1.81) than those who were led to believe that they had not forgiven (M = 4.57, SD = 1.63), F(1, 82) = 22.99, p < .001, η2 p = .22, 90% CI [0.10, 0.34]. The main effect of amends feedback condition was not significant, indicating that, when collapsing across forgiveness feedback conditions, there was not a reliable difference in happiness about forgiveness level between participants who were led to believe that they had received strong amends (M = 5.87, SD = 2.01) and those who were led to believe that they had received weak amends (M = 5.12, SD = 1.80), F(1, 82) = 2.63, p = .109, η2 p = .03, 90% CI [0.00, 0.11]. The forgiveness × amends interaction was significant, F(1, 82) = 4.86, p = .03, η2 p = .06, 90% CI [0.003, 0.15].
H1 results
H1 predicted that people would feel worse about their forgiveness level when they grant a high level of forgiveness to a perpetrator who has made weak (vs. strong) amends, thereby heightening their risk of exploitation. A simple effects test of amends supported this hypothesis. Among those in the high forgiveness feedback condition, participants who were led to believe that they had received weak amends reported feeling worse about their forgiveness level (M = 5.60, SD = 1.83) than those who were led to believe that they received strong amends (M = 6.99, SD = 1.55), F(1, 82) = 7.68, p = .007, η2 p = .09, 90% CI [0.01, 0.19]. This effect is illustrated by comparing the two points on the right side of Figure 2.
H2 results
H2 predicted that people would feel worse about their forgiveness level when they grant a low (vs. high) level of forgiveness to a perpetrator who has made strong amends, thereby putting the value of their relationship with the perpetrator at risk. A simple effects test of forgiveness supported this hypothesis. Among those in the strong amends feedback condition, participants who were led to believe that they had granted less forgiveness reported feeling worse about their forgiveness level (M = 4.46, SD = 1.60) than those who were led to believe that they had granted more forgiveness (M = 6.99, SD = 1.55), F(1, 82) = 24.33, p < .001, η2 p = .23, 90% CI [0.11, 0.35]. This effect is illustrated by the solid line in Figure 2. 4,5
Auxiliary analysis with manipulation checks as predictors
To explore the associations among forgiveness, amends, and feelings about forgiveness further, we conducted a regression analysis predicting participants’ feelings about their forgiveness level from their perceptions of their forgiveness and their perpetrators’ amends as assessed on the manipulation checks. We standardized all variables before analysis. Then, to create Figure 3, in which feelings about one’s forgiveness level is presented in its raw metric, we ran a second regression analysis in which the dependent variable was not standardized.

Feelings about one’s forgiveness level as predicted from the forgiveness and amends manipulation check measures (Study 2 auxiliary analysis).
This analysis did not provide additional support for H1. Among participants who reported high (+1 SD) forgiveness, reporting weaker amends was not reliably associated with feeling worse about one’s forgiveness level, β = .12, t = 0.90, p = .37, 95% CI [−0.15, 0.39]. However, this analysis did provide additional support for H2. Among participants who reported strong (+1 SD) amends, reporting less forgiveness was associated with feeling worse about one’s forgiveness level, β = .70, t = 3.87, p < .001, 95% CI [0.34, 1.06]. This association is illustrated by the solid line in Figure 3.
Discussion
Study 2 used false feedback to manipulate participants’ perceptions of their own forgiveness and their perpetrators’ amends for a severe transgression to examine the effects of forgiveness and amends on feelings about one’s forgiveness level. Both hypotheses were supported in the primary analysis. H1 was supported in that, among those who were led to believe that they had offered high forgiveness, participants who were led to believe that they had received only weak amends reported feeling worse about their forgiveness than did participants who were led to believe that they had received strong amends. H2 was supported in that, among those who were led to believe that they had received strong amends, participants who were led to believe that they had offered low forgiveness reported feeling worse about their forgiveness than did participants who were led to believe they had offered high forgiveness. However, in auxiliary analyses using the manipulation check assessments of forgiveness and amends as predictors, H2 received additional support but H1 did not. Nonetheless, the controlled laboratory setting and experimental manipulations in Study 2 give it strong internal validity, lending support for the causal effects of forgiveness and amends on feelings about one’s forgiveness.
It is important to note that Study 2 was conducted in early 2011, before the replication crisis in psychology began and when social psychologists frequently used relatively small samples, which were often determined by factors such as the number of participant pool subjects available in a given academic term. We now know that an N of 94 (86 after exclusions) is quite small for a 2 × 2 between-subjects design. For this reason, we conducted a higher powered, preregistered replication of Study 2.
Study 3
Study 3 was a preregistered replication of Study 2. To procure a much larger sample, we switched the testing context from the laboratory to Mechanical Turk (MTurk). As in Study 2, the experimental design allowed us to examine the causal effects of forgiveness and amends on happiness about forgiveness. Specifically, we were able to test whether combinations of forgiveness and amends that would increase one’s risk of exploitation or risk of spoiling a valuable relationship have a negative effect on happiness about forgiveness. Such effects would suggest that feelings about one’s forgiveness serve a functional purpose by signaling that one has not forgiven in accordance with one’s forgiveness system, thereby creating an adaptation risk.
Method
Power analysis
We conducted an a priori power test using G*Power (Faul, Erdfelder, Lang, & Buchner, 2007) to determine the required sample size for an adequately powered test of our hypotheses. We set the parameters at a “small to medium” effect size f of .2, error probability of .05, power to .95, numerator df of 1, and number of groups to 4. Given these parameters, the required total sample size is 327. We expected approximately 15% of participants to meet one or more of the exclusion criteria listed in the study’s preregistration (i.e., the participant did not describe a specific transgression incident, did not list a specific offender name, provided an offender name that was used as a filler name, did not complete the dependent variable and manipulation checks, did not correctly recall forgiveness test false feedback, did not recall amends false feedback with 10 percentile points, and/or expressed suspicion). Therefore, we recruited 388 participants, expecting approximately 329 participants’ data for analysis.
Participants
Three hundred eighty-eight MTurk workers (226 women, 161 men, 1 other) participated in Study 3. A higher-than-expected number of participants met one or more of our exclusion criteria (136 participants or 35% of the sample). The remaining 252 participants (166 women, 85 men, 1 other) were 35.06 years of age on average (SD = 11.69); 77% were White/Caucasian, 7% Black/African American, 6% Hispanic/Latino/a, 4% Asian/Asian American, 6% multiracial, and 1% other. Participants received US$1.
Procedure
All procedures were identical to Study 2 with the following three exceptions. First, Study 2 participants were undergraduates, whereas Study 3 participants were MTurk workers. Second, Study 2 took place in a laboratory with an experimenter present, whereas Study 3 took place wherever the participants completed the survey without an experimenter present. Third, Study 2 was programmed using MediaLab and DirectRT, whereas Study 3 was programmed using Qualtrics. The feelings about one’s forgiveness level measure exhibited good reliability (α = .90).
Preregistration
This study’s preregistration (https://osf.io/83gk5/?%20view_only=ea9d2a82f543454797a48dbfb9aee54e) included materials related to data collection (i.e., Qualtrics survey file) and data analysis (i.e., code for statistical analysis). As is often the case, we updated the theoretical framework and made corresponding changes to our hypotheses during this article’s review process. Therefore, the statistical analysis code in the preregistration was designed to test two hypotheses based on a previous version of the theoretical framework. Specifically, we originally predicted a main effect of forgiveness and a forgiveness × amends interaction effect, such that the effect of forgiveness on happiness about forgiveness would be dampened when amends were weak. The first hypothesis was supported, whereas the second was not—much like H1 was supported whereas H2 was not in the results reported below. It should be noted that the data analytic strategy reported here was not included in this study’s preregistration but is available in a subsequent registration for the same project (https://osf.io/4q7xr/?view_only=d6be51a65b26491993ba08e2c8a14cd1).
Results
Descriptive analyses
Participants reported transgressions committed by current or former romantic partners (38%), family members (33%), friends (22%), and others (6%). Participants reported a variety of transgressions, including negative communication/lies (28%), disrespectful/rude behavior (21%), irresponsible behavior (14%), lack of communication/secrets (13%), unfaithfulness/flirtation (10%), lack of support/emotional distance (7%), breakups (4%), and other transgressions (3%). Transgressions occurred an average of 9.10 months before the study (SD = 29.80).
Manipulation checks
Between-subjects t-tests showed that the manipulations were successful. Participants in the high forgiveness condition reported that they had forgiven to a greater extent (M = 5.69, SD = 2.52) than those in the low forgiveness condition (M = 3.75, SD = 2.22), t(250) = 6.46, p < .001, d = 0.81, 95% CI [0.56, 1.07]. Participants in the strong amends condition reported that the perpetrator had made stronger amends (M = 4.22, SD = 2.71) than those in the weak amends condition (M = 2.26, SD = 1.86), t(250) = 6.61, p < .001, d = 0.84, 95% CI [0.58, 1.09].
Primary analysis of the experimental effects
We conducted a 2 (forgiveness: low vs. high) × 2 (amends: weak vs. strong) between-subjects ANOVA on feelings about forgiveness level, followed by simple effects tests to examine our hypotheses. Results are presented in Figure 4.

Feelings about one’s forgiveness level as a function of forgiveness and amends feedback conditions (Study 3 primary analysis of the experimental effects).
Preliminary results
The main effect of forgiveness feedback condition was significant, indicating that, when collapsing across amends feedback conditions, participants who were led to believe that they had forgiven reported feeling better about their forgiveness level (M = 6.08, SD = 2.05) than those who were led to believe that they had not forgiven (M = 4.67, SD = 1.88), F(1, 248) = 32.87, p < .001, η2 p = .12, 90% CI [0.06, 0.18]. The main effect of amends feedback condition was not significant, indicating that, when collapsing across forgiveness feedback conditions, there was not a reliable difference in happiness about forgiveness level between participants who were led to believe that they had received strong amends (M = 5.57, SD = 2.09) and those who were led to believe that they had received weak amends (M = 5.17, SD = 2.08), F(1, 248) = 2.82, p = .094, η2 p = .01, 90% CI [0.00, 0.04]. The forgiveness × amends interaction was not significant, F(1, 248) = 0.01, p = .909, η2 p = .00, 90% CI [0.00, 0.001].
H1 results
H1 predicted that people would feel worse about their forgiveness level when they grant a high level of forgiveness to a perpetrator who has made weak (vs. strong) amends, thereby heightening their risk of exploitation. A simple effects test of amends did not provide support for this hypothesis. Among those in the high forgiveness feedback condition, participants who were led to believe that they had received weak amends did not report feeling worse about their forgiveness level (M = 5.88, SD = 2.06) than those who were led to believe that they received strong amends (M = 6.27, SD = 2.03), F(1, 248) = 1.24, p = .267, η2 p = .01, 90% CI [0.00, 0.03].
H2 results
H2 predicted that people would feel worse about their forgiveness level when they grant a low (vs. high) level of forgiveness to a perpetrator who has made strong amends, thereby putting the value of their relationship with the perpetrator at risk. A simple effects test of forgiveness supported this hypothesis. Among those in the strong amends feedback condition, participants who were led to believe that they had granted less forgiveness reported feeling worse about their forgiveness level (M = 4.88, SD = 1.91) than those who were led to believe that they had granted more forgiveness (M = 6.27, SD = 2.03), F(1, 248) = 16.72, p < .001, η 2 p = .06, 90% CI [0.02, 0.12]. This effect is illustrated by the solid line in Figure 4. 6,7
Auxiliary analysis with manipulation checks as predictors
To explore the associations among forgiveness, amends, and feelings about forgiveness further, we conducted a regression analysis predicting participants’ feelings about their forgiveness level from their perceptions of their forgiveness and their perpetrators’ amends as assessed on the manipulation checks. We standardized all variables before analysis. Then, to create Figure 5, in which feelings about one’s forgiveness level is presented in its raw metric, we ran a second regression analysis in which the dependent variable was not standardized.

Feelings about one’s forgiveness level as predicted from the forgiveness and amends manipulation check measures (Study 3 auxiliary analysis).
This analysis provided marginal support for H1. Among participants who reported high (+1 SD) forgiveness, reporting weaker amends was marginally associated with feeling worse about one’s forgiveness level, β = .07, t = 1.91, p = .058, 95% CI [−0.005, 0.28]. This analysis also provided additional support for H2. Among participants who reported strong (+1 SD) amends, reporting less forgiveness was associated with feeling worse about one’s forgiveness level, β = .83, t = 7.39, p < .001, 95% CI [0.61, 1.05]. This association is illustrated by the solid line in Figure 5.
Discussion
Study 3 was a preregistered, higher powered replication of Study 2. In this study, H1 was not supported in the primary analyses examining of the experimental effects, but it received marginal support in the auxiliary analysis using the manipulation check measures as predictors. H2 was strongly supported in both the primary and auxiliary analyses.
Given the seemingly trivial differences between Study 2 and Study 3, what might explain why H1 was supported in Study 2 but not in Study 3? Although we cannot answer this question definitively, it is instructive to examine the methodological attributes of the two studies and consider how these differences resulted in different strengths and weaknesses. Study 2 had a sample of 94 undergraduates who completed the experiment in a laboratory setting with an experimenter present. Study 2 was programmed using the MediaLab and DirectRT software programs. Study 3 had a sample of 388 MTurk participants who completed the experiment online without an experimenter presenter. Study 3 was programmed using Qualtrics, which allowed participants to skip a portion of the IAT forgiveness test by pressing the space bar before being instructed to do so. As illustrated in Figure 6, these methodological differences resulted in Study 2’s strengths of greater experimental control, greater internal validity, and a lower exclusion rate than Study 3. The differences also resulted in the Study 3’s strength of higher statistical power than Study 2. Each study’s unique strengths and weaknesses may have contributed to their differing H1 results.

Methodological attributes and resulting strengths and weaknesses of Studies 2 and 3.
Which study, then, provides the more definitive results regarding H1? That is a difficult question to answer in general terms (Finkel, Eastwick, & Reis, 2017). For example, if the issue surrounds estimating the likely replicability of the experimental results, Study 3 should probably carry more weight than Study 2. Study 3’s large sample and preregistered procedures give particularly strong confidence in the precision of the central p-values. In addition, even in the most optimistic interpretation, Study 3 puts major boundary conditions on the generality of the Study 2 findings. If the issue surrounds evaluating the internal validity of the results, Study 2 should probably carry more weight than Study 3 because of the greater experimental control afforded by laboratory studies versus MTurk studies. Our own interpretation vis-à-vis H1 is that the correlational results (Study 1) provide reasonably strong support, but the experimental results (Studies 2 and 3) provide weak support.
Mini meta-analysis
Given that H1 was supported in two of the three studies in this program of research, we conducted a small meta-analysis of our studies to assess the overall reliability of the hypothesized effects (Goh, Hall, & Rosenthal, 2016). We calculated four meta-analytic effects across the three studies. We calculated the first pair of effects using results from the primary hypothesis tests of the experimental manipulations in Studies 2 and 3. We calculated the second pair of effects using results from the auxiliary analyses using the manipulation checks as predictors in Studies 2 and 3. Because Study 1 supported both hypotheses both when examining mild transgressions and when examining severe transgressions, we collapsed across transgression severity, thereby calculating a single Study 1 effect for each hypothesis.
We standardized all predictor and outcome variables in all analyses. To calculate each meta-analytic beta, we weighted the beta for each effect from each study by the inverse of its variance. To calculate each meta-analytic standard error, we took the square root of the reciprocal of the sum of the weights. To conduct hypothesis tests on our meta-analytic effects, we divided the meta-analytic beta by the meta-analytic standard error, which yielded a z statistic. This approach is flexible enough to allow us to calculate meta-analytic effects across our studies and their analytic methods (i.e., multilevel analyses of nested longitudinal data in Study 1 and regression analyses and ANOVAs of experimental effects in Studies 2 and 3).
H1 was strongly supported when using the primary hypothesis tests of the experimental manipulations in Studies 2 and 3. Across the three studies, when participants granted a relatively high level of forgiveness, receiving weaker amends was associated with feeling worse about one’s forgiveness level, β = 0.16, z = 5.64, p < .001. The meta-analysis also supported H2. Across the three studies, when participants received relatively strong amends, granting less forgiveness was associated with feeling worse about one’s forgiveness level, β = 0.86 z = 21.56, p < .001. We obtained similar results for both hypotheses when using results from the auxiliary analyses in Studies 2 and 3, β = 0.16, z = 5.49, p < .001 and β = 1.01 z = 23.39, p < .001, respectively.
General discussion
A longitudinal study and two experiments examined the joint influence of forgiveness and amends on people’s happiness about their forgiveness level. We predicted that people would feel happiest about their forgiveness when they and their perpetrators act in a prosocial manner by forgiving and making amends, respectively. We derived two hypotheses about the circumstances under which people feel worse about their forgiveness. H1 predicted that people would feel worse about their forgiveness level when they grant a high level of forgiveness to a perpetrator who has made weak (vs. strong) amends, thereby heightening their risk of exploitation. H1 was supported in Study 1, in the primary analysis in Study 2, and marginally in the auxiliary analysis in Study 3. It was not supported in the auxiliary analysis in Study 2 or in the primary analysis in Study 3. Despite these inconsistencies, a meta-analysis of the three studies indicated that the H1 effect was reliable when considering the research program as a whole. H2 predicted that people would feel worse about their forgiveness level when they grant a low (vs. high) level of forgiveness to a perpetrator who has made strong amends, thereby putting the value of their relationship with the perpetrator at risk. H2 was supported across all three studies and in the meta-analysis.
Strengths and limitations
This program of research has a number of methodological strengths. First, we used both correlational and experimental methods, providing convergent support for our hypotheses. Second, we conducted a preregistered replication, which replicated H2 but not H1, and a meta-analysis to provide evidence for both hypothesized effects. Third, our participants included a community sample of dating, engaged, and married couples in Study 1; university students in Study 2; and MTurk workers in Study 3. Finally, because we investigated a range of transgressions committed by several types of relationship partners, our findings seem to generalize across a variety of interpersonal transgression situations.
At the same time, this program of research has several limitations that one should consider when interpreting its findings. First, we used artificial manipulations of forgiveness and amends in Studies 2 and 3. Although these manipulations affected participants’ perceptions of forgiveness and amends as designed, we do not know how the magnitude of these manipulations aligns with naturally occurring variations of forgiveness and amends. Accordingly, we should not assume that the effect sizes found in these experiments reflect real-life effect sizes. Study 1, which examined actual transgressions, provides a better basis for inferring effect sizes. The H1 effect sizes were β = .163, .370, and .093 for Studies 1, 2, and 3, respectively. The Study 1 effect size was in between the Study 2 and 3 effect sizes, suggesting that the actual H1 effect size may be in that range. The H2 effect sizes were β = 1.089, .658, and .334, respectively. The Study 1 effect size was the largest of the three, suggesting both that the actual H2 effect may be larger than the experimental studies suggest and that the H2 effect may be substantially larger than the H1 effect.
However, a second consideration may also affect interpretation of effect sizes: Both the measure of forgiveness and the measure of happiness about forgiveness are about forgiveness. As such, it may be unsurprising that the two measures are highly correlated. The size of H2 effect, which examines the simple effect of forgiveness on happiness about forgiveness, may be artificially inflated given that both of these measures are about forgiveness. Another explanation of the large H2 effect is that because people tend to view forgiveness positively, they tend to feel good about themselves when they forgive.
A third consideration is that we asked participants in Studies 2 and 3 to recall a severe and unresolved transgression. We would expect that severe transgressions highlight exploitation risk because being the victim of a severe transgression could be more costly than being the victim of a mild transgression. On the other hand, we would expect that mild transgressions highlight the risk of spoiling a valuable relationship because failing to forgive someone for a mild transgression, especially when they have made amends, may seem unreasonable to others and put the relationship on unstable ground. If our speculations are correct, our instructions to recall a severe offense may have skewed the results by artificially inflating the H1 effect, which is about exploitation risk, at the expense of the H2 effect, which is about the risk of spoiling a valuable relationship. Consistent with these ideas, Study 1 results showed that the H1 effect was stronger for severe transgressions than it was for mild transgressions, although it was significant for both. Future research could explore the role of severity in emphasizing exploitation risk versus relationship value.
Contributions, implications, and directions for future research
This program of research makes several innovative contributions, both because of its theoretical framework and because of its practical implications. We combined principles of evolutionary psychology (Buss & Duntley, 2008; Petersen et al., 2010; Sell et al., 2009), the forgiveness system analysis of Burnette et al. (2012), and ideas from sociometer theory (Leary & Baumeister, 2000; Leary et al., 1995). Combining these theoretical frameworks enabled us to derive novel hypotheses about the joint influence of forgiveness and amends on happiness about forgiveness. Specifically, evolutionary psychology and Burnette et al.’s forgiveness system analysis indicate that people are most likely to forgive when their adaptation risks are low, that is, when exploitation risk is low and relationship value is high. Sociometer theory adds to this by suggesting that affective experiences can function as an alert for social adaptation risks. Building on this idea from sociometer theory, we proposed that feelings about the extent to which one has forgiven are an indicator of the degree to which one has forgiven in accordance with one’s forgiveness system. Feeling unhappy or regret about one’s forgiveness could prompt one to follow their forgiveness system more closely in the future. In other words, evolutionary psychology and the forgiveness system analysis identify when people are most likely to forgive, whereas sociometer theory indicates what happens when people forgive in alignment or misalignment with their forgiveness system. Neither theoretical approach alone does both of these things. In addition to integrating these theoretical approaches, we identified amends as a cue for both exploitation risk and relationship value, thereby extending prior analysis of the forgiveness system and connecting it to the literature on amends.
Another contribution is that we examined happiness about forgiveness—a construct that, to our knowledge, has not been examined in prior research. Together, this article’s theoretical framework and findings suggest that happiness about forgiveness is an important construct because it can serve a functional purpose. Feeling unhappy or regret about one’s forgiveness could alert one to an adaptation risk caused by a mismatch between the extent to which one should forgive according to one’s forgiveness system and the extent to which one has actually forgiven. Future research could test the functional value of happiness about forgiveness by examining whether feeling badly about one’s forgiveness of one transgression predicts an adjustment in one’s forgiveness level and a corresponding change in happiness about forgiveness for a subsequent transgression. It could also be fruitful to identify and explore the potential functional role of feelings about other behaviors that affect one’s fitness and adaptation risks.
Finally, given the different H1 results between Study 2 and Study 3, a broader direction for future research is to examine the circumstances under which the results from studies using laboratory and MTurk samples and procedures are and are not generalizable. This is especially important given that the replication crisis has led many scholars to prioritize large samples which are easily recruited via MTurk. Although some types of studies may be unaffected by having participants hurry through a survey or simple procedures—perhaps as part of a multi-hour study binge—the more complex procedures of Study 3, which included an IAT and false feedback manipulations, may not be as amenable to MTurk samples. This speculative suggestion seems like a worthwhile topic for future research.
Conclusion
Psychologists have known that people are more likely to forgive when they have received strong amends or perceive the combination of low exploitation risk and high relationship value (e.g., Burnette et al., 2012; Hannon et al., 2010; Pansera & La Guardia, 2012). Until now, however, psychologists have not known much about what happens when people do not follow this pattern. In this article, we have theorized that when people forgive despite not having received strong amends, they heighten their exploitation risk; when people do not forgive despite having received strong amends, they heighten their risk of spoiling a valuable relationship. In both cases, we have shown that people tend to feel worse about their forgiveness level than if they received strong amends and matched those amends with a high level of forgiveness. Thus, it takes the prosocial behavior of both parties—victims and perpetrators—to create the situation in which victims can forgive without experiencing a decrement in their happiness about doing so.
Footnotes
Authors’ note
Any views expressed herein are those of the authors and do not necessarily reflect those of the funding agencies.
Acknowledgements
The authors would like to thank Jeff Coppola, Steve Geissinger, Michelle Irby, Julie Kittel, Tara Joy Knibbe, Blake Staat, and Lindsey Zamarripa for their assistance with data collection and coding.
Funding
The author(s) disclosed receipt of the following financial support for the research, authorship, and/or publication of this article: This work was supported by a grant from the University of Chicago’s Defining Wisdom Project and the John Templeton Foundation to Green, Davis, and Finkel.
