← All articles

Conflict patterns

8 min read

The First Three Minutes of an Argument

How a disagreement opens carries most of the information about where it ends up. The famous prediction percentage attached to that finding is the part to be careful with.

A half-open doorway with warm light falling in a stripe across a rug

There is a claim that circulates about couples research, usually in the form of a statistic. Watch a couple argue for a few minutes and you can predict, with some impressive percentage, whether they will still be married in six years.

The underlying research is real, it was carefully done, and it found something worth knowing. The percentage is the part you should not repeat, and the reason why is more interesting than the claim.

What they did

A 1999 study took 124 newlywed couples and recorded each of them having one fifteen-minute conversation about a live disagreement in their marriage. Ordinary lab procedure for this field: the couple sits down, picks a real issue, and talks while the cameras run.

The tapes were then coded second by second for facial expression, tone of voice and content, by people trained to do it who knew nothing about the couples or what the study was testing. The conversations were divided into five three-minute blocks so the researchers could look at how the emotional tone moved across the quarter of an hour.

Then they waited six years and checked who was still married.

What they found

The information was concentrated at the front.

One fifteen-minute conversation about a real disagreementminutes 1–3minutes 4–6minutes 7–9minutes 10–12minutes 13–15The opening carried itHow the issue was raised, andwhether it opened hard or soft.Everything after this point waslargely already in motion.Coded second by second for face, voice and content, by people who knew nothing about the couples.
The whole conversation was coded. What separated the couples who later divorced was concentrated at the front, in how the issue got raised rather than in how it got resolved.Design of a 1999 study of 124 newlywed couples, followed for six years. See the sources below.

Outcomes over the following six years could be modelled from the first three minutes on their own. The seventeen couples who went on to divorce had opened their conflict conversations with noticeably more negativity and less warmth than the couples who stayed together. Not louder, not more dramatic. They opened harder.

An earlier study from the same programme, following 130 newlywed couples over the same period, pointed in the same direction and added a second detail. What distinguished the couples who stayed happy was not skill at resolving anything. It was how issues got raised, whether someone made a move to bring the temperature down when it climbed, whether their bodies settled during the conversation, and whether the husbands in that sample took their wives’ positions seriously enough to give ground on something.

That study also tested the technique that dominated couples education at the time, in which partners take turns paraphrasing and validating each other before responding. It did not hold up. Couples who stayed happy almost never did anything like it while they were upset. It was largely absent from the tapes.

The part you should be suspicious of

Both of those studies report classification accuracy figures, and those figures are why this research is famous. They are also the weakest thing in it.

The models were built by looking at which behaviours separated the couples who divorced from the couples who did not, in that specific set of couples. Then the accuracy of those models was measured on the same couples.

What was doneone group of couplesmodelbuilthereandscoredhereSame couples, both times.The accuracy figure thisproduces is not a forecast.What should be donebuilt herescored hereCouples it has never seen.This is the number thatwould mean something.Anyone quoting a divorce-prediction percentage at you should be asked which of these they did.
A model scored on the couples it was built from is being asked to recognise faces it has already memorised. Tested properly, on couples it has not seen, the accuracy in this literature falls a long way.The methodological critique in Heyman & Slep (2001). No accuracy figures are quoted here, in either direction, because the honest ones vary by model and the flattering ones should not circulate.

In 2001, two researchers reanalysed archival observational data from this literature and showed what happens when you do it properly. Build the model on one set of couples, then test it on couples it has never seen. Accuracy drops substantially. Their conclusion was that results reported without cross-validation should be treated with extreme caution no matter how impressive the headline number looks.

We think that critique is correct, which is why no percentage appears anywhere in this article. The mechanisms the research describes are not in doubt. The forecasting claim built on top of them is.

An honest reading of the limits

Only seventeen divorces. With an outcome group that small, a model has very few cases to learn from, and small changes in the sample would move the result around considerably.

One conversation, in a lab, on a day everyone knew they were being filmed. People behave somewhat differently under those conditions. The direction of that bias is unknown.

The sample is narrow. Newlywed couples in the American Pacific Northwest in the early 1990s, overwhelmingly heterosexual and married. The finding about husbands accepting influence was specifically about husbands in that sample and should not be carried anywhere else without care.

The active-listening result is a failure to find something. That is weaker evidence than finding something. It does not prove that structured listening is useless, only that it was not what these particular couples were doing when things went well.

Nobody can forecast your relationship. Not from three minutes, not from fifteen, not from a questionnaire. Anyone offering to is selling something.

What this means for the fight you keep having

Strip out the prediction claim and a genuinely useful finding is left standing: the opening of a conversation does a large share of the work, and the opening is the part most under your control.

Consider how the chores conversation usually starts. “You never think about anyone but yourself” and “I’m tired and the kitchen was like that again this morning” are about the same underlying grievance. The first is a verdict on a person. The second is a description of a situation and a feeling. Only one of them can be answered with anything other than a defence.

The same is true of the in-laws conversation, which almost always opens with an accusation of divided loyalty and almost never opens with the specific thing that happened on Sunday.

This is a small, learnable, unglamorous skill. Say what happened, say how it landed, say what you would like instead. Do it in the first thirty seconds. It is probably the highest-leverage change available to most couples, and it takes about a week of deliberate practice to start doing it under pressure.

There is a second finding worth taking, which is about giving ground. Accepting influence is not conceding the argument. It is locating the part of what your partner said that is fair, and saying that part out loud, before you make your own case.

Where a session with Elena fits

Elena is an AI coach, and the specific thing a session does with this finding is rehearsal.

Before an issue gets raised with your partner, she has you say it to her first, and works on it with you until it comes out as a situation and a feeling rather than a verdict. The sentence you settle on goes into the note, so the next session starts from the version that worked rather than from memory. When one of you makes a move to lower the temperature, she stops and marks it, because those moves are easy to miss from inside the conversation.

The honest boundary: our structured turn-taking is scaffolding to keep two people from talking over each other. The research above does not say that turn-taking is the mechanism of change, and the one technique in this area that was directly tested here did not survive the test. We use structure because it makes a conversation possible, not because we have evidence it is the active ingredient.

The papers

  1. (1999). Predicting divorce among newlyweds from the first three minutes of a marital conflict discussion. Family Process, 38(3), 293–301.

    doi.org/10.1111/j.1545-5300.1999.00293.x124 newlywed couples, and the information sitting in the opening.

  2. (1998). Predicting marital happiness and stability from newlywed interactions. Journal of Marriage and the Family, 60(1), 5–22.

    doi.org/10.2307/353438Soft start-up, accepting influence, and the active-listening null result.

  3. Heyman, R. E., & Slep, A. M. S. (2001). The hazards of predicting divorce without crossvalidation. Journal of Marriage and Family, 63(2), 473–479.

    Open access · PMC1622921Why no accuracy percentage appears anywhere in this piece.

Checked against the publisher record before this piece went up. If a number here cannot be traced to the paper above, it should not be here.


Try it with someone

Reading about the pattern is not the same as having it named while it happens.

Your first session is free, 45 minutes, both of you, no card.