Research Note
More candor will make it worse
By James Carter · July 2026
Third in a series. The first examined what decision paralysis costs; the second, why almost no executive team runs a real after-action review.
The standard prescription for a leadership team that tolerates mediocrity is more candor. Say the hard thing. Debate openly. Have the conversation nobody wants to have.
Applied to most executive teams, that advice will degrade performance rather than improve it. This note is about the condition that determines which way it goes — and about why almost every article on this subject cites only the half of the research that supports the advice it is already giving.
Where the research disagrees
Two meta-analyses sit at the center of this literature. They reach different conclusions, nine years apart, and most writing in this market cites whichever one is convenient.
De Dreu and Weingart (Journal of Applied Psychology, 2003) aggregated the research on intragroup conflict and found both types negatively correlated with team performance and member satisfaction. Relationship conflict, the interpersonal kind, was negative as expected. Task conflict — the disagreement-about-the-work kind that academic research and textbooks had been recommending — was also strongly negative. The authors said plainly that this contradicted what the field had been teaching. They also found conflict more damaging in complex tasks, specifically decision-making and project work, than in production work.
De Wit, Greer and Jehn (Journal of Applied Psychology, 2012) revisited the question across 116 empirical studies covering 8,880 groups. They confirmed the stable negative relationships for relationship conflict and for process conflict. They did not replicate the strong negative association for task conflict. Instead they found the picture was contingent, and three of the contingencies matter here: task conflict related more positively to performance when its correlation with relationship conflict was weak, in top management teams rather than other teams, and when performance was measured as financial results or decision quality rather than overall performance.
Where they agree, which is the part that matters
Read those two side by side and the shared finding is easy to miss, because it is a moderator rather than a headline.
Both meta-analyses found the same hinge. De Dreu and Weingart reported that task conflict was less damaging when task and relationship conflict were weakly rather than strongly correlated. De Wit and colleagues reported that task conflict was more beneficial under exactly the same condition. Two independent aggregations, nearly a decade apart, over largely different bodies of primary research, landing on the same variable from opposite directions.
The variable is not how much a team argues. It is whether arguing about the work stays separate from arguing about each other.
Simons and Peterson (Journal of Applied Psychology, 2000) supply the mechanism, and they did it in the right population. Studying 70 top management teams, they found that intragroup trust moderates whether task conflict turns into relationship conflict. Their evidence supported what they call the misattribution explanation: in a low-trust room, a challenge to the plan gets received as a challenge to the person. Trust is what keeps the two apart.
Sull, Homkes and Sull (Harvard Business Review, 2015), surveying 7,600 managers across 262 companies, describe what the failure looks like at scale. A majority of the companies they studied delayed action on underperformance, addressed it inconsistently, or tolerated it outright. Asked what happens to a manager who hits his numbers but fails to collaborate across units, one in five said it would be handled promptly, three in five said inconsistently or late, and one in five said it would be tolerated. Separately, a third of the distributed leaders in their sample believed factions existed within the C-suite.
On segment. Both meta-analyses aggregate studies of groups broadly, and a substantial share of that primary research involves student groups and non-executive work teams. The top-management-team finding in de Wit and colleagues is a moderator within that larger set, not a study of executives. Simons and Peterson is the exception, and it is a single study of 70 teams. All of this is correlational. Nobody has established the direction of causation, and it is entirely plausible that failing teams generate conflict rather than the reverse. Sull’s sample had median annual sales around $430 million — overlapping the mid-market band — alongside average headcount near 6,000 and a sector concentration that does not.
What follows, and what doesn’t
The research names the condition. It stops there, and the advice industry fills the gap with the wrong instruction.
Here is the claim we make that the research does not.
Telling a team to be more candid is an intervention on volume. The research says the operative variable is separation. Those are not the same lever — and in a room where they are already fused, turning up the volume moves you further along the curve the 2003 meta-analysis measured, in decision-making work, where the relationship was most negative. You would expect the conversations to get more heated and the decisions to get worse. That is a specific, falsifiable prediction, and it is the opposite of what the person who ran your last offsite predicted.
We think this explains something CEOs report constantly and rarely get a good account of: the candor initiative that worked for a quarter and then made everything harder. The room did not fail to follow instructions. It followed them, and the instructions were aimed at the wrong variable.
Underneath it sits the thing that makes avoidance feel unfixable. Silence is comfortable. Naming a problem in front of eight peers is not, and it stays uncomfortable for exactly as long as the team experiences it as an event. The Flag Model treats this as the Standard, one of four disciplines that hold a team upright, and the rebuild is not more courage. It is making the naming of a problem routine enough that it stops registering as an accusation. Frequency is what produces separation. A team that surfaces small problems weekly is not braver than yours — it has simply removed the signal value from speaking up, so that raising something no longer implies the thing was worth the risk.
We flag that last step as interpretation. No study cited here measured frequency of problem-naming as a route to separating conflict types. It is our reading of what the moderator implies in practice.
The test
If the Standard is what broke on your team, you would predict a specific pattern rather than a general absence of candor. The team can name the problem. It just does it after the meeting.
So check the hallway. If the substantive assessment of a colleague’s plan happens in twos and threes afterward, and the meeting itself produced agreement, the problem is not that your people lack the words. They have the words. They are choosing an audience where saying them costs less.
Then check the second thing. When someone does challenge a plan in the room, watch what happens over the following week, not in the moment. If the challenge follows the person out of the room — if it shows up later as a comment about their judgment rather than about the plan — the two conflict types are fused, and adding candor will make it worse before it makes it better. If the challenge dies at the door and the two of them are fine on Thursday, your Standard is intact and your execution problem is somewhere else.
The second test is the one worth running. It costs nothing and it takes a week.
Sources
All primary. Every figure was verified against the original study or article text.
- De Dreu, C. K. W., & Weingart, L. R. (2003). Task Versus Relationship Conflict, Team Performance, and Team Member Satisfaction: A Meta-Analysis. Journal of Applied Psychology, 88(4), 741–749. ↗
- De Wit, F. R. C., Greer, L. L., & Jehn, K. A. (2012). The Paradox of Intragroup Conflict: A Meta-Analysis. Journal of Applied Psychology, 97(2), 360–390. ↗
- Simons, T. L., & Peterson, R. S. (2000). Task Conflict and Relationship Conflict in Top Management Teams: The Pivotal Role of Intragroup Trust. Journal of Applied Psychology, 85(1), 102–111. ↗
- Sull, D., Homkes, R., & Sull, C. (2015). Why Strategy Execution Unravels — and What to Do About It. Harvard Business Review, March 2015. ↗
“This truly resonated and will have a lasting influence on us as we work to create a technology organization where the best and brightest choose to be.”
Larry Quinlan · Global CIO, Deloitte
About the author
James Carter
Founder of Be Legendary and creator of the Flag Model™. Twenty-five years inside executive teams; co-author alongside Stephen Covey, Ken Blanchard, Deepak Chopra & Brian Tracy, and featured on CNN and in Business Insider. More about James →
See where your team breaks first.
A Calibration Call is 15 minutes — you leave with a concrete read on your team, whether or not we work together. It's a calibration, not a pitch.
Not ready to talk? Start free
Take the free Break-Point Self-Assessment
See which of the five components your team is most likely to lose first — in ~4 minutes. No email required to see your result.
Field notes, by email
Straight thinking on executive-team execution.
One short note, roughly monthly, on the disciplines that decide whether a leadership team executes. No fluff, no pitch. Unsubscribe anytime.