Let's Talk CultureTake the Culture Strength Check →
Methodology

The science behind the Culture Strength Check

How we measure a team's culture, the research each choice rests on, and, just as importantly, what we deliberately do not claim.

The whole thing, in a paragraph

Culture is built, on purpose, one conversation at a time. The Culture Strength Check measures how strongly and deliberately a team has the six conversations that shape it, on a foundation of psychological safety, and how well that translates into engagement. It measures strength and intention, how much your culture is by design rather than by default, not what type of culture you have. Every score is built from behavioural questions grounded in published research and scored transparently, never by asking you to rate yourself.

01 · The claim

What we measure, and what we don't

There are two very different things you can measure about a culture. The first is its content, or type: whether it is collaborative or competitive, cautious or bold. Most culture tools sort you into a type. The second is its strength and intentionality: how clearly and deliberately the culture has been built, and therefore how coherent and resilient it is. We measure the second.

We do this for three reasons. First, it is actionable: a type label leaves you nowhere to go, whereas measuring the practice of building culture hands you the change levers. Second, there is no single right culture, what works for one team fails for another, so ranking teams against one ideal is the wrong move. What matters is how clearly a culture is being built and whether people can see it and choose to belong to it. Third, it is honest: strength and intention can genuinely be read from behaviour, whereas claiming to X-ray a team's soul in eight minutes cannot.

What this instrument does not claim. It is not a personality or culture-type test. It does not prove causation for your team. And on its own, taken by a single leader, it is an informed estimate, not a validated measurement, a distinction we return to in Section 06.

02 · The model

Five layers, read as one system

The model is not a flat checklist. Its parts sit at different depths, and reading them as layers is what makes it a model rather than a list.

Foundation · Psychological safety

Psychological safety is the shared belief that a team is safe for interpersonal risk-taking, that you can ask a question, admit a mistake, or disagree without it being held against you.1 Amy Edmondson's foundational work showed that teams higher in safety learn and perform better,1 and Google's study of 180 of its own teams (Project Aristotle) found it the single largest differentiator between its best and worst teams.2 We treat it as the substrate, not a seventh conversation: when safety is low, people do not voice the truth, so the rest of the report is systematically inflated. In scoring, low safety therefore acts as a gate on confidence in everything above it (Section 04).

Practice · The six conversations

The six conversations are drawn from Shane Hatton's book Let's Talk Culture, the behaviours through which culture is built by design. Each also maps to a published construct, so the framework is grounded in both practice and research.

#ConversationWhat it readsGrounded in
1ExpectationWhether expectations are surfaced and shared, not privately assumedRole clarity and ambiguity;3 shared mental models;4 group norms5
2ClarificationWhether values are specific, observable behavioursGoal specificity;6 behaviourally anchored standards7
3CommunicationWhether culture lives in shared language and visible reasonsOrganisational sensemaking;8 procedural justice and transparency9
4CapabilityWhether people have the skills and support to meet the standardSelf-efficacy;10 competence in SDT;11 the knowing-doing gap;12 training transfer13
5ConfrontationWhether accountability is kind and early, not avoidedFeedback intervention theory;14 task versus relationship conflict15
6CelebrationWhether the behaviours you want get noticed and namedReinforcement; recognition and engagement;16 the progress principle17

The order is not arbitrary: each conversation builds on the last, and capability sits before confrontation deliberately, because holding someone to a behaviour they have not been equipped for is unfair and erodes trust.

Lens · The three gaps

The three gaps are not additional things we measure; they are three ways of reading the same conversations. Each has a distinct research anchor, and this framing, culture as a set of gaps rather than a grade, is the part no other instrument does.

Outcome · Engagement

Engagement, rendered here in plain language as energy, purpose and focus, is our read of the validated dimensions of work engagement, vigour, dedication and absorption.20 It is an outcome, not a lever: you do not build engagement directly, you build the conditions and the conversations, and engagement tells you whether they are landing. Kahn's original conditions for engagement, meaningfulness, safety and availability,21 map almost one to one onto this model, with safety as the foundation, availability as capability, and meaningfulness as the purpose we read at the top.

03 · The questions

Behavioural, not evaluative

The single most important design choice is that we never ask people to rate the construct directly. “Rate your team's psychological safety out of 100” triggers ego and social desirability, and almost everyone answers around 85.22 Instead we ask about specific, recent, observable behaviour and compute the construct from it. Asking “when someone makes a mistake here, does it tend to get held against them?” is answered honestly, because it does not feel like a judgement of the respondent, and it tells you far more.

Three moves make this work:

Every item maps to one facet of one construct, so results are specific enough to act on rather than vague. There are 32 items in total, across the six conversations, psychological safety and engagement. In the team version, items written from the leader's point of view are reframed to the team member's seat, so both are reading the same team through comparable questions, which is what makes the perception gap valid.

04 · The scoring

From answers to a score

05 · The team benchmark

How the perception gap is measured

A single leader can give an informed read, but two of the three gaps, perception and spread, do not exist until the team also answers. In the team benchmark, teammates take the same read anonymously, and:

06 · The honest part

Validity, limits, and what comes next

A methodology that will not name its own limits does not deserve to be trusted, so here are ours plainly.

None of this is a reason to distrust the read; it is the basis for trusting it, at exactly the level it earns. The citations below are the grounding for each design choice.

References

What each choice stands on

  1. Edmondson, A. C. (1999). Psychological safety and learning behavior in work teams. Administrative Science Quarterly, 44(2), 350–383.
  2. Google re:Work / Project Aristotle (2015). The five keys to a successful Google team; psychological safety as the top factor.
  3. Rizzo, J. R., House, R. J., & Lirtzman, S. I. (1970). Role conflict and ambiguity in complex organizations. Administrative Science Quarterly, 15(2), 150–163.
  4. Cannon-Bowers, J. A., Salas, E., & Converse, S. (1993). Shared mental models in expert team decision making. In Individual and Group Decision Making.
  5. Feldman, D. C. (1984). The development and enforcement of group norms. Academy of Management Review, 9(1), 47–53.
  6. Locke, E. A., & Latham, G. P. (2002). Building a practically useful theory of goal setting and task motivation. American Psychologist, 57(9), 705–717.
  7. Smith, P. C., & Kendall, L. M. (1963). Retranslation of expectations: behaviourally anchored rating scales. Journal of Applied Psychology, 47(2).
  8. Weick, K. E. (1995). Sensemaking in Organizations. Sage.
  9. Colquitt, J. A. (2001). On the dimensionality of organizational justice. Journal of Applied Psychology, 86(3), 386–400.
  10. Bandura, A. (1997). Self-Efficacy: The Exercise of Control. Freeman.
  11. Deci, E. L., & Ryan, R. M. (2000). Self-Determination Theory and the facilitation of intrinsic motivation. American Psychologist, 55(1).
  12. Pfeffer, J., & Sutton, R. I. (2000). The Knowing-Doing Gap. Harvard Business School Press.
  13. Baldwin, T. T., & Ford, J. K. (1988). Transfer of training: a review and directions for future research. Personnel Psychology, 41(1).
  14. Kluger, A. N., & DeNisi, A. (1996). The effects of feedback interventions on performance. Psychological Bulletin, 119(2), 254–284.
  15. Jehn, K. A. (1995). A multimethod examination of the benefits and detriments of intragroup conflict. Administrative Science Quarterly, 40(2), 256–282.
  16. Gallup. Employee recognition and its relationship to engagement (Q12 research programme).
  17. Amabile, T. M., & Kramer, S. J. (2011). The Progress Principle. Harvard Business Review Press.
  18. Atwater, L. E., & Yammarino, F. J. (1992). Does self-other agreement on leadership perceptions moderate the validity of leadership and performance predictions? Personnel Psychology, 45(1).
  19. Schneider, B., Salvaggio, A. N., & Subirats, M. (2002). Climate strength: a new direction for climate research. Journal of Applied Psychology, 87(2), 220–229.
  20. Schaufeli, W. B., Bakker, A. B., & Salanova, M. (2006). The measurement of work engagement with a short questionnaire (UWES). Educational and Psychological Measurement, 66(4).
  21. Kahn, W. A. (1990). Psychological conditions of personal engagement and disengagement at work. Academy of Management Journal, 33(4), 692–724.
  22. Paulhus, D. L. (1984). Two-component models of socially desirable responding. Journal of Personality and Social Psychology, 46(3); see also Edwards, A. L. (1957).
Let's Talk Culture · the Culture Strength Check. This methodology will be updated as validation data accumulates. Take the check →