happily.ai research All research
Measurement

The Science Behind DEBI

The Dynamic Engagement Behavioral Index measures engagement from what people do, not what they say on a survey. What it reads, why those five signals, and how the score behaves in our data.

5
behaviors → one 0–100 score
15×
DEBI gap across one behavior
633
managers behind the validation

The usual way to measure engagement is to ask. An annual survey goes out, people rate how they feel across a set of scales, and the averages become the number the organization steers by. It is a reasonable idea with a measurement problem sitting underneath it. The thing being measured, how engaged someone is, and the instrument measuring it, a self-report survey, come from the same person, in the same sitting, with the same reasons to round up. When a predictor and an outcome share a source, that shared method inflates and distorts what you see (Podsakoff et al., 2003).

DEBI takes the other route. The Dynamic Engagement Behavioral Index does not ask people how engaged they are. It reads what they do: five everyday behaviors on the platform, scored continuously and combined into a single number from 0 to 100. It is behavioral because it is built from actions rather than answers, and dynamic because it moves as the behavior moves instead of freezing between survey windows.

This page is the reference for what that number means. It covers why behavior is a more faithful signal than self-report, the five behaviors DEBI reads and how they are weighted, the theory that connects those behaviors to engagement, and, most importantly, how the score behaves when we test it against our own data. Where the evidence is thin, it says so.

15× the gap in team engagement, measured by DEBI, between managers who never complete the daily check-in and those who do it most days. A score built to move with behavior should move with behavior. This one does.
Why this matters

Survey scores lag, and they can be answered the way people think they should be. By the time an annual number falls, the people it was meant to warn you about are often already leaving. In our study of the engaged exit, employees who later quit stayed active while quietly growing less happy, so the usual disengagement alarm never tripped. Behavioral warning signs, by contrast, tend to show up early: complaints about a manager precede departures by about 90 days. A measure read from behavior, updated continuously, can surface withdrawal while there is still time to act.

What DEBI is
Measure
A behavioral engagement index scored 0 to 100 per person, aggregable to team and company.
Signals
Five behaviors: daily check-in rate, open-feedback rate, open-feedback effort, recognition participation, and values-aligned recognition.
Construction
A weighted composite; weights 0.34 / 0.22 / 0.11 / 0.22 / 0.11 sum to 1.0, as defined by Happily.
Bands
Below 40 poor, 40 to 70 adequate, above 70 highly engaged, as defined by Happily.
Validation data
633 managers across 60 companies, replicated on an independent 542 managers across 100 companies. 365-day windows.
Methods
Dose-response bins, Cohen's d, and Pearson correlation against known engagement drivers.

Why watching beats asking

Several well-documented problems live inside self-report engagement measurement, and reading behavior sidesteps most of them. People answer surveys in ways that look good, the tendency social-desirability research has measured for decades (Crowne & Marlowe, 1960). They also have limited access to their own mental states, so a rating is often a story told after the fact rather than an accurate readout (Nisbett & Wilson, 1977). And what people say diverges from what they do: the attitude-behavior gap is one of the oldest findings in social science (LaPiere, 1934), and its economic form, revealed preference, holds that choices are a more reliable guide to what someone values than stated intent (Samuelson, 1938).

Cadence is the second problem. A once-a-year survey is a snapshot, and response rates erode as surveys pile up, which biases the sample that remains (Rogelberg & Stanton, 2007). Decades of ecological momentary assessment show that frequent, in-the-moment measurement beats retrospective recall, because it captures experience as it happens instead of asking people to reconstruct it later (Csikszentmihalyi & Larson, 1987; Shiffman, Stone & Hufford, 2008). A daily behavioral read is on the right side of that evidence. An annual questionnaire is not.

A survey is a snapshot. Behavior is a stream. The same year of engagement, read two ways. One of them misses the middle. Engagement → Jan Jun Dec One year A dip the survey never records DEBI · behavioral, continuous Survey · self-report, once or twice a year Illustrative schematic. Cadence grounded in ecological momentary assessment.
Figure 1 A survey samples one or two moments a year; a behavioral index reads the whole arc. The straight line between two survey points floats above a mid-year dip it never records.

None of this means the construct is wrong. Engagement as Kahn (1990) defined it, the investment of one's full self in a role, is real and matters for outcomes, and the instruments that measure it, the Utrecht Work Engagement Scale (Schaufeli, Bakker & Salanova, 2006) and Gallup's Q12 (Harter, Schmidt & Hayes, 2002), are well validated. They are also self-report. DEBI keeps the construct and changes the instrument: the same target, read from behavior instead of a questionnaire.

The five behaviors DEBI reads

DEBI blends five behavioral signals, each weighted, into one score. The weights sum to 1.0 and fall into three families that carry almost equal weight: showing up, speaking up, and lifting others up.

The five behaviors DEBI reads Each score is a weighted blend of five everyday behaviors. The weights sum to 1.0. Daily check-in response rate 0.34 Open feedback — how often 0.22 Open feedback — how much effort 0.11 Recognition — giving and getting 0.22 Recognition — aligned to values 0.11 Presence 0.34 · showing up Voice 0.33 · speaking up Recognition 0.33 · lifting others Weights as defined by Happily.ai. The three behavioral families are weighted almost equally.
Figure 2 The five signals and their weights, colored by family. Presence carries 0.34, voice 0.33, and recognition 0.33, so no single behavior dominates the score. Weights as defined by Happily.
The five components of DEBI
Behavioral signalFamilyWeight
Daily check-in response ratePresence0.34
Open-feedback rateVoice0.22
Open-feedback effortVoice0.11
Recognition participationRecognition0.22
Values-alignment of recognitionRecognition0.11
Total1.00

Why these five

Each family maps to a behavior the engagement literature already treats as a marker of engagement, which is what makes the composite more than a convenient sum.

Presence is showing up to the daily pulse, day after day. Consistent participation is the behavioral trace of the attentional and physical investment Kahn (1990) placed at the center of personal engagement, and reading it daily rather than annually is exactly what the momentary-assessment evidence recommends. It carries the largest single weight, 0.34.

Voice is open-ended feedback, split into how often people give it (0.22) and how much effort they put in (0.11). Feedback-giving is voice behavior, the discretionary act of speaking up with ideas and concerns, and it is a validated extra-role behavior distinct from required tasks (Van Dyne & LePine, 1998). Its absence, silence, is a recognized signal of withdrawal (Morrison, 2011). Weighting rate above effort says that speaking up at all matters more than how polished it is.

Recognition is taking part in peer recognition (0.22) and whether that recognition maps to company values (0.11). Recognition is a documented antecedent of engagement through social exchange (Saks, 2006), and well-recognized employees are markedly more likely to stay (Workhuman & Gallup, 2024). The values-alignment component asks not just whether recognition happens but whether it points at the behaviors the company says it cares about.

Read together, the three families line up with self-determination theory's basic needs, autonomy, competence, and relatedness (Ryan & Deci, 2000): voice expresses autonomy, recognition builds relatedness, and steady participation reflects competence in the daily work of a team. The score is a formative index, meaning the behaviors define it rather than reflect a hidden trait (Diamantopoulos & Winklhofer, 2001), and it is assembled the way composite indicators are built and weighted (Nardo et al., 2008). The premise underneath all of it, that behavioral traces carry real signal about latent psychological states, is well supported (Kosinski, Stillwell & Graepel, 2013).

How each family maps to the engagement literature
Family (weight)What DEBI readsWhat it marks
Presence (0.34)Daily check-in participationAttentional investment; daily cadence (Kahn 1990)
Voice (0.33)Open feedback, rate and effortDiscretionary voice; silence signals withdrawal (Van Dyne & LePine 1998; Morrison 2011)
Recognition (0.33)Peer recognition, values-alignedRelatedness and social exchange (Saks 2006; Ryan & Deci 2000)

Does the score track engagement?

A definition is only as good as its behavior in the field. A valid measure should do three things: move with what should drive it, correlate with the right things and ignore the wrong ones, and hold up in a second dataset. DEBI does all three.

1. It moves with behavior

Across 633 managers, team DEBI rises monotonically with how often the manager completes the daily check-in, from 3.4 for those who never check in to 51.9 for those who do it most days. That is a fifteen-fold gap and a very large effect (Cohen's d = 2.09). A score that stayed flat while a known engagement driver ranged from zero to daily would not be measuring engagement. This one climbs at every step.

A behavioral score should move with behavior. DEBI does. Mean team DEBI by the manager's daily check-in rate. n=633 managers, 365 days. 0% 3.4 n=206 1–25% 33.0 n=223 26–50% 44.5 n=98 51–75% 51.9 n=106 0 20 40 60 Mean team DEBI 15× gap between never and consistent check-ins · Cohen's d = 2.09 Source: Happily Research, Manager DEBI Drivers (2026). 633 managers, 60 companies.
Figure 3 Mean team DEBI by the manager's daily check-in rate. The climb is monotonic: even occasional check-ins produce roughly ten times the engagement of none, and the curve keeps rising from there.

Recognition shows the same shape. Managers who give no recognition in a year sit at a mean team DEBI of 8.8; those who give more than fifty reach 61.0, and every bin in between is higher than the one before it.

Team DEBI by recognitions given per year (n=633)
Recognitions givenMean team DEBIn
08.8274
1–530.484
6–2042.0128
21–5047.2101
51+61.046

2. It tracks the right things and ignores the wrong ones

A good measure is defined as much by what it does not respond to as by what it does. The six behaviors an engaged workforce shows, a manager's own well-being, recognition given and received, replying to feedback and replying well, and checking in, all correlate with team DEBI in a tight band between 0.54 and 0.59. Two things engagement should not depend on, how long a manager has been at the company and how many people report to them, sit at essentially zero (0.04 and -0.04). A score that rewarded tenure or head count would be suspect. DEBI does neither. This is the pattern Campbell and Fiske (1959) named convergent and discriminant validity.

What DEBI tracks, and what it ignores Correlation (r) of each variable with team DEBI. n=633 managers. r = 0 0.3 0.6 Manager happiness 0.59 Recognition received 0.58 Reply rate to feedback 0.57 Reply quality 0.56 Check-in frequency 0.55 Recognition given 0.54 Manager tenure 0.04 Team size -0.04 Convergent: what an engaged workforce does Discriminant: what engagement shouldn't depend on Source: Happily Research, Manager DEBI Drivers (2026). Pearson r, 633 managers, 60 companies.
Figure 4 Correlation of each variable with team DEBI across 633 managers. Engagement behaviors cluster at r = 0.54 to 0.59; tenure and team size sit near zero. The measure responds to behavior, not to seniority or span.

3. It replicates

The pattern holds in a separate sample of 542 managers across 100 companies. Managers who never reply to the feedback they receive score a team DEBI of 31.0, against 45.7 for those who do, a medium effect (d = 0.599). And the theoretically best behavior, high-quality replies given within a few days, reaches 54.1, against 28.5 for rushed, low-quality ones, a large effect (d = 0.988). DEBI separates engaged from disengaged managerial behavior in a dataset it was not tuned on.

Read it as a strong signal, not a verdict

Two things temper these numbers, and both point the same way. Much of the gap at the bottom is participation rather than disengagement: about a quarter of managers register almost no platform activity and score near zero, which stretches the effect sizes. And company culture is a large confound. Within a single company, individual behavioral differences predict DEBI far more weakly than the cross-company figures suggest, because a big share of the score reflects the organization's baseline rather than the individual. The direction of the evidence is clear; the exact magnitudes are not.

Reading the score

DEBI runs from 0 to 100, and Happily reads it in three bands.

DEBI bands, as defined by Happily
BandRangeReading
PoorBelow 40Behavioral engagement is low; participation, voice, or recognition is thin.
Adequate40–70A working level of engagement, with room to move.
Highly engagedAbove 70Strong, consistent engagement behavior across the families.

Two rules make the number more useful. First, a DEBI score is a level, and a level reads best next to a trend. A team at 55 and climbing is a different story from a team at 55 and sliding, and the sliding one is where attention belongs. Second, a very low score, especially a zero, often means "not enough behavior to read" rather than "actively disengaged." Before treating a zero as a crisis, check whether the person is on the platform at all.

What this means

Used well, DEBI is a continuous early-warning layer that complements periodic surveys and hard outcomes rather than replacing them. The table below is how to read common patterns.

Reading DEBI in practice
What you seeLikely meaningWhat to do
DEBI near 0, almost no activityNon-participation, not measured disengagementFix coverage and onboarding before you interpret the score
Low DEBI, active userGenuine low engagementLook at which family is low: presence, voice, or recognition
Falling team DEBIBehavioral withdrawal formingAct now, before it would ever reach an annual survey
High, stable DEBIEngaged teamProtect the conditions producing it
High DEBI, other flags highEngaged but possibly at riskDo not read engagement as safety; pair with well-being and flight-risk signals

Limitations

  • Participation confound. Inactive users score near zero, and about 25% of managers in the drivers study sat at DEBI 0. A behavioral index cannot separate "disengaged" from "absent from the platform" without checking activity first, and this inflates the effect sizes reported above.
  • Platform-bounded. DEBI only reads behavior that happens on Happily. Engagement expressed in other places, and disengagement hidden elsewhere, are invisible to it.
  • Culture confound. A meaningful share of a DEBI score reflects the company's baseline rather than the individual. Within a single company, individual behavioral differences predict it more weakly, and the raw driver correlations shrink after adjusting for company.
  • External validation is still open. In our data DEBI tracks behavioral drivers and self-reported manager happiness, but we have not yet published its correlation against independent gold-standard outcomes such as eNPS, WHO-5 well-being, or retention. Until that is done, the convergent case rests on internal drivers, which is a real but partial form of validation.
  • Fixed, proprietary weights. The component weights are set by Happily and have not been independently re-derived from outcomes or audited.
  • Proxies, not the felt state. The five signals are behavioral indicators of engagement, not the internal experience itself. They are a good read on it, not a substitute for it.

References

  1. Podsakoff, P. M., MacKenzie, S. B., Lee, J.-Y. & Podsakoff, N. P. (2003). Common Method Biases in Behavioral Research. Journal of Applied Psychology, 88(5), 879–903.
  2. Crowne, D. P. & Marlowe, D. (1960). A New Scale of Social Desirability Independent of Psychopathology. Journal of Consulting Psychology, 24(4), 349–354.
  3. Nisbett, R. E. & Wilson, T. D. (1977). Telling More Than We Can Know: Verbal Reports on Mental Processes. Psychological Review, 84(3), 231–259.
  4. LaPiere, R. T. (1934). Attitudes vs. Actions. Social Forces, 13(2), 230–237.
  5. Samuelson, P. A. (1938). A Note on the Pure Theory of Consumer's Behaviour. Economica, 5(17), 61–71.
  6. Rogelberg, S. G. & Stanton, J. M. (2007). Understanding and Dealing with Organizational Survey Nonresponse. Organizational Research Methods, 10(2), 195–209.
  7. Csikszentmihalyi, M. & Larson, R. (1987). Validity and Reliability of the Experience-Sampling Method. Journal of Nervous and Mental Disease, 175(9), 526–536.
  8. Shiffman, S., Stone, A. A. & Hufford, M. R. (2008). Ecological Momentary Assessment. Annual Review of Clinical Psychology, 4, 1–32.
  9. Kahn, W. A. (1990). Psychological Conditions of Personal Engagement and Disengagement at Work. Academy of Management Journal, 33(4), 692–724.
  10. Schaufeli, W. B., Bakker, A. B. & Salanova, M. (2006). The Measurement of Work Engagement with a Short Questionnaire. Educational and Psychological Measurement, 66(4), 701–716.
  11. Harter, J. K., Schmidt, F. L. & Hayes, T. L. (2002). Business-Unit-Level Relationship Between Employee Satisfaction, Employee Engagement, and Business Outcomes. Journal of Applied Psychology, 87(2), 268–279.
  12. Ryan, R. M. & Deci, E. L. (2000). Self-Determination Theory and the Facilitation of Intrinsic Motivation, Social Development, and Well-Being. American Psychologist, 55(1), 68–78.
  13. Van Dyne, L. & LePine, J. A. (1998). Helping and Voice Extra-Role Behaviors: Evidence of Construct and Predictive Validity. Academy of Management Journal, 41(1), 108–119.
  14. Morrison, E. W. (2011). Employee Voice Behavior: Integration and Directions for Future Research. Academy of Management Annals, 5(1), 373–412.
  15. Saks, A. M. (2006). Antecedents and Consequences of Employee Engagement. Journal of Managerial Psychology, 21(7), 600–619.
  16. Campbell, D. T. & Fiske, D. W. (1959). Convergent and Discriminant Validation by the Multitrait-Multimethod Matrix. Psychological Bulletin, 56(2), 81–105.
  17. Diamantopoulos, A. & Winklhofer, H. M. (2001). Index Construction with Formative Indicators. Journal of Marketing Research, 38(2), 269–277.
  18. Kosinski, M., Stillwell, D. & Graepel, T. (2013). Private Traits and Attributes Are Predictable from Digital Records of Human Behavior. Proceedings of the National Academy of Sciences, 110(15), 5802–5805.
  19. Nardo, M., Saisana, M., Saltelli, A., Tarantola, S., Hoffman, A. & Giovannini, E. (2008). Handbook on Constructing Composite Indicators. OECD Publishing.
  20. Workhuman & Gallup (2024). From Thank You to Thriving: Recognition, Wellbeing, and Retention. Industry report.
  21. Happily.ai (2025). The Science Behind the Dynamic Engagement Behavior Index (DEBI). Company blog. Source for the index definition, components, weights, and interpretive bands.
  22. Happily Research (2026). Manager DEBI Drivers (633 managers, 60 companies); Reply Quality Beats Reply Speed (542 managers, 100 companies); The Engaged Exit. Internal analyses cited inline.
Measure engagement from behavior, not belief

DEBI turns five everyday behaviors into a continuous engagement signal, so you can see disengagement forming while there is still time to act. See what it reads for your team.

Get in touch
Free pilot for qualifying teams