Research Fit, teams & culture

The office was part of the instrument.

Remote and hybrid work measurably work — and autonomy changes who performs: conscientiousness predicts far more steeply when nobody is watching. The evidence, and the re-weighting hiring owes a distributed role.

The office has never appeared in a job description, but it has always been part of one. A building supplies management that no one has to write down: fixed hours, a shared working rhythm, colleagues in peripheral vision, a manager close enough to notice a stalled week before it becomes a stalled quarter. The office, in other words, was part of the instrument — part of how the job was measured and structured. Hiring for remote work looks like a new problem because removing the building removes all of that at once: the same duties, stripped of ambient supervision and imposed routine, become a high-autonomy job.

The evidence says two things about that transformation, and they pull in the same direction. Remote and hybrid arrangements themselves work: the flagship randomized experiments found a 13% performance gain with attrition cut in half at full remote, and a third fewer quits at hybrid with no detectable performance loss. And autonomy changes who performs: the predictive weight of conscientiousness rises sharply when a job leaves people to structure their own work, and experts who staff virtual teams rate self-directed action and the analysis and interpretation of information as significantly more important than their co-located counterparts do.

Neither finding says remote work needs a new kind of worker, and the expert-ratings evidence is explicit that broad trait requirements do not differ. What changes is the weighting. The dispositions an office compensated for (the structure a routine-poor employee borrowed from the building, the visibility that made supervision effortless) now have to come from the person. Remote employee selection is the practice of measuring, before the offer, what the building used to supply after it.

Remote and hybrid work passed their randomized trials

For years the remote-work argument ran on surveys, self-reports, and competing anecdotes. Bloom, Liang, Roberts, and Ying (2015) replaced it with a nine-month randomized experiment, published in the Quarterly Journal of Economics and run at Ctrip, a 16,000-employee Chinese travel agency. From 249 qualifying call-center volunteers, the company randomized by birthdate: the treatment group worked four days a week at home with one day in the office; controls kept commuting.

The study's design, more than its setting, changed the debate. Observational comparisons of remote and office workers confound the arrangement with the people who chose it, and every claim built on them inherits that confound. Randomization severs the link: whatever differences emerged between the two Ctrip groups were caused by the arrangement itself, since the volunteers were sorted by birthdate. It is the same logic that separates selection science from hiring folklore, and it is the reason two studies of one travel company carry more evidential weight than shelves of workplace surveys.

Home workers' measured performance rose 13%. The gain decomposes: roughly 9 points came from more minutes worked per shift, mostly fewer breaks and sick days, and roughly 4 points from more calls per minute, which the authors attribute to a quieter working environment. Call quality did not decline. Reported work satisfaction rose. Attrition among home workers fell by roughly half.

The trial also recorded a penalty that outlasted it. Conditional on performance, home workers' promotion rates roughly halved. The metrics could see the output; the people deciding advancement could not see the workers. Any organization running remote arrangements inherits that calibration problem, and it returns at the end of this article.

The subtlest result came after the experiment ended, when employees were allowed to re-select their location. Over half switched, and the average performance gain almost doubled, to 22%. Bloom and colleagues attribute the improvement to learning and selection together: nine months of experience had taught people where they worked well, and the re-sorting kept the arrangement for those it suited. The 22% figure is not a purer estimate of working from home; it is an estimate of working from home plus informed matching, which for a hiring audience is the more interesting quantity.

Nine years later the same company, by then Trip.com, hosted the hybrid replication at scale. Bloom, Han, and Liang (2024) randomized 1,612 employees during 2021–2022 into two home days per week and reported the results in Nature: quit rates fell by a third and satisfaction rose, while null-equivalence tests, statistical tests designed to affirm the absence of an effect, found no impact on performance grades or promotions over the following two years. For engineers, lines of code written did not move. The retention gains were concentrated among non-managers and employees with long commutes. And one number tracked beliefs rather than outcomes: before the experiment, the company's managers estimated that hybrid work would cut productivity by 2.6%; afterward, their average estimate was a gain of 1.0%.

Both trials, plotted in Figure 1, are policy experiments on incumbent employees: volunteers already hired, trained, and socialized before anyone randomized their working arrangement. They establish that remote and hybrid arrangements can work, and work well. They establish nothing about whom to hire into them. That question needs evidence about how the changed job changes prediction, and it starts with what the change actually is.

The active ingredient is autonomy

Gajendran and Harrison (2007) pooled 46 studies covering 12,883 employees in a meta-analysis in the Journal of Applied Psychology, asking what telecommuting does to the psychology of a job. The clearest signature was perceived autonomy: telecommuting correlated with employees' sense of control over their own work at a corrected mean correlation (ρ) of .22, and the analysis positions that autonomy as the psychological path through which the arrangement's benefits travel. (The correction machinery behind such estimates is described in the series' foundation article on what predicts job performance.)

The downstream consequences are small and consistently signed: job satisfaction .10, turnover intent −.10, role stress −.13, work-family conflict −.13, supervisor-relationship quality .12. Each is modest; together they point one way.

The performance figures sit side by side in the analysis, and the contrast between them carries the point: supervisor-rated and objectively measured performance correlated with telecommuting at .19, while self-rated performance came in at .01, statistically indistinguishable from zero. Whatever telecommuting does for output, it is not happening in workers' imaginations: it shows up in the ratings they do not write.

The meta-analysis also marks the arrangement's edge. At high intensity, defined as telecommuting 2.5 or more days per week, coworker relationships suffered (r = −.19), even as the individual-level benefits held. Intensity is a real variable: a role that is remote two days a week and a role that is remote five are different roles, and later sections treat them that way.

Mediation matters because it names the ingredient. If the benefits of telecommuting traveled through, say, commuting time alone, remote work would be a logistics story with no selection content. Because they travel through perceived autonomy, the arrangement reaches into the structure of the job itself: what is discretionary expands, what is imposed contracts, and the traits that govern self-imposed structure move closer to the criterion. The meta-analytic path is the reason the rest of this article is about selection at all.

Put the mediator beside the experiments and the mechanism comes into focus: remote work is, by measurement, high-autonomy work. Removing the building is experienced chiefly as gaining control over time, place, and method. That is what makes a selection literature written decades before the pandemic directly relevant to hiring remote teams; the field was asking the right question long before remote work made it urgent.

Autonomy changes who performs

Barrick and Mount (1993) is not a remote-work study. It is a Journal of Applied Psychology study of 146 managers whose jobs varied in the discretion they allowed, and discretion is precisely the variable the telecommuting literature says remote arrangements move. Across the full sample, conscientiousness predicted supervisor-rated performance at r = .25 (corrected ρ = .35) and extraversion at r = .14 (ρ = .20), ordinary numbers for personality in managerial work.

The finding that matters here is the interaction. For conscientiousness, extraversion, and agreeableness, the strength of the trait-performance relationship depended significantly on the job's autonomy. For conscientiousness the difference was stark in slope terms: roughly seven times steeper under high autonomy, meaning the same difference between two candidates on the trait corresponded to roughly seven times the difference in rated performance when the job left the manager to structure the work. Figure 2 shows the geometry with the study's anchors attached.

Each significant interaction added about three percentage points of explained variance (ΔR² = .03), and the scale of the moderation deserves as much care as its direction. Barrick and Mount reported that increment as modest, and it is: autonomy does not turn conscientiousness from a minor predictor into a dominant one; it steepens the slope where hiring decisions actually operate, among shortlisted candidates who differ moderately on the trait.

The study's secondary result ran against intuition: the agreeableness interaction was negative. Managers lower on agreeableness performed better under high autonomy. Whatever the mechanism, it is a caution against assuming that every socially attractive trait gains predictive value with independence, and a reason to let measured relationships set the weights.

The general case for conscientiousness, its record across occupations and criteria, is made in the companion article on conscientiousness at work; this article needs only the moderation. If remote work's psychological signature is autonomy, the moderation applies: converting a role to remote steepens the conscientiousness slope, and a candidate pool that was roughly interchangeable on the trait under supervision stops being interchangeable without it. That is the sense in which hiring for remote work re-weights its criteria instead of replacing them.

Where the demands rise in virtual teams

Personality is one axis of remote employee selection; competencies are the other, and the evidence there comes from the people who staff both kinds of team. Krumm, Kanthak, Hartmann, and Hertel (2016) collected expert ratings of the knowledge, skills, abilities, and other characteristics required in virtual (n = 175) and traditional (n = 205) teams, organized against the Great Eight competency framework, a standard taxonomy of work competencies.

Two categories were rated significantly more important in virtual teams. Leading and Deciding covers self-directed action: initiating, deciding, and taking responsibility without waiting for direction. Analyzing and Interpreting covers working with information, and it is the category that contains written reporting, the channel a distributed team lives on. The claims hold at the category level and should be used at that level: heavier demand on the cluster, without a certified ranking of the individual skills inside it.

Expert requirement ratings are job analysis, not validation: they record what experienced observers say a team needs, and say nothing about how strongly any measure of that predicts performance. For this question, they are still the right tool. Validity coefficients generalize poorly to arrangements that barely existed when the classic estimates were made, while requirement judgments can be collected from people currently running the arrangement. Requirement ratings and the autonomy moderation check each other here: the moderation says the trait side re-weights, the expert comparison says which demand clusters get heavier, and both agree that the underlying list is stable.

The null result is as useful as the positive one. Broad Big Five requirements did not differ between virtual and traditional teams. The experts were not describing a different species of employee; they were describing the same employees under heavier demands in particular categories. Krumm and colleagues tie the findings explicitly to selection and training, and it pays to keep the division of labor in the evidence straight: the competency study says nothing about conscientiousness mattering more remotely. That claim rests entirely on the autonomy moderation.

Allen, Golden, and Shockley (2015), reviewing the scientific status of telecommuting research, add the constraint the enthusiasm needs: benefits are real but contingent on extent and design, and high-extent telecommuters risk professional isolation and reduced knowledge sharing. Selection can carry what the office supplied to the individual role; it cannot by itself replace what the office supplied to the network. Figure 3 maps the substitution the evidence does support — and the top row has the largest selection consequence: the structure an office imposes and the structure a person generates are substitutes, and only the second survives the move home.

Hiring for remote work starts with the arrangement, not the candidate

Everything above converges on a design sequence, and it starts before any candidate is assessed: specify the arrangement. A fully remote role, a hybrid role, and an office role that tolerates occasional home days are different measurement targets, and the 2.5-day boundary in the meta-analytic record says the differences are not cosmetic. An organization that writes remote-friendly in the advertisement while meaning three office days has not defined the job's autonomy level, and every downstream decision about hiring for remote work inherits the vagueness.

With the arrangement fixed, the personality half of remote work assessment becomes a weighting decision. No new instrument is required. The evidence locates the action in conscientiousness, and more precisely in the facets that generate structure from inside: self-discipline and achievement orientation. As the role's autonomy rises, those facets should carry more weight in the composite, while the trait list itself stays put, exactly as the expert-ratings evidence says it should. This is where facet-level measurement stops being a psychometric nicety: weighting a facet instead of a whole domain is the difference between adjusting the relevant dial and dragging the entire dashboard.

The competency half needs work-relevant measures aimed at the category level. Self-directed action can be exercised directly: structured tasks in which the candidate must scope, prioritize, and commit without being handed a plan. Analytic, written communication can be sampled: a short brief produced under realistic conditions says more about Analyzing and Interpreting than any self-description, and it samples the channel the team will actually run on. The same category logic extends to the structured interview: past-behavior questions about stretches of unsupervised work, written artifacts the candidate actually produced, deadlines met without an external cadence. What matters is that every element is scored against the arrangement the advertisement promised, so that the file reaching the decision reads on the job that exists. Both kinds of measure sit alongside personality in a configurable assessment platform, and the research base these design choices draw on is set out in the science behind the approach.

The management side of the design is what keeps selection from being asked to do too much. Build the cadence and visibility the office used to provide: written status norms, predictable checkpoints, deliberate exposure of remote work to the people who calibrate performance. The alternative, screening people for tolerance of monitoring software, mistakes surveillance for structure and selects for compliance the evidence never asked for. And the Ctrip promotion result should be read as a standing warning rather than a curiosity: conditional on measured output, promotion rates for home workers roughly halved. If visibility leaks into advancement while output holds constant, a performance system is drifting into a presence system, and calibration reviews need the arrangement itself on the table as a known source of bias.

Publish the arrangement, and assess for it

The implications sort by audience. For the organization, the retention findings are the operational headline: attrition halved at full remote, quits down by a third at hybrid, with no measured performance penalty. What early exits do to the economics of a hiring program is quantified in the companion article on predicting early attrition; the experiments here supply the other half of that argument, an arrangement lever that moves retention directly rather than predicting it. Even the choice between full remote and hybrid is evidence-informed: the 2.5-day boundary and the isolation risks Allen and colleagues review sit on one side, the attrition and quit results on the other, and the balance between them belongs to role design.

For the hiring system, the weight shifts up a level as well. When supervision happens across a screen, first-line manager selection becomes less forgiving: no one can manage by walking the floor, and the deliberate cadence that replaces the office is only as good as the person running it. The evidence on selecting first-line managers applies with extra force where the span of control is distributed.

For candidates, the fix is symmetry. Publish the actual arrangement in the advertisement, in days per week, and assess against it. A candidate who opts into a truthfully described high-autonomy role and then clears an assessment weighted for that autonomy has passed matched filters on both sides of the table. The Ctrip re-selection result, learning plus selection nearly doubling the gain, is an incumbent-side hint of what informed matching can add. It arrived after nine months on the job, so the hiring version has to be built deliberately, out of truthful role definitions and measures aimed at them.

A role converted from office-based to remote is, for selection purposes, a new role, and it deserves a new definition instead of a reprinted one. The requirements do not disappear along with the office. They move onto the candidate profile, and hiring for remote work is the act of measuring them there, where the building can no longer cover for what was never assessed.

Where 5Profiler stands

A fully remote role and its office-based twin can share every line of the job description and still make different demands of the person doing the work. 5Profiler treats them as different roles: role-referenced scoring carries the arrangement in the role profile itself, so the weight on self-discipline, achievement, and self-directed work rises with the autonomy the role actually grants, and a candidate read for the office version is not assumed identical for the remote one. Because personality is measured at facet level, the shift lands where the evidence puts it, on specific facets of conscientiousness, instead of pretending the whole trait list changed when the desk moved.

Read the science behind the platform · See it on your roles

References

  1. Allen, T. D., Golden, T. D., & Shockley, K. M. (2015). How effective is telecommuting? Assessing the status of our scientific findings. Psychological Science in the Public Interest, 16(2), 40–68.
  2. Barrick, M. R., & Mount, M. K. (1993). Autonomy as a moderator of the relationships between the Big Five personality dimensions and job performance. Journal of Applied Psychology, 78(1), 111–118.
  3. Bloom, N., Han, R., & Liang, J. (2024). Hybrid working from home improves retention without damaging performance. Nature, 630(8018), 920–925.
  4. Bloom, N., Liang, J., Roberts, J., & Ying, Z. J. (2015). Does working from home work? Evidence from a Chinese experiment. Quarterly Journal of Economics, 130(1), 165–218.
  5. Gajendran, R. S., & Harrison, D. A. (2007). The good, the bad, and the unknown about telecommuting: Meta-analysis of psychological mediators and individual consequences. Journal of Applied Psychology, 92(6), 1524–1541.
  6. Krumm, S., Kanthak, J., Hartmann, K., & Hertel, G. (2016). What does it take to be a virtual team player? The knowledge, skills, abilities, and other characteristics required in virtual teams. Human Performance, 29(2), 123–142.

© 2026 Future Proof. All rights reserved. 5Profiler™ and the 5Profiler bloom mark are trademarks of Future Proof.

See it on your roles

Evidence over intuition, on your next hire.

A 30-minute walkthrough of 5Profiler with your roles, not a canned deck — and a sample report to keep.