No consistent voice-gender effect on engagement or conversion
Adjusted for direction and portfolio, the two voices differ in engaged time by 0.1 seconds (95% CI −0.4 to +0.6). On promise to pay, the outcome a collections operation is actually run on, they differ by 0.1 percentage points (−0.2 to −0.0), a tenth of a point favoring the male voice, with an interval that reaches zero.
That pooled figure is the headline, not the finding. Forty-four thousand live conversations, voice randomly assigned within every strategy, inbound and outbound, two portfolios, one voice per gender, channel and book mix removed, put the average difference within a second and a fraction of a point of zero.
The pooled figure averages strata that disagree. The four direction × portfolio strata are heterogeneous on every outcome, and the inverse-variance weights are dominated by the two outbound strata, which hold 39,000 of the 44,000 conversations. The inbound Book C stratum, where the female voice holds the line roughly 37 seconds longer, carries almost no weight. The stratum-level results, not the pooled figure, are the substantive finding of this note.
A natural defense of the claim is that averages conceal what happens once a conversation gets going. On the pooled numbers, they do not. Within long conversations the female voice holds 87.9 seconds against 90.7 for the male voice, and promise to pay is nearly identical.
The female voice shows no consistent advantage on any outcome; where it leads in one stratum, it trails in another.
Adjusted difference, female voice minus male voice
Stratified by direction × portfolio, 95% CI. Zero means no difference.
Engaged time
+0.1s
95% CI −0.4 to +0.6
Promise to pay
−0.1pts
95% CI −0.2 to −0.0


