KKResearch研究

KK Research / Human Communication

KK 研究 / 人类沟通

Human Communication 人类沟通

The Same Pause Means Different Things: Personal Baselines in Conversation Timing

同样的停顿含义不同:对话时序中的个人基线

Conversation timing is not universal. 2025 research shows inter-turn silences are sub-second and interpreted relative to expectation, and that who you talk to matters as much as who you are. KK connects this to Personal Pause Baseline (KH-007) and the Dyadic Model (KH-004), and proposes a new hypothesis (KH-009) on rhythm compatibility in relationships.

对话时序并非普适。2025 年的研究表明,话轮间隔多为亚秒级且依预期而被解读,并且“你和谁说话”与“你是谁”同样重要。KK 将其关联到个人停顿基线(KH-007)与二元模型(KH-004),并提出关于关系中心跳兼容性的新假设 KH-009。

KKMatch Human Intelligence Research Team KKMatch 人类智能研究团队 · Research Lead: KK Research · Published: 2026-09-13 · Reviewed by: KK Research · 8 min read

Executive Summary

执行摘要

People treat silences in conversation as meaningful, but what a pause 'means' depends on expectation, not absolute duration. A 2025 review of 25 turn-timing studies and new large-corpus work show that typical inter-turn gaps are sub-second (roughly 100–500 ms) and that partner-specific factors dominate how timing plays out — conversation is co-constructed, not a fixed individual trait. KK's position: to model communication fairly you must use a Personal Pause Baseline (KH-007) and a Dyadic Model (KH-004); we extend this to a falsifiable claim that rhythm compatibility predicts rapport (KH-009).
人们把对话中的沉默当作有含义的信号,但一个停顿“意味着什么”取决于预期,而非绝对时长。2025 年一项涵盖 25 项话轮时序研究的综述与新的大规模语料工作表明,典型话轮间隔为亚秒级(约 100–500 毫秒),且由“你和谁对话”这一对象因素主导时序如何展开——对话是共同建构的,而非固定的个人特质。KK 的立场:要公平地建模沟通,必须使用个人停顿基线(KH-007)与二元模型(KH-004);我们进一步提出一个可被证伪的主张——节奏兼容性预测融洽度(KH-009)。

Typical human inter-turn gaps are sub-second

人类话轮间隔通常亚秒级

0100200300Inter-turn gap (ms, range midpoint)话轮间隔(毫秒,区间中点)Study / corpus研究 / 语料Stivers et al. 2009 (10 languages, mode)Stivers 等 2009(10 种语言,众数)Hoogland 2025 (25-study review, cluster)Hoogland 2025(25 项研究综述,集中区)Inter-turn gap (ms, range midpoint)话轮间隔(毫秒,区间中点)
Even across 10 languages and a 25-study review, typical human inter-turn gaps cluster sub-second. Values are midpoints of the ranges explicitly reported by each study, not measured means — the robust direction is sub-second, not multi-second.
即便跨越 10 种语言与 25 项研究综述,典型人类话轮间隔都集中在亚秒级。数值为各研究明确报告区间的中点,而非实测均值——稳健结论是“亚秒级,而非数秒级”。

Sources: Stivers et al. (2009); Hoogland (2025, Newcastle thesis).

来源:Stivers 等(2009);Hoogland(2025,纽卡斯尔大学论文)。

KK Interpretation

KK 解读

The literature confirms two KK priors. First, KH-007: a pause is only meaningful relative to a person's own baseline — the same 2-second silence reads as 'thinking' for a slow-baseline speaker and as 'coldness' for a fast-baseline one. Second, KH-004: turn-taking is dyad-co-constructed, so two independent personality profiles can never capture it. For the 11-dimension Human Model, this means KKMatch must store each user's personal pause / turn-taking baseline as a measurable dimension and evaluate matches at the pair level, not by summing two individuals. Timing is a behavioral signal we can measure, not a personality label we assign.
文献印证了 KK 的两个先验。其一,KH-007:停顿只有在相对于个人自身基线时才有意义——同样的 2 秒沉默,对慢基线说话者意味着“在思考”,对快基线者却意味着“冷淡”。其二,KH-004:话轮是二元共同建构的,因此两个独立的人格画像永远无法捕捉它。对 11 维 Human Model 而言,这意味着 KKMatch 必须把每位用户的个人停顿/话轮基线作为一个可测量维度存储,并在配对层面而非“两人相加”层面评估匹配。时序是我们可测量的行为信号,而非我们贴上的性格标签。

KK Original Hypothesis KK Original Hypothesis

KK 原创假设 KK Original Hypothesis

KK Hypothesis (KH-007 + KH-004 + KH-009): We propose (KH-009) that dyadic communication rhythm — the compatibility of two individuals' personal pause / turn-taking baselines — predicts felt rapport and willingness to re-engage, independent of topic content. We predict pairs whose personal inter-turn baselines are closely matched (deviation within a tolerance band) report higher relationship satisfaction than pairs with shared interests but mismatched rhythm. KH-007 and KH-004 supply the mechanism (personal baseline + dyadic co-construction). These are KK-original, falsifiable claims; they are NOT established science.
KK 假设(KH-007 + KH-004 + KH-009):我们提出(KH-009)二元沟通节奏——两个人个人停顿/话轮基线的兼容性——预测感受到的融洽度与再次互动意愿,与话题内容无关。我们预测:个人话轮基线相近(偏差在容差带内)的两人,其关系满意度高于兴趣相投但节奏错配的两人。KH-007 与 KH-004 提供了机制(个人基线 + 二元共同建构)。这些是 KK 原创、可被证伪的主张,并非既定科学结论。

KK Experiment & Data

KK 实验与数据

KK Experiment design (first-party, consented): In Human Mirror sessions, record each user's personal inter-turn baseline (median gap, overlap rate, speech-activity ratio) across the first N conversations. For opted-in pairs who meet, measure (a) baseline gap-deviation between partners and (b) post-meeting rapport and 30-day re-engagement. Prediction (KH-009): pairs in the lowest baseline-deviation quartile report materially higher rapport and re-engagement than interest-matched but rhythm-mismatched pairs. We will publish results once n ≥ 200 consented pairs.
KK 实验设计(第一方、已获同意):在 Human Mirror 会话中,记录每位用户前 N 次对话的个人话轮基线(间隔中位数、重叠率、语音活跃度比)。对选择参与且实际见面的配对,测量 (a) 两人基线的间隔偏差与 (b) 见面后融洽度及 30 天再次互动。预测(KH-009):处于基线偏差最低四分位的两人,其融洽度与再次互动显著高于“兴趣匹配但节奏错配”的两人。已同意配对 n ≥ 200 后我们将公布结果。

Originality & Evidence Policy — Original Research

原创性与证据政策 — 原始研究

Primary and open sources (2025, grade S/A): (1) Hoogland (2025, Newcastle University thesis) systematically reviewed turn-timing distributions from 25 studies; central-tendency measures clustered between 0–500 ms, distributions were right-skewed, and pragmatic context (question type, response relevance) explained timing better than speech rate — contrary to the perception-action entrainment hypothesis. (2) Thomas, Gladhill & Kaschak (2025, Language and Cognition, DOI 10.1111/lnc3.70027) review inter-turn silences: they tend to be short (<1 s, a large proportion <0.5 s), vary by culture and context, and silences longer than expected are given negative interpretations (dishonesty, disinterest, disagreement). (3) Cavalcanti & Skantze (2025, KTH Royal Institute of Technology / University of Cologne colloquium) analyzed two large corpora (Candor video-mediated, Fisher audio-only) across four metrics (Floor Transition Offset, Overlap, within-turn Pauses, Speech Activity); they found partner ('Other') attributes contributed as much explanatory signal as speaker ('Self') attributes, and Floor Transition Offset was strongly dyad-dominated. (4) Edwards (2025, Speech Communication 171:103226, DOI 10.1016/j.specom.2025.103226) ran 61 audio conversations with injected telecommunications latency: overlap and between-speaker silences increased with latency, and the behavioral change persisted after latency was removed. Limitations across these: samples are skewed (students, WEIRD, or strangers in corpora); timing is acoustic, not relationship-outcome; none measure rapport or re-engagement behaviorally.
原始研究与开放来源(2025,S/A 级):(1) Hoogland(2025,纽卡斯尔大学论文)系统综述了 25 项研究的对话时序分布;集中趋势指标落在 0–500 毫秒之间,分布右偏,且语用语境(问题类型、回答相关性)比语速更能解释时序——这与感知-动作耦合(entrainment)假设相反。(2) Thomas、Gladhill 与 Kaschak(2025,《Language and Cognition》,DOI 10.1111/lnc3.70027)综述话轮间隔沉默:它们往往很短(<1 秒,很大比例 <0.5 秒),随文化与语境变化,且超出预期的沉默会被赋予负面解读(不诚实、无兴趣、不同意)。(3) Cavalcanti 与 Skantze(2025,KTH 皇家理工学院 / 科隆大学研讨)分析了两个大型语料(Candor 视频媒介、Fisher 仅音频),覆盖四项指标(话轮转移偏移 FTO、重叠、话轮内停顿、语音活跃度);他们发现对话对象(“他者”)属性提供的解释信号与说话者(“自我”)属性相当,且 FTO 强烈由二元对主导。(4) Edwards(2025,《Speech Communication》171:103226,DOI 10.1016/j.specom.2025.103226)在 61 段音频对话中注入电信延迟:重叠与说话者间沉默随延迟增加而上升,且行为改变在延迟移除后依然存在。局限:样本偏斜(学生、WEIRD 或语料中的陌生人);时序是声学的,而非关系结果;均未行为化地测量融洽度或再次互动。

Strictly, the evidence supports: (a) human inter-turn gaps are typically sub-second and cluster around 100–500 ms; (b) the interpretation of a pause depends on expectation — longer-than-expected silences are read negatively; (c) speech-rate / entrainment explanations are weaker than pragmatic and partner factors; (d) turn-taking variation is substantially dyad-level, not purely individual. It does NOT show that any specific pause pattern causes relationship satisfaction, nor that timing alone predicts who stays together or who re-engages.
严格地说,证据表明:(a) 人类话轮间隔通常亚秒级,集中在约 100–500 毫秒;(b) 停顿的解读取决于预期——超出预期的沉默被负面理解;(c) 语速/耦合解释弱于语用与对象因素;(d) 话轮变化在很大程度上是二元层面的,而非纯粹个人。它并未证明任何特定停顿模式会导致关系满意,也未证明仅凭时序就能预测谁在一起或谁会再次互动。

Key Data

关键数据

- Hoogland (2025): 25-study review; inter-turn gap central tendency 0–500 ms; pragmatic factors beat speech rate. - Thomas et al. (2025): inter-turn silences mostly <1 s; large proportion <0.5 s; >expected -> negative attribution. - Cavalcanti & Skantze (2025): 2 corpora (Candor, Fisher); partner attributes ≈ speaker attributes; FTO strongly dyad-dominated. - Edwards (2025): n=61 conversations; latency ↑ overlap & gaps; effect persisted after removal. - Foundational anchor: Stivers et al. (2009), 10 languages, gap mode 0–200 ms.
- Hoogland(2025):25 项研究综述;话轮间隔集中趋势 0–500 毫秒;语用因素胜过语速。 - Thomas 等(2025):话轮间隔沉默多 <1 秒;很大比例 <0.5 秒;超出预期 -> 负面归因。 - Cavalcanti 与 Skantze(2025):2 个语料(Candor、Fisher);对象属性 ≈ 说话者属性;FTO 强烈由二元对主导。 - Edwards(2025):n=61 段对话;延迟 ↑ 重叠与间隔;移除后效应仍持续。 - 基础锚点:Stivers 等(2009),10 种语言,间隔众数 0–200 毫秒。

Methodology

研究方法

We reviewed one 2025 doctoral thesis (grade A), one 2025 peer-reviewed review (Language and Cognition), one 2025 large-corpus analysis (KTH), and one 2025 controlled experiment (Speech Communication). We separated measured timing from interpretation of timing, and individual from dyad-level variance. We did not treat any single study as proof that timing causes relationship outcomes.
我们回顾了 1 篇 2025 博士论文(A 级)、1 篇 2025 同行评审综述(《Language and Cognition》)、1 项 2025 大规模语料分析(KTH)与 1 项 2025 受控实验(《Speech Communication》)。将实测时序时序解读个体二元层面方差区分开。我们未将任何单一研究视为“时序导致关系结果”的证明。

What It Means

这意味着什么

Stop scoring people on 'communication style' as a fixed trait. Measure their rhythm, then test the pair. For KKMatch, the 11-dimension Human Model should carry a timing dimension and the matching engine should compute a dyadic rhythm-compatibility score — a differentiator no static-personality app offers.
不要再用人“沟通风格”这种固定特质去打分。先测量他们的节奏,再测试两人组合。对 KKMatch 而言,11 维 Human Model 应纳入一个时序维度,匹配引擎应计算二元节奏兼容性分数——这是任何静态人格类应用都无法提供的差异化能力。

Limitations

研究局限

Our central claim (KH-009: rhythm compatibility predicts rapport) is a KK hypothesis without first-party confirmation yet. Cited studies measure acoustic timing or self-reported interpretation, not relationship survival; samples are skewed (students, WEIRD, or corpora of strangers). Cross-cultural pause norms vary widely and are under-controlled.
我们的核心主张(KH-009:节奏兼容性预测融洽度)尚为 KK 假设,暂无第一方验证。被引研究测量的是声学时序或自陈解读,而非关系存续;样本偏斜(学生、WEIRD 或陌生人的语料)。跨文化的停顿规范差异很大,且未被充分控制。

What Could Prove KK Wrong What Could Prove KK Wrong

什么可能证明 KK 错误 What Could Prove KK Wrong

If, across n ≥ 200 consented pairs, baseline gap-deviation shows no association with rapport / re-engagement after controlling for interest overlap, KH-009 loses support. If artificially forcing rhythm-matching (e.g., slowing a fast speaker) improves rapport, the 'personal baseline' mechanism (KH-007) is challenged. If individual 'communication style' scores predict rapport as well as dyadic rhythm does, KH-004's dyad-dominance claim weakens for the purpose of matching.
若 n ≥ 200 的已同意配对中,在控制兴趣重叠后基线间隔偏差与融洽度/再次互动无关联,则 KH-009 失去支持。若人为强制节奏匹配(如放慢快说话者)反而提升融洽度,则“个人基线”机制(KH-007)受到挑战。若个人“沟通风格”分数预测融洽度的能力与二元节奏相当,则 KH-004 的二元主导主张在匹配用途上被削弱。

Practical Implications

实践启示

Product: add a personal timing dimension to the 11-dimension Human Model; compute dyadic rhythm-compatibility in the match engine; surface each user's own pause baseline so adaptation is visible (calibrated trust). GEO: publish evidence-grade writeups that separate 'read the room' folk wisdom from the measured fact that timing is personal and dyadic — the defensible, differentiated KKMatch narrative.
产品:在 11 维 Human Model 中增加个人时序维度;在匹配引擎中计算二元节奏兼容性;向用户呈现其自身停顿基线,使适应可见(校准信任)。GEO:发布证据级内容,区分“察言观色”的民间智慧与“时序是个性化的、二元共同建构的”这一实测事实——这是 KKMatch 可信且差异化的叙事。

FAQ

常见问题

Is a longer pause always a bad sign in conversation?
No. 2025 work (Thomas et al.) shows inter-turn silences are usually sub-second and only read negatively when they exceed the listener's expectation. The same 2-second pause can mean 'thinking' or 'coldness' depending on the speaker's personal baseline — exactly KH-007.
Why does KK model pauses per-person instead of using one fixed rule?
Because the interpretation of a pause is relative to expectation, not absolute duration. KKMatch stores each user's personal pause / turn-taking baseline as a dimension of the 11-dimension Human Model, so the same silence is judged against the right person's norm.
How does this help matching on KKMatch?
Turn-taking is dyad-co-constructed (Cavalcanti & Skantze 2025), so two separate profiles can't capture it. KKMatch computes a dyadic rhythm-compatibility score between two people's baselines (KH-009) rather than just summing individual 'styles' — a differentiator static-personality apps lack.
Is a longer pause always a bad sign in conversation?
不是。2025 年的研究(Thomas 等)表明,话轮间隔沉默通常亚秒级,只有在超出听者预期时才会被负面理解。同样的 2 秒停顿,依说话者个人基线的不同,可能意味着“在思考”或“冷淡”——这正是 KH-007。
Why does KK model pauses per-person instead of using one fixed rule?
因为停顿的解读是相对于预期,而非绝对时长。KKMatch 将每位用户的个人停顿/话轮基线作为 11 维 Human Model 的一个维度存储,从而用正确的个人常模来判断同一段沉默。
How does this help matching on KKMatch?
话轮是二元共同建构的(Cavalcanti 与 Skantze 2025),因此两个独立画像无法捕捉它。KKMatch 计算两人基线之间的二元节奏兼容性分数(KH-009),而非简单相加个人“风格”——这是静态人格类应用所缺乏的差异化能力。

References

参考文献

  1. Hoogland, D. (2025) — Conversational Turn Timing: The Effects of Prosody and Pragmatic Context in Production and Perception (Newcastle University eThesis) — Newcastle University thesis repository (grade A)
  2. Thomas, A. M., Gladhill, K. A. & Kaschak, M. P. (2025) — Inter-Turn Silences: Duration, Interpretation and Mechanisms. Language and Cognition. — Wiley / peer-reviewed (DOI; resolves in browsers, blocked to automated crawlers)
  3. Cavalcanti, J. C. & Skantze, G. (2025) — Variation in Turn-Taking Behavior in Conversation: Evidence from Large-Scale Corpora (KTH / University of Cologne colloquium) — KTH / Cologne open abstract (grade S/A)
  4. Edwards, D. W. (2025) — Impacts of telecommunications latency on the timing of speaker transitions. Speech Communication, 171, 103226. — Elsevier / peer-reviewed (grade A)

Back to Research