SYNTHESIS NOTE
Topics›Conversation Topics Dialog›this note

Can LLMs truly update shared conversational common ground?

Explores whether large language models can participate symmetrically in Stalnaker's picture of communication, where speakers mutually revise shared assumptions. The question matters because it reveals whether human-LLM dialogue is genuinely interactive or structurally asymmetrical.

Synthesis note · 2026-05-01 · sourced from Conversation Topics Dialog

On Stalnaker's picture, communication is a process of mutually proposing and accepting updates to shared assumptions. Each assertion is a candidate for incorporation into common ground; participants accept, query, or reject. The common ground evolves as conversation proceeds, and that evolution is itself the substance of communication.

LLMs cannot participate in this process symmetrically. The prompt establishes the model's working context, and the model interprets subsequent turns within that frame. Even when a user pivots — shifting from climate policy to historical precedent, or revealing they are not actually a five-year-old after asking for a five-year-old explanation — the LLM cannot smoothly absorb the revision into a jointly held common ground. It either ignores the pivot, fabricates continuity, or requires the user to re-scaffold from scratch. The asymmetry is structural: humans propose, the LLM either adopts or routes around, but the LLM cannot itself propose updates that change what counts as background.

This is a deeper deficit than failures of memory or inference. It means that the conversational scoreboard — Lewis's mechanism for tracking what counts as a felicitous next move — is one-sidedly maintained by the user. The user is keeping score for both players. The model is producing moves that look responsive but cannot reciprocally update the score in the way the conversational practice requires. What looks like dialogue is structurally closer to oracle-consultation, where the questioner provides all context and the oracle returns a response framed within it.

Inquiring lines that read this note 110

This note is a source for these research framings, grouped by the broader line of inquiry each explores. Scan the bold lines of inquiry; follow any specific question forward.

What enables conversational agents to guide rather than just respond? Can AI systems participate in genuine communication or only simulate it? What structural patterns sustain successful multi-turn dialogue and prevent breakdown? Do language models reason through disagreement or only accommodate it? Can LLMs distinguish between linguistic form and semantic meaning? Is embodied interaction necessary for language meaning and agency? What distinguishes genuine communicative competence from surface language performance? How do curriculum design and feedback approaches affect model learning? Can language models reason beyond surface pattern matching? Can real-time working alliance measurement improve therapy outcomes? Can language models reliably simulate personas and predict behavior? How do interpretive frames override surface features in text comprehension? What limits language model accuracy in evaluating ideas? Why do planning and grounding require opposing optimization strategies? What causes coordination failures in multi-agent language model systems? How susceptible are language models to conversational persuasion and belief change? Why do language models fail at sustained therapeutic relationships despite understanding techniques? Do language models encode knowledge that influences generation, or primarily imitate surface patterns? Does preference optimization undermine conversational grounding in language models? What design features sustain romantic bonds with AI companion systems? What prevents LLMs from applying their reasoning knowledge to improve outputs? When do multi-agent systems improve over single frontier models? Does augmenting symbolic reasoning improve LLM logical reasoning ability? What capabilities differentiate diffusion from autoregressive language models?

Related papers in this collection 8

Papers most semantically related to this note, ranked by cosine similarity in the embedding space.

Original note title

Common ground in human-LLM conversation cannot be jointly updated because the LLM treats prompts as static frames