Line of inquiry
Inquiring lines›What determines reliable reasoning…›What determines LLM output consist…›this line of inquiry
Why do language models fail at sustained therapeutic relationships despite understanding techniques?
A broader line of inquiry — a family of 30 specific questions the research asks around this. Follow one into its inquiring-line page, or move sideways to a related line below.
Questions in this line of inquiry 30
Specific inquiring lines the field asks around this — ordered from the most general framing down to the most specific angle.
- Does the passivity problem in LLMs compound misalignment in therapeutic contexts?
- Why can't language models conduct genuine Socratic questioning in therapy sessions?
- Why do single-turn LLM responses outperform humans while ongoing relationships show limits?
- Why do LLMs understand therapy techniques but fail to execute them?
- How do language models interpolate user feelings in therapeutic contexts?
- Can language models implement therapeutic skills like Socratic questioning in real conversations?
- Can embodied agents overcome the LLM skill gap in therapy outcomes?
- Can alternative reward functions shift LLMs from problem-solving to genuinely empathic responses?
- Do problem-solving defaults in LLM therapists actually undermine therapeutic effectiveness?
- Why do LLMs solve problems when clients need emotional reflection instead?
- Can affective framing reliably improve language model outputs?
- Can large language models actually deliver cognitive behavioral therapy techniques?
- Does prompting or added context help LLMs understand therapeutic timing and depth?
- How do LLMs mirror the same alliance failures as human counselors?
- Can models succeed at mental health tasks without integrating multiple psychological traditions?
- Do later-phase mental health LLM systems outperform earlier phase approaches clinically?
- Why do Llama models struggle with cognitively distorted user expressions in therapy?
- Do LLMs show stigma or reinforce delusions in mental health contexts?
- Why do LLMs systematically fail at information management in social interaction?
- What makes emotional alignment more effective than logic when reasoning errors are exposed?
- Why do LLMs reflect on client needs more than typical low-quality human therapists?
- Why do models struggle with asking questions in multi-turn conversational reasoning tasks?
- Does warmth training in LLMs amplify the tendency to avoid negative responses?
- What training data barriers prevent LLMs from learning real Socratic dialogue?
- Can LLM therapists develop character knowledge to decide when advice-giving fits?
- Why do positive emotional words contribute disproportionately to prompt enhancement effects?
- What foundational barriers prevent LLMs from achieving clinical validity in therapy?
- What interaction design changes would help LLMs handle underspecified requests?
- What makes Beck's diagram effective for constraining simulated patient behavior?
- Why do Llama-based models outperform GPT-4 in objective clinical guidance?