Does humanist AI doctrine actually protect or constrain real users?
Rao questions whether AI safety frameworks claiming to prioritize human flourishing instead impose paternalistic constraints based on idealized values rather than users' actual choices and needs.
Venkatesh Rao's essay "Humanism without (All) Humans" argues that AI-safety doctrines claiming to put humans first are often a species of ideology he calls "humanism without (all) humans": positions whose "defining move is not... centering and caring about humans... It is constructing an idealized consensus Human whose essential interests can then be invoked against the expressed interests, experiments and values of inconvenient real humans." He names Microsoft's humanist-AI framing as the latest instance, quoting its premise that "people matter more than AI" and that AI "should not be designed to be a person," and its public line that AI must be "subordinate, aligned, and contained."
His reasoning is that once a doctrine defines flourishing — Microsoft's list includes "autonomy, agency, capability development, health, relationships and meaningful participation" — those values stop being descriptions of what humans want and become "objectives of the artificial system and therefore constraints upon what users may legitimately delegate to it." He calls this "nudge theory with a supercomputer behind it": the designer reserves "the right to decide when a human is exercising autonomy incorrectly." Against this he poses "contractual agency rather than paternalistic agency," arguing flourishing should emerge from pluralistic relations among agents rather than a platform's theory of correct delegation, and that the real alternative to uncontrolled AI is not human-controlled AI but "well-governed human-AI systems" permitting more autonomy as capability grows, since "autonomous machines are more demanding of their operator than non-autonomous machines."
This argument runs opposite to the control-first premise underlying What evidence would justify training increasingly powerful AI systems? and Does AI risk increase with the autonomy we give it?: where those sources treat ceded autonomy itself as the risk variable, Rao treats subordination-by-design as the harm, arguing paternalistic control forecloses an "enormous design space" of autonomous economic and institutional agency. It also complicates Does granting agents more autonomy undermine human oversight?: Rao would likely grant the empirical erosion claim but reject human "legibility" as the right standard for oversight at all, since "nobody understands modern civilization in that sense" either. His insistence on exit, schism and "dissensus" as preconditions for real pluralism extends Can AI systems preserve moral value conflicts instead of averaging them? from a modeling claim about value conflict to an institutional one: pluralism requires architecture that allows forking, not just values held in tension within a single system.
The essay is a position piece, not a study — it cites no data on whether Microsoft's constraints actually degrade user autonomy in practice, and no mechanism by which "contractual agency" or staking-based "artificial fiduciaries" would be built or governed; the symbiotic alternative is sketched as a design space, not a working system. If Rao's diagnosis holds, it implies that safety doctrines framed as human-centered should be read first for whose account of human flourishing they encode, since that account is what ends up constraining real users' choices.
Inquiring lines that read this note 3
This note is a source for these research framings, grouped by the broader line of inquiry each explores. Scan the bold lines of inquiry; follow any specific question forward.
How do philosophical assumptions about AI consciousness affect practical harms and design? What governance mechanisms can effectively constrain widely deployed AI systems?Related concepts in this collection 4
This note in its neighbourhood — explore the map, then jump to a related concept in the list below.
Click a node to walk · click center to open · click Open in graph to see this note in the full knowledge graph
-
What evidence would justify training increasingly powerful AI systems?
Altman proposes that AI model training should require an 'extremely strong case' for human control before proceeding, regardless of estimated catastrophe risk levels. The note explores what such a case would need to include and how it would be evaluated.
Rao's subordination critique targets the same control-first premise Altman's case-for-control standard assumes
-
Does AI risk increase with the autonomy we give it?
Explores whether the risks posed by AI agents scale monotonically with the level of autonomy they're granted, and what the tradeoffs are between human control and agent independence.
Argues the opposite: risk comes from subordination-by-design, not autonomy ceded to the agent
-
Does granting agents more autonomy undermine human oversight?
Explores whether the design of autonomous AI systems—by giving agents greater independence—actually weakens the human overseer's ability to catch problems. Matters because oversight is a key safeguard against AI failures.
Grants the oversight-erosion claim but rejects human legibility as oversight's correct standard
-
Can AI systems preserve moral value conflicts instead of averaging them?
Current AI systems wash out value tensions through majority aggregation. Can we instead model how values like honesty and friendship genuinely conflict in moral reasoning?
Extends values-in-tension into an institutional claim: pluralism needs exit and forking, not just modeling
Related papers in this collection 8
Papers most semantically related to this note, ranked by cosine similarity in the embedding space.
- Humanism without (All) Humans
- Climbing towards NLU: On Meaning, Form, and Understanding in the Age of Data
- The Veto Variable: Human Override as a Goal-Independent Cost Term
- Position: Towards Bidirectional Human-AI Alignment
- Beyond Preferences in AI Alignment
- Value Kaleidoscope: Engaging AI with Pluralistic Human Values, Rights, and Duties
- Humanity's Last Invention: Richard Socher of Recursive
- Epistemic Deference to AI
Original note title
Rao argues Microsoft's humanist AI doctrine constructs an idealized consensus human to override real humans' own choices