Andrew Critch introduces 'Schelling goodness' — a framework for thinking about moral coordination among diverse intelligent agents who share no history, where agents try to converge on moral verdicts using only common knowledge and civilizational survival pressures. This connects game theory, moral philosophy, and multi-agent coordination.
Category intelligence
Research Briefing — March 1, 2026
9 current items analyzed and ranked.
Executive synthesis
Research Summary
An unusually thin day for research output. The sole standout is Andrew Critch's (BERI/CHAI) conceptual framework of Schelling goodness — proposing shared morality as a coordination equilibrium among diverse intelligent agents, including AI systems, with direct relevance to multi-agent alignment.
- A conceptual model frames LLM behavior as navigation through semantic topology with attractor basins, offering intuition for prompt engineering and failure modes
- The "AI slop as vegan hamburger" analogy provides a useful model for why AI-generated content fails despite surface-level pattern-matching to human output
- Remaining items cover epistemics, rhetorical analysis, and non-AI topics (lithium/Alzheimer's, meditation) with minimal research substance
Overall, today's pool lacks empirical papers, benchmarks, or technical contributions; only the top three items offer frameworks with any bearing on AI research or practice.
Key Themes
Primary evidence
Top Ranked Signals
Presents an intuitive mental model for understanding LLM behavior as navigation through a semantic space with attractors (helpful responses, language matching) and repellers (refusals, safety boundaries), using topological metaphors to explain prompt engineering phenomena like jailbreaks and mode-switching.
Uses the analogy of vegan meat substitutes to argue that AI-generated content ('slop') pattern-matches to real content on superficial inspection but fails on deeper engagement, potentially relating this to Kolmogorov complexity differences between genuine and generated content.
Argues that when evaluating practical proposals (not abstract arguments), the credibility and track record of the proposer matters significantly. Draws on Dan Davies' 'One Minute MBA' framework to suggest that known fibbers' forecasts and proposals should be heavily discounted, even in communities that prize idea-over-identity evaluation.
Linkpost: "Lithium Prevents Alzheimer’s—Here’s How to Use It"
By Jackson Wagner
Linkpost summarizing evidence that very low-dose lithium supplementation (hundreds to thousands of times below psychiatric doses) may reduce Alzheimer's risk by 20-50%, supported by converging evidence from mouse models, observational studies, psychiatric patient data, and mechanistic research.
Describes a rhetorical manipulation technique where a speaker substitutes a secondary aspect of a concept for the whole, builds layers of theory on the altered definition ('buffer overflow'), then uses the conclusions to steer audiences while retaining the emotional weight of the original concept.
Explores the concept of 'mindscapes' — how people visualize the organization of ideas, memories, and emotions in their minds — through informal surveys. Finds most people's mindscapes are disorganized, physical, and primarily visual, and raises concerns about compartmentalization.
Personal reflection on spending a year practicing jhana meditation without achieving any of the eight jhana states. Discusses the technique, the challenges of generating positive feelings as prerequisites, and reflections on the practice.
A very brief question post asking what mental tools rationalists could use to perform exceptionally well at novel tasks, using beating a difficult video game on a first try as an example.