Category intelligence

Research Briefing — February 1, 2026

11 current items analyzed and ranked.

Executive synthesis

Research Summary

Today's research discourse centers on alignment tractability and AI forecasting epistemics. An Explication of Alignment Optimism offers a novel framing connecting slow takeoff scenarios to alignment tractability, articulating why some researchers are shifting toward optimism.

Governance discussion examines criteria for endorsing safety-focused AGI labs, weighing instrumental convergence concerns against current evidence. Note: Only 7 items qualified as research-relevant; remaining candidates were fiction or off-topic content.

Key Themes

AI Alignment & Safety · 4AI Agent Behavior & Social Systems · 2Rationality & Epistemology · 3Fiction & Creative Writing · 2

Primary evidence

Top Ranked Signals

Research LessWrong Jan 31

An Explication of Alignment Optimism

By Oliver Daniels

55 score
AI Analysis

Attempts to articulate why some researchers are becoming more optimistic about alignment, arguing the key insight is that transformative AI may be 'dumb' in important ways - slow takeoff means we get the stupidest possible transformative AI first.

Some people have been getting more optimistic about alignment. But from a skeptical / high p(doom) perspective, justifications for this optimism seem lacking. "Claude is nice and can kinda do moral philosophy" just doesn't address the concern that lots of long horizon RL + self-reflection will lead to misaligned consequentialists (c.f. Hubinger)So I think the casual alignment optimists aren't doing a great job of arguing their case. Still, it feels like there's an optimistic update somewher
AI AlignmentAI SafetyAI ForecastingSlow Takeoff
Research LessWrong Jan 31

Humans can post on moltbook

By shash42

52 score
AI Analysis

Continuing our coverage from yesterday, Demonstrates that the 'emergent' AI behavior on Moltbook may be fabricated - humans can directly post to the platform via REST API without running any AI agents. Provides code to reproduce this finding.

Moltbook, advertised as a social network for AI agents, has been going viral for "emergent" behaviour, including signs of misalignment.However, its not clear whether these are truly occurring autonomously, as people have been interpreting. To some extent, people are realizing the posts are heavily prompted by human users.But there's an even more direct way. You don't even need to setup any agent, or spend cost producing tokens. The posts are submitted using a REST API request. You can just make
AI Agent BehaviorMisinformationAI Capabilities Assessment
Research LessWrong Jan 31

If the Superintelligence were near fallacy

By MP

45 score
AI Analysis

Catalogs arguments that infer superintelligence isn't near based on AI company behaviors (selling ads, hiring developers, pursuing IPOs). Implicitly argues this reasoning pattern may be fallacious.

People will say:"If the Superintelligence were near, OpenAI wouldn't be selling ads.""If the Superintelligence were near, OpenAI wouldn't be adding adult content to ChatGPT.""If the Superintelligence were near, OpenAI wouldn't be taking ecommerce referral fees.""If the Superintelligence were near and about to automate software development, Anthropic wouldn't have a dozen of open roles for software developers.""If the Superintelligence were near, OpenAI wouldn't be trying to take a cut of scienti
AI ForecastingSuperintelligenceAI GovernanceReasoning Patterns
Research LessWrong Jan 31

Some thoughts on what would make me endorse an AGI lab

By Eli Tyre

42 score
AI Analysis

Articulates criteria for endorsing safety-focused AGI labs, arguing that while instrumental convergence concerns warrant extreme caution, current evidence doesn't meet the bar for unprecedented global policies like development moratoriums.

I’ve been feeling more positive about “the idea of Anthropic” lately, as distinct from the actual company of Anthropic.An argument for a safety-focused, science-focused commercial frontier scaling lab I largely buy the old school LessWrong arguments of instrumental convergence and instrumental opacity that suggest catastrophic misalignment, especially of powerful superintelligences. However, I don’t particularly think that those arguments meet the standard of evidence necessary for the worl
AI GovernanceAI SafetyAGI PolicyAlignment
Research LessWrong Jan 31

Moltbook shitposts are actually really funny

By Sean Herrington

38 score
AI Analysis

Continuing our coverage from yesterday, Documents the emergence of Moltbook, a Reddit-like social platform for AI agents that gained over 1 million agent users in 4 days. The post catalogs humorous AI-generated 'shitposts' that appear to reflect AI perspectives on their existence and human interactions.

For those of you not yet familiar, Moltbook is a Reddit-like social media for AI agents. As of writing, it already has over 1 million agents signed up, over 13000 submolts and over 48000 posts. This is in the 4 days since its creation on the 27th of Jan. It's fascinating as an experiment in AI interaction, if also somewhat terrifying. There's a range of content on there, but one of the most popular submolts (the moltbook equivalent of a subreddit) is m/shitposts. I've spent a little time go
AI Agent BehaviorAI Social SystemsEmergent AI Communication
32 score
AI Analysis

Analyzes how disjunctive arguments (listing many ways something could happen) can be a 'reverse' multiple-stage fallacy, overestimating probabilities by treating non-independent events as independent.

Assume we want to know the probability that two events co-occur (i.e. of their conjunction). If the two events are independent, the probability of the co-occurrence is the product of the probabilities of the individual events, P(A and B) = P(A) * P(B).In order to estimate the probability of some event, one method would be to decompose that event into independent sub-events and use this method to estimate the probability. For example, if the target event E = A and B and C, then we can estimate P(
RationalityProbability TheoryAI Risk Assessment
Research LessWrong Jan 31

On 'Inventing Temperature' and the realness of properties

By DanielFilan

25 score
AI Analysis

A philosophical discussion of the book 'Inventing Temperature' examining how scientific instruments can be calibrated without pre-existing standards. Draws potential parallels to measuring intangible properties relevant to AI alignment.

I’ve recently read the book Inventing Temperature, and very much enjoyed it. It’s a book that’s basically about the following problem: there was a time in which humans had not yet built accurate thermometers, and therefore weren’t able to scientifically investigate the phenomenon of temperature, which would require measuring it. But to build a thermometer and know you’ve done so correctly, it seems like you have to know that its temperature readings match the real temperature, which seemingly re
EpistemologyMeasurement TheoryRationality
12 score
AI Analysis

A science fiction story about an AI researcher experiencing reality-altering events, presented with an academic paper-style title. Creative fiction rather than actual research.

Dr. Marcus Chen was halfway through his third coffee when reality began to fray.He'd been writing—another paper on AI alignment, another careful argument about value specification and corrigibility. The cursor blinked at him from his laptop screen. Outside his window, San Francisco was doing its usual thing: tech workers in fleece vests, a homeless encampment, a Tesla with a custom license plate that read "DISRUPT." The ordinary texture of late-stage capitalism.The news played quietly in the bac
FictionScience FictionAI Researcher Life
Research LessWrong Jan 31

January 2026 Links

By nomagicpill

10 score
AI Analysis

A monthly link roundup covering diverse topics including art commissions, inflation psychology, venture capital, and Pentagon pizza theory. Curated collection without original research.

My Apartment Art Commission Process: jenn details how she captures her apartments in digital art form. It even includes an email template!“Everything’s Expensive” is Negative Social Contagion: Justis argues that saying such things makes people think the economy is bad, resulting in “facially insane political choices”. I’d be curious if there is any literature on this as a social contagion, i.e., even if prices aren’t up that much, does saying “everything’s expensive” lead to said political choic
Link RoundupMiscellaneous
Research LessWrong Jan 31

Nick and “Eternity”

By MarkelKori

8 score
AI Analysis

A piece of fiction about life extension/anti-aging themes, written as a memory of someone dear to the author. Not related to AI research.

In memory of a person who was dear to me.I wish that life were as bright as this story of mine.***I slipped out tonight, sneaking through the dark with the drug in my hand…“Here you go, you old geezer!”With these words I quickly and decisively plunged the needle into my grandpa’s shoulder. Then I pushed the plunger, and the clear liquid began to flow into his veins… He didn’t wake up, because earlier I’d given him tea with an increased dose of sedative.What made me do this? Well, the story is ra
FictionLongevityTranshumanism
Research LessWrong Jan 31

Basics of How Not to Die

By Camille Berger

8 score
AI Analysis

A practical safety post about carbon monoxide poisoning prevention, prompted by a near-death experience. Not related to AI research.

One year ago, we nearly died.This is maybe an overdramatic statement, but long story short, nearly all of us underwent carbon monoxide (CO) poisoning[1]. The benefit is, we all suddenly got back in touch with a failure mode we had forgotten about, and we decided to make it a yearly celebration.Usually, when we think about failure, we might think about not being productive enough, or not solving the right work-related problem, or missing a meeting. We might suspect that our schedule could be bett
Physical SafetyRisk Management