Daily News

AI research, safety, product, and engineering links

Latest AI Reading

Source-dated posts from the last 14 days

Updated 2026-08-07 23:35 UTC

AI Alignment Forum AI Safety and Governance

Why do models task game?

TL;DR How can we study misalignment with today's models as proxies? They're clearly not paperclip maximizers, but they also often do things the user doesn't want. A strong contender for a real misaligned propensity is t...

Read original
AI Alignment Forum AI Safety and Governance

Returning to ARC

I've returned to the Alignment Research Center (ARC) as executive director. My main focus for the next six months will be driving forward ARC's research agenda—building techniques to find mechanistic explanations for ne...

Read original
AI Alignment Forum AI Safety and Governance

Thousand-dimensional structure

Summary: One area we plan to explore at Resolution is personas and character training, operationalized as finding and controlling low-dimensional structure in models that emerges in pretraining and flows through post-tr...

Read original

Sources

Frontier Labs

Anthropic Research

Anthropic research on alignment, interpretability, evaluations, and societal impacts.

AI Safety and Governance

LessWrong

AI alignment, rationality, and AI risk discussion.

MIRI Blog

Machine Intelligence Research Institute blog.

METR

Model evaluation and frontier-risk research.

Epoch AI Blog

AI trends, compute, data, economics, and forecasting.

Epoch AI Latest

Unified stream for papers, newsletters, data insights, and podcasts.

AI Companies and Research Labs
Academic Labs
Personal Blogs and Newsletters

Import AI

Jack Clark's AI research and industry newsletter.

Interconnects

Nathan Lambert's frontier AI research and industry newsletter.

Distill

Classic visual and explanatory machine learning articles.