Explaining complex AI safety topics in short, engaging forms, including demos, writeups, and videos.
About the project
We will produce a series of digestible explainers, demos, and videos on AI safety topics for the general public, politicians and policymakers, journalists, and educators. We will host and distribute them across relevant channels, including AI safety Slacks and Discords, YouTube, X, LinkedIn, Hacker News, and university AI safety groups.
Mentees will write, edit, and produce content. Possible topics include sleeper agents, eval awareness, CoT unfaithfulness, AI Control, bio and cybersecurity uplift, compute verification, security levels, safeguards, and self-fulfilling misalignment.
Mentees work independently on their own explainers, videos, or demos, and pool effort on a more ambitious group build. Over three months of part-time work, the team would ship a body of published, distributed work: individual pieces under 1,000 words, short videos, and at least one substantial interactive artifact, plus a distribution plan targeting university AI safety groups, policy newsletters, and social channels.
Theory of change
Many valuable people do not have the time, bandwidth, or context to understand long and technical AI safety materials. Short, engaging, and easy-to-distribute content can reduce the barrier to understanding AI safety, improve the salience of relevant topics, and help shift the Overton Window in AI safety's favor.
Your role
Mentees will write, edit, and produce content, working independently on their own explainers, videos, or demos while pooling effort on a more ambitious group build. Expected output over three months is a body of published, distributed work: individual pieces under 1,000 words, short videos, and at least one substantial interactive artifact, plus a distribution plan.
Prerequisites
Mentees should have broad knowledge of technical or policy AI safety, or the ability to learn quickly. Useful skills include writing short and simple pieces, video editing or production, running simple technical experiments or scraping operations, and proficiently using coding agents.
Application question(s)
Explain a complicated technical concept in fewer than 200 words. Your explanation should be legible to all demographics, engaging, and easy to read.
No AI usage is permitted in writing or editing. Submit a Google Doc with full editing history.
About the mentors
Kaustubh Kislay
Wisconsin AI Safety Initiative
Kaustubh Kislay directs the Wisconsin AI Safety Initiative.
Christine Corry
Second Look Research and XLab @ UChicago
Christine Corry is an AI safety replication researcher at Second Look Research and XLab at UChicago.