Create an annotated collection of valuable LessWrong/EA Forum comment threads that explains the main disagreements, arguments, and cruxes in AI safety debates.
About the project
Some of the best discussion of AI safety appears in comment threads rather than standalone posts, making it difficult to index. Mentees will identify particularly valuable threads on topics such as the difficulty of alignment, sources of misalignment, and whether specific alignment techniques are likely to work.
The final collection will summarize the major positions, highlight the strongest arguments, explain important points of disagreement, and provide enough context for readers to understand why each thread is useful.
Theory of change
Making high-quality discussions easier to find and understand could help researchers and newcomers build better models of important debates, identify unresolved cruxes, and avoid repeating arguments that have already been explored in depth.
Your role
Mentees will search for promising threads, assess their quality and relevance, trace the surrounding debate, and write concise annotations and syntheses. The main output will be a public collection organized by topic, eventually posted on LessWrong and circulated among intro AI safety resource channels.
Prerequisites
Strong reading, synthesis, and editorial judgment. Familiarity with AI safety and LessWrong would be useful.
Application question(s)
Share a link to a LessWrong comment thread you think is unusually valuable. Explain the main takeaway, why you find it interesting/useful, and provide further context if necessary. (100 words, 300 words max)
About the mentor

Helena Tran
Constellation
Helena Tran runs the Generator Residency as a program coordinator at Constellation. Previously, she founded AI Safety Collective Irvine and contracted for Kairos to help run SPAR.