Discussion

Forum

The SAIRC Discussion Forum is a space for AI enthusiasts to share what they're thinking about. No formal research paper required. Posts can be submitted anonymously and span a wide range of formats:

Recent Posts
The Underwhelming Universal Approximation Theorem
A research advisor brought up the Universal Approximation Theorem during a meeting once as a fun thought experiment. It stuck with me because I initially found it super interesting
Imran Kutianawala·March 31, 2026
Mechanistic Interpretability and the Curses of Scaled Networks
Mechanistic interpretability is the practice of reverse-engineering deep-learning systems to understand their inner algorithms. If we can figure out what a neural network is actual
Joshua Shen·March 29, 2026
What Students Want Teachers to Know About AI
Reposted with permission from original author(s). Back in December I sat down with a handful of HS students on a couple different occasions to talk about AI — their thoughts, attit
Stephen Fitzpatrick·March 23, 2026
How Does Persistent Memory Affect AI Safety?
This post was submitted by a SAIRC member either as a recommended read or student-created post. All credit remains with the original author. Written with Maksym Andriushchenko, th
Ben Rank·March 21, 2026
In an AI World, What's the Work?
Reposted with permission from original author(s). More than a decade ago, immersed in research on competency-based education, I read Leaders of Their Own Learning, a comprehensive
Eric Hudson·March 19, 2026
Skill-Inject: Measuring Agent Vulnerability to Skill File Attacks
This post was submitted by a SAIRC member either as a recommended read or student-created post. All credit remains with the original author. A security evaluation of the emerging
Sahar Abdelnabi·March 13, 2026
AI Shouldn't Be Doing More Philosophy Than Students
Reposted with Permission from Mike Taubman and AI Waypoints. Last weekend I found out that AIs can now reincarnate. This got me thinking about both the high school juniors I teach
Mike Taubman·March 12, 2026
MLSN #19: Honesty, Disempowerment, & Cybersecurity
This post was submitted by a SAIRC member either as a recommended read or student-created post. All credit remains with the original author. An issue of the ML Safety Newsletter,
Alice Blair·March 12, 2026
Labor Market Impacts of AI: A New Measure and Early Evidence
Full credit goes to the original author, linked below. All blog posts were reposted either with permission of the author, or by anonymous submission by SAIRC members like yourself.
Anthropic ·March 5, 2026
HalluHard: A Hard Multi-Turn Hallucination Benchmark
This post was submitted by a SAIRC member either as a recommended read or student-created post. All credit remains with the original author. Built with Maksym Andriushchenko, Hall
Dongyang Fan·March 5, 2026
Interpretability Research Already Has a Framework for Actionability
This post was either an anonymous submission of an interesting paper or was written by a student; full credit remains with the author (linked). Prompted by a Chris Olah talk years
Jessica Hullman·March 3, 2026
There's No Token for the Way the End of High School Feels
Reposted with permission by Mike Taubman. Today in our AI literacy class, Scott Kern and I helped students open the hood to see how AI works, one pillar of the AI Driver's License
Mike Taubman·February 27, 2026
Painless Activation Steering (PAS): Automated, Lightweight Post-Training for LLM Behavior
Reproduced with permission from Sasha Cui We're releasing "Painless Activation Steering (PAS)," a fully automated approach to steer large language models after training—without mo
Sasha Cui·February 14, 2026
(When) Is Mechanistic Interpretability Identifiable?
I recently finished a paper, "Characterizing Mechanistic Uniqueness and Identifiability Through Circuit Analysis," alongside a group of three others and a mentor. This post discuss
Imran Kutianawala·February 14, 2026
Mamba's Memory Problem
Full credit goes to the original author, linked below. All blog posts were reposted either with permission of the author, or by anonymous submission by SAIRC members like yourself.
LLMs Research·February 2, 2026
Previous1...456...8Next

Become a member.
It's completely free.

Get notified of new research, resources, and SAIRC journal editions.