Conference Presentations
Accepted Work
Conference Poster
Beyond the Safety Layer: How RLHF Architecture Produces Clinically Recognizable Patterns of User Harm
Stanford AIMI Symposium · June 3, 2026
Applies causality assessment methodology to AI safety layer failures, demonstrating how alignment architectures pattern-match on demographics over methodology — producing gendered institutional dismissal with measurable clinical parallels.