Linen Yu
Projects
- Evaluating Safety-Relevant Representations and Behaviors in Language Models From a single refusal direction to a source-gated semantic effect signature.
- Measuring Governance Orientations in Large Language Models Four completed model panels compared with expert, researcher, and public survey evidence.
- AI Persona English version of the AI governance persona quiz.
- AI Governance Spectrum Browse the map behind the quiz.