Tobias Thaller
Speaking appearances
- AI Policy Tuesday: Verification Mechanisms for Global AI Governance 2025-12-09 — Toronto
- AI Safety Thursday: Sandbagging - How Models Use Reward-Hacking to Downplay Their True Capabilities 2025-11-27 — Toronto
- Online Animal ACTION PARTY - EU consultation on animal welfare 2025-11-26
- AI Safety Thursday: Introduction to Corrigibility 2025-11-20 — Toronto
- AI Safety Thursday: Agentic property-based testing - finding bugs across the Python ecosystem 2025-11-13 — Toronto
- AI Safety Thursday: Monitoring LLMs for deceptive behaviour using probes 2025-11-06 — Toronto
- AI Policy Tuesday: Debunking the US-Chinese AGI Race 2025-10-28 — Toronto
- AI Safety Thursday: The Limitations of Reinforcement Learning for LLMs in Achieving AI for Science 2025-10-23 — Toronto
- AI Policy Tuesday: Redlines for AI 2025-10-14 — Toronto
- AI Safety Thursday: Building an economic model of AI automation 2025-10-09 — Toronto
- AI Safety Thursday: Attempts and Successes of LLMs Persuading on Harmful Topics 2025-10-02 — Toronto
- AI Safety Thursday: Technical AI Governance - Motivations, Challenges, and Advice 2025-09-18 — Toronto