Install our extension to search inside any video instantly.

Live Interview - AI Safety with Joel Rorseth

Added:
142 views0likes59:45i-thinktogetherOriginal Release: 2026-07-17

AI safety ensures that artificial intelligence systems are ethical, aligned with human values, and do not harm users. A critical challenge in AI safety is the 'black box' problem, where complex AI models like large language models make decisions without transparent reasoning. Explainable AI (XAI) addresses this by developing tools that help users understand how AI models arrive at their decisions, enabling bias detection, auditing, and building trust. Two practical tools demonstrate this: RAGE shows how different source documents affect answers in large language models, while Credence allows users to test hypotheses about why search engines rank documents differently.