NEWESTPODCAST

Can agents sabotage code and safety research?

What unprompted evaluations, continuation tests, refusal patterns, and model-organism audits actually establish

T3E3Agentic Safety & Alignment: From Predictors to Governed Agents15:24
Show notes & transcript

Blog

    All posts (0)

    Podcasts

    All podcasts (131)

    Projects

    • fidx

      local hybrid semantic search that beats QMD on recall and precision at ~300–1000× lower latency.

    • token-count-compare

      a side-by-side tokenizer comparison across frontier models.

    • token-usage-analyzer

      finds what's burning your token budget in AI coding CLIs.

    All projects (3)