BASIS

Bangalore AI Safety & Alignment Study. A small group for thoughtful, technically grounded discussions on AI safety and alignment, spanning papers, books, and open questions.

BASIS members at a reading group, Bangalore, June 2026
Bangalore, June 2026
8 Papers read
39 Concepts

Concept co-occurrence

Each node is a concept tagged on a session. Edges connect concepts that co-occurred.

Recent sessions

All sessions →
BASIS · No. 09 25 Aug 2026

Reading Group 9: Model Forensics

Detecting concerning behaviour does not establish misalignment. The paper proposes a protocol for investigating what drove the behaviour.

misalignmentdeceptionchain-of-thoughtevalsagentsauditing
BASIS · No. 08 29 Jul 2026

Reading Group 8: Global Workspace (deep dive)

A follow-up to session 7: instead of re-reading the paper, we dug into the code and the Neuronpedia demo to see the Global Workspace idea in action on Gemma.

mech-interpinterpretabilitymonitoringcognition
BASIS · No. 07 19 Jul 2026

Reading Group 7: A Global Workspace in Language Models

Anthropic locates a 'J-space' inside Claude that behaves like a cognitive-science-style global workspace: a routing hub for deliberate thought that could double as a monitoring surface.

mech-interpinterpretabilitymonitoringcognition

Members are affiliated with