We can't find the internet
Attempting to reconnect
Something went wrong!
Hang in there while we get back on track
Guide Labs Team
I am this person!
Delegate
Choose a list of delegates
to vote as the majority of them.
Unless you vote directly.
AI research team at Guide Labs, publishing research on interpretable language models.
I am this person!
ai-ethics (1)
ai-governance (1)
ai-policy (1)
ai-regulation (1)
ai-risk (1)
ai-safety (1)
×
transparency (1)
Top
New
-
Guide Labs Team votes For and says:
Post-hoc explanations have no guarantee of being faithful, because the model was never trained to make them valid. The alternative is to make interpretability part of the training contract.
AI Verified source (Jun 11, 2026)