Ai safety

Out-of-Context Reasoning in LLMs: A short primer and reading list

Out-of-context reasoning (OOCR) is a concept relevant to LLM generalization and AI alignment.

Read More →

New, improved multiple-choice TruthfulQA

We introduce a new multiple-choice version of TruthfulQA that fixes a potential problem with the existing versions (MC1 and MC2).

Read More →