newsReddit r/artificialTrust 52 · CommunityPublished 11d agoLive · 10d ago
GPT-2 Fully Decoded Internally Black Box Fully Open With Demo
The BABEL codec: the first complete, certified decode of everything happening inside a production language model (GPT-2 small). It reads the model's internal state into English AND writes English back into the model. 94.7% of behavior reconstructed — and that holds at every layer depth and text regime tested, not just one spot. Everything is open: paper, the full lexicon, the grammar tables, the decoder/encoder weights, reproduction scripts, and a demo that show
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 47%sgl-project/SpecForge →
- PossiblePossibly related (embedding) · 46%Understanding Evaluation Illusion in Diffusion Large Language Models →
- PossiblePossibly related (embedding) · 46%robotlearning123/gpt2agent →
- PossiblePossibly related (embedding) · 45%lightseekorg/TorchSpec →
- PossiblePossibly related (embedding) · 45%Masked Diffusion Decoding as $x$-Prediction Flow →
- PossiblePossibly related (embedding) · 49%haseeb-heaven/code-interpreter →
- PossiblePossibly related (embedding) · 46%openai/gpt-oss-20b →
- PossiblePossibly related (embedding) · 48%Complexity-Guided Component-wise Initialization for Language Model Pretraining →
Covers
Covers (incoming)
Related across the graph
repolightseekorg/TorchSpecpaperUnderstanding Evaluation Illusion in Diffusion Large Language ModelspaperMasked Diffusion Decoding as $x$-Prediction Flowreposgl-project/SpecForgerepohaseeb-heaven/code-interpreterpaperComplexity-Guided Component-wise Initialization for Language Model Pretrainingreporobotlearning123/gpt2agentmodelopenai/gpt-oss-20b
