Reading up on Ouro
2 deep · digging since sep 09
- GPT-6 Astra, Looped Transformers, and Hidden Reasoning
The article examines GPT‑6 Astra’s performance, explains looped transformers as weight‑shared depth increase, and evaluates claims that they hide chain‑of‑thought reasoning.
- GPT-6 Astra, Looped Transformers, and Hidden Reasoning
GPT-6 Astra achieves top coding and multimodal scores via looped transformers that reuse weights for extra depth, while recent work shows this architecture does not inherently conceal chain‑of‑thought reasoning.