MiniMax researcher Olive Song laid out the architectural decisions behind the lab's M3 model in a talk on the AI Engineer podcast: a ...
For two years, interpretability researcher Eric Bigelow has been mapping exactly where large language models make their choices. His conclusion, ...