Post
Post
Sparser, Faster, Lighter Transformer Language Models Author's Explanation: x.com/SakanaAILabs/s… Overview: Sparser, Faster, Lighter Transformer Language Models introduces a sparse packing format and custom CUDA kernels to execute unstructured sparsity within LLM feedforward
Manifold Steering Reveals the Shared Geometry of Neural Network Representation and Behavior Author's Explanation: x.com/GoodfireAI/sta… Overview: Manifold steering investigates the causal link between internal representations and behavioral outcomes by intervening along activation manifolds rather than assuming Euclidean geometry. This approach demonstrates that steering along the activation manifold produces behavioral trajectories that align with output geometry, whereas linear interventions fail to recover natural dynamics. By verifying this bidirectional relationship across LLMs and video world models, this work identifies geometric structure as the primary mechanism for principled control over neural model behavior. Paper: arxiv.org/abs/2605.05115
That's a wrap for last week, thanks for reading. You can see in-depth explanations on my newsletter too, stay tuned here: mail.bycloud.ai and check out our latest learning materials, where you can learn LLMs intuitively: intuitiveai.academy
Don't miss what's happening
People on X are the first to know.


