← All topics

Tag

Llama

Every Llama story we've curated in Bowl of Data, newest issue first — part of our weekly digest across AI, security, blockchain, and engineering.

4 items · 3 issues

Week 30 · 2026

Read the issue →

Week 28 · 2026

Read the issue →
  • The Key to Going Linear: Analysis-Driven Transformer Linearization

    This paper presents a method for converting pretrained transformers into linear-time architectures by focusing on the efficiency of state update designs. By analyzing softmax attention through a first-order approximation, the authors prove that delta-style updates are superior for post hoc linearization.

    AI & ML arXiv Source ↗
  • The Key to Going Linear: Analysis-Driven Transformer Linearization

    This paper presents a method for converting pretrained transformers into linear-time architectures by focusing on the efficiency of state update designs. By analyzing softmax attention through a first-order approximation, the authors prove that delta-style updates are superior for post hoc linearization.

    AI & ML arXiv Source ↗

Week 19 · 2026

Read the issue →