Tag
Llama
Every Llama story we've curated in Bowl of Data, newest issue first — part of our weekly digest across AI, security, blockchain, and engineering.
Week 30 · 2026
Read the issue →-
Cursor, Ramp, and Meta are all building model routers — but two have major model ambitions themselves
The AI industry is seeing a surge in 'model routers' designed to intelligently distribute LLM queries to the most cost-effective and capable models. Companies including Cursor, Ramp, and Meta are developing these systems to optimize performance and reduce the high costs associated with frontier-grade models.
Week 28 · 2026
Read the issue →-
The Key to Going Linear: Analysis-Driven Transformer Linearization
This paper presents a method for converting pretrained transformers into linear-time architectures by focusing on the efficiency of state update designs. By analyzing softmax attention through a first-order approximation, the authors prove that delta-style updates are superior for post hoc linearization.
-
The Key to Going Linear: Analysis-Driven Transformer Linearization
This paper presents a method for converting pretrained transformers into linear-time architectures by focusing on the efficiency of state update designs. By analyzing softmax attention through a first-order approximation, the authors prove that delta-style updates are superior for post hoc linearization.
Week 19 · 2026
Read the issue →-
Bleeding Llama: Critical Unauthenticated Memory Leak in Ollama (CVE-2026–7482)
A critical vulnerability in Ollama allows unauthenticated attackers to trigger an out-of-bounds heap read via malicious GGUF files. This exploit can expose sensitive information like user messages and system prompts by leaking them into newly created model files.
Free weekly digest
Get next Saturday’s issue in your inbox
The week’s most relevant AI, security, blockchain, and engineering stories — curated, summarised, and reviewed by humans. No spam, unsubscribe anytime.
Subscribe — it’s free