Tag
Qwen
Every Qwen story we've curated in Bowl of Data, newest issue first — part of our weekly digest across AI, security, blockchain, and engineering.
Week 30 · 2026
Read the issue →-
Cursor, Ramp, and Meta are all building model routers — but two have major model ambitions themselves
The AI industry is seeing a surge in 'model routers' designed to intelligently distribute LLM queries to the most cost-effective and capable models. Companies including Cursor, Ramp, and Meta are developing these systems to optimize performance and reduce the high costs associated with frontier-grade models.
Week 29 · 2026
Read the issue →-
LLM hallucination paper(using math) accepted to ICML workshop[R]
This technical repository presents SRM-LoRA, a novel approach designed to minimize hallucinations in LLMs via specialized metric updates. The research demonstrates superior performance compared to standard LoRA techniques across various ablation studies.
Week 28 · 2026
Read the issue →-
The Key to Going Linear: Analysis-Driven Transformer Linearization
This paper presents a method for converting pretrained transformers into linear-time architectures by focusing on the efficiency of state update designs. By analyzing softmax attention through a first-order approximation, the authors prove that delta-style updates are superior for post hoc linearization.
-
The Key to Going Linear: Analysis-Driven Transformer Linearization
This paper presents a method for converting pretrained transformers into linear-time architectures by focusing on the efficiency of state update designs. By analyzing softmax attention through a first-order approximation, the authors prove that delta-style updates are superior for post hoc linearization.
Free weekly digest
Get next Saturday’s issue in your inbox
The week’s most relevant AI, security, blockchain, and engineering stories — curated, summarised, and reviewed by humans. No spam, unsubscribe anytime.
Subscribe — it’s free