Tag
GPU
Every GPU story we've curated in Bowl of Data, newest issue first — part of our weekly digest across AI, security, blockchain, and engineering.
Week 35 · 2026
Read the issue →-
Import AI 470: No rights for machines; automating environment generation with SPADE; and building better GPU kernels with Hawkeye
This article details how AI is unevenly accelerating scientific progress, with major impacts in cybersecurity and minor effects in mathematics. It also highlights two new frameworks, SPADE for synthetic environment generation and Hawkeye for automated GPU kernel optimization.
Week 33 · 2026
Read the issue →-
OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse Prefetching
OasisKV addresses the memory wall in LLM inference by implementing a sparse prefetching mechanism for the KV cache. By leveraging speculative decoding to predict future token importance, it moves less critical KV data to cheaper memory tiers without stalling the decode process.
Week 30 · 2026
Read the issue →-
PagedWeight: Efficient MoE LLM Serving with Dynamic Quality-Aware Weight Quantization
PagedWeight optimizes the serving of MoE LLMs by implementing a dynamic quantization strategy that adapts to runtime memory pressure. It balances hardware efficiency with model accuracy by monitoring expert routing statistics and prompt-specific sensitivities.
Free weekly digest
Get next Saturday’s issue in your inbox
The week’s most relevant AI, security, blockchain, and engineering stories — curated, summarised, and reviewed by humans. No spam, unsubscribe anytime.
Subscribe — it’s free