Tag
Diffusion Transformer (DiT)
Every Diffusion Transformer (DiT) story we've curated in Bowl of Data, newest issue first — part of our weekly digest across AI, security, blockchain, and engineering.
Week 38 · 2026
Read the issue →-
VC-Attention: Value Smoothing and Softmax Casting for Low-bit Attention
VC-Attention is a new low-bit attention kernel that optimizes video diffusion transformer inference by addressing value quantization errors and softmax bottlenecks. It utilizes value smoothing via k-means clustering and a fused probability casting method to achieve high fidelity and significant hardware acceleration.
Week 30 · 2026
Read the issue →-
Self Gradient Forcing: Native Long Video Extrapolation
The researchers present Self Gradient Forcing (SGF) to solve the lack of gradient flow in historical KV caches during autoregressive video generation. This method enables much more stable and consistent long-form video extrapolation without the massive memory overhead of full backpropagation.
Week 28 · 2026
Read the issue →-
From RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image Models
This paper proposes ReChannel, a novel architecture that transforms text-to-image models from RGB generators into efficient dense prediction engines. By treating transformer tokens as spatial carriers for task-specific data rather than RGB pixels, the method achieves new state-of-the-art performance with much higher computational efficiency.
Free weekly digest
Get next Saturday’s issue in your inbox
The week’s most relevant AI, security, blockchain, and engineering stories — curated, summarised, and reviewed by humans. No spam, unsubscribe anytime.
Subscribe — it’s free