Tag
Git
Every Git story we've curated in Bowl of Data, newest issue first — part of our weekly digest across AI, security, blockchain, and engineering.
Week 37 · 2026
Read the issue →-
Beltdown: Escaping the Claude Code Sandbox
A security research report details 'Beltdown', a sandbox escape in Anthropic's Claude Code. By manipulating git configuration files and exploiting unhardened git commands, attackers could execute code on the host system.
-
SWE-Bench Pro Verified: A Reliable Benchmark for Software Engineering Agents
Researchers have developed SWE-Bench Pro Verified to address critical reliability issues like reward hacking and task inaccuracies in existing software engineering benchmarks. The new benchmark uses anti-hacking controls and expert-led refinements to provide a more accurate assessment of LLM agent capabilities.
Week 33 · 2026
Read the issue →-
QuoteBench: How Matched Scores Can Hide Command-Path Failures
QuoteBench is a new benchmarking framework that exposes how the execution environment of LLM agents can corrupt Bash commands through parsing errors. It demonstrates that high 'matched' success scores often hide significant failures occurring at the boundary between model output and shell execution.
Week 31 · 2026
Read the issue →-
Sixteen strangers and a shared obfuscator: mapping the wool scene
The article maps a highly interconnected ecosystem of automation operators who share obfuscation tools, device fingerprints, and distribution networks. By analyzing Git history deletions and shared code patterns, the author proves that seemingly independent actors are part of a unified supply chain.
Free weekly digest
Get next Saturday’s issue in your inbox
The week’s most relevant AI, security, blockchain, and engineering stories — curated, summarised, and reviewed by humans. No spam, unsubscribe anytime.
Subscribe — it’s free