← All topics

Tag

LLM (Large Language Models)

Every LLM (Large Language Models) story we've curated in Bowl of Data, newest issue first — part of our weekly digest across AI, security, blockchain, and engineering.

3 items · 3 issues

Beats AI & ML
Often covered with Anthropic 1 Docker 1 vLLM 1 JSON 1

Week 36 · 2026

Read the issue →

Week 33 · 2026

Read the issue →
  • QuoteBench: How Matched Scores Can Hide Command-Path Failures

    QuoteBench is a new benchmarking framework that exposes how the execution environment of LLM agents can corrupt Bash commands through parsing errors. It demonstrates that high 'matched' success scores often hide significant failures occurring at the boundary between model output and shell execution.

    AI & ML arXiv Source ↗

Week 28 · 2026

Read the issue →