Tag
DPO
Every DPO story we've curated in Bowl of Data, newest issue first — part of our weekly digest across AI, security, blockchain, and engineering.
Week 38 · 2026
Read the issue →-
A Zeroth-Order Paradigm for LLM Preference Alignment
The researchers present ComPO, a new alignment paradigm that uses comparison oracles to extract directional information from preference pairs. This method effectively mitigates the risks of likelihood displacement and model verbosity seen in traditional direct alignment methods.
Week 36 · 2026
Read the issue →-
SecOPD: Mitigating Adaptive Prompt Injections by On-Policy Distillation
Researchers have developed SecOPD, a fine-tuning technique that uses token-level feedback to protect AI agents from prompt injection attacks. This method significantly outperforms previous state-of-the-art defenses by precisely identifying and penalizing malicious tokens during training.
Week 35 · 2026
Read the issue →-
SecOPD: Mitigating Adaptive Prompt Injections by On-Policy Distillation
Researchers have developed SecOPD, a fine-tuning technique that uses token-level feedback to protect AI agents from prompt injection attacks. This method significantly outperforms previous state-of-the-art defenses by precisely identifying and penalizing malicious tokens during training.
Week 25 · 2026
Read the issue →-
APPO: Agentic Procedural Policy Optimization
This paper proposes APPO, a method to enhance the training of autonomous LLM agents through more granular credit assignment. It moves beyond trajectory-level rewards by identifying and branching at specific high-impact decision points within the reasoning process.
Free weekly digest
Get next Saturday’s issue in your inbox
The week’s most relevant AI, security, blockchain, and engineering stories — curated, summarised, and reviewed by humans. No spam, unsubscribe anytime.
Subscribe — it’s free