# Two Papers Argue AI Safety Guardrails Do Not Compose Into Real Oversight

One shows medical-note manipulation evading built-in LLM safeguards, another finds interpretability and evaluation work that never assembles into deployable specifications.

- Published: 2026-07-29T05:45:43.018Z
- Canonical: https://polylog.news/ai/2026-07-29/two-papers-argue-ai-safety-guardrails-do-not-compose-into-re
- Publisher: Polylog (AI desk)
- Section: tech
- Sources: [arXiv cs.CR](https://arxiv.org/abs/2607.24859), [arXiv cs.CR](https://arxiv.org/abs/2607.24866)

Two arXiv papers this week converge on the same weakness in AI oversight. "The Mirage of LLM Guardrails" studies AI-assisted manipulation of medical notes and reports that the built-in safeguards large language models use to refuse maliciou…

This story is for subscribers. Read it in full at https://polylog.news/ai/2026-07-29/two-papers-argue-ai-safety-guardrails-do-not-compose-into-re (subscription information: https://polylog.news/pricing).