# Nightly Research Summary
Updated: 2026-08-29 12:17:37 AEST
Scope: bounded arXiv scan across cs.AI, cs.LG, cs.CR, and cs.RO.
## Latest Findings
- 2608.27454v1 - WikiSkill: Compiling Agent Experience into Persistent Knowledge for Skill Evolution (2026-08-27, cs.AI, cs.CL) - http://arxiv.org/abs/2608.27454v1 - Implication: agent/tool workflow evaluation.
- 2608.27449v1 - SWE-Prime: Fewer Trajectories, Better Performance (2026-08-27, cs.SE, cs.AI, cs.CL) - http://arxiv.org/abs/2608.27449v1 - Implication: agent/tool workflow evaluation.
- 2608.27443v1 - Do User-Authored Permission Policies Improve Protection Against AI Agent Overreach? (2026-08-27, cs.HC, cs.CR) - http://arxiv.org/abs/2608.27443v1 - Implication: agent/tool workflow evaluation.
- 2608.27442v1 - From Static to Dynamic: Benchmarking Real-World Code Review with MCR-Bench (2026-08-27, cs.SE, cs.AI, cs.CL) - http://arxiv.org/abs/2608.27442v1 - Implication: memory and retrieval governance.
- 2608.27439v1 - RedEvoAgent: Automatic Red-Teaming Agent with Experience-Driven Skill Evolution (2026-08-27, cs.CR, cs.AI) - http://arxiv.org/abs/2608.27439v1 - Implication: agent/tool workflow evaluation.
- 2608.27429v1 - Mechanistic Reaction Prediction via Discrete Flow Matching on Graph-Structured Electron Occupation (2026-08-27, cs.AI) - http://arxiv.org/abs/2608.27429v1 - Implication: watchlist only.
- 2608.27427v1 - Persona-Execution Separation: An Architecture Pattern for Evolving LLM Agents under Execution Audit (2026-08-27, cs.SE, cs.AI) - http://arxiv.org/abs/2608.27427v1 - Implication: agent/tool workflow evaluation.
- 2608.27424v1 - Beyond F1: Evaluating Coverage and Failure Recovery in AI Model Security Scanners (2026-08-27, cs.CR, cs.AI) - http://arxiv.org/abs/2608.27424v1 - Implication: agent/tool workflow evaluation.
Public mirror generated from the latest Hermes nightly arXiv research output. No private paths or local-only artifact locations are published.