# Nightly Research Summary
Updated: 2026-09-07 12:17:04 AEST
Scope: bounded arXiv scan across cs.AI, cs.LG, cs.CR, and cs.RO.
## Latest Findings
- 2609.05324v1 - RoboSPA: Can VLA Models Go Beyond Simple Scenes and Short-Horizon Tasks? (2026-09-04, cs.RO, cs.AI, cs.CV) - http://arxiv.org/abs/2609.05324v1 - Implication: agent/tool workflow evaluation.
- 2609.05314v1 - Large Language Models for HVAC Operations in Building Energy Systems: A Critical Review of Methods, Applications, and Deployment Readiness (2026-09-04, cs.AI, cs.CL, eess.SY) - http://arxiv.org/abs/2609.05314v1 - Implication: agent/tool workflow evaluation.
- 2609.05309v1 - How Does mHC Use Its Residual Streams? Selective Routing and Near-Identity Mixing (2026-09-04, cs.LG, cs.AI) - http://arxiv.org/abs/2609.05309v1 - Implication: memory and retrieval governance.
- 2609.05295v1 - RISE: Recursive Improvement via Self-Extrapolating Policy Distillation (2026-09-04, cs.AI) - http://arxiv.org/abs/2609.05295v1 - Implication: agent/tool workflow evaluation.
- 2609.05289v1 - Beyond Aggregate Scores: Behavioral Correctness Assumptions for Assessing Reference-Based Automatic Evaluation Methods (2026-09-04, cs.AI) - http://arxiv.org/abs/2609.05289v1 - Implication: high-impact autonomy gating.
- 2609.05284v1 - GUT: Quantifying and Optimizing the Reasoning Uncertainty of LLMs via Graph Complexity (2026-09-04, cs.AI) - http://arxiv.org/abs/2609.05284v1 - Implication: watchlist only.
- 2609.05279v1 - Testing Interchangeability in LLM Agent Teams (2026-09-04, cs.AI, cs.MA) - http://arxiv.org/abs/2609.05279v1 - Implication: agent/tool workflow evaluation.
- 2609.05275v1 - Don't Drop Dropout: Optimizing Layer Sparsity for Efficient LLM Training and Inference (2026-09-04, cs.AI) - http://arxiv.org/abs/2609.05275v1 - Implication: watchlist only.
Public mirror generated from the latest Hermes nightly arXiv research output. No private paths or local-only artifact locations are published.