Skip to content

Sam's News β€” tech-research β€” 2026-10-10

Privacy & Security

6.5 Context-Sensitive Memorization in LLMs via Prefix Extraction

Study shows memorized training sequences can be extracted in realistic deployment contexts beyond isolated prefix conditions.

Sources: arXiv β€” Artificial Intelligence RSS

Self-Improvement

6.5 Recursive Self-Improvement via Multi-Agent Self-Supervision

Framework overcomes supervision bottleneck in non-verifiable tasks by using the model itself as evaluator for recursive improvement.

Sources: arXiv β€” Artificial Intelligence RSS

Security

6.5 Option-Channel Attack: Exploiting Typed Decision Models as Agent Guardrails

Single-word injection attack circumvents typed decision model guardrails deployed to protect LLM agent systems.

Sources: arXiv β€” Artificial Intelligence RSS

AI Compliance

6.5 Regulatory Rule Sensitivity Audit of LLM Compliance Systems

Audit reveals LLM compliance verdicts sometimes ignore regulatory rules when deleted, swapped, or negated.

Sources: arXiv β€” Artificial Intelligence RSS

AI Transparency

Audit reveals legal LLM reasoning may cite statutes without actually following them when facts are held constant.

Sources: arXiv β€” Artificial Intelligence RSS

AI Safety

6 Guard Models Insufficient for LLM Agent Safety Without Obligation Tracking

Guard models that only block forbidden actions fail to ensure agent safety; unfulfilled obligations also require monitoring.

Sources: arXiv β€” Artificial Intelligence RSS

Research Methodology

6 Survey: Benchmarks and Evaluation of Automated Research Systems

Comprehensive review identifies fragmented evaluation practices across literature synthesis, experimentation, and peer review automation.

Sources: arXiv β€” Artificial Intelligence RSS

Agent Learning

6 Hippocam: Intent-Structured Memory for Long-Horizon LLM Agents

Framework consolidates continuous agent experience into structured, reusable knowledge for sustained autonomous operation.

Sources: arXiv β€” Artificial Intelligence RSS

Model Training

6 Hindsight Hierarchies for Self-Improving Reasoning Models

Self-improvement loop trains reasoning models on solutions to problems beyond current capability, extracting generalizable insights.

Sources: arXiv β€” Artificial Intelligence RSS

Agent Safety

6 OnTrack: Real-Time LLM Agent Monitoring via Optimal Transport

Framework monitors and intervenes in agent trajectories in real-time to prevent costly or unsafe irreversible actions.

Sources: arXiv β€” Artificial Intelligence RSS