Sam's News β tech-research β 2026-09-21¶
Research¶
6 HERMES: knowledge graph reasoning from clinical notes for patient outcome prediction¶
Researchers introduced HERMES, a method leveraging knowledge graphs and clinical notes to predict patient outcomes more accurately than sequence-based approaches.
Sources: arXiv β Computation and Language RSS
6 TALON: temporally aware framework for longitudinal radiology report generation¶
Researchers presented TALON, a model that generates radiology reports with temporal awareness to enable meaningful longitudinal comparisons between patient examinations.
Sources: arXiv β Computation and Language RSS
6 Study uses LLMs to identify Kubernetes misconfigurations in cloud-native environments¶
Researchers demonstrate how large language models can detect security misconfigurations in Kubernetes deployments to improve cloud-native infrastructure security.
Sources: arXiv β Computation and Language RSS
AI Research¶
6 Deep Research Agent Training via Self-Generated Rollout Traces for Extended Context¶
A technique using reinforcement learning on self-generated reasoning traces improves agentic deep research and long-context handling.
Sources: arXiv β Computation and Language RSS
6 ArenaFlow: Hierarchical Credit Assignment for Open-Ended LLM Agent Reinforcement Learning¶
A trajectory ranking and hierarchical credit propagation method extends reinforcement learning to open-ended agent tasks without scalar rewards.
Sources: arXiv β Computation and Language RSS
5.5 dSTAR: Straggler Tolerant and Byzantine Resilient Distributed SGD¶
A new distributed training algorithm addresses stragglers and Byzantine failures in multi-node gradient aggregation.
Sources: arXiv β Artificial Intelligence RSS, arXiv β Cryptography and Security RSS
5.5 BI-Agent and BI-Bench: Automating End-to-End Business Intelligence¶
Researchers introduce an agent and benchmark for automating data preparation and analysis workflows in business intelligence tools.
Sources: arXiv β Artificial Intelligence RSS, arXiv β Machine Learning RSS
5.5 Scaling Discovery through Test-Time Communication¶
Research demonstrates that allowing AI agents to communicate at test time substantially improves collaborative scientific discovery.
Sources: arXiv β Artificial Intelligence RSS, arXiv β Machine Learning RSS
5.5 ASGARD: Action-Space Guard for UAV Resilience via Reinforcement Learning¶
A defense mechanism protects reinforcement-learning-controlled drones from action-space attacks that intercept and override policy commands.
Sources: arXiv β Machine Learning RSS, arXiv β Cryptography and Security RSS
AI Safety¶
6 MME-Safety: Fine-Grained Safety Benchmark for Multimodal Language Models¶
A new fine-grained safety benchmark evaluates vulnerabilities in multimodal LLMs that bypass unimodal content filters.
Sources: arXiv β Computation and Language RSS