전체 AI 논문 - 2026-08-06

1. Argus: A General-Purpose Agentic Runtime for Long-Horizon Reasoning


2. OctoLong: Mid-Training On Cross-Repository Code Contexts Enhances Long-Context Modeling


3. CoPlan: A Trustworthy Co-Intelligence Interface for Care Planning through Role-Based Contestable Argument Graphs


4. ABSeeker: Training Long-Horizon Search Agents via Answer-Backtracked Credit Assignment


5. Hierarchical Graph Memory for LLM Agents with Path-level Localization and Rewrite


6. Item Response Theory for AI Safety


7. From Score Matrices to Football-Aware Match-State Simulation: An Auditable LLM Harness for Exact-Score Reranking


8. Short-term load forecasting under EU-AI Act Requirements in Safety-Critical Environments: Results from a 41-day live challenge on the aggregated German transmission-grid load


9. WorldCycle: Self-Verifiable Reinforcement Learning for Long-Horizon Video World Models


10. When Shared Rollouts Fail in Defensive Driving Evaluation: A NAVSIM Score Basis Audit


11. ContextWeave: A Real-World Workflow Benchmark


12. Privileged, but Biased: How PI-Conditioned Teachers Break Self-Distillation


13. NSF-HRPT: Neural Semantic Field meets Hierarchical Risk Perception Tree for Safety-Critical Scenario Assessment


14. Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning


15. EviGraph: Evidence-Guided Autonomous Research Agents


16. Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings


17. When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning


18. Diagnosing Tool-Selection Reasoning in LLM Agents with Canary Tools


19. Traceable LLM-Generated Hazard Scenarios for Operational Safety Analysis of Aviation Systems Using ASRS Reports


20. Calibrating Artificial Guilt: Neurally Grounded Reward Shaping for Prosocial Multi-Agent Reinforcement Learning



22. A/B Agent: A Self-Evolving Agent for Strategy Iteration in Industrial A/B Testing


23. Agreement Before Diversity: Verification-First Complementarity for Heterogeneous Language-Model Coordination


24. Joint UAV Flight and Opportunistic Routing under Reinforcement Learning for Delay-Tolerant Networks


25. What Is a Skill Worth? Structure-Aware Shapley Valuation of Agent Skills


26. Leak-Resistant Unlearning: A New Benchmark for Evaluating Multi-Hop Reasoning Consistency and Recovery Robustness


27. CARGO-VL: Counterfactual Arbitration with Risk-Constrained Group Optimization for Vision-Language Models


28. Architectural Implications of Agentic AI Workflows


29. Improving Auto-Design of Neural PDE Solvers with a Domain-Specific Language


30. NeuMoSync: End-to-End Neuromodulatory Control for Plasticity and Adaptability in Continual Learning


31. SafeCommit: Certifying When Memory-Grounded Agents May Safely Act


32. The RAIL Principles for Neurosymbolic AI: Reasoning, Assurances, Interfacing and Learning


33. Interoceptive Attention as Dynamic Homeostatic Prioritization in a Foraging Agent


34. MatrAIx: Simulating the World with 8.3 Billion Persona Agents


35. Adversarially Robust Abductive Fusion of Pre-trained Transformer-based Perception Models


36. BrainBench: Benchmarking Large Language Models for Comprehensive EEG Understanding


37. FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM Agents


38. FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables


39. Monte Carlo Tree Search for Table-to-Multimodal Report Generation


40. The LLM Proposes, the Executive Disposes: A Self-Verifying Agent Instrument that Dissociates Commitment Drift from Binding Drift in Long-Horizon Agents


41. A Long-Run Persistence Theory for AI Systems under the Redundancy-Adjusted Artificial Age Score (AAS)


42. Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation for Modern Greek across Specialist Domains


43. OPD-V: Visual On-Policy Self-Distillation with Modality Balance


44. SSTQ:Privacy-Preserving Vector Quantization via Subsampled Stochastic TurboQuant


45. Chained Recursive Language Models for Multi-Iteration Reasoning


46. Robust and Efficient Motion Reasoning for Privacy-Aware Classroom Incident Recognition


47. Representational separation between unitary and channel quantum generative models via shared classical randomness at shallow depth


48. Capability-Gated Planning: Cost-to-Goal Discovery and the Limits of Myopic Experiment Selection


49. MultiPathFormer: Towards a Foundation Model for Multipath Wireless Propagation


50. VQ-VAD: Vector-quantized Motion Representation Learning for Human-centric Video Anomaly Detection


51. Provable Limits and Certified Deferral for Verbalized Uncertainty in Small Language Models


52. Hardware Design and Security in the Era of Chiplets and LLMs


53. RepairFormer: Automated Repair of Structured Inputs Using Transformers


54. MarsCast: Transfer Learning of AI Weather Foundation Models to Planetary Atmospheres


55. The Effect of Perceived Race and Gender on Police Language Use: Experimental Evidence from VR Simulations


56. Gradient Immunity: Null-Space Resistance to Malicious Fine-Tuning


57. ArtAnno: Annotating Implicit Semantics in Artworks through LLM Agent-Driven Bidirectional Human-AI Augmentation


58. Revealed Rationality: Label-Free Evaluation and Regularization from Representation Theorems


59. OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents


60. ORACLE: A Multi-Objective Reinforcement Learning-Based Analog Circuit Design Optimizer with Large Language Models-Guided Exploration


61. Protoreasoning in Tiny Transformers


62. SciCode-Verified: How Benchmark Defects Underestimated the Scientific-Coding Ability of Language Models


63. A General Sufficient Condition for Rewriting Horn-ALCHI Atomic Queries into GQL


64. CheMLFlow: An Open-Source Platform for Cheminformatics and Materials Informatics Applications


65. SVI-DAG: A Structured Variational Inference Approach to Bayesian Causal Discovery


66. Consistency-Driven Co-Evolution for Self-Supervised Cross-Representation Learning



68. When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LLMs


69. A-SR: Self-Evolving Agentic LLMs for Symbolic Regression via Hierarchical Coordination


70. Towards a satellite image manipulation and deepfake localization benchmark dataset


71. Scrouting: Cost-Aware Routing of Coding Agents by Scouting the Repository First


72. Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation


73. RepoProbe: Benchmarking Architecture-Aware Repository Comprehension with Checklists


74. IMFACT: Counterfactual Explanations for Time Series via Intrinsic Mode Function Substitution


75. Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage Agent


76. FUSEP: A Multi-Center Benchmark for Diverse Tasks in Early Pregnancy Fetal Ultrasound Screening


77. Explicit Language Memory for Long-Horizon Planning in Vision-Language-Action Models


78. InsightEmb: Learning Action-Intent Embeddings for Agentic Insight Retrieval


79. PURPOSE: Poisoning Conflict Resolution in RAG via Proxy-Fact-Grounded Updates


80. Toward Integrating Adaptive Experience Replay and Online Uncertainty Estimation in Safe Actor-Critic Optimal Control


81. What We Observe as LLM Behavior Can Be a Side-effect of Inference Backend


82. A 6G Integrated Sensing and Communication Framework for Railway Intrusion Detection and Collision Prediction


83. Design Choices That Matter: A Functional ANOVA Analysis for Remote Sensing Multi-Label Classification


84. Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO


85. Personalized Federated Sparse Adaptation of Time-Series Foundation Models


86. Active-SWE: Benchmarking Coding Agents for Proactive Bug Fixing without Issue Reports


87. Easy to Complete, Hard to Choose: Investigating LLM Performance on the ProverbIT Benchmark


88. CSGen: A Multi-Domain Curvilinear Structure Generation Model via Hierarchical Multimodal Diffusion


89. DisMix: Order-Aware Mixup for Medical Imaging via Disentangling Ordinal and Non-Ordinal Features


90. Masked diffusion enables coherent beat tracking


91. The Order Is the Guarantee: Verifier-Budgeted Code Deletion with Static-First Learned Proposals


92. Rethinking Reservoir Pruning: A Dynamical Perspective for Echo State Networks


93. When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large Language Models


94. The First EgoCross Challenge at EgoVis 2026: Cross-Domain Egocentric Video Question Answering


95. EASy: Towards Efficient LLM-Based Agentic System


96. Breaking the Curse ofMultilinguality inMany-to-Many Speech-to-Text Translation via a Resource-AwareMixture of Speech Encoders


97. PhysMind: From Video to Executable Worlds for Training-Free Physical Reasoning


98. Breadcrumbing Search Agents


99. EuroExec: Frontier Language Models Fall Short of Expert Judgment on European Executive Decision Tasks


100. A Model Merging Approach for Continual MLLM Unlearning


101. CARVE: Cross-Slice Anisotropic Reallocation of Visual Evidence for Efficient 3D Medical Volume Understanding


102. GUARD: Grounding Uncertainty and Ablation-Based Risk Detection for Diffusion-Based VLAs


103. GeoReward: Mitigating Contextual Variable Overestimation in Vision-Language Models for Cross-Market Preference Prediction


104. AFD-Ledger: Deployment Provisioning for Attention–FFN Disaggregation


105. AudioScape-TTA: A Structured Soundscape Benchmark for Fine-Grained Text-to-Audio Evaluation


106. EndoVLM: An Endoscopy Vision-Language Pre-training Model via Anatomy-Guided Sparsity and Progressive Alignment


107. Beyond Linear Dynamics: Neural Bilinear Dynamical Models for Time Series Forecasting


108. Tropical Algebraic Geometry for Neuronal Representations: An Arakelov-Green Measure Based Descriptor for Graph Learning


109. Eigenius: A Typed Knowledge-Graph DBMS with Epistemic Stratification and Institution-Mediated Reasoning


110. TwinIR: Coordinated Invisible Dual-Point Attacks on Online HD Map Construction


111. Q-CueGraph: Query-Conditioned Visual Evidence Graphs for Multimodal Reasoning


112. When does training on downscaled images yield the same gradients?


113. D$^2$F-ReAG: Dynamic Decomposition and Filtering for Multi-Hop Reasoning-Augmented Generation


114. ExeCRE: Execution-Consistency Guided Reliability Estimation for Self-Correcting Code Generation


115. MERaLiON-GR: Speech Gender Recognition Model for English and SEA Languages


116. Generative Optimization for Incentivized Advertising with Global Level Constraints


117. SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation


118. Not Every Divergence Should Be Suppressed: Counterfactual Recoverability in On-Policy Distillation


119. MESH: Memory-Efficient Sinkhorn Optimization for Mixture-of-Experts Training


120. Training-Free Hashing-Based Attention via Binary Principal Components


121. Approximate Multi-Objective Search Under Rulebooks


122. NodeJEPA: Structure-Conditioned Latent Prediction for Node-Level Graph Self-Supervised Learning


123. Image Classification Using CNN-QNN Hybrid Model with Optimized Correlated Features


124. Towards Trustworthy Hypergraph Neural Networks under Label Noise


125. FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation


126. Combating Knowledge Corruption in Agent Systems: A Byzantine-Tolerant Secure Collaborative RAG Framework


127. HyPASE: Hyperbolic Geometry for Parameter-Efficient Speech Emotion Fine-Tuning Framework for Large Audio-Language Models


128. iStructTab: Structured Feature Sequencing for Multimodal Learning of Image and Tabular Data


129. Equitable System-Prompt Selection via Constrained Mixed-Strategy GroupDRO


130. COMPAS: Difficulty-Aware Joint Search for Optimizing Code Generation


131. ATLAS: Adaptive Topological Learning with Abstract Successors for Continual Learning


132. Efficient Online Lexicographic Generalized Low-Rank Matrix Bandits


133. Trident : How to Break Deep Reinforcement Learning Cyber Defenses (Agentic)


134. MIDAS: Multi-LLM Iterative Data-Adaptive Summarization


135. EA-Graph: Artifact-Anchored Verification Memory for Coding Agents under Upstream Drift


136. Compass: Continuously Aligning Social Media Feeds via In-Situ Reflections


137. Strategic Evaluation of Planning Strategies for LLM Agents in Cyber-Physical Systems


138. Hallucinations on the Board: Tool-Augmented Evaluation of LLM Chess Commentary


139. A Unified Model for Cross-Domain Clone Detection via Model Merging


140. Patients-like-me: A Variational LM–GNN Framework for Explainable Clinical Prediction


141. Behavioral Skill Reconstruction: Reconstructing Hidden Functionality from LLM Agent Skills


142. Visualizing Graph-to-Answer Mechanism Recovery in Materials-Science Hypothesis Generation


143. TRNet: Topography-Guided Frequency Rectification and Structure-Aware Decoding for Multimodal Paddy Rice Segmentation


144. AgentForge: An Immersive Role-Playing Platform for Learning Agentic Software Engineering


145. LiNC: Lightweight Noise Correction via Per-Sample Trust and Gaussian Mixture Modeling


146. InvFlowFD: Reference-Free and Background-Set-Free Perceptual Music Quality Metric with Flow Matching Inversion


147. Perception Before Reasoning: Dynamic Latent Reasoning for Video Understanding and Question Answering


148. Interpretable Fuzzy Inference for UAV Target Tracking Using Bounding-Box Geometry


149. Out-Of-The-Loop Multi-Fidelity Bayesian Optimization


150. Spatiotemporal Graph Transformer for Traffic Intelligence in Edge Computing


151. Spend Bits Where Queries Look: KV Cache Vector Quantization with Attention-Preserving Transforms


152. FBID: Adaptive Personalized Federated Learning for Robust Out-of-Distribution Attack Detection in IoT Networks


153. An Inline Control Architecture for Language Models in Intelligent Transportation Systems


154. SJEPA: Learning Elegant Latent Dynamics with Hybrid Symbolic-Neural Predictors


155. LaPrune: Controllable Differentiable Sparsity at Million Scale


156. Modality Agreement- and Conflict-Aware Prototype Hypergraph Learning for Multimodal Intent Understanding


157. AgentAntibody: An Adaptive Immune System for Defending LLM Agents against Prompt Injection


158. Recurrent Residual Quantization: A Progressive Multi-Precision Representation for LLMs


159. Beyond the QBER Threshold: A Temporal QBER Based Machine Learning Framework for Multi Attack Detection in BB84 QKD


160. Robust and Personalized Federated Learning for Aircraft-Engine Prognostics under Benign and Adversarial Client Heterogeneity


161. Reconstructing Persistent Worlds from Narratives for Narrative-Grounded Interactive Experiences


162. CheckOne: Lightweight Fault Detection and Mitigation for Vision Transformers


163. EDATracer: An Agentic Framework for Large-Scale EDA Artifact Analysis


164. NuclearDiffusion: Text-to-Image Foundation Models for Learning Nuclear Energy Concepts


165. Lindblad-Inspired Multi-Timescale Reservoir Computing with Separable Rotation and Dissipation


166. A Trust-region Framework for Moment Estimation


167. Governing Execution Risk in Agentic AI Systems: A Trajectory-Guided Framework for Red Teaming


168. AI-driven Multimodal Representation Learning for Latent Mediation Structure Discovery of Socioeconomic Disadvantage, Psychosocial Factors, and Cardiometabolic Multimorbidity: Insights from the All of Us Research Program


169. On Hamming-Lipschitz Type Stability of the Subdominant (Minmax) Ultrametric: Theory and Simple Proofs


170. C$^2$MOE: Consistency and Complementarity-guided Mixture of Experts for Incomplete Multimodal Emotion Learning



172. RAG-Stack: Co-Optimizing RAG Serving Performance and Quality


173. Wiring Beats Blending: What Transfers Between Transformer Sizes – and What Doesn’t


174. Temporal Context Awareness: A Defense Framework Against Multi-turn Manipulation Attacks on Large Language Models


175. TourSynbio-Search: A Large Language Model Driven Agent Framework for Unified Search Method for Protein Engineering


176. AutoProteinEngine: A Large Language Model Driven Agent Framework for Multimodal AutoML in Protein Engineering