전체 AI 논문 - 2026-06-15

1. Towards Direct Latent-Space Synthesis for Parallel Branches in LLM-Agent Workflows


2. Abstracting Cross-Domain Action Sequences into Interpretable Workflows


3. A Temporal Planning Framework for Disruption Aware Dynamic Route Optimization in Heterogeneous Railway Systems


4. VISTA: View-Consistent Self-Verified Training for GUI Grounding


5. StreamMemBench: Streaming Evaluation of Agent Memory for Future-Oriented Assistance


6. Every Eval Ever: A Unifying Schema and Community Repository for AI Evaluation Results


7. Dense Coordinate-List Fine-Tuning Induces a Controllable Interference Surface in Vision-Language Models


8. From Chatbot to Digital Colleague: The Paradigm Shift Toward Persistent Autonomous AI


9. When the Tool Decides: LLM Agents Defer Blindly to Graph Neural Network Tools, and Stronger Backbones Defer More


10. GitOfThoughts: Version-Controlled Reasoning and Agent Memory You Can Replay, Diff, and Merge



12. CSPO: Constraint-Sensitive Policy Optimization for Safe Reinforcement Learning


13. Communication Policy Evolution for Proactive LLM Agents


14. HarnessX: A Composable, Adaptive, and Evolvable Agent Harness Foundry


15. AFFORDANCE20Q: Evaluating Affordance Reasoning from Physical Properties


16. SkillAudit: Ground-Truth-Free Skill Evolution via Paired Trajectory Auditing


17. Closing the Reflection Gap: A Free Calibration Bonus for Agentic RL


18. When Should Agent Trust Be Conditional? Characterizing and Attacking Skill-Conditional Reputation in Agent Swarms


19. VeriGeo: Controllable Geometry Question Generation with Numerical and Analytical Verification


20. FactoryLLM: A Safe and Open-Source AI Playground for Evaluating LLMs in Smart Factories


21. Applicability Condition Extraction for Therapeutic Drug-Disease Relations


22. Formalizing Numerical Analysis: An Agent Pipeline and Quality Audit Beyond Kernel Acceptance


23. Minim: Privacy-Aware Minimal View for Agents via Trusted Local Sanitization


24. Adversarial Concept Search: Predicting Compositional Errors From Feature Geometry


25. Sorries Are Not the Hard Part: An Expert-Review Case Study of a Semi-Autonomous Formalization


26. A Multi-Agent AI System for Automated High School Transcript Processing: Collaborative Document Analysis at Scale


27. Capability Minimization as a Safety Primitive: Risk-Aware Causal Gating for Least-Privilege LLM Agents


28. Hyperdimensional computing for structured querying on tabular data embeddings


29. Poker Arena: Multi-Axis Profiling of Strategic Reasoning and Memory in LLMs


30. MA-ProofBench: A Two-Tiered Evaluation of LLMs for Theorem Proving in Mathematical Analysis



32. When Sample Selection Bias Precipitates Model Collapse


33. TwinBI: An Agentic Digital Twin for Efficient Augmented Interactions with Business Intelligence Dashboards


34. YeasierAgent: Agentic Social Sandbox as a Canvas for Intent-Driven Creation of Platform-Agnostic Symbiotic Agent-Native Applications


35. Refusal Beyond a Single Direction: A Preliminary Comparison of Diff-in-Means and INLP


36. WorkBench Revisited: Workplace Agents Two Years On


37. Hybrid Open-Ended Tri-Evolution Makes Better Deep Researcher


38. Orchestra-o1: Omnimodal Agent Orchestration


39. History of the Muddy Children Puzzle


40. UP-NRPA: User Portrait based Nested Rollout Policy Adaptation for Planning with Large Language Models in Goal-oriented Dialogue Systems


41. A Deep Reinforcement Learning (DRL)-Based Transformer Method for Solving the Open Shop Scheduling Problem


42. ClinHallu: A Benchmark for Diagnosing Stage-Wise Hallucinations in Medical MLLM Reasoning


43. Learning Coordinated Preference for Multi-Objective Multi-Agent Reinforcement Learning


44. Flood and Harvest: The Provable Necessity of Trivia for Generating Valuable Mathematics via the Lens of Language Generation in the Limit


45. CottonLeafVision: An Explainable and Robust Deep Learning Framework for Cotton Leaf Disease Classification


46. Giving AI a Headache: Acoustic Adversarial Attacks to Computer Vision Applications


47. Listening with Attention: Entropy-Guided Explainability for Transformer-Based Audio Models


48. From Self-Supervised Speech Models to Mixture-of-Experts for Robust Anti-Spoofing


49. When Good Verifiers Go Bad: Self-Improving VLMs Can Regress on New Tasks


50. Moonlight in Latent Space: Chirality and Structural Correspondence Between Beethoven’s Op. 27 No. 2 and Machine Learning Mechanisms


51. Expert-Driven Survival Machines: Improving Stratification and Interpretability in Multiple Clinical Cohorts


52. A Comparative Study of Deep Learning Architectures for Multi-Horizon Behavioural Forecasting for Mobile Health


53. Regulating the Machine Contributor: Governance and Policy Alignment in Open Source


54. AudioDER: A Deduplication-Enhanced Reasoning Dataset for Post-Training Large Audio-Language Models


55. When Errors Become Narratives: A Longitudinal Taxonomy of Silent Failures in a Production LLM Agent Runtime


56. Sensitivity Shaping for Latent Modeling


57. CARE: Controlling LLM-Generated Policies through Auditable Review of Evidence in Scientific Experimentation


58. SIMMER: Benchmarking Latent Failures in LLM Executable Planning with a World Model


59. Regional Climate Model Emulation with Diffusion Approaches: What is the Added Value of Generative Machine Learning?


60. Rethinking Global Average Pooling: Your Classifier Is Secretly a Multi-Instance Learner


61. TRACE: Trajectory-Routed Causal Memory for Delayed-Evidence Visuomotor Imitation


62. From Shield to Target: Denial-of-Service Attacks on LLM-Based Agent Guardrails


63. Securing the Future of IoMT in the Post-Quantum Era: An Edge-Native Federated Learning Approach


64. Fodor and Pylyshyn’s Systematicity Challenge Still Stands


65. A Fixed-Point Neural Operator for Size- and Functional-Transferable Hamiltonian Prediction


66. The Perceived Fragility of Explanations in Audio Models: Manipulation of Attribution with Unchanged Predictions


67. MoDiCoL: A Modular Diagnostic Continual Learning Dataset for Robust Speech Recognition


68. tap: A File-Based Protocol for Heterogeneous LLM Agent Collaboration


69. CADET: Physics-Grounded Causal Auditing and Training-Free Deconfounding of End-to-End Driving Planners


70. Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack


71. Learning to Hear Hesitation: Continual Learning for Disfluency-Aware ASR


72. Discovery under Hypothesis Redundancy: A Geometric Theory of Discovery Bottlenecks


73. Elastic Queries Reinforcement Learning: Self-Aware Policy Execution for VLA Models


74. No Accidental Software Agent First Canonical Code for Human Code Entropy Reduction and 30 to 500 times Lower Frontier Model Requirements


75. PLAIground: SLO-Driven Runtime Model Selection for Compound AI Systems in the Edge-Cloud-Space Continuum


76. Design Methodology and Performance Trade-offs Management for Distributed and Compound AI Systems


77. Squeeze-Release: Iterative Pruning with Exact Structural Minimization


78. I’m Sorry Driver, I’m Afraid I Can’t Do That: Appraising the Safety of LLMs within Automotive Contexts


79. Achieving Precise Text-To-Cypher Via Grounded Knowledge Graph Data Generation


80. Transforming Shape Schemas with Composable Property-Graph Queries (Extended Version)


81. Thinking Outside the [Chat]Box: Bridging Computer Science and Industrial Design for Cognitive-Inclusive Generative AI


82. Pix2Pix-Hybrid: Structure-Guided Conditional Synthesis of Hajj Crowd Images with Multi-Channel Conditioning and Weak Attribute Supervision


83. AgentCyberRange: Benchmarking Frontier AI Systems in Realistic Cyber Ranges



85. DIFF-ERO: A Conformance-Aware Loss for Deep Learning in Process Mining


86. Robust Fall Recovery for Armless Bipedal-Wheeled Robots Via Force-Guided Learning


87. ChronoID: Infusing Explicit Temporal Signals into Semantic IDs for Generative Recommendation


88. When and How Severely: Scenario-Specific Safety Envelopes for Driving VLAs


89. Selective Agentic Recovery for UAV Autonomy with a Persistent Mission Runtime


90. Universal Manipulation Exoskeleton: Learning Compliant Whole-body Policies with Real-time Torque Feedback


91. From Prompts to Responses: Dual-Sided Data Leakage and Defense in Split Large Language Models


92. MeEvo: Metacognitive Evolution Combined with Natural Evolution for Automatic Heuristic Design


93. OdysSim: Building Foundation Models for Human Behavior Simulation


94. Robustness without Wrinkles: Parallel Simulation and Robust MPC for Certified Deformable Manipulation


95. Learning Urban Access Costs from Origin-Destination Flows via Inverse Optimal Transport


96. Learning High Coverage Discriminative Parsimonious Rulesets


97. Implicit Reasoning for Large Language Model-based Generative Recommendation


98. Spatio-Temporal Audio Language Modeling for Dynamic Sound Sources


99. Conditioning Matters: Stabilizing Inversion and Attention in Diffusion Image Editing


100. Recovering Stranded Discrimination in Knowledge Tracing: Per-Item Bias Correction via Empirical-Bayes Shrinkage


101. FAConformer: Frequency-Aware Convolutional Transformer for Auditory Attention Decoding


102. A Two-Stage Statistical Framework for Evaluating Associative Interference in Large Language Models


103. Numbers Already Carry Their Own Embeddings


104. FEMOT: Multi-Object Tracking using Frame and Event Cameras


105. Clay-CNN Hybrids: Leveraging Geo-Foundational Models as Auxiliary Context for Landslide Detection


106. Rethinking Backdoor Adversarial Unlearning through the Lens of Catastrophic Forgetting in Continual Learning


107. Knowledge Graph Enhanced Memory-Augmented Retrieval for Long Context Modeling


108. Same-Origin Policy for Agentic Browsers


109. Hidden in Plain Sight: Benchmarking Agent Safety Against Decomposition Attacks with DECOMPBENCH


110. Mask, Sample, Revise: A Revisable CTMC Inference Stack for Guided Discrete Flow Matching Text-to-Speech


111. STREAM: Multi-Tier LLM Inference Middleware with Dual-Channel HPC Token Streaming


112. The Silent Cost of Artificial Intelligence Assistance: A Theory of Autonomy Surrender, the Recovery Mechanism, and the Restoration of Human Agency


113. GMN4AD: Graph Matching Network for Alzheimer’s Disease Diagnosis with Test-Time Domain Adaptation using Multi-centered Structure Magnetic Resonance Imaging


114. SANA: What Matters for QA Agents over Massive Data Lakes?


115. HiLo-Token: Input-Adaptive High-Low Frequency Token Compression for Efficient Image Editing


116. How do Self-Supervised Remote Sensing Vision Models Transfer to Downstream Tasks?


117. Gefen: Optimized Stochastic Optimizer


118. Crypto x AI, AI x Crypto: A Survey


119. Mirage Probes: How Vision Models Fake Visual Understanding


120. SuperThoughts: Reasoning Tokens in Superposition


121. Mood-Aware Music Recommendation: Integrating User Affective Signals into Ranking Systems


122. SpheriCity: Designing Trustworthy Conversational AI for Sustainability Decision Support


123. Explaining RhythmFormer: A Systematic XAI Analysis of Periodic Sparse Attention for Remote Photoplethysmography


124. When Plausible Is Not Realistic: Evaluating Human Mobility in LLM-Based Urban Simulation


125. Safety-Contract Graph Multi-Agent Reinforcement Learning for Autonomous Network Security Response


126. AI can help scientists publish less


127. Aligning Quantum Operators with Large Language Models


128. A Benchmark and Framework for Evaluating Next Action Predictions in Spreadsheets


129. An integrated interpretable control effectiveness learning and nonlinear control allocation methodology for overactuated aircrafts


130. CineOrchestra: Unified Entity-Centric Conditioning for Cinematic Video Generation


131. Beyond LoRA: Is Sparsity-Induced Adaptation Better?


132. SEVRA-BENCH: Social Engineering of Vulnerabilities in Review Agents


133. Position: Align AI to Our Aspirations, Not Our Flaws


134. The Weight Norm Sets the Grokking Timescale: A Causal Delay Law


135. A fully GPU-based workflow for building physics emulators of hypersonic flows


136. A Virtuous AI is an Existential Risk


137. FreoStream:Enhancing Stream Guardrails via Future-Aware Reasoning and Safety-Aligned Optimization


138. VHDLSuite: Unified Pipeline for LLM VHDL Generation with Data Synthesis and Evaluation


139. Morphology-Aware Sample Assignment: Overcoming IoU Insensitivity for Surface Defect Detection


140. CisTransCell: Single-Cell Perturbation Prediction via Gene Function, Regulatory Control, and Cellular Context


141. HierSVA: A Data Synthesis Pipeline, Dataset, and Benchmark for LLM-Driven Hierarchical Hardware Formal Verification


142. Can Editing 1 Neuron Fix Repetition Loops in LLMs?


143. Position: AI Must Become Planet-Centered, Not Just Human-Centered


144. Active Inference for Adaptive Traffic Signal Control in Noisy Nonstationary IoT Environments


145. Korzhinskii-Net: Physics-Informed Neural Network for Sub-Surface Mineral Prospectivity Modelling


146. Efficient Temporal Modeling for Mobile Sleep Staging via Lightweight Random Attention


147. An Agentic Retrieval Framework for Autonomous Context-Aware Data Quality Assessment


148. The Coin Flip Judge? Reliability and Bias in LLM-as-a-Judge Evaluation


149. Cross-Dataset Bloom Question Classification: Supervised Models and Prompted LLMs


150. Simplex-Constrained Sparse Bagging: Transitioning from Uniform Priors to Sparse Posteriors in Ensemble Learning


151. GAGPO: Generalized Advantage Grouped Policy Optimization