전체 AI 논문 - 2026-07-27

1. Explainable Reinforcement Learning for assisting Air Traffic Controllers


2. The Regression Tax: Decomposing Why Skills Help and Hurt LLM Agents


3. TRACE-ROUTER: Task-Consistent and Adaptive Online Routing for Agentic AI


4. Dynamic Capability Scoping for Enterprise AI Agents: A Synthetic Dataset and Three-Source Permission Architecture


5. SceneActBench: Can Agents Act on the 3D Scenes They See?


6. Agentic Root Cause Analysis through Evidence-Grounded Reasoning


7. IDEAgent: Agentic Quality-Diversity Search for Research Idea Generation


8. Do Agent Benchmarks Measure Capability? Protocol Validity in the Age of Agentic AI


9. Learning Structural Convergence: A Neuro-Symbolic Benchmark for Temporal Reasoning


10. A Roadmap to Impactful Pluralistic Alignment Research


11. AI4PLE: A Methodology for Integrating AI into Product Line Engineering


12. Deconstructing Off-Policy Ratios: Entropy-Scaled Trust Regions for Asynchronous Reinforcement Learning


13. Learning on the Job: Continual Learning from Deployment Feedback for Frozen-Weights Agents


14. Industrial Tokenization for LLM-Based Health Intelligence: A Federated Architecture for Industrial Evidence Integration


15. Reasoning Denoiser: Denoising Reasoning Traces for Hallucination Detection in Large Reasoning Models


16. Nanbeige4.2-3B: Unlocking Agentic Capabilities in a Compact Mode


17. Zero-Shot Mission-Level Evaluation for Aerial MLLM Agents


18. Learning as Reasoning Unfolds: Progressive Rollout Allocation for Efficient Reinforcement Learning


19. Semiotic logical hexagon theory for LLM logical reasoning


20. TRW: TRACE-RealWorld—An Auditable Consistency Contract for World Models as Materialized Views


21. Multi-Agent System-driven Digital Twins for predictive maintenance: architectures, technologies and open research challenges


22. When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual Auditing of Frozen Locomotion Policies


23. DAGForge: Auditable Causal DAG Authoring with Biomedical Literature


24. QLPO: Quadrant-weighted Sampling for Length-aware Policy Optimization


25. From Seasonality to Semantics: Benchmarking a Hybrid Probabilistic Forecasting System for Roadblocks in Bolivia


26. What AI Red-Team Evaluations Can and Cannot Prove


27. Persistent Computational State: A Session-Centric Runtime for Generative World Models


28. Defining AI-Native Systems: Autonomy as Revision Authority


29. Wavelet Phase Diffusion for Structurally and Semantically Consistent Sim-to-Real Translation


30. Do Modules Stay in Their Lane? Role Drift in Compound LLM Systems


31. Discrete Action Space as a Prerequisite for GRPO Convergence in Small-Model Continuous Control


32. Trajectory-Aware Retrieval Agents for Temporal Decision- Making


33. FBLayout: Optimizing Memory Layout for Efficient LLM Finetuning on Mobile GPUs


34. From Profiles to Steering Vectors: Global Sparse Priors and Local Semantic Calibration for Personalized Text Generation


35. LeafData: An Agentic System for Data Migration


36. Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models


37. Lost in Context: Addressing Context Anxiety in Large Language Models


38. FrED: External Data Influence Estimation via Domain Knowledge Graph Grounding


39. Household Movement Detection in Mixed-Format Occupancy Data Using LLM-Based Entity Resolution


40. The Hard Decision Layer: Evidence for Committed Inference in Transformers


41. Procedural Knowledge Is Not Low-Rank: Why LoRA Fails to Internalize Multi-Step Procedures


42. SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text


43. Coupled Hierarchical Search over Topology and Execution for Agentic Workflow Synthesis


44. Spectral Flow Certificates for Depth-Aware Long-Range Propagation in Graph Neural Networks


45. TILT: Improving Compositional Generation in Diffusion Models with a Model-Intrinsic Reward


46. AgentKVShift: Efficient KV Cache Reuse for Agentic Memory Systems


47. Transferable Latency Prediction for Fast LLM Screening on Heterogeneous Edge Devices


48. From Frame-Level Recognition to Event-Level Confirmation: Repair Traces and Runtime Failure Analysis of Public-Space Gesture Interaction


49. Securing Multimodal AI through Internal Information Decomposition


50. Risk Is Not the Target: A Monotonic Framework for Evaluating Wildfire Operational Risk Signals


51. FlowEvo: Self-Evolving Agents through the Co-Evolution of Workflows and Executable Skills


52. SM4RT: Learning Structured Motion Geometry for 4D Reconstruction


53. Quantum Spectral Model: Data Reuploading with Input-Conditioned Frequency Support


54. Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science


55. CausalForge: A Formally Grounded, Self-Improving Agentic Framework for Automated Research in Causal Inference


56. \k{appa}-LoRA: Condition Numbers Reveal Which LoRA Matrices Worth Updating


57. MineValiCoder: Reliable Code Generation with Test Case Quality Mining and Bipartite Graph-Based Mutual Validation


58. Learning to Prepare Molecular Ground States with Transformer Models


59. Beyond Perspectives: A Trio-Ethnography of Interpretation Evolution in LLM-Supported Programming Education


60. Phylogenetic signal in marine mammal and bird vocalizations captured by audio foundation models: the limited benefit of domain-specific pretraining


61. Hyperball May Not Be a Free Lunch


62. Robot Learning to Communicate through Projected Visual Abstractions


63. Unboxing Diffusion Models for the Arts: Interactive Model Bending and Practice-Based Explainability


64. PRIMS: Physics-guided Representation for Fluid Identification in Multimodal Sensing


65. A Self-Calibrating Agentic AI Framework for Autonomous Edge Resource Allocation


66. HiKV: Hierarchical Importance-Aware KV Cache with Hardware Acceleration for LLM Decoding


67. Interior interpretability with attention rollout: contraction and propagation profiles in Transformers


68. Indexing: the Beginning and the End


69. SiPhy: Single-Image Physical Property Reasoning


70. Time-Reversed Imaging: A Multimodal Benchmark and Framework for Reconstructing Past Human-Environment Interactions


71. Teachy Mini: Development and Preliminary Evaluation of a Knowledge-Based Generative Social Robot for Higher Education


72. Cross-Tokenizer On-Policy Distillation via Byte-Prefix Marginalization


73. Towards Trustworthy and Cost-Efficient Data Integration: From Naïve RAG to Agentic RAG


74. Neptuna: A Comprehensive Machine Learning Framework for Benchmarking Complex Multiphase Flows


75. Explicit Iteration Complexity of Exact Data-Driven Inverse Optimization for Integer Linear Programs


76. IFCLoRA: Topology-Aware Rank Allocation for Parameter-Efficient Fine-Tuning


77. Optimization of time-consuming experimental conditions using pseudo-experimental data guided by adaptive polynomial regression


78. TRaM-VSR: Importance-Aware Token Routing and Merging for One-Step Diffusion Video Super-Resolution


79. Why Large Language Models and Humans Converge and Diverge in Evaluating Creativity


80. Filling Before Advancing: Capability-Gap-Driven Post-Training for Scenario-Specialized Remote Sensing MLLMs


81. From Isolated Tasks to Structured Capabilities: A Multilayer Taxonomy for Large Language Models


82. Learning Spatiotemporal Decision Priors for Efficient Path Planning under Partial Observability


83. DBA-Bench: A Production-Fidelity Benchmark for LLM-Based Database Operations Agents


84. dRAE: Representation Autoencoder with Hyper-Spherical Codes


85. TriGlue: a Biology-Inspired Generative Model for Generating Molecular Glue-Induced Ternary Complex


86. CARDIAG: A Dense Segment Classification Benchmark of Deep Learning Architectures for Coronary Angiography


87. One Hand Watches The Other: Dynamic Multi-Agent Cooperation for Sample-Efficient Bimanual Manipulation in Dynamic Environments


88. Benchmarking Text-to-SQL under Role-Based Access Control


89. MEUSLI: a Multilingual Projector for LLM-based ASR and Beyond


90. A Leakage-Free Stacked Ensemble Method for Multiclass Classification


91. FSE: Continual Learning for Named Entity Recognition by Fast-Slow Experts


92. Benchmarking Fine-tuning and Retrieval Strategies for a Multimodal Language Model on the NRC Reactor Operator Licensing Examination


93. Multiplicity of Stable Attractors in Disordered Neural Models


94. CEL: Comprehensive Counterfactual Explanations Library and Benchmark


95. Sparse by Command: Task-Conditional Compute Skipping for Multi-Task Inference Accelerators


96. Agent Security Needs Redefinition through a Holistic Framework


97. EVL-MCoT: Enhanced Vision-Language Multi-CoT for Harmful Meme Detection


98. Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning


99. Practical Graph Optimisation and AI-Driven Models for Active Directory Security Hardening


100. From Perturbation Correction to Geometry-Aware Sampling: Sharpness-Guided Equilibrium Sampling for Balanced Flat Minima in Long-Tailed Learning


101. Unified Static-Dynamic Pruning for Efficient LLM Inference


102. J-CoT: Chain-of-Thought in J-Space


103. Teaching LLMs to Self-Evolve: Cultivating Core Meta-Skills with Reinforcement Learning


104. TextSLIP: Text Self-Supervised CLIP for Medical Report Generation


105. ACME: A Multi-Cultural, Multi-Embodiment Social-Navigation Dataset


106. MA-DAR: Manifold-Aligned Dynamic Adaptive Routing for Continual Temporal Knowledge Graph Reasoning


107. Generalized Neural Operator for Parametric and Boundary-Value Problems


108. Interventional Score Geometry for Causal Inference


109. ISPCloak: Weaponizing ISP for Optimization-Free Physical Camouflage against Deepfake Detectors


110. Data Quality over Capacity: Internalizing Documents into LoRA Adapters for Closed-Book QA


111. LeAct: Learning to Reason from Expert Actions


112. SCALE: Self-Supervised Constraint-Aware Layout GEneration for Local P&R DRV Fixing at Advanced Nodes


113. ToolGuardian: Declarative Security for AI Agent-Tool Interactions


114. Probing Speaker Identity Sensitivity in Audio Deepfake Detectors


115. MosaicJoin: Compact Semantic Sketches for Value-Level Join Discovery


116. Khondo: A Multimodal Benchmark for Document Packet Splitting of Bangla Forms


117. Graph-Theoretic Neural Network Fragmentation with Covariant Direct Molecular Force Learning: Enabling Coupled-Cluster Accuracy AIMD for Fluxional Systems


118. AI-Integrated Scientific Inquiry: A Practice-Centered Vision for Science Education


119. Probing Latent Colombian Identity Inferences in Qwen2.5-7B with Natural Language Autoencoders


120. Every Model Cheats: Prompt-Level Mitigation of Cheating on Offensive Cyber Tasks


121. Co-design of LLM-based preference agents: participation may drive overtrust


122. Deep Sigma Point Processes for RCS Modeling in Spaceborne SAR Imagery


123. A Defense of the Quadratic Model


124. Enhancing SLMs for Sustainable Code Optimization in Radio-Astronomy


125. Self-Poisoning in Adaptive Out-of-Distribution Detection: A Sharp-Threshold Theory and Certified Label-Free Calibration


126. Pixels for Programs? A Cross-Provider Case Study of Input-Token Accounting for Source Code as Text and Images


127. Ordered Action Tokens for Visuomotor Policy Learning


128. Generative and multimodal AI for materials prediction and design: Progress, challenges, and perspectives


129. Cross-Model LLM Code Review: Should you use Claude to review Codex or vice versa?


130. A Drift Stable Quantum Federated Learning for Intelligent Services


131. Multi-Horizon Consistency as Geometry: When Latent Dynamics Contract, and When They Do Not


132. Tool-Guided Retrieval-Augmented Repair for Securing LLM-Generated C Code


133. Measuring the Dependency Gap: Diagnosing Inter-Column Fidelity in Tabular Generative Models


134. MotifRole-Diff: Risk-Optimal Role-Aware Corruption for Masked Molecular Graph Diffusion


135. On the Depth Scalability of Logic Gate Networks


136. Local Synaptic Rules Can Implement a SIGReg Gradient Without Backpropagation


137. Adversarial Style Optimization: Enhancing VLM Jailbreaks by GRPO-based Stylistic Triggers Optimization


138. A Systematic Survey on Image Description Techniques for STEM Domains


139. From Obligation to Specification: A Survey on Validating EU AI Act Requirements in RE


140. Analyzing Middle School Students’ Dialogue and Behaviors during Collaborative AI Chatbot Development Using Ordered Network Analysis


141. Decoupled Attention Fusion: Accelerating RAG with Efficient KV Cache Reuse


142. Control panels to clarify user intent with Large Language Models


143. Do emulated quantum circuits change what CNNs look at? Performance and explainability comparison in medical image classification


144. AI-Driven Surrogate Models for Predicting Electrode-Scale Discharge Behavior in Lithium-Ion Batteries