전체 AI 논문 - 2026-06-11

1. Nonslop: A Gamified Experiment in Human-AI Collaborative Writing


2. PROJECTMEM: A Local-First, Event-Sourced Memory and Judgment Layer for AI Coding Agents


3. A Five-Plane Reference Architecture for Runtime Governance of Production AI Agents


4. The Impossibility of Eliciting Latent Knowledge


5. Towards Responsibly Non-Compliant Machines


6. IntElicit: Eliciting and Assessing Contextualized Creativity via Dialogue Policy Optimization


7. Automating Geometry-Intensive Compliance Checking in BIM: Graph-Based Semantic Reasoning Framework


8. A Lightweight Multi-Agent Framework for Automated Concrete Barrier Design


9. Existential Indifference: Self-Nonpreservation as a Necessary Architectural Condition for Aligned Superintelligence (or: The Suicidal AI)


10. Human-Enhanced Loop Modeling (HELM): Agent-Based Finite Element Modeling of Concrete Bridge Barriers


11. MODF-SIR: A Multi-agent Omni-modal Distilled Framework for Social Intelligence Reasoning


12. The Art of Interrogation: Consistency Amplifies Factuality in Spatial Reasoning


13. Embodied-BenchClaw: An Autonomous Multi-Agent System for Embodied Spatial Intelligence Benchmark Construction


14. AutoMine Solution for AV2 2026 Scenario Mining Challenge


15. StatefulDiscovery: Evidence-Calibrated Claim Formation in Open-Ended Scientific Discovery


16. Skill-Augmented AI Agents for Medical Research Analysis: An Exploratory Multi-Model Human Evaluation in an NSCLC Transcriptomic Biomarker Task


17. Toward Trustworthy AI: Multi-Target Adversarial Attacks and Robust Defenses for Continuous Data Summarization


18. SVoT: State-aware Visualization-of-Thought for Spatial Reasoning via Reinforcement Learning


19. When Do Data-Driven Systems Exhibit the Capability to Infer?


20. Mind the Perspective: Let’s Reason Recursively for Theory of Mind


21. Organize then Retrieve: Hierarchical Memory Navigation for Efficient Agents


22. Lung-R1: A Knowledge Graph-Guided LLM for Pulmonary Diagnostic Reasoning



24. TouchThinker: Scaling Tactile Commonsense Reasoning to the Open World with Large-scale Data and Action-aware Representation


25. Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning


26. HERO: Hindsight-Enhanced Reflection from Environment Observations for Agentic Self-Distillation


27. SkillJuror: Measuring How Agent Skill Organization Changes Runtime Behavior


28. MoCA-Agent: A Market-of-Claims Code Agent for Financial and Numerical Reasoning


29. Search Discipline for Long-Horizon Research Agents


30. Forecasting Future Behavior as a Learning Task


31. INFRAMIND: Infrastructure-Aware Multi-Agent Orchestration


32. Automated Mediator for Human Negotiation: Pre-Mediation via a Structured LLM Pipeline


33. Knowing When to Ask: Self-Gated Clarification for Hierarchical Language Agents


34. Can AI Agents Synthesize Scientific Conclusions?


35. Position: Hippocampal Explicit Memory Is the Cornerstone for AGI


36. From Explicit Elements to Implicit Intent: A Predefined Library for Auditable Behavioral Inference


37. Reroute, Don’t Remove: Recoverable Visual Token Routing for Vision-Language Models


38. FACTR 2: Learning External Force Sensing for Commodity Robot Arms Improves Policy Learning


39. DIRECT: When and Where Should You Allocate Test-Time Compute in Embodied Planners?


40. Redesign Mixture-of-Experts Routers with Manifold Power Iteration


41. System Report for CCL25-Eval Task 5: New Dataset and LoRA-Fine-Tuned Qwen2.5


42. TAHOE: Text-to-SQL with Automated Hint Optimization from Experience


43. ATLAS: Active Theory Learning for Automated Science


44. APPO: Agentic Procedural Policy Optimization


45. SPEA2$^+$: Improved Density Estimation in SPEA2 with Provable Runtime Guarantees


46. Illumination-Robust Camera-Based Heart-Rate Estimation for Physiological Sensing in Robots


47. Ambient Diffusion Policy: Imitation Learning from Suboptimal Data in Robotics


48. Latent World Recovery for Multimodal Learning with Missing Modalities


49. CHORUS: Decentralized Multi-Embodiment Collaboration with One VLA Policy


50. Atlas H&E-TME: Scalable AI-Based Tissue Profiling at Expert Pathologist-Level Accuracy


51. ALIGNBEAM : Inference-Time Alignment Transfer via Cross-Vocabulary Logit Mixing


52. Harness In-Context Operator Learning with Chain of Operators


53. Natural-Language Temporal Grounding in Hour-Long Videos is a Search Problem: A Benchmark and Empirical Decomposition


54. The Standard Interpretable Model: A general theory of interpretable machine learning to deductively design interpretable methods using Lagrangian mechanics


55. SpikeDecoder: Realizing the GPT Architecture with Spiking Neural Networks


56. CCKS: Consensus-based Communication and Knowledge Sharing


57. Mathematical perspective on genetic algorithms with optimization guided operators



59. Using Explainability as a Training-Time Reliability Signal for Efficient ECG Classification


60. Reinforcement Learning Disrupts Gradient-Based Adversarial Optimization


61. DiffCold: A Diffusion-based Generative Model for Cold-Start Item Recommendation


62. VIA-SD: Verification via Intra-Model Routing for Speculative Decoding


63. Multi-Rate Mixture of Experts for Accelerating Liquid Neural Network Training


64. Rule Taxonomy and Evolution in AI IDEs: A Mining and Survey Study


65. Adapting Prithvi-EO for Fallow Detection for Food-Water Nexus: ViT-Adapter Necks and Parameter-Efficient Backbone tuning of Geospatial Foundation Model


66. Making Foresight Actionable: Repurposing Representation Alignment in World Action Models



68. Implicit Neural Representations of Individual Behavior


69. Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application


70. OpenMedReason: Scientific Reasoning Supervision for Medical Vision-Language Models


71. nD-RoPE: A Generalized RoPE for n-Dimensional Position Embedding


72. Unstable Features, Reproducible Subspaces: Understanding Seed Dependence in Sparse Autoencoders


73. Soft-Prompt Tuning for Fair and Efficient LLM Benchmark Evaluation


74. Augmenting Molecular Language Models with Local $n$-gram Memory


75. Bridging the Morphology Gap: Adapting VLA Models to Dexterous Manipulation via Intent-Conditioned Fine-Tuning


76. MSUE: Multi-Modal Soccer Understanding Expert


77. Non-frontal face recognition using GANs and memristor-based classifiers


78. “That’s AI Slop, You Bot!” Studying Accusations, Evidence, and Credibility in Online Discourse Towards LLM-Generated Comments


79. On the Limits of LLM-as-Judge for Scientific Novelty Assessment


80. Metadata-Aware Multi-Prompt Reasoning for Zero-Shot Accident Understanding


81. Runtime Enforcement of Hybrid System Properties


82. Generalization Hacking: Models Can Game Reinforcement Learning by Preventing Behavioral Generalization


83. Tabular Foundation Models for Clinical Survival Analysis via Survival-Aware Adaptation


84. Time-Series Foundation Model Embeddings for Remaining Useful Life Estimation


85. Exploration Structure in LLM Agents for Multi-File Change Localization


86. Categorical Prior Lock-in: Why In-Context Learning Fails for Structured Data


87. Frozen Multimodal Embeddings for Personality and Cognitive Ability Assessment in Asynchronous Video Interviews


88. Toward Generalist Autonomous Research via Hypothesis-Tree Refinement


89. Lung-SRAD: Spectral-Aware Regularized Audio DASS with Dual-Axis Patch-Mix Contrastive Learning for Respiratory Sound Classification


90. Characterizing Software Aging in GPU-Based LLM Serving Systems


91. Quality Adaptive Angular Margin Learning for Respiratory Sound Classification


92. DuoBench: A Reproducible Benchmark for Bimanual Manipulation in Simulation and the Real World


93. Beyond representational alignment with brain-guided language models for robust reasoning


94. Task-Aligned Stability Analysis of Vision-Language Models for Autonomous Driving Hazard Detection


95. Agents All the Way Down; A Methodology for Building Custom AI Agents from Substrate to Production


96. Fine-tuning Multi-modal LLMs with ART: Art-based Reinforcement Training


97. Task-Aware Structured Memory for Dynamic Multi-modal In-Context Learning


98. LASA: A Weak Supervision Method for Open-Vocabulary Scene Sketch Semantic Segmentation


99. Towards Data-free and Training-free Compression for Speech Foundation Models Using Parameter Clustering


100. Designing AI-Supported Focus Groups: A Role x Modality Playbook


101. From Uniform to Learned Graph Priors: Diffusion for Structure Discovery


102. Feature-Aligned Speech Watermarking for Robustness to Reconstruction Distortions


103. Grammar-Constrained Decoding Can Jailbreak LLMs into Generating Malicious Code


104. WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning


105. Sparsified Kolmogorov-Arnold Networks for Interpretable Quantum State Tomography


106. TextHOI-3D: Text-to-3D Hand-Object Interaction via Discrete Multi-View Generation and Joint Mesh Optimization


107. Multimodal Ordinal Modeling of Alzheimer’s Disease Severity Using Structural MRI and Clinical Data


108. AI4Land: Scalable Deep Learning for Global High-Resolution Land Use Reconstruction


109. MultiToP: Learning to Patch Visual Tokens to Mitigate Hallucinations in Video Large Multimodal Models


110. What Limits Does Quantization Place on Dense Top-$k$ Retrieval? A Theoretical Study


111. Blind Dexterous Grasping via Real2Sim2Real Tactile Policy Learning


112. Fast Speech Foundation Model Distillation Using Interleaved Stacking


113. Automated Creativity Evaluation of Language Models Across Open-Ended Tasks


114. AnchorEdit: Maintaining Temporal Consistency in Multi-turn Image Editing via Causal Memory


115. From Prompts to Tokens: Internalizing Causal Supervision in Vision-Language Model for Multi-Image Causal Reasoning


116. Hey Chat, Can You Teach Me? Structuring Socratic Dialogue for Human Learning in the Wild


117. Multi-View In-Cabin Monitoring System for Public Transport Vehicles


118. ICA Lens: Interpreting Language Models Without Training Another Dictionary


119. Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning


120. Substrate Asymmetry in User-Side Memory: A Diagnostic Framework


121. MedCTA: A Benchmark for Clinical Tool Agents


122. T2S: A Rehearsal-Based Approach for Extraction-Resistant Model Watermarking


123. Noise-Aware Framework for Correcting Corrupted Labels


124. Goal-Autopilot: A Verifiable Anti-Fabrication Firewall for Unattended Long-Horizon Agents


125. Layer-Isolated Evaluation: Gating the Deterministic Scaffold of a Production LLM Agent with a No-LLM, Regression-Locked Test Harness


126. Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning


127. Can Open-Source LLM Agents Replace Static Application Security Testing Tools? An Empirical Assessment


128. Runtime Skill Audit: Targeted Runtime Probing for Agent Skill Security


129. ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation


130. Sparse probes and murky physics: a case study of interpretability challenges in a foundation model for continuum dynamics


131. TAROT: Task-Adaptive Refinement of LLM-prior Graphs for Few-shot Tabular Learning


132. Are LLMs Bad at Moral Reasoning?


133. Sovereign Assurance Boundary: Certificate-Bound Admission for Agentic Infrastructure


134. LUCID: Learning Embodiment-Agnostic Intent Models from Unstructured Human Videos for Scalable Dexterous Robot Skill Acquisition


135. When Context Returns: Toward Robust Internalization in On-Policy Distillation


136. Information-Theoretic Decomposition for Multimodal Interaction Learning


137. Physics-Distilled Neural Network enabled by Large Language Models for Manufacturing Process-Property Predictive Modeling


138. Model-Based and Data-Driven Hierarchical Control and Topology Co-Design for Robust Networked Systems


139. AVIS: Adaptive Test-Time Scaling for Vision-Language Models


140. ConsistencyPlanner: Real-time Planning with Fast-Sampling Consistency Models


141. LLMs+Graphs: Toward Graph-Native, Synergistic AI Systems


142. Privacy-Preserving Federated Autoencoder for ECG Anomaly Detection on Edge Devices


143. End-to-End Machine Learning for Depressive State Classification via EEG and fNIRS


144. Pretrained self-supervised speech models can recognize unseen consonants


145. AI Researchers Must Help Lead Arms Control to Mitigate Military AI Risks


146. ISE: An Execution-Grounded Recipe for Multi-Turn OS-Agent Trajectories


147. SirenFNO: Efficient and Full Frequency Learning of Fourier Neural Operators


148. On the Study of Biometric Spoofing Detection using Deep Learning


149. When Roleplaying, Do Models Believe What They Say?


150. Hubs or Fringes: Pretraining Data Selection via Web Graph Centrality


151. Towards Fully Automated Exam Grading: Fairness-Aware Recognition of Handwritten Answers with Foundation Models


152. CRUMB: Efficient Prior Fitted Network Inference via Distributionally Matched Context Batching


153. LSTM-Based Detection of Structural Breaks in Property Insurance Loss Reserving: A Climate-Informed Approach


154. APEX: Automated Prompt Engineering eXpert with Dynamic Data Selection


155. AI Coding Agents in Social Science: Methodologically Diverse, Empirically Consistent, Interpretively Vulnerable


156. The Power of Test-Time Training for Approximate Sampling


157. Towards a Bridge Layer Between Bibliographic and Formalized Mathematical Knowledge


158. JailbreakOPT: Tool-Assisted Iterative Jailbreak Prompt Optimization


159. Signed Compression Progress on a Sealed Audit is Goodhart-Resistant


160. MPC-Patch-Bench: Security-Aware LLM Code Patch for Multi-Party Computation


161. Risk Under Pressure: Compute-Aware Evaluation of Adversarial Robustness in Language Models


162. Steering Where to Listen: Instruction-Based Activation Steering Redirects Temporal Attention in Large Audio-Language Models


163. Small Experiments, Cheaper Decisions: A Case Study in Staged Promotion for Micro-Pretraining


164. Overcoming State Inertia in Full-Duplex Spoken Language Models via Activation Steering


165. When Probing Accuracy Saturates, Fragility Resolves: A Complementary Metric for LLM Pre-Training Analysis


166. The Dynamics of Human and AI-Generated Language: How Semantics Fluctuates across Different Timescales


167. TileFuse: A Fused Mixed-Precision Kernel Library for Efficient Quantized LLM Inference on AMD NPUs


168. Quantized Stochastic Primal-Dual Methods for Distributed Optimization under Relaxed Global Geometry


169. Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models


170. FlowBank: Query-Adaptive Agentic Workflows Optimization through Precompute-and-Reuse


171. FreeBridge: Variational Schrödinger Bridges for Cellular Transition Dynamics


172. RoVE: Rotary Value Embeddings Attention for Relative Position-dependent Value Pathways


173. Federated continual learning: A comprehensive survey on lifelong and privacy-preserving learning over distributed and non-stationary data


174. Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation


175. When Poison Fails After Retrieval: Revisiting Corpus Poisoning under Chunking and Reranking Pipelines


176. OmniBioTwin: A System-of-Twinned-Systems Framework for Health Digital Twins


177. PermDoRA – Understanding Adapter Interference in Language Models: Limits of Parameter-Space Geometry


178. RAIL: Rethinking Auditory Intelligence in Large Audio-Language Models with a CHC-Grounded Benchmark


179. Physics-informed generative AI for semiconductor manufacturing: Enforcing hard physical constraints in generative models by construction


180. SPEAR: A System for Post-Quantization Error-Adaptive Recovery Enabling Efficient Low-Bit LLM Serving


181. Artificial Intelligence in Ship Finance: Applications, Opportunities, and a Case Study in AI-Augmented Loan Origination


182. Every Act Has Its Price: Compressed Moral Composition in Frontier LLMs


183. Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents


184. An Ethical eValuation Agent (EeVA): Results of a Proof-of-Concept Test on a Prototype Agentic-like Workflow to Assist Ethical Deliberations


185. Preregistration for Experiments with AI Agents


186. The Environmental Cost of LLMs in AIED: Reporting and Practices


187. From Awareness to Action: Understanding and Overcoming the Research-Practice Gap in Algorithmic Fairness for Public Health


188. Calibration Drift Under Reasoning: How Chain-of-Thought Budgets Induce Overconfidence in Large Language Models


189. T2MM: An LLM Supported Architecture For Inquiry-Based Modeling


190. ProcessThinker: Enhancing Multi-modal Large Language Models Reasoning via Rollout-based Process Reward


191. BioDivergence: A Benchmark and Evaluation Framework for Hidden Contextual Contradictions in Biomedical Abstracts


192. Dual-Stance Evaluation of Sycophancy: The Structure of Agreement and the Limits of Intervention


193. To Intervene or Not: Guiding Inference-time Alignment with Probabilistic Model Blending


194. NightFeats @ MMU-RAGent NeurIPS 2025: A Context-Optimized Multi-Agent RAG System for the Text-to-Text Track


195. The Structural Attention Tax: How Retrieval Format Hijacks In-Context Learning Independent of Content


196. MA-DLE: Speech-based Automatic Depression Level Estimation via Memory Augmentation


197. PoQ-Judge: A Multi-Architecture Evaluation Framework for Cost-Aware Proof-of-Quality in Decentralized LLM Inference


198. From Consumption to Reflection: Designing Human-AI Relations for Stable Reasoning


199. From Architecture to Output: Structural Origins of Hallucination in Large Language Models and the Amplifying Role of Data