전체 AI 논문 - 2026-06-26

1. Language-Based Digital Twins for Elderly Cognitive Assistance


2. When Does Combining Language Models Help? A Co-Failure Ceiling on Routing, Voting, and Mixture-of-Agents Across 67 Frontier Models


3. Prompt Injection in Automated Résumé Screening with Large Language Models: Single and Multi-Injection Settings


4. Simulation-based inference for rapid Bayesian parameter estimation in epidemiological models: a comparison with MCMC


5. EO-WM: A Physically Informed World Model for Probabilistic Earth Observation Forecasting


6. Ask, Don’t Judge: Binary Questions for Interpretable LLM Evaluation and Self-Improvement


7. Vulnerability of Natural Language Classifiers to Evolutionary Generated Adversarial Text


8. A Process Harness for Uplifting Legacy Workflows to Agentic BPM: Design and Realization in CUGA FLO


9. TOPS: First-Principles Visual Token Pruning via Constructing Token Optimal Preservation Sets for Efficient MLLM Inference


10. OpenRCA 2.0: From Outcome Labels to Causal Process Supervision


11. Joint Learning of Experiential Rules and Policies for Large Language Model Agents


12. How to evaluate clustering with ground truth?


13. Semantic Early-Stopping for Iterative LLM Agent Loops


14. Adaptive Utility driven Resource Orchestration for Resilient AI (AURORA-AI)


15. Einstein World Models


16. Look-Before-Move: Narrative-Grounded World Visual Attention in Dynamic 3D Story Worlds


17. Where Do CoT Training Gains Land in LLM based Agents?


18. Diagnosing Task Insensitivity in Language Agents


19. Learning to Recover Task Experts from a Multi-Task Merged Model


20. Generative Retrieval via Diffusion Transformer with Metric-Ordered Sequence Training and Hybrid-Policy Preference Optimization


21. A Pipeline for Generating Longitudinal Synthetic Clinical Notes Using Large Language Models


22. TAVR-VLM: Risk-Conditioned Causal Grounding for Hallucination-Resistant Report Generation


23. AgentX: Towards Agent-Driven Self-Iteration of Industrial Recommender Systems


24. LCAi: Life Cycle Assessment with big data fusion and retrieval-augmented generation-assisted interpretation


25. Context-Aware Synthesis of Optimization Pipelines for Warehouse Optimization


26. The Capability Frontier: Benchmarks Miss 82% of Model Performance


27. Computational Analysis of Heart Rate Variability in Healthy Adults


28. KARLA: Knowledge-base Augmented Retrieval for Language Models


29. Memory Depth, Not Memory Access: Selective Parametric Consolidation for Long-Running Language Agents


30. ResilPhase: Plug-and-Play Phase Mapping and Noise-Resilient Macro-Trajectory Extrapolation for Diffusion Acceleration


31. EGG: An Expert-Guided Agent Framework for Kernel Generation


32. Scientific discovery as meta-optimization: a combinatorial optimization case study


33. Socratic agents for autonomous scientific discovery in high-dimensional physical systems


34. A Latent ODE Approach to Spatiotemporal Modeling of Cine Cardiac MRI


35. LithoDreamer: A Physics-Informed World Model for Multi-Stage Computational Lithography


36. Kalman Prototypical Networks for Few-shot Fault Detection in Combined Cycle Gas Turbines


37. Do Safety Guardrails Need to Reason? LeanGuard: A Fast and Light Approach for Robust Moderation


38. NebulaExp-8B: An Empirical Post-Training Pipeline via Full-Scale Ablation Research


39. SKILL-DISCO: Distilling and Compiling Agent Traces into Reusable Procedural Skills


40. Autoformalization of Agent Instructions into Policy-as-Code


41. LLM-based Models for Detecting Emerging Topics in Service Feedback


42. Content-Based Smart E-Mail Dispatcher Using Large Language Models


43. A Multi-Level Validation and Traceability Framework for AI-Generated Telescope Scheduling Decisions


44. EvoOptiGraph: Weakness-Driven Coevolution via Graph-Based Structural Generation for Optimization Modeling


45. Explainable Ensemble-Based Machine Learning Models for Detecting the Presence of Cirrhosis in Hepatitis C Patients


46. PMDformer: Patch-Mean Decoupling Information Transformer for Long-term Forecasting


47. Radical AI Interpretability


48. Boundary-Aware Context Grounding for A Low-Channel EEG Agent


49. NeuraDock Visual Cognitive Load Agent Tutorial: A Quality-Gated Open-Source EEG Workflow for Alpha Dynamics and Real-Time Applications


50. Humans Disengage, Reasoning Models Persist: Separating Difficulty Registration from Deliberation Allocation


51. Clinical Harness for Governable Medical AI Skill Ecosystems


52. auto-psych: Automating the science of mind using agent-driven theory discovery and experimentation


53. MKG-RAG-Bench: Benchmarking Retrieval in Multimodal Knowledge Graph-Augmented Generation


54. Data-driven Machine Learning Cannot Reach Symbolic-level Logical Reasoning – The Limit of the Scaling Law


55. Estimating Uncertainty in Classifier Performance with Applications to Large Language Models and Nested Data


56. Unbiased Canonical Set-Valued Oracles Via Lattice Theory


57. When Agents Meet Electric Bus Fleet Operations: Pricing Behavior, Trade-offs, and Policy Implications in an Aggregator Framework


58. Geometry-Aware MCTS for Extremal Problems in Combinatorial Geometry


59. Narration-of-Thought: Inference-Time Scaffolding for Defeasible Ethical Reasoning in Large Language Models


60. Accelerating Returns and the Qualitative Engine for Science


61. Instruction Bleed: Cross-Module Interference in Prompt-Composed Agentic Systems


62. OpenFinGym: A Verifiable Multi-Task Gym Environment for Evaluating Quant Agents


63. What We are Missing in Multimodal LLM Evaluation?


64. How Do Tool-Augmented LLM Agents Perform on Real-World Energy Analytics Tasks?


65. The Verification Horizon: No Silver Bullet for Coding Agent Rewards


66. COrigami: An AI Pipeline for Co-Designing Flat-Foldable Visually Recognisable Origami


67. Governing Actions, Not Agents: Institutional Attestation as a Governance Model for Autonomous AI Systems


68. Accelerating Skill Assessment in Chess: A Drift-Diffusion-Enhanced Elo Rating System


69. Knowledge-augmented Agentic AI for Mental Health Medication Information Seeking


70. Agentic Analysis for Agentic Infrastructure: An LLM-Powered Pipeline for Comparative Governance of DAO and Corporate AI Protocols


71. AlgoEvolve: LLM-driven Meta-evolution of Algorithmic Trading Programs


72. Refusal Lives Downstream of Persona in Chat Models


73. Life After Benchmark Saturation: A Case Study of CORE-Bench


74. Detecting and Controlling Sycophancy with Cascading Linear Features


75. Autoregressive Boltzmann Generators


76. Error-Conditioned Neural Solvers


77. Understanding Domain-Aware Distribution Alignment in Budgeted Entity Matching


78. Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning


79. Beyond the Hard Budget: Sparsity Regularizers for More Interpretable Top-k Sparse Autoencoders


80. AI Healthcare Chatbots as Information Infrastructure: A Large-Scale Study of User-Reported Breakdowns


81. E-TTS: A New Embodied Test-Time Scaling Framework for Robotic Manipulation


82. Advancing Omnimodal Embodied Agents from Isolated Skills to Everyday Physical Autonomy


83. From Celebrities to Anyone: Characterizing AI Nudification Content, Technology, and Community Dynamics on 4chan


84. Bridging Talk and Thought: Understanding Dialogue Dynamics Across Collaborative Problem-Solving Contexts


85. CARVE: Content-Aware Recurrent with Value Efficiency for Chunk-Parallel Linear Attention


86. Automating Potential-based Reward Shaping with Vision Language Model Guidance


87. Learning to Fold: prizewinning solution at LeHome Challenge 2026 (1st place online, 2nd offline)


88. Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks


89. Efficient foundation decoders for fault-tolerant quantum computing


90. Heavy-Ball Q-Learning with Residual Weighting Correction


91. Application of LLMs to Threat Assessment of Foreign Peacekeeping Missions


92. Data-Free Reservoir Features for Efficient Long-Horizon Cold-Start Continual Learning


93. Inherited Circuits, Learned Semantics: How Fine-Tuning Creates Evasion Vulnerabilities Invisible to Standard Evaluation


94. Beyond Global Divergences: A Local-Mass Perspective on Bayesian Inference


95. Parametric Open Source Games


96. NuclearQAv2: A Structured Benchmark for Evaluating Domain-Science Competence in Large Language Models


97. The Spec Growth Engine: Spec-Anchored, Code-Coupled, Drift-Enforced Architecture for AI-Assisted Software Development


98. State Representation Matters in Deep Reinforcement Learning: Application to Energy Trading


99. ShareLock: A Stealthy Multi-Tool Threshold Poisoning Attack Against MCP


100. On-board Remote-Sensing Foundation Models for Unsupervised Change Detection of Disaster Events


101. Inverse Design of Compact and Wideband Inverted Doherty Power Amplifiers Using Deep Learning


102. Event-Aware Instructed Assistant for Referring Video Segmentation


103. Decision-Aligned Evaluation of Uncertainty Quantification


104. Where Do Models Find Happiness? Emotion Vectors in Open-Source LLMs


105. ReaORE: Reasoning-Guided Progressive Open Relation Extraction Empowered by Large Reasoning Models


106. Auditing Framing-Sensitive Behavioral Instability in Large Language Models for Mental Health Interactions


107. In-Context Model Predictive Generation: Open-Vocabulary Motion Synthesis from Language Models to Physics


108. XMSE-Aware Adaptive Empirical Bayes Estimation


109. Scaling Multi-Reference Image Generation with Dynamic Reward Optimization


110. Chai: Agentic Discovery of Cryptographic Misuse Vulnerabilities


111. A Deterministic Control Plane for LLM Coding Agents


112. Risk-Aware Selective Multimodal Driver Monitoring with Driver-State World Modeling


113. GEOALIGN: Geometric Rollout Curation for Robust LLM Reinforcement Learning


114. Confidence-Aware Tool Orchestration for Robust Video Understanding


115. SamaVaani: Auditing and Debiasing Multilingual Clinical ASR for Indian Languages


116. Bridging Vision and Language Concepts through Optimal Transport Semantic Flow


117. Information-Aware KV Cache Compression for Long Reasoning


118. Fortress and Gatekeeper: Theorizing Transitive Trust in Third-Party Cybersecurity Risk Governance


119. NaviCache: Test-Time Self-Calibration Caching for Video Generation


120. ReasonCLIP-58M: Visually Grounded Commonsense Reasoning Supervision for CLIP


121. MIRROR: Novelty-Constrained Memory-Guided MCTS Red-Teaming for Agentic RAG


122. AIGP: An LLM-Based Framework for Long-Term Value Alignment in E-Commerce Pricing


123. Anatomy-Guided Residual Motion Diffusion for Controllable 4D Cardiac MRI Synthesis


124. Robust Onion: Peeling Open Vocab Object Detectors Under Noise


125. MLFFM-SegDiff: A Multi-Level Feature Fusion Diffusion Model for Skin Lesion Segmentation


126. Algorithmic Foundations of Deep Learning: Complexity-Theoretic Rates and a Characterization of Universal Approximation


127. Learning Motion Feasibility from Point Clouds in Cluttered Environments


128. Beyond Logical Forms: LLM-Extracted Patterns for Fallacy Classification


129. Disco-LoRA: Disentangled Composition of Content, Style, and Motion for Multi-concept Video Customization


130. TGHE: Template-based Graph Homomorphic Encryption for Privacy-Preserving GNN Inference in Edge-Cloud Systems


131. Zero-Shot Size Transfer for Neural ODEs on Sparse Random Graphs: Graphon Limits and Adjoint Convergence


132. LAMP: Lane-Aligned Motion Primitives for Feasible Trajectory Prediction


133. CAT-Q: Cost-efficient and Accurate Ternary Quantization for LLMs


134. Agents That Know Too Much: A Data-Centric Survey of Privacy in LLM Agents


135. Discovering Millions of Interpretable Features with Sparse Autoencoders


136. HiLSVA: Design and Evaluation of a Human-in-the-Loop Agentic System for Scientific Visualization


137. SharQ: Bridging Activation Sparsity and FP4 Quantization for LLM Inference


138. IDEA: Insensitive to Dynamics Mismatch via Effect Alignment for Sim-to-Real Transfer in Multi-Agent Control


139. scBench-Long: Verifiable Benchmarking of Long-Horizon Single-Cell Biology


140. SpaceRipple: Lightweight Semantic Delivery for Mission-Oriented LEO Earth Observation Satellite Networks


141. Perception, Verdict, and Evolution: Hindsight-Driven Self-Refining Forensics Agent for AI-Generated Image Detection


142. CascadeFormer: Depth-Tapered Transformers Motivated by Gradient Fan-in Asymmetry


143. From Hallucination to Grounding: Diagnosing Visual Spatial Intelligence via CRISP


144. VoiceTTA: Enhancing Zero-Shot Text-to-Speech via Reinforcement Learning-Based Test-Time Adaptation


145. \textsc{DiARC}: Distinguishing Positive and Negative Samples Helps Improving ARC-like Reasoning Ability of Large Language Models


146. The Inattentional Gap: Task-Conditioned Language and Vision Models Omit the Safety-Critical Signals They Can Otherwise Report


147. Multipath Adaptive Gated Bottleneck Latent ODE with Raman Data Fusion for Cell Culture Process Forecasting


148. Temporal Validity in Retrieval Memory: Eliminating Stale-Fact Errors for AI Agents over Evolving Knowledge


149. Evaluation-Strategy Gap in Fault Diagnosis of Deep Learning Programs


150. An Empirical Study of LLM-Generated Specifications for VeriFast


151. Speaking Numbers to LLMs: Multi-Wavelet Number Embeddings for Time Series Forecasting


152. Adaptive Evaluation of Out-of-Band Defenses Against Prompt Injection in LLM Agents


153. Retrieval-Warmed Energy-Based Reasoning: A Five-Arm Ablation Methodology for Diffusion-as-Inference on Structured Reasoning Tasks


154. Localizing RL-Induced Tool Use to a Single Crosscoder Feature


155. 3D Spatial Pattern Matching


156. Active Adversarial Perturbation-driven Associative Memory Retrieval for RGB-Event Visual Object Tracking


157. ProvenAI: Provenance-Native Traces of Evidence in Generated Answers


158. Closing the Loop to Discover Psychological Theories with an Automated Cognitive Scientist


159. WatchAct: A Benchmark for Behavior-Grounded Robot Manipulation


160. AXLE: A Cloud Infrastructure for Lean 4 Theorem Proving Utilities


161. ConflictScore: Identifying and Measuring How Language Models Handle Conflicting Evidence


162. Play2Perfect: What Matters in Dexterous Play Pretraining for Precise Assembly?


163. CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation


164. Beyond Feedforward Networks: Reentry Neural Systems as the Fundamental Basis of Subjecthood and Intrinsic Safety of Next-Generation AGI


165. Deterministic Pareto-Optimal Policy Synthesis for Multi-Objective Reinforcement Learning


166. Sampling sea state using a diffusion model


167. SOLAR: AI-Powered Speed-of-Light Performance Analysis


168. Charting the Growth of Social-Physical HRI (spHRI): A Systematic Review Pipeline Augmented by Small Language Models


169. Hybrid privacy-aware semantic search: SVD-truncated document geometry and CKKS-encrypted query reranking under a restricted threat model


170. EVOM: Agentic Meta-Evolution of Actor-Critic Architectures for Reinforcement Learning


171. Parametric Generalized Adaptive Moment Features (PG-AMF) for Bearing Fault Diagnosis and Machine Health Monitoring


172. The Red Queen Gödel Machine: Co-Evolving Agents and Their Evaluators


173. SSM Adapters via Hankel Reduced-order Modeling: Injection Site Determines Task Suitability in Long-Context Fine-Tuning


174. TEMPO-Diffusion: Temporally Exposed Malicious Poisoning of Diffusion Models


175. From Clicks to Intent: Cross-Platform Session Embeddings with LLM-Distilled Taxonomy for Financial Services Recommendations


176. A multi-task spatiotemporal deep neural network for predicting penetration depth and morphology in laser welding


177. Lacuna: A Research Map for Machine Learning


178. CyberChainBench: Can AI Agents Secure Smart Contracts Against Real-World On-Chain Vulnerabilities?


179. Statistical and Structural Approaches to Algorithmic Fairness


180. From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models


181. LiMoDE: Rethinking Lifelong Robot Manipulation from a Mixture-of-Dynamic-Experts Perspective


182. KG-TRACE: A Neuro-Symbolic Framework for Mechanistic Grounding in Antimicrobial Resistance Prediction


183. LCG: Long-Context Consistent Image Generation with Sparse Relational Attention


184. Neural Architecture Search for Generative Adversarial Networks: A Comprehensive Review and Critical Analysis


185. Reducing Redundancy in Whole-Slide Image Patching for Scalable Indexing and Retrieval


186. Unsupervised Memory-Enhanced Video Transformers: Obstacle Detection for Autonomous Agricultural Rover


187. Multiscale Exit-Join Dynamics: Tactical Consensus and Strategic Coalition Formation


188. Thinking Like a Scientist? A Structural Study of LLM-Generated Research Methods


189. Geometric Fairness-Aware Routing for Federated Edge Networks


190. Privacy-Aware Agent Collaboration for Dynamic VR Slice Management in 6G SD-RAN


191. Dot-Flik: A Scalable Edge AI Architecture for Distributed Insect Monitoring


192. The Open Source Economic Index of AI Adoption and Capability


193. The Governance Inversion Hypothesis: Why More AI Regulation May Produce Less Organisational Control


194. Divergent Recommendations, Convergent Diagnoses: Cross-Provider Failure-Mode Convergence in AI Commercial Recommendation


195. A Multi-Layer AI Framework for Information Landscape Analysis


196. Dream machine – the next creative economy


197. From Lexicon to AI: A Structured-Data Pipeline for Specialized Conversational Systems in Low-Resource Languages



199. Low Resource Multimodal Translation of Nepali Spoken Words into Emotion-Conditioned Sign Language Avatars


200. Reducing Conversational Escalation in Large Language Model Dialogue with Nonviolent Communication Constraints


201. Context Recycling for Long-Horizon LLM Inference


202. Assert, don’t describe: Linguistic features that shift LLM reasoning about animal welfare


203. Investigating LLM’s Problem Solving Capability – a Study on Statics Questions


204. Helpfulness Hurts: Domain-Dependent Degradation of Mid-Trained Compassion Values Under Post-Training


205. Know2Guess: A Contamination-Aware Multi-Zone Benchmark for Knowledge-Boundary Evaluation in Large Language Models


206. Benchmarking Open-Weight Foundation Models for Global AI Technical Governance


207. Patent Representation Learning via Self-supervision