전체 AI 논문 - 2026-07-02

1. AutoMem: Automated Learning of Memory as a Cognitive Skill


2. Theoria: Rewrite-Acceptability Verification over Informal Reasoning States


3. Optimal Resource Utilization for Autonomous Laboratory Orchestrators


4. Can Agents Generalize to the Open World? Unveiling the Fragility of Static Training in Tool Use


5. Agentic generation of verifiable rules for deterministic, self-expanding reaction classification


6. PedNStream: Scalable Network Flow Simulation for Pedestrian Traffic Management


7. Bayesian Uncertainty Propagation for Agentic RAG Pipelines: A Proof-of-Concept Study on Multi-Hop Question Answering


8. Graph-Native Reinforcement Learning Enables Traceable Scientific Hypothesis Generation through Conceptual Recombination


9. Two AI Metrics Diverged: Will it Make All the Difference?


10. Self-Evolving Agents with Anytime-Valid Certificates


11. Self-GC: Self-Governing Context for Long-Horizon LLM Agents


12. Coachable agents for interactive gameplay


13. AGI Maze as a Benchmark Framework for World-Modeling Agents


14. HARC: Coupling Harmfulness and Refusal Directions for Robust Safety Alignment


15. AI Native Games: A Survey and Roadmap


16. Multi-scale Mixture of World Models for Embodied Agents in Evolving Environments


17. Agri-SAGE: Simulation-Grounded Multi-Agent LLM for Context-Aware Agricultural Advisory Generation


18. PHREEQC-MCQ-200: A Diagnostic Benchmark for Tool-Augmented Scientific Simulator Agents


19. Personalization as Inverse Planning: Learning Latent Design Intents for Agentic Slide Generation via Structural Denoising


20. Managed Autonomy at Runtime: Gear-Based Safety and Governance for Single- and Multi-Agent Cyber-Physical Systems


21. Mnemosyne: Agentic Transaction Processing for Validating and Repairing AI-generated Workflows


22. Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity


23. From Signals to Structure: How Memory Architecture Drives Language Emergence in LLM Agents


24. Constructing Epistemic AI Literacy: Detecting Epistemic Aims and Processes in Student-AI Co-Programming


25. A Contextual-Bandit Oversight Game with Two-Sided Informational Asymmetry


26. RareDxR1: Autonomous Medical Reasoning for Rare Disease Diagnosis Beyond Human Annotation


27. Solution space path planning for supporting en-route air traffic control


28. Making Failure Safe: A Constrained, Verifiable Agent Framework for Open-Web Data Collection


29. The MMM Data Model – A Normative Specification for Knowledge Interoperability in a Decentralisable Knowledge Commons


30. Bounded Morality: Defining the Space of Moral Computation


31. Constructive Alignment: Governing Preference Dynamics in Human-AI Interaction


32. Measuring the Gap Between Human and LLM Research Ideas


33. Language-Critique Imitation Learning from Suboptimal Demonstrations


34. The State-Prediction Separation Hypothesis


35. FurnitureVLA: Learning Long-Horizon Bimanual Furniture Assembly with Vision-Language-Action Model


36. Are Performance-Optimization Benchmarks Reliably Measuring Coding Agents?


37. Distill to Detect: Exposing Stealth Biases in LLMs through Cartridge Distillation


38. GPU-Parallel Linearization Error Bounds for Real-Time Robust Optimal Control of Nonlinear and Neural Network Dynamics


39. World from Motion: Generative Dynamic Gaussian Reconstruction from Monocular Video


40. Right in the Right Way: LM Training with Verifiable Rewards and Human Demonstrations


41. Diffusion-GR2: Diffusion Generative Reasoning Re-ranker


42. Adversarial Pragmatics for AI Safety Evaluation: A Benchmark for Instruction Conflict, Embedded Commands, and Policy Ambiguity



44. Skills Are Not Islands: Measuring Dependency and Risk in Agent Skill Supply Chains


45. Autonomous Scientific Discovery via Iterative Meta-Reflection


46. Muon as a Residual Connection


47. Towards Developing a Multimodal Chat Assistant for University Stakeholders: RAG-based Approach


48. FAR: Failure-Aware Retry for Test-Time Recovery and Continual Policy Improvement


49. CausalMix: Data Mixture as Causal Inference for Language Model Training


50. Cheap Code, Costly Judgment: A Case Study on Governable Agentic Software Engineering


51. LongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language Models


52. Staleness-Learning Rate Scaling Laws for Asynchronous RLHF


53. MemSyco-Bench: Benchmarking Sycophancy in Agent Memory


54. DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation


55. EchoRisk: A Multicentre Echocardiography Dataset and Benchmark for Cardio-Oncology


56. Behavior-Adaptive Conversational Agents: Toward a Fluid Personality Framework


57. Reading Order Inference for Complex Document Layouts


58. Logit-Contribution Scoring Identifies Non-Literal Retrieval Heads


59. SWE-Doctor: Guiding Software Engineering Agents with Runtime Diagnosis from Multi-Faceted Bug Reproduction Tests


60. SenseWalk: Agent-Based Semantic Trajectory Simulation Powered by Large Language Models in Zoned Environments


61. TRCGL-Net: A Long-Tailed Multi-Label Chest X-Ray Classification Framework with Generative Data Augmentation and Label Co-Occurrence Modeling


62. Aionoscope: Debugging Latent-State Accessibility in Time-Series Representations


63. Learning Cardiac Motion Priors for Implicit Neural Representations


64. Post-Training Pruning for Diffusion Transformers


65. Human-Machine Collaboration on Generative Meta-Learning: Model and Algorithm


66. From Personas to Plot: Character-Grounded Multi-Agent Story Generation for Long-Form Narratives


67. Valdi: Value Diffusion World Models


68. DeWorldSG: Depth-Aware 3D Semantic Scene Graph Generation via World-Model Priors


69. Improving Sparse-View 3DGS Generalization via Flat Minima Optimization


70. CAT: Confidence-Adaptive Thinking for Efficient Reasoning of Large Reasoning Models


71. Meta-Transfer Learning for mmWave Beam Alignment


72. Recovering Input Text from Hidden States: Study of Gradient-Based Inversion of Decoder-Only Language Models


73. From World Models to World Action Models: A Concise Tutorial for Robotics


74. Pano2World: End-to-End 3D Generation via Unified Multi-View Sequences


75. Exploring the Semantic Gap in Agentic Data Systems: A Formative Study of Operationalization Failures in Analytical Workflows


76. LRAT-Catcher: Importing SAT Solver Certificates into Lean4 by Reflection


77. LeVLJEPA: End-to-End Vision-Language Pretraining Without Negatives


78. Active Learning for Cascaded Object Detection: Balancing Coverage and Uncertainty in Table Extraction Pipelines


79. GaussianFusion: Unified 3D Gaussian Representation for Multi-Modal Fusion Perception


80. Prototype Memory-Guided Training-Free Anomaly Classification and Localization in Prenatal Ultrasound


81. Phantom References: Hallucinated Citations That Survive Peer Review at Top-Tier Conferences


82. ConRTF: Edge-Constrained Boundary Distribution Refinement for Realtime TransFormer Table Structure Recognition


83. LLM-Guided ODE Discovery and Parameter Inference from Small-Cohort Aggregate Data


84. Detecting the Undetectable: Enhancing Unsupervised time series Anomaly Detection via Active Learning


85. Partial Skeleton Visibility for Action Recognition: A Constrained Field-of-View Approach


86. Self-conditioned Flow Map Language Models via Fixed-point Flows


87. Creating Impactful Autonomous Driving Datasets: A Strategic Guide from Research Gap to Benchmark


88. LLVM-Bench: Benchmarking and Advancing Large Language Models for LLVM Compiler Issue Resolution


89. LUMA: Benchmarking Segmentation via a Lightweight Universal Mask Adapter


90. Multi-Label Node Classification with Label Influence Propagation


91. Faithful by Definition: Emotion Analysis via Natural Semantic Metalanguage Explications


92. Loss Smoothing for Stable Adaptation Under Distribution Shift


93. Identifying Latent Concepts and Structures for Generalized Category Discovery


94. Auditing Forgetting in Limited Memory Language Models


95. A Methodology for Investigating AI Patterns Prevalence in Software Repositories


96. Group-Equivariant Poincaré Convolutional Networks


97. Cross-Domain Generalization Failure in Lightweight Intrusion Detection Models for IIoT Networks


98. EgoGapBench: Benchmarking Egocentric Action Selection in Multi-Agent Scenes


99. Flow-Map GRPO: Reinforcement Learning for Few-Step Flow-Map Generators via Anchored Stochastic Composition


100. Active-GRPO: Adaptive Imitation and Self-Improving Reasoning for Molecular Optimization


101. From Technical Metrics to User Perception: A User Study of a Multimodal Human-Robot Interaction System for Object Detection and Grasping


102. AI, Trust, and Teaming: The Humans-as-Handlers Approach for Autonomous and Opaque AI Systems


103. Cross4D-JEPA: Dense Cross-modal Correspondence Distillation for 4D Point Cloud Representation Learning


104. BaseRT: Best-in-Class LLM Inference on Apple Silicon via Native Metal


105. MindEdit-Bench: Benchmarking Object-Level Counterfactual Spatial Reasoning in VLMs from In-the-Wild Photos


106. PAPA: Online Personalized Active Preference Alignment


107. Beyond the Prompt: Jailbreaking Function-Calling LLMs via Simulated Moderation Traces


108. Predicting Lethal Outcome (Cause) And Understanding Key Biomarkers Linked With Acute Myocardial Infarction Using Deep Artificial Neural Network And Ensemble Of Machine Learning Methodologies


109. A Multi-Resolution Finite-Volume Inspired Deep Learning Framework for Spatiotemporal Dynamics Prediction


110. Gauging, Measuring, and Controlling Critic Complexity in Actor-Critic Reinforcement Learning


111. Real-Time Hard Negative Sampling via LLM-based Clustering for Large-Scale Two-Tower Retrieval


112. VideoSearch-R1: Iterative Video Retrieval and Reasoning via Soft Query Refinement


113. Search-Based Spatiotemporal and Multi-Robot Motion Planning on Graphs of Space-Time Convex Sets


114. Learning Gait-Aware Quadruped Locomotion with Temporal Logic Specifications


115. EO-VGGT: Orbital Ray-Conditioned 3D Foundation Models for Satellite Multi-View Reconstruction


116. The Illusion of High Utility in Safety Alignment of Text-to-Image Diffusion Models


117. Holographic Quantum Transformer: A Generalist Neuro-Symbolic Architecture for Solving Frustrated Systems via Generative Attention


118. NeuroCogMap Reveals Cognitive Organization of Large Language Models


119. Learning Generalizable Skill Policy with Data-Efficient Unsupervised RL


120. MalariAI: A Label-Resilient Decoupled Framework for Universal Cell Segmentation and Explainable Stage Classification in Dense Malaria Blood Smears


121. Learning to Compose: Revisiting Proxy Task Design for Zero-Shot Composed Image Retrieval


122. MEPA: Multi-Scale Representation Alignment for Visual Autoregressive Modeling with Mixture of Experts


123. When AI meets quantum information: A comprehensive review


124. Enhancing Flow Matching with A Unified Guidance Framework for Efficient and Robust Speech Synthesis


125. SoK: Attack and Defense Landscape of Mobile On-device AI Systems


126. DiscoLoop: Looping Discrete Embeddings and Continuous Hidden States for Multi-hop Reasoning


127. K-Inverse-RFM: A Modified RFM that Bridges the Gap to Neural Networks for Data-Corrupted Mathematical Tasks


128. RetailSMV: Exocentric vs. Egocentric Adaptation of Foundation Video World Models in Retail


129. Mapping the Evaluation Frontier: An Empirical Survey of the Bias-Reliability Tradeoff Across Eleven Evaluator-Agent Conditions


130. Learning When to Listen: Gated Affect Fusion for Human Motion Prediction


131. An LLM-Based Framework for Intent-Driven Network Topology Design


132. What’s Hidden Matters: Identifying Planning-Critical Occluded Agents using Vision-Language Models


133. Testing Frontier Large Language Models’ Physics Literacy in Parallel Physical Worlds


134. Entropy-Regularized Probabilistic Gates for Sparse Model Discovery in Scarce-Data Federated Learning


135. SEFORA: Student Essays with Feedback Corpus and LLM Feedback Evaluation Framework


136. ASPIRE: Agentic /Skills Discovery for Robotics


137. Validating Causal Abstraction Metrics on Simulated Complex Systems


138. Multi-Hypothesis Test-Time Adaptation to Mitigate Underspecification


139. Leveraging Phase Information to Boost Unrolled Network Learning for Image Deblurring


140. Adaptive Perturbation Selection for Contrastive Audio Decoding


141. A Category Theory Account of AI Identity


142. EgoSafetyBench: A Diagnostic Egocentric Video Benchmark for Evaluating Embodied VLMs as Runtime Safety Guards


143. SLIM-RL: Risk-Budgeted Random-Masking RL for Diffusion LLMs Without Trajectory Slicing


144. HydraCollab: Adaptive Collaborative-Perception for Distributed Autonomous Systems


145. Play Like Champions: Counterfactual Feedback Generation in Latent Space


146. Scaling Up Thermodynamic AI Models


147. EVOTS: Evolutionary Transformer Search for Time Series Forecasting


148. GRPO, Dr. GRPO, and DAPO Are Three Operations on One Number: The Group-Standard-Deviation Identity


149. A Mechanism-Driven Theory of Phase Transitions in Active Learning


150. Hate Speech Detection in Turkish and Arabic Languages: A Comprehensive Study


151. Would You Marry Superintelligence?


152. SNAP-FM: Sparse Nonlinear Accelerated Projection for Physics-Constrained Generative Modeling


153. Lost in the Tail: Addressing Geographic Imbalance in Urban Visual Place Recognition


154. Harnessing the Latent Space: From Steering Vectors to Model Calibrators for Control and Trust


155. Optimal any-angle path planning in static and dynamic environments


156. Spectral Geometry and Bosonic-Bloch Probes: Explorations in Quantum Learning


157. AlgoBench: Benchmarking Algorithmic Adaptation in Code Generation


158. Enhancing Oracle Bone Inscription Recognition via Multi-Scale Layer Attention


159. Active Sensing for RIS-Aided Tracking and Power Control: A Hybrid Neuroevolution and Supervised Learning Approach


160. SWE-Router: Routing in Multi-turn Agentic Software Engineering Tasks


161. AGE: Adaptive-masking for Graph Embedding in Graph Retrieval-Augmented Generation


162. Prompting GPT-5 on Scrum Certification Questions: An Empirical Accuracy Study


163. Comparing Large Language Models on Scrum Certification-Style Questions: Accuracy, Stability, and Error Patterns


164. Destination-Labeled Self-Looping Systems with Dwell: Intrinsic Characterization, Realization Cost, and Recognition


165. ATM: CID-Brokered Pre-Write Admission for Multi-Agent Code Co-Synthesis


166. Learning Dexterous Manipulation Using Contact Wrench Guidance From Human Demonstration


167. Memory-Native Non-Terrestrial Networks for Embodied Intelligence


168. FLYNN: Robust Neural Network for Robot Navigation using Fly Brain Topology


169. Aligning Sentence Embeddings to Human Concepts via Sparse Autoencoders


170. LLMs in the Real World: Evaluating “AI” in Emergency Contexts


171. Learning User-Aware Recall: Personalized Retrieval in Long-Term Conversational Memory


172. Libra: Training the Environment for Agentic Information Retrieval


173. Towards an automated AI-based framework for floor plan compliance checks for residential buildings


174. GRACE-RAG: Governed Retrieval Architecture for Canonical Evidence Synthesis, Enabling Lightweight Deployment in Closed-Domain Institutional Settings


175. PRA-RAG: Provably Robust Aggregation in Retrieval-Augmented Generation against Retrieval Corruption


176. SkillSelect-Serve: Budget-Controllable and QoS-Aware Skill Service Recommendation and Composition for Small LLM Agents


177. Prompt Optimization for User Simulation in Conversational Recommender Systems: A Multi-Objective Framework


178. Controllable Narrative Rendering for Enhanced Assisted Writing


179. SchemaRAG: Dynamic Large Schema Reduction for LLM-driven Structured Information Extraction


180. BaRA: BFS-and-Reflection Web Data Collection Agent


181. Persona Without Substrate: Regime-Dependence and the LLM Individuation Problem


182. Topological Void Analysis A Mathematical Framework for Systematic Technical Innovation Discovery in Knowledge Spaces


183. Why Advanced Encoders Lag on Sparse Retrieval? The Answer and an Approach to Bridging Vocabulary Gaps


184. From “Strings” to “Things” for Personal Knowledge Graphs: Evaluating LLM Triple Extraction for Recommendation Systems


185. DigitalCoach: Communication and Grounding Gaps in Human and Agentic Computer Use Coaching


186. UltraFlux: Data-Model Co-Design for High-quality Native 4K Text-to-Image Generation across Diverse Aspect Ratios