전체 AI 논문 - 2026-09-16

1. ScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents


2. Verifiable Social Reasoning for LLM Assistants


3. LimiX-2: A Contextual Mechanism Network Towards General Structured-Data Intelligence


4. JustFit: 200K-Token LLM Serving on a 24 GiB Laptop with Just-in-Time State Management


5. Talking Head Synthesis with Facial Landmark Guidance via 3D Gaussian Splatting


6. Transformer-Based Token Fusion and Dynamic Graph Planning for Audio-Visual Navigation


7. World Model Science: Self-Organized Criticality, Weak Chaos, and Metastable Belief Dynamics in Long-Horizon LLM Agents


8. Never Stop Thinking: Continuous-Time Language Agents


9. FlashVector: Agent for Hierarchical Model Serving Stack Optimization


10. Self-Emergence Agent Architecture:Behavior-Inertia HMM, Reflexive Metacognition,and Social-Contrastive Self-Modeling


11. From Transient Prompts to Persistent Control: Scientific Poster Generation via Recursive Semantic-Geometric Contracts


12. Intrinsic Motivation in Reinforcement Learning: A Research Agenda for Adaptive Self-Organisation


13. Extracting ontology-compliant knowledge from scientific text describing irradiated materials using large language models


14. End-to-End Latency-Minimizing and Load-Balanced Request Scheduling for Edge LLM Inference in Agentic AI Services


15. MOCC-R1: Reinforcing Reasoning-Response Consistency for Multimodal Counselor Response Generation


16. FirmCORe: A Benchmark for Structured Reasoning about Inter-Firm Collaboration Opportunities


17. Shared-Prefix KV Reuse Across Standard LoRA Adapters: Quality and Serving Tradeoffs


18. Symbolic Separation: Grounding Deep Agents in Knowledge Graphs for Trustworthy Operational Data Analytics


19. Semi-Supervised Learning-Based Genetic Biomarkers Dataset for Multiple-Stage Hepatocellular Carcinoma Prediction


20. Scaling-Score Conformal Prediction for Multi-Target Regression


21. Interactive Memory Learning for Long-Term Conversations


22. Sample-Conditioned Representation Selection for Audio Few-Shot Learning


23. Neuro-Symbolic Hierarchical Intention Anticipation in Human Behavior


24. Sparse MLLM Anchors, Dense Adaptation: Breaking the Self-Referential Loop in Wild Test-Time Adaptation


25. SKIP: a Self-knowledge-guided Step-wise Preference Learning Framework for Concise Reasoning


26. ORDER: Task-Conditioned Routing for Retrieval-Augmented Generation


27. ThinkFlow: Self-Evolving Probabilistic Latent Memory for Lifelong Conversational Agents


28. FlexEE: Self-Speculative and KV-Compatible Early Exiting for Offloading-Aware LLM Inference


29. Affect-Prototype Guided Fusion for Open-Vocabulary Incomplete Multi-modal Emotion Recognition


30. AntennaFlow: A Generative Flow Model for Offset Correction in Phaseless Antenna Testing


31. QART: A Quantum-Classical Hybrid Architecture for Long-Horizon Reasoning – Exploring a Conditional Path toward Quantum Scaling


32. Bridging Learned Visual Perception and Symbolic Belief-Space Planning


33. CoAdapt: An LLM-based Framework for Adaptive Collaborative Perception in IIoT Robotic Swarms


34. Execution Flexibility in Automated Planning: A Comparative Evaluation of Deordering and Reordering Strategies


35. Can We Do Interpretable NLI with Graphs Based on Atomic Propositions?


36. Layers, Sinks, and Scaling: Adaptive Evidence Selection for Multimodal Large Language Models


37. Integrating the Analytic Hierarchy Process with Large Language Models for Transparent Multi-Criteria Decision-Making


38. Coverage-Aware Virtual IMU Augmentation for Low-Resource Human Activity Recognition


39. Turn-level Multiscale Density Ratio Estimation for LLM Agents


40. Beyond Episodic AI: Cognitive Field Networks for Biologically Inspired Persistent Cognition


41. LSREP: A Longitudinal State-Replay Protocol for Evaluating Conversational Memory, with ICE v2 as an Audited Local-First Architecture


42. VideoMM: Adaptive Macro-Micro Inference for Efficient Video MLLMs


43. little m: An AI Agent for Industrial Process Optimization


44. AI for Games in the Foundation Model Era


45. ANIMASK: What the Model Contributes to Role Play in Simulated Story Worlds


46. ReDraft, Don’t Just Distill: Reference-Driven Revision for Continual VLLM Post-Training


47. EchoPath: Execution-Level Replayable Memory for GUI Agents


48. A Framework for Generating Valid Context-Specific Benchmarks through Expert Guidance


49. Do LLMs Have Values? A Quantitative Analysis and Alignment Framework for Values in Large Language Models


50. Query-Aware Source-Risk Triage for Retrieval-Augmented Generation


51. QueryFormer: Winning Solution for KDD Cup 2026 Tencent UniRec Challenge


52. AquiLLM: Evaluating Faithfulness in Open-Weight RAG-LLM Systems for Scientific Research


53. From Manual Construction to AI-Driven Scenario Emergence: Rethinking Catastrophe Risk Modeling


54. Skill-based Agentic Evaluation for Real-time Data Science Tasks


55. Fine-Tuning Fixes Mode Collapse and Over-Dispersion in LLMs


56. Breaking the 1.58-bit Barrier for Ternary LLMs


57. Cross-Anatomy Transfer Versus Sparse Interpolation in Digital-Twin-Oriented Aortic Fluid-Structure Interaction Surrogates


58. BLINDSPOT: A Benchmark for Safety and Refusal Calibration in Long-Horizon Tool-Using Agents


59. CLEAR: Cross-Source Evidence Adjudication for Large Language Models in Medicine


60. Closing the Loop: Branch-and-Bound for Scalable Verification of Nonlinear Neural Feedback Systems


61. The AI-Enabled Scientific Frontier


62. CADWorld: Computer-Use Benchmark for Long-Horizon Computer-Aided Design


63. The Pain Axis: LLMs Represent Self-Directed Harm and Act to Relieve It


64. Metacognitive Steering: Learning the Structure of Scientific Judgment


65. Toward Governance-Aware Autonomous GIS: A Narrative Review of Ethical and Privacy Risks in LLM-Enabled GeoAI


66. Where Should the KV Cache Live? Placement Policies Across GPU, CPU, and SSD for Long-Lived Sessions


67. Artificial intelligence and biosecurity: capabilities, threat pathways, and defense-in-depth governance


68. Calibrate, Then Route: A Measured Study of Learned Request Routing for Disaggregated LLM Serving


69. Position: AI Is Not Ready for Strategic Conflicts


70. GPEvac: GNN-Based PPO for Adaptive Evacuation Routing During Shooting Events


71. Safe Error Correction for Language Models: Frozen-Base Adjustment with Capability Preservation


72. Optimal Pruning for Neural Architectures using Fisher Information Distances


73. Agentic Societies Need a Social Harness


74. PhysStream: Streaming Physics-Grounded Video Generation with Structured Scene Memory and Fine-Grained Motion Control


75. When Should LLMs Abstain? Chain-of-Self-Questioning for Selective Risk Control


76. LACE: Layer-Wise Compression for Dynamic Frame Rate Codecs


77. ENCP: Episode-Normalized Conformal Prediction for Vision-and-Language Navigation


78. Det-LIME: Detector-Aware, Multi-Instance Local Interpretable Model-Agnostic Explanations for Automated Marine Mammal Detection


79. Coupled Calibration and Learning: Mitigating Teacher Bias in LLM Distillation without Target-Domain Reward Feedback


80. Decomposition Buys Integrity, Not Yield


81. Evaluating Verified Autonomy in Quantum Engineering


82. CareMirror: Bringing Caregiver Wellbeing into the Dementia Care Ecosystem


83. Learning-Guided Planning in Large Dynamic Action Spaces: Budgeted Tree Search for One-to-Many Mobile Charging


84. Tracking the Unseen: An Occlusion-Robust Framework for Target Tracking Under Full and Long-Term Occlusion


85. CTAN: Cycle-Temporal Attention Network for Embodied Audio-Visual Navigation


86. Coding Agents Have Converged: Why the SWE-bench Leaderboard Can No Longer Order Its Top Entries, and What to Measure Instead


87. Where Should a Document Live: Context, Representations, or Parameters?


88. Vroom-Vroom at SHROOM-Visions: A Multi-Judge Committee for Detecting Hallucinated Spans in Vision-Language Outputs


89. Mo’ Models, Mo’ Problems: How to best select model pools when designing Multi-Agent Systems


90. After the Party: Governing What a Viral Agent-Skill Ecosystem Left Behind


91. FROD: Feature Matching Residual Denoising Oracle Bone Decipher


92. Easy to Catch a Liar, Hard to Clear an Honest One: Language Models Diagnosing a Corrupted Reward Channel from a Verified Record


93. Grounding SWE-Agent Decisions in Architecture-0 Design: Navigating Unknown Unknowns through Physical Mapping


94. FluxVLA Engine: A One-Stop VLA Engineering Platform for Embodied Intelligence


95. Multimodal Cultural Heritage Architectural Style Classification for Residential Buildings in the UAE Based on CLIP Embeddings and SVM


96. A unified framework for global and local interpretability using adaptive derivative-ordered random explanation


97. MUMINS: Metadata-conditioned Uncertainty-aware Medical Image Next-state Synthesis


98. ResLRP: The Role of Residual Cancellation in Attribution Instability in Vision Transformers


99. Kernel-Based Metrics Learning for Uncertain Opponent Vehicle Trajectory Prediction in Autonomous Racing


100. Continual Learning for Traversability Prediction with Uncertainty-Aware Adaptation


101. AI for Science with GPT-6 Astra: Thermal Design and Electrothermal Analysis of 2D CFET


102. Finding Common Mistakes In Modelling With Mathematical Formalisms Using LLMs


103. Beyond In-Distribution Metrics: A Systematic Out-of-Distribution Evaluation of Congenital Heart Disease Segmentation


104. Beyond “ChatGPT Can Make Mistakes”: Designing Interventions to Support Metacognitive Monitoring in AI-Assisted Work


105. Repurposing Unified Topological Signatures for Graph Representation Learning


106. Diagnosing the Fact-Grounding Gap in Multi-Hop Question Answering


107. Distributed JEPA: A Self-Supervised Framework for Energy Forecasting


108. The Role of Implicit and Explicit Demographic Signals in Large Language Model-based Student Assessment


109. AeroLat: Channel-Aware Latent Space Semantic Communication for Decentralized UAV Swarms


110. Beyond Token-Local Imitation: Reward-Compatible Temporal Credit Assignment for On-Policy Distillation


111. RepoAtlas: Guiding Coding Agents via Evolving Multimodal Repository Views


112. When Confidence Signals Disagree: Local and Global Confidence in Autoregressive Language Models


113. Causal Discovery via Transformed Low-Rank Quantile Surfaces


114. Repurposing Deep Limit Order Book Forecasting for Scenario-Conditioned Market Impact Modeling


115. Verbalizing Subliminal Learning Effects Using Text Optimization


116. VOR-Bench: A Human Perception-Driven Benchmark for Video Object Removal


117. The Evolution of Coordination in a Collective Intelligence System: 25 Years of English Wikipedia and the Emergence of Generative AI


118. RegRet: Enhancing Region-Level Retrieval in Large Multimodal Models


119. StackTok: Accelerating VLMs Inference with Budget-Adaptive Visual Token Selection


120. What Breaks Local Watermarks? A Robustness Benchmark for Local Invisible Image Watermarking


121. SOTER: A Generative Time-Series Foundation Model for Wearable Human Physiological Signals


122. Available but Unclaimed: An Empirical Study of Human-AI Synergy


123. TAME: Token Attribution and Masking for Emergent misalignment


124. Seeing What Matters: Visual Cue Guided Video Planning for Generalizable Robot Navigation


125. Continuous-Time Machine Learning: A Unified Mathematical Perspective


126. World Models for Embodied Intelligence: From Plausible to Controllable to Actionable


127. Weave: Learning Whole-Body Dexterous Loco-Manipulation from Human-Object Interactions


128. Right Direction, Wrong Step: Geometric Analysis of Finite-Step Failure in Looped Transformers


129. AURA: Agentic Diagnosis and Refinement for Production Recommender Systems at Scale


130. RoleBreak: Benchmarking Long-Horizon Role-Playing Robustness in Spoken Dialogue


131. Structure Across Voices: Comparing acoustic-event type accumulation and sequence dependence across four vocal repertoires using frozen audio encoders


132. Large Language Models in the Loop: A Stability- and Network-Aware Survey in Networked Control, Cyber-Physical, and Multi-Agent Systems


133. A Vision-Language Foundation Model for Precise and Comprehensive Brain Tumor Diagnosis from Preoperative Multimodal Data


134. ProxiDex: Learning Dynamics-Guided Proximity Policy for Dexterous Manipulation


135. Efficient Text-to-Image Generation: An Adaptive Step Schedule Controller for Diffusion Models


136. Vision And Text Transformer For Predicting Answerability On Visual Question Answering


137. The MAL Simulator: Cyber Operations Simulation based on Attack & Defense Graphs


138. A Cyber Range Evaluation of Autonomous Network Incident Response Agents


139. On the Importance of Gating: Memorization vs. In-Context Learning in State Space Models


140. What Does Layer-Importance Reveal About Transformers and State-Space Models?


141. QALPA: Property-guided diffusion modeling for efficient exploration of chemical spaces of flexible molecules


142. Competence-Preserving Resume Perturbations Expose Presentation Sensitivity in LLM Screening


143. Beyond the Name: Demographic Leakage in De-Identified Résumés and Evaluation Artifacts in LLM Bias Audits


144. Geospatial Metadata Improves Discoverability by Connecting Datasets Across Scientific Disciplines


145. AI Policies: Help or Hindrance? A Software Developer’s Perspective


146. Decoder Design Matters for ECG Delineation


147. Protocol-Preserving Context Trimming for Agentic Workflows: Benefits, Failure Regimes, and Budget Guardrails


148. OPD-Aha: From Linguistic Momentum to Visual Reflection in Multimodal On-Policy Distillation


149. Early-Bird Decoding: Accelerating Diffusion LLMs with Learnable Block Sizes and Parallel Sampling


150. Interpreting and Steering LLM Agents for Social Simulations


151. How Good Are Time-Series Foundation Models for Pedestrian Crowd Count Forecasting? A Cross-Dataset Comparative Study


152. On the Expressive Power of Implicit Line-Graph Higher-Order Weisfeiler–Leman


153. Balancing Trial and Reorder: A Hybrid Sequential Transformer-GBDT Ranker for On-Demand Delivery


154. UDAV: Uncertainty-Driven Adaptive VLM Waypoint Planner


155. How Humans and LLMs Read Gender into Gender-Neutral Physical Descriptions


156. Auto-HSI: Personalized human control of a robot swarm on demand by using LLMs for online automatic code generation


157. From Momentary Emotion Inference to Sustained Emotion Support: Evaluating a Companion Agent in a Longitudinal Study


158. FairLint-DL: An IDE-Native Tool for Fairness Debugging of Deep Learning Software


159. Cognitive Admission Control: Risk-Conditioned Assurance for Consequential Actions in Agentic Distributed Systems


160. Efficient One-to-Many Translation with Joint Multi-Stream Diffusion


161. Assurance Envelopes for Autonomous Coding Agents: Minimum-Cost Evidence for Software Change


162. Intelligent Interaction Techniques (IIxT) - Proposal


163. Symmetric solution of the Bellman optimality equation for repeated harmony game


164. ProtoLIP: From Sentence-Level to Object-Level Evidence Disentanglement


165. Recovery Rates Are Not Comparable Across Transcription Factors: Chance Correction for Attribution Evaluation


166. Mapping U.S. Federal AI Governance Against Sector Vulnerability


167. Efficient Reasoning Distillation: Small Video-Language Models via Synthetic CoT and Difficulty-Aware Fine-Tuning


168. Models as Governed Interfaces for AI-Native MBSE: Read-Side Adequacy and Write-Side Admissibility


169. SceneBench: A Hierarchical Benchmark for Vision-Language Understanding of 3D Scenes


170. Permutation-Based Stegomalware in Large Language Models: Threats and Countermeasures


171. AI-Driven Feedback Systems, Digital Labour, and Silent Quitting: Transforming African Workplaces


172. LLMs as Master Forgers: Generating Synthetic Time Series Data for Manufacturing


173. A Decision-Support Audit Protocol for Supervision Drift in Proxy-Labeled Credit-Risk Prediction


174. Universal Defenses for Tool-Integrated LLM Agents Against Adversarial Attacks


175. Coaching Qwen3 Coder 30B to Think Like a CodeClash Arena Agent


176. RAG-CT: Mitigating Privacy Risks on Retrieval-Augmented Generation Systems via Scanning Prompt Distribution


177. MUUNRiver-Bench: Diagnosing Relation-Dependent Music Retrieval with Multimodal Instructions


178. Pseudo-Label Augmentation for Affect Sensing in Small Collaborative Groups


179. The Imitation Game: When LLMs Learn to Reason Like Programs via Code-Centric Reasoning Data Synthesis


180. AssemblyGrid v1: A Benchmark for Multi-Robot Production with Temporary Coalitions, Local Information, and Geometric Constraints


181. The Immutable Past: Formalizing State Mutability and Conflict Resolution in Mutable RAG


182. Schema-Adaptive Action-Conditioned JEPA for Cross-Machine CNC Transfer under Partial Sensor Overlap


183. Efficient Multimodal Generative Recommendation with Latent Narrative Reasoning


184. Beyond Distribution Matching: Semantics-Consistent Tabular Diffusion with Weak Semantic Priors


185. You Don’t Need To Train: Agentic Heuristic Learning Studio for Executable Human Activity Recognition


186. POSPAN: Position-Constrained Span Masking for Language Model Pre-training


187. HintMiner: Automatic Question Hints Mining From Q&A Web Posts with Language Model via Self-Supervised Learning


188. Driver Behavior Estimation at Signalized Intersections Using a Physics-Constrained Decision-Conditioned Autoregressive Transformer


189. OmniHarness: Harnessing Generalizable Visual Generation via Symbolic Policy Learning


190. Managing Action Preconditions in Neuro-Symbolic RL: Three Placement Strategies for Embodied Agents


191. State of Thought Enables Endogenous Reasoning


192. Causal neural set filtering for online multi-target tracking


193. Retrieval-Driven Memory Reconsolidation for Long-Term LLM Agents


194. “Looking for Something Weird to Happen”: How Humans Sustain AI Agent Novelty Amid Semantic Collapse


195. SemanticAdv: Generating Adversarial Examples via Attribute-conditional Image Editing