전체 AI 논문 - 2026-08-10

1. Interaction Creates Dynamical AI Behavior Absent in Isolation


2. SkillProx: Self-Evolving Agent Skills via Proximal Textual Gradient Descent


3. Blast Radius


4. PsychoAgent: An Affect-Sensitive Cognitive Architecture for Conflict-Aware Memory in LLM Agents


5. Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing


6. Post-Grokking Collapse at the Representation-Readout Interface in Muon-Trained Transformers


7. TEPA: Revoking Stale Memories for Conflict-Robust Language Agents


8. A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs While Improving Accuracy


9. CoBa: Cost-Effective Test-Time Scaling via Compute-Balanced Routing


10. ResidencyRL: Reinforcement Learning in Simulated Clinical Environments



12. FinRank: An Evidence-Grounded Benchmark for Financial Question Answering and Retrieval over SEC Filings


13. People Are Not Just Their Countries. Disentangling Social Determinants of LLM Value Alignment Across Europe


14. Curriculum as Code: An AI-Assisted Architecture for Instructional Design in STEM Education


15. QFCQT: A Chaotically Gated Quantformer Framework for Volatile Time-Series Forecasting


16. An End-to-End Agent Auditing Engine


17. Winning by Peeking: Unenforced Budgets and Test-Set Selection Inflate Short-Budget AutoML Comparisons


18. WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN


19. Recipes for Creativity: Iterative Generation and Evaluation in Large Language Models


20. From probability to causality in probabilistic logic programming


21. Beyond the Black Box: Interpretable Models of Human Randomisation Failures


22. Authoring and Management of Transparent Research Integrity Assessments of Randomised Clinical Trial Publications Using LLM-assisted Tools and Provenance Knowledge Graphs


23. EMAS: Stabilizing Multi-Agent System Evolution through Evidence-Guided Revision


24. SetEasy: A Multi-Modal Classroom Engagement Assessment and Seating Optimization Framework


25. Agent Memory Distillation: Empowering Small LLM Agents with Hierarchical Teacher Memory


26. NiyamAI - An Intent-Bound AI Agent with Cryptographically Verifiable Guardrails using Zero-Knowledge Proofs


27. A MARL Centered Reference Architecture for Large Language Model Augmentation in Smart Manufacturing


28. DiDPO: Diff-in-Diff Policy Optimization for Coding Agent Training


29. How Much, Then Where: Credit-Conserving Action-to-Token Allocation for Multi-Turn Agent Reinforcement Learning


30. MemWM: Memory-Augmented Text-Based World Model


31. Transformers Struggle to Use Their Emergent World Models: Revisiting the Tower of Hanoi, and the Illusion of Thinking


32. MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents


33. DocMemo: Dynamic Evidence Discovery via Probabilistic Memory-Guided Retrieval for Multi-Modal Document Understanding


34. PTQ4SNN: Membrane-Aware Post-Training Quantization for Spiking Neural Networks


35. BONSAI: Evolvability-Guided Tree Search over Skills


36. Unsupervised Adaptation of PDE Foundation Models


37. Not All Problems Are Best Modeled as MILP: A DSL-Centric Framework for Flexible and Accurate Optimization Modeling


38. ZIPBrain: Can EEG Foundation Models Be Faster, Locally Deployable, but Accurate?


39. ReQuant: Fixed-Grid Discrete Refinement for Post-Training Quantization


40. FedLBW: A Loss-Based Weighting Strategy for Federated Learning on Non-IID Data in Wireless Networks


41. Finding Usable Weight Mechanisms with Tiled SVD


42. Learning in Deep Networks under Dale’s Constraint


43. CAi Copilot: Reducing Operational Workload in Molecular Design through Intent-Driven Agentic Workflows


44. Critical Acclaim Orientation in Large Language Models: Evidence from Film Preference Elicitation


45. Does Splitting a Triage Decision Across Agents Hide Bias or Help Catch It? A Multi-Agent Simulation Study of LLM-Based Resource Allocation Under Audit Capacity Constraints


46. LMM Modality Transfer: A Pre-requisite for Autonomous GIS Agents


47. Blind to the Pivotal Vote: Aggregate Independence Metrics Miss Where Verification Actually Helps


48. Science Edge Evaluation: SEE the Missing Step Toward Real Scientific Discovery


49. TRIBE: Predicting Team Performance via Communication Behavior Ensembles


50. Deal Me Maybe: The Role of Emotions in Multi-Agent Negotiation


51. ReGraph: Learning to Generate Recipe Graphs from Food Images


52. Fast LapSum: Exact Differentiable Top-k at Million Scale


53. Long-Horizon Agent Trajectory Attribution: A Unified Benchmark and Fine-Grained Annotation Framework


54. From Points to Edges: Edge-Conditioned Spectral Operators for Physics-Sensitive PDE Learning


55. SkillEval: Decomposing Agent Skill Quality into Interpretable Signals


56. CEDAR: Agent-Orchestrated Tree Search for Goal-Directed Optimization of Complex Systems


57. Gated-BEPO: Confidence-Gated Bellman Credit Assignment for Large Language Model Agents


58. Evolving Parallel Algorithm Portfolios via Potential-Aware Instance Generation with LLMs


59. Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts


60. LiFTER: A Grounded Neuro-Symbolic Microscope for Continuous-Time Dynamic Graph Forecasting


61. Capek 0.5: An Execution-Centric Vision-Language Model for Embodied Intelligence


62. Mind the Gap: A Dual Knowledge Graph Framework for Unified Multi-task User Intent Inference


63. MemPrism: Task-Conditioned Relational Memory Views for Long-Horizon Agents


64. IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents


65. From Cheap Fakes to Pure Synthesis: Addressing the New Era of T2V Fake News Videos


66. bioMoR: Biology-Guided Mixture-of-Recursions for Effective Genomic Learning


67. The Optimizer Is the Agent: Reasoning-Driven Search across Prompts, Programs, and ML Workflows


68. MolBioKG: Grounding Out-of-Graph Molecules in Biomedical Knowledge Graphs via Multi-Resolution Structural Anchoring


69. WebRider: Persona-Conditioned Intent Controllers for Live-Web Assistance


70. AgentPatch: Coarse-to-Fine Weak-Task Repair for Merging Agentic Multimodal Large Language Models


71. A Multi-Agent Framework for Automated Coarse-Grained Molecular Dynamics of Polymers


72. Vehicle routing problem using deep reinforcement learning - A case study about truck planning in the industry


73. CellWorld: From Gene-Level Reconstruction to Latent Cell Prediction in Spatial Transcriptomics Foundation Models


74. TRACE: A Multi-Layer Benchmark for Human AI Controller Coordination Under Drift and Failure


75. Shape Your Feed: An LLM-based Agentic System for Conversational Recommendation


76. NxN E-valuation: Hypothesis Certification via a Conformal CRT Null


77. Automated item evaluation: Predicting item acceptance and rejection using LLM-generated critiques


78. Divergent Response Modes in Frontier Language Models Under Steering Pressure


79. TaskSense: Focusing on What Matters in World Models


80. KNOWPLAN: Knowledge-Driven AI Agents for Smart Degree Pathway Planning


81. Can MLLMs Decode the Creative Leap? Introducing C4 for Cross-Concept Understanding


82. WebGrader: Training LLMs for Web Development with Self-Evolving Programmatic Grader


83. Learning to Predict Middle-Layer Attention in MLLMs for Visual Token Prunin


84. ADIAS: Automated Design of Interactive Agentic Systems


85. Interpretable Unsupervised Community Detection with LLM-Symbolized Structured Processes


86. Beyond Routing Weights: Faithful Response-Level Interpretation of Mixture-of-Experts Reward Models via Contribution Contrast


87. EntropyMoE: Entropy-Aware Sparse Expert Routing for Tokenizer-Free LLMs


88. Towards Multi-Label Graph Foundation Models: from Single-Vector Representation Learning to Multi-Semantic Basis Learning


89. CreativeInstruct: Scalably Teaching LLMs to Balance Quality, Creativity, and Diversity


90. CoinRAG: Contextualized Information Nugget KV Cache Reuse for Long-Context RAG


91. Strategy-first synthesis planning for complex natural products


92. Taxonomy-Driven Analysis of Open-Source AI Risk Mitigation Tools


93. SABRE: Scalable and Automated Benchmarking of VLMs under Stress


94. Diffusion LLMs as Targets and Adversaries: Mechanistic Safety Exploits


95. I Seek You in Videos: Identity-Conditioned Queries for Person-Centric Video Reasoning


96. GeoDistill-Refine: Silhouette-First Geometry Distillation for Annotation-Free Spacecraft Segmentation


97. PACE: Primitive-Aware Code Evolution for Automated Algorithm Design


98. Omni-modal decomposition autoencoders learn full-stack wearable disentangled representations


99. LSEAD: A Privacy-Preserving LLM-Based Speech Analysis Framework for Early Alzheimer’s Disease Screening


100. Measurements Automatically Extracted from Zero Echo Time MRI Using Deep Learning Image Segmentation and Geometric Modeling Agree with Expert Manual Readings


101. Assessing AI-generated music detection in real-world broadcast monitoring


102. Geo-Spatial Concept Probing of Large Language Models: Abstraction, Compositionality, and Grounding


103. Zero Gap Is Not Restoration: Stratified Per-Question Probability Evaluation and Step-wise Mitigation of Benchmark Contamination


104. H2AL: Hyperbolic Hierarchy-aware Aggregative Learning for Registration-based Few-shot Medical Image Segmentation


105. Aftab: A Comprehensive Benchmark of CNN Encoders and Advanced Value Functions in Parallelized Q-Networks


106. Towards Assurance Closure in AI-Native Large-Scale Agile Software Development


107. Natural Language Processing Psychometrics


108. Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination


109. EliSeg: Verified Target Construction for Report-Grounded Abnormality Segmentation


110. FUSE: Feature-Wise Unified Specialization with Cross-Column Exchange for Mixed-Type Tabular Flow Matching


111. How Much AI Is in This Track? Quantifying the Proportion of AI-Generated Stems in Hybrid Music Mixtures


112. A Finite E-Group of Nilpotency Class Three


113. TOFD: Target-Oriented Feature Decoupling against Poisoning Attacks in Split Federated Learning


114. SCALE: Scientific Concept Aggregation via LLMs and Embeddings for Fine-Grained Taxonomy Extension


115. Reading Copom’s Tone: A Weighted LLM Framework for Hawkish-Dovish Sentiment, Forward Guidance, and Uncertainty


116. Artificial Intelligence Can Match Domain Experts in Evidence Extraction and Critical Appraisal of Microbial Oncogenesis Research Publications


117. Toward a Causal Data Management Ecosystem for Decision Making and Agentic AI


118. Measuring Concept Content in Text from LLM Activations: ESG Evidence from Concept Vectors and Linear Probes


119. Momba: Network Modernization Improves Multi-Objective Reinforcement Learning


120. Representation-driven Endoscopic Visual Embedding Alignment for Latent Generation


121. Fluid-DiT: Graph-Free Diffusion Transformers for Fluid Flow Simulations Learning


122. Representation Handoffs for OpenArm-Based Laboratory Mobile Manipulation


123. Interpretable reinforcement learning with decision-tree pruning


124. Autonomous discovery of accelerator commissioning algorithms


125. PHOENIX: Fine-Tuned SLM-Powered Autonomous Satellite Lifetime Extension via Predictive Self-Healing and Multi-Agent AI Recovery


126. Geometry-Aware Camera Localization for Bronchoscopy


127. International Transfer of Stochastic Cortical Self-Reconstruction


128. Human-Centered Explainable AI for TinyML Edge Devices: A Pareto-Based Selection Framework with LLM-Guided Design


129. RoRA: Role-Oriented Regional Allocation for Visual Token Pruning in MLLMs


130. Beyond Isolation: Unlocking Reinforcement Learning Component Synergy for Sample-Efficient Continuous Control


131. LifelongCrossNav: Persistent 3D Semantic Memory for Cross-Floor Multi-Object Navigation


132. Scalable High-Fidelity Macromolecular Docking for GPU-Accelerated Supercomputers


133. AutoIntervene: Calibrated Intervention for Action-Chunking Imitation Learning Policies


134. Soft Redaction of Image Provenance via Zero-Knowledge Proofs


135. Beyond Text Matching: Towards Reference-Free Evaluation for Human-Oriented Binary Reverse Engineering


136. Accounting Graph Transformer for Short-History Multi-KPI Forecasting in Small Businesses


137. An Agentic Hybrid Top-Down and Bottom-Up Approach to Knowledge Graph Generation


138. Decoupling Intention from Trajectory: A Representational Deduction Framework for World Action Models


139. Beyond Foundation Models: Dimension-Aware Neural Architecture Search with Small-Data Representation Models for Cryocooler Lifetime Prediction


140. GPTKB 2.0: Browsing, Querying, and Auditing a Disambiguated LLM-Derived Knowledge Base


141. Density-aware Hierarchical Clustering Based on Element-Categorized Connection Subgraphs


142. HarnessSafe: Evaluating Safety Across Persistent Carriers in Agent Harnesses


143. PHASE-Tree: Modeling Character-State Evolution in Long-Horizon Role-Playing Dialogue


144. Same physical state, different collective dynamics: state encodings select synchronization outcomes in language-model agents


145. Explicit, Not Longer: What Makes Epistemic Stance Survive Memory Compression


146. Debias in Text, Believe Your Eyes: Text-Anchored Cross-Modal Transfer for Visual Counter-Commonsense Reasoning


147. Ask-E: An Environment for Calibrated Question Generation


148. MaskFlow: Precise, Consistent and Seamless Regional Image Editing


149. Calibrating WEAT Against Anisotropy: ZCA Whitening as a Geometric Pre-Processing Step for Embedding Association Tests


150. Georeferencing Non-Gazetteered Place Names using Biological Specimen Records


151. FedVAR: Prototype-Aligned Federated Framework for Video Anomaly Recognition


152. Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection


153. Bridging the Gap Between Hyperdimensional Computing and Kernel Methods via the Nyström Method


154. Autonomy-of-Heads: Data-Free Sparse Attention from Frozen Query-Key Geometry


155. Investigating Quantum-Embedded Transformers on Classical Datasets for Cross-Modality Classification


156. Control-Anchored Residual Flow Matching Conditioned on Gene Geometry for Virtual Cell Perturbation Modeling


157. FutureBridge: Token Selection Beyond Local Preference in Collaborative Decoding


158. Coupling Planning with Episodic Memory in LLM Agents for Software Issue Resolution


159. LoRAScan: Detecting Backdoor Prompts in Low-Rank Adapters for Large Language Models via Down-Projection Activation Spikes


160. Progressive Alignment of Recommender Foundation Model through Multi-Phase Post-Training


161. HLSmith: An Expert-Guided Agentic Framework for C/C++-to-HLS Translation


162. Genotypic Triggers: Exposing Pharmacogenomic Blind Spots via Host-Specific Backdoors in Generative Antimicrobial Peptide Models


163. Hidden Gauge Controls Feature Specialization in ReLU Networks


164. Beyond Starry Night: Shortcut-Aware Control-State Planning for Artist-Grounded Text to Image Generation


165. Progressive Content Refinement with Decaying Reward Joint LinUCB


166. KReF: Training-Free Retrieval for Long-Term Time-Series Forecasting and Predictive Uncertainty


167. Multi-Level Modeling of Large Language Model Inference Latency and Energy via Hybrid Analytical–Machine-Learning Predictors


168. Dueling World Models: Advantage-Style Action Channels for Common-Mode Distractor Rejection


169. Scalable Long-Horizon Planning with Staggered Updates for Lifelong MAPF


170. Online Monitoring and Corrective Steering of Programming Agents


171. Policy-Masked Private Experts: Auditable and Reversible Capability Access Control in Sparse MoE Models


172. SoRoMoX: Fast, Differentiable, and Parallelizable Soft Robot Models


173. Characterizing the Quality Profile of AI-Generated C++ in Production


174. MI-MIDI: Mechanistic Interpretability of Text-to-MIDI Generation Models via Probing, Lenses and Steering


175. Bypassing Krum: Selection-Aware Backdoor Attacks in Federated Learning


176. Cryptanalytic Extraction of Isolated Bias-Free GLU Feed-Forward Blocks by Antipodal Separation


177. Factorized Hypothesis Search for Evidence-to-Taxonomy Retrieval


178. Do 3D Medical Foundation Models See Through MRI Artifacts? A Controlled Study of Representation Robustness


179. SLED: Scalable Location Encoding via Distillation


180. Flowing Through States: Neural ODE Regularization for Reinforcement Learning


181. Beyond “AI Language”: The case for the idiolectal nature of LLM output


182. SyncSBC: Decentralized Swarm Behavior Prediction for Synchronized Autonomous Control


183. TradeVerse: A Longitudinal Benchmark of Political Negotiation in International Trade


184. CertBind from Multimodal Connectivity to Certifiable Retrieval Decisions


185. Agentic AI: User Empowerment or Enclosure?


186. Do AI Personas Grow? Analyzing and Benchmarking Personality Evolution in LLM Agents After Life Events


187. LyEvO: Lyapunov-Guided Evolutionary Optimization for Safe and Robust Sim-to-Real Policy Learning


188. StepJack: Benchmarking Computer-Use Agent Safety Against Multi-Step Indirect Prompt Injection


189. CyberForge: Verified Vulnerability Injection at Repository Level for Cybersecurity Agent Training


190. ED-CSP: Crystal Structure Prediction from Electron Diffraction


191. Risk-Aware Decision Policies for Agents Under Noisy Perception


192. WorldMark: A Plug-and-Play World Knowledge Interface for Cross-Host Language Model Watermarking


193. Separating Decision-Rule Misalignment from Readout-Coverage Limitations in Speech Language Models


194. TransSLR: A Lightweight Transformer for Sign Language Recognition


195. Recovering Explanations from Transformed Rule-Based Ontologies


196. Agentic Planning for Symbolic Execution


197. TEXAS: Task-Expert-Aware Supervision for Downstream Mixture-of-Experts LLM Adaptation


198. Evaluating XAI Support From A Hierarchical Reinforcement Learning Policy in Human-Agent Collaboration


199. Mobile Interaction for Assessing Fatigue, Sleep, and Activity in Neurodegenerative and Chronic Diseases


200. Multimodal Drivers’ Emotion Recognition and Safety-Oriented Intervention for Intelligent Transportation Systems


201. Mitigating Scoring Bias in LLM-as-a-Judge via Random Number Generation