전체 AI 논문 - 2026-07-28

1. ERUnderstand: Evaluating Vision-Language Models on Structured ER Diagrams


2. Eviction as Estimation: A Fixed-Lag Smoothing View of Test-Time Memory, and When Measuring Beats Accumulating


3. Reason-Mediated Behavioral Models for Auditing LLM Social Simulators


4. Efficiency Matters in Autonomous Research


5. Artificial Intelligence and Innovation Ecosystem: Evolutionary Developments, Challenges, and Future Directions


6. SIREN: Towards End-to-End Extreme-Weather Early Warning with Experience-Grounded LLM Agents


7. LLM-SoccerArena: Benchmarking LLMs on Real-World Predictions in Sports


8. DSCH-Loss: A Dynamic Semantic Channel Objective for Deep Semantic Hashing


9. TRACE-CTI: Auditable Post-Extraction Governance of TTP Claims with Knowledge Graphs


10. Hierarchical Group-Conditional Conformal Risk Control for Selective Prediction in Language Models



12. Task-Conditional Faithfulness Auditing of Multimodal LLMs for Grid Diagnosis


13. Making Mathematical Knowledge Explainable, Accessible and Interoperable Through Large Language Model Integration


14. From Execution to Capability: Scientific Experience Consolidation via Procedural Knowledge Synthesis


15. Failures Reveal What Metrics Miss: An Evidence-Driven Agent for Recursive Refinement of ECG Classifiers


16. Are Prompt Optimizers Blind? Cross-Modal Visual Feedback for Automatic Prompt Optimization


17. Simulating Tenant Responses to Energy Policy Interventions with Transaction-Cost-Aware LLM Age


18. Gubernaut: A Deterministic Homeostatic Controller for Affect-Regulated LLM Agents, Validated Across Independent Model Families


19. Unequal Trips, Unequal Places: Diagnosing and Mitigating Delay Inequity in Autonomous Vehicle Fleet Coordination



21. Generative Artificial Intelligence (GenAI) to convert images of queuing networks into verifiable simulation models: an open-weight LLM workflow approach


22. Epistemic Norms for AI Safety and Alignment Research


23. Integrating Factual and Normative Industrial Knowledge via Constraint-Aware Graph Attention for Process Plan Recommendation


24. Myopia Prevention and Control 3.0: Artificial Intelligence–Driven Risk Stratification, Proactive Monitoring, and Personalized Intervention


25. Falsifiable Commitment Planning for Self-Correcting Web Agents


26. Agent-UCT: Upper Confidence Bounds Applied to Trees for Agentic Workflow Optimization with Cost-Awareness


27. A Motion-Aware Vector Quantization Framework with Centroid Reuse for Efficient VLA Inference


28. Grading the Narrators: An Isnad-Rijal Framework for Claim-Level Provenance in Multi-Agent Knowledge Systems


29. Scaling GUI Agents with Visual State Transitions


30. MemChain: Learning Interpretable Memory Traces for Memory-Augmented LLM Agents


31. Towards High-Level Semantic Intelligence


32. MiSS: A Logic-Driven Explanation of Minimal Sufficient Coalitions for Point Cloud Classifiers


33. The Cost of Knowing: A Resource-Aware Protocol for Benchmarking Hallucination Beyond Static Leaderboards


34. Success Is Not Self-Explanatory: Auditing Success Provenance in Agent Evaluation


35. Quantum-Inspired Evolutionary Neighborhood Search for Arrival-Departure Track Utilization Adjustment under Short-Term Disturbances


36. The Half-Lives of Generative-AI Evidence: A 40-Record Audit, a Claim-Currency Framework, and a Reflexive Case of Frontier-Model-Assisted Research


37. A Cyclic Adaptation-Generalization Framework with Uncertainty-Guided Self-Paced Learning for Long-Term Brain-Machine Interfaces


38. Self-Supervised Consistency Enhanced Disentangled Learning for Neural Decoding Generalization in Brain-Machine Interface


39. Exploring Budgeted Image Classification with Content-Sensitive Resource Allocation


40. Plato-Bio: verification-first biological novelty screening with temporal rediscovery and structural benchmarks


41. Grokking on the Weight-Decay Clock: A Rate Hierarchy from Softly Broken Symmetries


42. EviBack: Search-Agent Reinforcement Learning via Evidence-Constrained Teacher Backoff


43. DICA: Dual-Indicator Guided Contrastive Alignment in Multimodal Large Language Models


44. From Cognitive Architectures to Language Agents: A Mechanism-Level Review of Lineage, Convergence, and Migration Gaps


45. MemTX: Transactional Belief Commit for Stateful Agent Memory


46. Reality Monitoring in Large Language Models: Self-Knowledge That Transforms with Conversation Memory


47. GOTS: Greedy Orthogonal Token Selection for High-Resolution Vision-Language Models


48. Cost-Aware Recovery-Pathway Identification and Bayesian Optimization for Autonomous Materials Discovery



50. Do Visual Features Improve Other-Initiated Repair Detection? A Dyadic Multimodal Approach


51. ACM: Agentic Context Management for Long Horizon Tasks


52. From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement


53. Training Language Models to Cooperate with Inference-Time Controllers


54. E-Bench: Benchmarking Multi-Step Tool-Use Agents in Real-World Product Scenarios


55. Offline-Online Curriculum RL for Multimodal Reasoning


56. Offline-to-Online Creative Optimization with Generative Models and Adaptive Testing


57. Compute Globally, Materialize Locally: The Memory Contract of Sparse Event-KV


58. Focus Is All You Need: Adaptive Goal-aware Attention Orchestration for Multi-Agent Graph Systems


59. SpecAHD: Localize to Specialize for Automated Heuristic Design in Large-Scale Routing Problems


60. Hybrid Advantage Estimation with Unified Critic for VLM Agentic Reinforcement Learning


61. Are You Still the Agent I Authorized? Earned Authority under a Fixed Ceiling for Evolving Agents


62. Verification-Notebook Learning for Source-Aware Multimodal Misinformation Detection


63. ObsDriveBench: Benchmarking Multimodal Understanding under Adverse Weather with Observability Awareness


64. Delegation Intelligence in Deep Search: A Controllable Framework for Disentangled Capability Diagnosis


65. Do LLMs Know Their Vulnerable Scenarios?


66. Separating Capability from Permission: A Governance Framework for Agentic AI Autonomy Levels


67. NeurGO: Learning to Generate Elite Candidates for Meta-Black-Box Expensive Optimization


68. Inference-Time Consensus for Mitigating Hidden Behaviors from LLM Fine-Tuning


69. Key-Interval A*: Accelerating Grid Pathfinding via Structural Abstraction


70. Confidently Wrong: Exception Chain Collapse in Frontier LLM Rule Evaluation


71. ESF-Bench: Benchmarking Challenging Slot-Filling Scenarios for Real-World Enterprise Applications


72. Ordered Network Analysis of Epistemic Emotions during Collaborative Problem Solving


73. RareLens: Towards End-to-End Rare Disease Care via Aligning Divergent Large Language Model Reasoning


74. TopoFE: topology-aware LLM-guided Automated Feature Engineering


75. SeekJudge: A Practical Reward Framework for Reinforcement Learning in Computer-Use Agents


76. CAPT: A Multi-task Continuous Autoregressive Transformer enabling Cross-dataset and Cross-species Transfer for Calcium Population Dynamics


77. Characterisation of Density-based FM generation methods in the context of Information Fusion


78. An Ontology for Machine Learning Interatomic Potentials


79. CachedSearch: Training-Free Cached Exploration for Test-Time Search in Video Diffusion


80. AgentOmnia: Scaling Agentic Models for Full-Scenario Applications


81. SQBench: A Benchmark for Evaluating Task Delivery by Language-Model Agents in Production-Oriented Workflows


82. Compiler-Grounded Hierarchical Diagnosis for LLM-Based Triton Kernel Optimization


83. Structure over Depth: A Single-Block Spatio-Temporal Transformer for Multi-Entity Reasoning


84. SymStep: Symbolic Step Verification for Logical Reasoning


85. Stress-testing large language model agents in a robotic chemistry laboratory


86. Reason Popper-ly: Patching In-Context Reasoning with Inductive Logic Programming


87. ConsistencyGate: Preventing Memory Contamination in LLM Agents via Self-Consistency Admission Control


88. Share No More Than the Request Requires: Federated Disclosure for Perspective-Aware AI


89. Let AI Agents Translate Networks, Not Reason About Them


90. Design Theater: A Benchmark for Generative UI


91. SAGE: Safety-First Defense-in-Depth Guardrails for Verified Lifecycle Control of High-Impact Generative AI


92. Agent Team Work Zone: An Automated, Persistent Workspace for Long-Lived Coding Agent Teams


93. How Well Can AI Generate Backlogs from App Mockups?


94. Physical AI Governance: From Theory to Practice Across Life Cycle


95. What Can Be Enforced? A Theory of Certified Runtime Safety for Tool-Using Agents


96. Coordinated Networking for On-Device Agent-Augmented Real-Time Communication


97. Disentangling Multi-View Scanning in Mamba for Network Traffic Anomaly Detection


98. Commitment To Cooperation With Self-Negotiated Contracts


99. Spatial Reasoning in LLM Game Agents: Impact of Causal Context and Multi-Step Planning


100. SEGRA: Structured Experience-Guided Graph Reasoning Agent for Gremlin Based Question Answering


101. MPR-CiteG: Enhancing RAG with Multi-Portfolio Retrieval and Citation-Grounded Generation


102. RoleMix: Unifying Sequential and Non-Sequential Features via Semantic Tokenization for Post-Click Conversion Rate Prediction


103. Similarity All The Way Up: Multilingual Generalization in LLMs Relies on Language-Level Similarity Structures


104. Test-Time Coverage: Test-Conditioned Data Curation for Deployment-Aware Learning


105. PANOPTICON: A PII-Based Assemblage of Naturalistic Output Tokens for Investigating Privacy Leakage Within LLM Context Window


106. Bayesian Repetition Penalty: A Principled Adjacent-Conditional Framework for Reversing Attention Collapse in Autoregressive Language Models


107. Risk Governance for Generative AI Mental Health Support: A Multi-Turn Safety Architecture


108. HiLLTS: Zero-Shot Hierarchical LLM-Guided Traffic Signal Control for Sustainable Transportation


109. LazyMem: Retrieve Broadly, Construct Selectively for Efficient Long-Term Agent Memory


110. Beyond Sequential Interaction: Benchmarking Parallel Execution and Coordination for GUI Agents


111. Co-Harness: Co-Evolving Harnesses and Model Weights for LLM Agents


112. Imprompt: A Language Framework for Prompt Programming


113. A Vocabulary for Multi-Agent Automated Research Systems


114. DOSA: A Tree-Guided, Self-Regressive Framework for Long Document Structure Analysis


115. How LLM Task-Adaptation Reshapes Alignment: A Multi-dimensional Study of Behavioral and Representational Drift


116. AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models


117. Reinforcement Learning for Heterogeneous Sensor Selection in Maritime Surveillance


118. Obliviate: Efficient Unlearning in Recommender Systems


119. Beyond Block Boundaries: Multi-Block Editing for Diffusion Large Language Models


120. CuraWeb: Joint Optimization of Quality, Redundancy, and Diversity for Web-Scale Pretraining Data


121. TRE: Training-Free Hallucination Detection for Diffusion Language Models


122. StanceBench: A Benchmark for Audio LLM-Based Interpersonal Stance Evaluation from Speech


123. EventOD: Event-Aware OD Flow Generation via LLM-Guided Semantic Modulation


124. MINT-V2X: A Mobility-Integrated Network Trajectory Dataset for Predictive Resource Management


125. Do Language Models Converge to Themselves? Recursive Self-Refinement as Textual Relaxation


126. KG2Code: Bridging Knowledge Graphs and Large Language Models via Executable Code for Question Answering


127. ARdena: Scenario-driven control of real-time LLM agents


128. STAIF: A Stage-wise Optimization for Complex Instruction Following


129. PTStore (Prefix Tensor Store): Distributed Prefix Caching and Replication for High Throughput Inference Serving


130. Extracting Algorithms in Pre-trained LLMs: A Case on Hidden Markov Models


131. DocHRL: A Hierarchical Reinforcement Learning Framework for Cost-Optimised Document Classification


132. Reason Before You Retrieve: Agentic Planning for Multi-modal RAG


133. CRAFT: Learn the Schema, Execute the Plan


134. TRACE: Business Rule-Grounded Reasoning Curriculum for Knowledge-Preserving Parametric Tool Retrieval in Enterprise LLMs


135. Fast Cross-Scenario Adaptation of CSI Models via Channel Conditional Parameter Generation


136. Answering Path Queries under Linear and Guarded Existential Rules


137. CallBench: A Benchmark for Dual-Goal Coordination in Phone Call Assistants


138. PRESTO: Prefix-Aligned Tree Drafting for Diffusion Speculative Decoding


139. Evolving from Lessons: Skill-Augmented Table Graph Reasoning for Operation-wise Table Question Answering


140. VlogReward: Learning Multi-Dimensional Evaluation for Vlog Editing


141. Masked Distillation: Internalizing the Chain-of-Thought in Language Models


142. TokenMem: Faithful Knowledge Injection for Frozen LLMs


143. CHS-SQL: A Text-to-SQL approach based on Confidence-Guided Heuristic Search Schema Linking process


144. Opti-Q: A Constraint-Based Optimization Framework for Multi-LLM Question Planning


145. DynaResize: Runtime GPU Reallocation for Disaggregated LLM Post-Training


146. Decentralized Granular Access Control for Agentic AI Systems in Critical Infrastructure


147. Tokengeist: Multi-Turn Attribution Tracing in Agentic Conversations


148. Evaluating LLMs as Interpretable Controllers for Dynamical Systems


149. Group Preference Collapse in Personalized Multimodal Large Language Models


150. DeepLook: Deeper Thinking with Lookahead


151. Chart Deception in Vision-Language Models: From Vulnerability to Mitigation


152. Differencing the Diffusion Trajectory toward Uncertain Components for Time Series Forecasting


153. HyCE-RAG: Hypergraph Chain-of-Evidence Retrieval-Augmented Generation for Explainable Multi-hop Question Answering


154. An Agentic Orchestration of Atomistic Simulations


155. xMIx: High-Performance Serving-Time Platform for Mechanistic Interpretability Apps


156. Structure Over Scale: Schema-Constrained Causal Graphs for RAG


157. Lexical discovery in unknown environments orchestrated by Large Language Models


158. ParBench: A Benchmark for Reliable Evaluation of LLM Parallel Code Translation


159. TriSP: Tri-Signal Structured Pruning for Large Language Models


160. MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models


161. The Scaffold Effect in Coding Agents: Harness Choice as a Hidden Variable in Coding-Agent Evaluation


162. Source-Aware Reranking for Retrieval-Augmented Generation: A Reliability Prior Approach


163. Multi-Objective Structured Pruning of LLMs for Latency and Model Size Optimization


164. HeraSys: Collaborative Serving of Multiple LLM Workflows via Fine-Grained End-to-End Optimization


165. cMoLLM at Scale: Horizontal Scaling Laws for Mixture-of-LLMs


166. Temporal Context Reinstatement Drives Episodic-Like Order Memory in Long-Context Language Models


167. Too much evidence, too little time: From text to actionable recommendations through multi-objective evidence reasoning


168. PhononBench-MP40: a spectrum-resolved benchmark dataset for phonon stability


169. Schema-Aware Localisation (SAL): Live Schema Grounding and Hallucination Validation for Oracle NL2SQL


170. SCAIR: Schema-Conditioned Agentic Iterative Reasoning for Enterprise Knowledge Graphs


171. Reference Feature Atlases for Mechanistic Auditing of Language Models


172. Execution-Grounded Security Testing for Coding Agents in Software Engineering Pipelines


173. Keyword Matters: Unveiling the Energy Sensitivity of On-Device LLM Prompting


174. MedLoCoMo: A Long-Context Multi-Session Medical Dialogue Benchmark for Large Language Models


175. DSTFView: Multi-View Cloud-Edge Workload Forecasting with Dual-Input Spatio-Temporal-Frequency Modeling


176. Loss-Aware Feature-Map Pruning in Convolutional Neural Networks Using Multi-Armed Bandits


177. Synthetic Scenario Generation for Evaluation of Industry 4.0 Agents


178. SF-AMS: Strategic Forgetting for Structured Memory in LLM Agent


179. Codifying the Judge: Scalable Evaluation via Program Distillation


180. MIITA: Memory-Induced Inference-Time Adaptation for Continual Learning with Small Language Models


181. DeepLens Diagnosis Agent: Agentic Workflow Design Lets a Small Reasoning Model Compete with Frontier LLMs


182. Same Question, Different Answers: Evaluating LLM Reliability Beyond Accuracy


183. QFoldAgent: An Autonomous Quantum Optimization Multi-Agent System for Protein Structure Prediction


184. SeT-Diff: Towards Semantic Foundation Models for HPC Telemetry and Time-Series


185. Concept-based Visual Counterfactual Explanations with Diffusion Models


186. ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding


187. Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation


188. KANEx: Translating Kolmogorov-Arnold Networks’ Interpretability to Medical Explainability


189. The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation


190. DataOrchestra: Learning to Orchestrate Per-Example Curation of Pretraining Data


191. Efficient LLM-Generated Shuttling Compilers for Complex Trapped-Ion Architectures


192. Denial of Deadline: Network-Driven Accuracy Collapse in Distributed Inference Pipelines


193. Co-Learning for Missing Arbitrary Modalities in Multi-modal Classification


194. A corrective agentic hybrid RAG and an operations-grounded evaluation for a scientific facility


195. Sparse Autoencoders Encode Both Concepts and Functions: The Downstream Geometry of Feature Effects


196. Agentic Permissions Policy Algebra for Taint Confinement in LLM Agents


197. Looping Is Not Reliability: State-Bound Evidence and Typed Revision Contracts for Agentic Code Repair


198. Evaluating the Impact of Explainable AI on Trust in AI-Assisted Code Review


199. D-Score: A Spectral Hidden-State Signal for Hallucination Detection in Large Language Models


200. CADER: Confidence-Aware Dynamic Evidence Reasoning for Long-Video Understanding


201. The Visual Bottleneck: Sparse-Frame Adaptation of MLLMs for Joint Spatial-Temporal Video Grounding


202. EgoPlay: Event-Triggered Video Editing for Egocentric Streams


203. BettiSplit: Topology-Guided Privacy-Aware Split Learning Against Feature Inversion and Gradient Leakage


204. LOCKS: Page-Local Compact Key Summaries for Efficient Long-Context Decoding


205. EchoBridge: Long-Tail-Aware ECG-Echocardiography Text Alignment for Echocardiography-Derived Cardiac Findings


206. Stress-Testing EEG Foundation Models for Clinical Decoding: Dataset Identity and Targeted Negative Controls


207. DecoupleMix: Decoupled Ratio Search and Convex Allocation for Scalable VLM Data Recipes


208. UNIFUSION: Adapting Autoregressive Language Models into Discrete Diffusion under a Unified Reverse-Rate Objective


209. ESRVS: Extreme Semi-Supervised Retinal Vessel Segmentation with a Single Annotated Image


210. Evaluating RAG for French immigration law: a benchmark and baseline study


211. LEX-EC: A Lexical Evidence-Channel Audit Framework for Zero-Shot LLM Personality Classification in Black-Box Settings


212. DraftExpert: Expansion-Aware Self-Speculative Decoding for End-Device MoE Inference


213. Multivariate Time Series Forecasting with Adaptive Non-Local Observables


214. The SpiNNaker2 chip: a many-core platform for flexible and scalable brain-inspired computing


215. Regulating for AI Legitimacy


216. MXAttention: Data-Free Optimal Scaling and Pre-Normalization Quantization for MXFP4 Attention


217. Closed-Loop Validation-Repair for Healthcare Interoperability: A Multi-Model Study of Schema Compliance in Clinical LLMs


218. DeepFaith: Evidence-Grounded LLMs for Faithful Incident Reporting in Multi-Stage APT Defense


219. Beyond Aggregate Risk: Role-Stratified Conformal Risk Control for LLM Tool Calls


220. Teacher Knows It Best: Spontaneous Symmetry Breaking and Tipping Points in Networked Langevin Dynamics AI Sycophancy


221. The Tokenizer Tax: Quantifying and Explaining the Cross-Lingual Cost of Subword Tokenization for Indian Languages


222. A Computational Ethical Framework for Financial Digital Phenotyping for Mental Health


223. Physics-Guided Generative AI for Property-Targeted 3D Porous Media Design


224. ML-based Predictive Models for Power Consumption in Virtualised O-RANs


225. FilmBench: A Film-Grade Benchmark for Cinematic Video Generation


226. Every Client Is an Environment: Federated De-confounding for Spatio-Temporal Forecasting


227. StanceFlip: A Comprehensive Multi-Dimensional Benchmark for Multimodal Conversational Stance Flipping Forecasting


228. Not Forgotten: Implementation and Evaluation of a Personalized Episodic Memory for the Humanoid Robot Head Kim


229. Monitoring Post-Disaster Urban Recovery Using High-Resolution SAR Time Series and Unsupervised Learning: Evidence from the 2023 Türkiye-Syria Earthquake


230. EEGForceFusion: Joint Tokenised-Continuous Representation Learning for Subject-Independent Grasp Force Decoding


231. A Case Study on the Acceptance of a Humanoid Robotic Head Employed in Three Public Spaces


232. LU-500: A Logo Benchmark for Concept Unlearning


233. Towards simultaneous decoding of kinetic and kinematic movement parameters during grasp and lift task by noninvasive brain imaging


234. MarineEVT: Advancing Event-Centric Marine Video Understanding via Visual Tool Reasoning


235. ACRL: Adaptive Control of Training-Inference Discrepancy for Stable Reinforcement Learning



237. HELIOS: An LLM-Driven Autonomous Indirect Trajectory Optimization Agent


238. Disentangling Semantic Attention from Structural Bias in the Attention Manifold


239. Agentic Cloud Decoys: A Deception-Driven Framework for Autonomous Intrusion Investigation


240. SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding


241. Adaptive Data Admission and Retention for Streaming Federated Learning


242. Moral Hazard in Multi-Agent Language Models


243. Multimodal Semantic-Probabilistic Objectness for Open World Object Detection


244. Tag Questions and the Generational Reversal of Sycophancy Across 45 Language Models


245. Understanding Machine Unlearning Through the Lens of Mode Connectivity


246. SpecBox: Speculative Sandbox Scheduling for Efficient LLM Agent Serving


247. DuoAD: Leveraging [CLS] Dual Characteristics for Training-Free Few-Shot Anomaly Detection


248. Understanding Tone-Dependent Inference Cost in Large Language Models


249. Embodied GPT-5.1: Evidence of a World Model?


250. Visible to the Court: How AI Is (and Isn’t) Litigated in U.S. Federal Court Opinions


251. Harnessing X-ray Absorption Spectroscopy Data through Multimodal Mining of Battery Literature


252. Physics-Informed Neural Networks for Predicting Nitrous Oxide Flux


253. MulRobBench: A Decision-Level Benchmark for Safe and Security-Policy-Compliant Multimodal UAV Agents


254. A Coulomb Particle Model for Learning Kernel Attention in Transformers


255. Limbomorphs


256. TriShieldRAG: A Three-Ring Defense-in-Depth Framework Against Knowledge Corruption in Retrieval-Augmented Generation


257. Kalypso: Relational LLM Serving


258. Earnings25: A Comprehensive 500-Hour Speech Benchmark for Finance


259. Indic DiarBench: A Multilingual Joint Diarization and ASR Benchmark for Indian Languages


260. A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever


261. How Context Attribution Handles What the Model Already Knows


262. PathScale-R1: Cross-scale Reasoning for Pathological Image Analysis


263. Maximum Satisfiability of Simple Temporal Problems


264. A Few Words Go a Long Way: Language Guided Robot Policy Synthesis


265. Scale Weight Decay and Train Better


266. Outcome-Fair Restless Multi-Armed Bandits for Stochastic Deadline Scheduling


267. WISERouter: LLM Routing with Workload Budget Constraint


268. Escaping the Euclidean Void: Manifold-Informed Flow Matching for Sequential Recommendation


269. AI Strategy: How to Choose What AI Product to Implement


270. An Exact Counterexample to Carlson’s Associated-Prime Depth Conjecture from a Group of Order 128


271. The Illusion of Secure LLM Code: Closing the Security Gap via Iterative Reprompting


272. An empirical investigation into the properties of standard word embeddings


273. Plans Work in Mysterious Ways: Evaluating a Plan Mode for Spreadsheet Agents


274. CALMRec: Causally Aligned Language Memory for Long-Horizon Recommendation


275. Extending Desbordante with Probabilistic Functional Dependency Discovery Support


276. Variational-Ising-Attention (VIA):TailoredAttentionMattersfor Science


277. Order in Desbordante: Techniques for Efficient Implementation of Order Dependency Discovery Algorithms


278. Where Is the Cost of Third-Party API Routers in Agentic Software Development?


279. DualityCert: Verifier-Gated Language-Model Repair of Broken Duality Claims in Quantum Field Theory


280. Action from Adjacent Set in Physical Space Outperforms the Best Prediction in World Models


281. D3O: Dynamic Distribution Distillation for Ordinal Regression


282. An Unofficial FastLAS Tutorial: A Programmer’s Guide


283. GTIN: A Unified Framework for Joint Event and Time Prediction in Temporal Graphs


284. Neonatal Hypoxic-ischaemic Encephalopathy Classification from the EEG and HRV Signals Using a Conformer based Masked Autoencoder


285. Mission-Level Runtime Assurance for LLM-Assisted ISR Swarms over a Verification-Aware Fabric


286. Auditing Alignment Controllability in LLMs via Political Axes


287. Novel Claim or Déjà Vu? Rethinking “Contamination-Free’’ Dynamic Evaluation for Multimodal Automated Fact-Checking


288. Do Diagrams Help Large Language Models Reason? Evidence from Syllogistic Reasoning


289. Choosing a Text Embedding Model: A Practical Benchmarking and Decision Framework


290. Impute On-Demand: Adaptive Correlated Time Series Imputation for Changing Environments


291. Formalizing Flag Algebras in Lean


292. Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation


293. ATLAS: Automated Approximation of Transformers for Efficient Homomorphic Inference in One Hour


294. When Every Simulation Counts: Value-Based Reinforcement Learning for Accelerated Photonics Inverse Design


295. Constraint-Bound Agnostic Bayesian Optimization: One Model for All Thresholds



297. Reasoning or Memorization: Can LLMs Understand and Generate Chinese Xiehouyu Riddles?


298. A Characterization of the Orthocomplement of the Tangent Space of Semiparametric Markov Models


299. TLA$^{+}$-Bench: An Execution-Grounded Benchmark and Dataset for Natural-Language to TLA+ Specification Generation


300. Blood Pressure Estimation from PPG: A Comparative Study of Direct and ECG-Mediated Deep Learning Pipelines


301. Directional Influence Function: Estimating Training Data Influence in Constrained Learning


302. Semantic Semi-Incremental Data-Association-Free Object SLAM


303. When Activation Oracles Learn Not to Read: Concept-Specific Blind Spots in Fine-Tuned Oracles


304. Explaining BiomedCLIP with Weighted Banzhaf Interactions Supported by Tree-Gram Parsing


305. Fair Division with Strictly Increasing Valuations: A Tight Threshold for Two-Agent EF1 and PO


306. On AI Safety and Security Technical Debt in Engineering AI-Enabled Systems


307. Patient-Agnostic Synthetic Pretraining for Efficient Patient-Specific Intraoperative 2D/3D Registration


308. Training with (Swap) Regret Loss in a Single-Layer Self-Attention Model: A Case Study on the Probability Simplex


309. Online Fair Division with Budget Constraints


310. FILLER: Feature Imputation via Latent Location Exploration and Retrieval


311. Statistically Supported LLM Ingredient and Recipe Data Collection in Computational Nutrition


312. What CLIP Knows but Cannot Say: Recovering Negation from Frozen Intermediate Features


313. X-Stage: An Overlooked Pipeline Stage for Communication-Computation Overlap in DiT Inference


314. Context-Aware Concept Distillation for Trustworthy Flood Prediction


315. FedSLIM: Privacy-Preserving Federated MDL-Based Descriptive Pattern Mining Across Data Silos


316. BoneAgeTW2: Automated Skeletal Maturation Assessment via the Tanner-Whitehouse 2 Method, Deep Learning, and Clinical Report Generation with Distribution Curves


317. Fashion-3DLR: A Controllable 3D Garment Generation Using Pairwise Fashion Elements for Intelligent Design


318. Beyond a Global Norm: Personalizing Toxicity Sensitivity in Language Models Without Retraining


319. In-Context Learning as Implicit Policy Gradient


320. False Prophets: On the Security of World Models in Agentic Systems


321. From Vibe to Code – and Back: Lexical Oscillation in the Formation of Design Intent with Generative AI


322. A scalable online machine learning approach for Stock Recommendation


323. KAYROS: An Anytime and Exact Open-Source Solver for Duration-Minimization Time-Dependent Vehicle Routing. A Technical Report and a Case Study in Human-AI Engineering


324. Poster: Rethinking Security in LLM Code Generation through Real-World Risk Scenarios


325. Scoping Review of AI, Metrology, and ESG in the Semiconductor Sector: Implications for Safe and Sustainable by Design (SSbD)


326. Traceable LLM Reasoning for Fake-Order Fraud Detection


327. Attention-Guided Layer Selection for Contrastive Decoding in Large Language Models


328. ADAGE: A Language-Agnostic Pipeline for Analogical Reasoning Evaluation


329. Through the Bottleneck: How Multi-head Latent Attention Separates Content from Position in Language Models


330. MixQuant: Adaptive Mixed-Precision Quantization for Large Language Models


331. Multi-Agent Privacy Game in Federated Learning: A Unified Mean-Field View


332. All in One: Generative Modeling as Mean-Field Game Design


333. VecTree-RAG: An Agentic Retrieval-Augmented Generation Framework Combining Vector and Tree Retrieval for Efficiency and Accuracy


334. Exact values and exact upper bounds for families of integers with arithmetic progression intersections (Erdős Problem #272)


335. Adversarial Test-Hardening for AI-Written Code: An Instrument Autopsy and a Pre-Registered Causal Estimate of the Critic Loop


336. WCM: World-Cognition Model for Generalizable Human-Robot Interaction


337. Real2Sim2Real for Vision-Language-Action Manipulation: An AMD ROCm-Based Pipeline


338. Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning


339. An Explicit Counterexample to Stanley’s Rankwise Lower-Bound Conjecture for Differential Posets


340. Label-free Industrial Fault Detection via Adversarial Inverse Reinforcement Learning: A System for Run-to-Failure Prognostics


341. HALLELUAI: A Hallucination-Aware AI System for Ultra-Realistic Image-to-Video Generation at Scale


342. Modeling Memory-Dependent Reliability of LLMs: A Hidden Markov Model


343. Building AI That Works: ESnet’s Pragmatic Approach to AI-Driven Operational Excellence


344. Invariant Discovery for Networked Systems


345. Not All LLM Reasoning is Visible in the Chain-of-Thought


346. Controlling Embedding Spaces with Text-Conditioned Transformations


347. Evaluating and Mitigating the Misguidance Effect of Buggy Code in LLM-Generated Unit Tests


348. Do Coverage and Mutation Scores of LLM-Generated Test Suites Correlate with Their Effectiveness? (Replicability Study)


349. Spatial Prediction of Soil Microplastics and Organic Matter Using Graph Attention Networks



351. AI-interpreted Optical Scattering for Robust and Focal Depth-Aware Imaging


352. Spatial-IQ: Deconstructing Spatial Intelligence via Hierarchical Capability Tests


353. Robustifying pathology foundation models via fine-tuning


354. Language-Routed RAG and Direct Option Scoring for Multilingual Financial QA: DS@GT at FinMMEval


355. Frustratingly Simple Black-Box Adaptation of Language Models via Logit Bias


356. Agentic Autoresearch for CT Reconstruction


357. From Hybrid Mechanistic–Data-Driven Modeling Toward Neuro-Symbolic AI: What, Why, and How


358. Hybrid Semantic and Spectral Ensemble for Robust Synthetic Image Source Attribution


359. OrchNAS: Orchestrated Neural Architecture Search Service for Personalised Federated Edge Intelligence


360. LithoFormer: A Robust Framework for Stratigraphic Inference via Transformers


361. Physically Verifiable Evidence and LLM-Based Reporting for Bearing Fault Diagnosis


362. Multimodal Domain Generalization for Depression Detection: An Attention-Based BiLSTM Network with Domain-Adversarial Training


363. FMOPF: Latent Flow Matching with Constraint-Aware Interaction Priors for AC Optimal Power Flow


364. Optimizing Transformer Neural Network for Real-Time Outlier Detection on FPGAs


365. What Softmax Throws Away: Mass-Aware Attention for Evidence Accumulation


366. Multimodal Surface EMG Hand Gesture Recognition Using Query-Based Transformers for Prosthetic Control


367. LC-SEPLM: long-range contact-supervised adaptation for sequence-only protein representation learning


368. Cheap Probes Predict Expensive Training in 3D-CT Vision–Language Models


369. DomainPilot: Domain-Level Loss-Guided Two-Stage Data Mixture Optimization for Efficient Language Model Fine-Tuning


370. Beyond Shapley: An Influence-Based Data Auditing Pipeline for LLM Alignment and Evaluation


371. Spectral Dynamics of Semantic Drift in Clinical Multi-Agent Language Model Networks


372. Hierarchical Grading in Large Language Models


373. Real-time Reconstruction of Human Visual Perception from fMRI


374. Post-Operative Glioma Segmentation via Loss Stabilization, Normalization and Subspace Attention


375. Learning to Access Computation: Accessibility Plasticity as a Principle of Adaptive Intelligence


376. Advancing All-Weather Building Damage Mapping to the Instance Level: Outcomes and Insights from the 2026 Bright Challenge


377. AI-generated Images Challenge Visual Trust in High-risk Scenarios


378. QFedPolyp: A Communication- and Inference-Efficient Federated Learning Framework for Polyp Segmentation


379. Cortex: Compact Behavior Cloning for Quake with Frozen Visual Features


380. Structural Preservation Governs Data Augmentation in Deep Learning-Based Laser Speckle Material Classification


381. Progress-conditioned Group Policy Optimization for Long-Horizon Agentic Tasks


382. A New Kind of Adversarial Example: Measuring the Human-Model Gap, and Its Relationship to OOD Detection


383. An Interactive Vision Language Platform for Cognitive Remediation in Schizophrenia


384. DAMamba-UNet3D: A Parameter-Efficient Mamba State Space U-Net with Dynamic Adaptive Scan for 3D Medical Image Segmentation


385. Real-Time Semantic Segmentation with Optimized RetinaNet Architectures for Embedded Automotive Systems


386. scMIR: a vision-language foundation model for single-cell light microscopy image representation


387. CORVUS: Context Optimization and Reduction Via Underlying Synchronization for LLM Coding Agents


388. RMS@CC-MMD 2026: Multimodal Misogyny Detection via Geometric Interaction and Multi-View Consensus


389. Visible-Light Imaging Diagnosis of Neutral Particle Emission Tomography in the Tokamak Divertor: An Efficient Transformer-based Surrogate Model


390. MegaSlide-DiT: Memory-Centric Adaptation and Deformable Local Attention for Efficient Video Diffusion


391. Towards Nexus-Score: Metadata Gaps Limit Scholarly AI Attribution


392. SetGo: Metadata Readiness for Scientific AI Datasets


393. Between Suppression and Collapse: Evaluating Narrative Unlearning with LENS



395. AI-Assisted Causal Inference and Mediation Analyses of Environmental and Psychosocial Determinants of Subjective Cognitive Difficulties in the All of Us Research Program


396. Learning When to Reason for Text-to-SQL via SFT and DPO


397. Masked Autoencoders Learn Perception-Relevant Representations from Resting State Neural Data


398. Revitalizing Public Urban Places through Cultural and Political Memory: A Technological Approach with LLMs and Augmented Reality


399. Quotient Tree Arithmetic: Deferred-Division Computation with Bounded Symbolic Depth and Cross-Subtree Cancellation


400. A didactical-driven teacher assistant for a dimensional modeling course


401. A Formal Kinetic Theory for Zeroth-Order Newton Dynamics:Stein-Corrected Hessian Estimation and Curvature–Variance Trade-offs


402. Evaluating the Impact of Reviewer Guideline Design on LLM-Based Automated Peer Review


403. Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B


404. Comparing Optimization Models for Radiotherapy Scheduling


405. Evaluating Large Language Models for Symbolic Security Protocol Analysis


406. Creative Integration: A Decidable Criterion of Creativity