전체 AI 논문 - 2026-09-15

1. Corrupt Plans, Clean Traces: Evading Chain-of-Thought Monitoring with Plan Injection


2. Stellar Colosseum: A Many-Agent Harness for Long-Horizon Research in Mathematics and Theoretical Computer Science


3. Recurrent GraphNeural NetworkswithSet-BasedAggregation


4. Pilot Early, Commit Late: A Real-Options Model of Enterprise AI Adoption under Rapid Technological Progress


5. LongAgent: History-Guided Agentic Search for Longitudinal Outcome Prediction


6. AlgoEvo: Self-Evolving Agentic Search for Automated Algorithm Discovery


7. Atria Dawn: The Dawn of Agentic Superintelligence


8. When Should a World Model Move? Loss-Conditioned State Execution


9. Navigating Sparse Evidence: Agentic Visual RAG via Explicit Context Selection and Consolidation


10. KnowBench: Effort Reduction as a Unified, Deployment-Grounded Benchmark for Clinical AI


11. EvoOntology: A Self-Evolving Ontology Layer for Data Agents


12. Design of a Deep Learning Credit Risk Early Warning System Integrating Multi-source Heterogeneous Data


13. Are LLMs Good Financial User Simulators? A Preliminary Study


14. Data storytelling meets interpretable machine learning: Decoding AI decisions for non-experts without revealing sensitive data and model details


15. Predicting build orientation for SLM dental parts: a comparison of rotation representations and direct vector regression


16. New Conditions for Philosophers to Catch the Wave of Citizen Deliberation in the Age of Artificial Intelligence in advance


17. Beyond Accuracy: Robustness, Cost, and Governance Trade-offs for Vision-Language Models in Templated Document Extraction


18. NoteVQA: Benchmarking VLMs on Real-Life Questions from Human Communities


19. EEG-Xplain: Decoding Neural Black-Boxes of EEG Foundation Models


20. Potential of Artificial Intelligence Algorithms for Identification of Relevant Diagnostic and Prognostic Biomarkers of Early-Stage Liver Cancer


21. Diversified and Perceptible Counterfactual Examples Leveraging Expert Knowledge


22. GRIN+: Towards Fast Yet Effective Machine Unlearning for Imbalanced Medical Data


23. Option-Aware Retrieval and Task-Specific VLM Adaptation for Medical VQA


24. Beyond Safe Answers: Segment-Aware Listwise Alignment for Reasoning Safety in Large Reasoning Models


25. The Troy Moment of AI: Why SomeWill Cheat and SomeWill Follow?


26. HISPO: Hierarchical Importance-Sampling Policy Optimization with Entropy-Derived Segments


27. Empirical Evaluation of Task-Based Permission Scoping Architecture for AI Agents


28. Can AI systems have free will?


29. Who Teaches Which Token? Verifier-Gated Multi-Expert On-Policy Distillation for Scientific Reasoning


30. When Tool Calls Succeed but Workflows Fail: Anomalies at the Agent-Tool Boundary


31. SkillLift: Learning Dense Rubrics from Sparse Oracles for Efficient Skill Evolution


32. RSIAgent: Autonomous Exploration for Recursive Self-improvement in New Environments


33. MAPS: Memory-Aware Predictive Scheduling Framework for Large Language Model Serving


34. Parameter-Efficient Adaptation of Pretrained Language Models for Time-Series Forecasting


35. Evaluation Metrics for Safe Reinforcement Learning


36. Reason What Matters: Retrieval-Grounded Reasoning for Universal Multimodal Embeddings


37. Why LLM Agents Collapse Without Oversight: The Enforcement Gap as the Mechanism Behind Emergence World Failures


38. ProIQA: A Process-Based Framework for Fine-Grained Math Item Quality Assessment


39. Empirical Evaluation of Open-Source Large Language Models for Retrieval-Augmented Generation in ESG Domain


40. CWM: Controllable White-Box Meta-Prompting for Adaptive Retrieval-Augmented Generation and Reasoning Ability


41. From Ideas to Actions: A Public-Data Decision-Support Toolchain Across the Venture Lifecycle


42. Issue Bias in Generative AI Writing Assistance: Political Issues and LLMs in the Swedish 2026 Election


43. VisInteract: Towards Dynamic Interactive Text-to-Visualization under Imperfect Queries


44. STHMoE: Hypergraph-Enhanced Heterogeneous Dependency Coordination for LLM-Based Urban Traffic Data Forecasting


45. T-LoopFormer: Token-Level Elastic-Depth Looped Transformers for Latent Reasoning With Dynamic Routing


46. HazardAuditor: From Executable Threats to Safer Computer-Use Agents


47. Medical Knowledge Simplification for Patients in the Era of LLMs: A Case Study on Diabetes


48. OpenAl4S: Code as Action, Science as Sessions


49. ER-EDF: A Psychology-Grounded Emotion Regulation Framework for Speech Empathetic Dialogue Generation in Large Audio-Language Models


50. Enabling Creative Exploration for Vibe Design Agents


51. BusMA: A Bus Communication Substrate for Multi-Agent Systems


52. The average-farmer illusion in language-model simulations of agricultural decisions


53. Horizon-specific Expert Fusion for Photovoltaic Power Forecasting


54. Four Ledgers, Not One Score: Responsible Communication of LLM-Judge Calibration in Biomedical ML


55. Overflip: Repetition-Induced Label Flips in Guardrail Models


56. Semantic-TVM: Structure-Preserving Trustworthy Virtual Memory for Memory-Augmented and Tool-Using Agents


57. CoMem: Collective-Individual Memory Synergy for Evolutionary Multi-Agent Systems


58. Shallow Beliefs: Synthetic document finetuning does not inoculate against emergent misalignment from reward hacking


59. Converting Sequenced Fuzzy Cognitive Maps to Causal Virtual Worlds with Large Video Generators


60. MemRiskBench: Trace-Aware Risk-Preserving Evaluation for Long-Horizon LLM Agents


61. Towards a knowledge-enhanced single-cell foundation model


62. Geometric Flow enhanced Graph Coarsening


63. Externalizing Requirement-to-Repair Artifacts as Observable Traces for LLM-Based Program Repair


64. GGUF-Metadata Prediction of Single-Sequence llama.cpp Throughput Across Three Systems


65. Domain Generalization for Smartphone-Based Human Activity Recognition: A Systematic Analysis of Components and Interactions


66. Self-Orchestrating Language Models: Leveraging Semantic Dependence for Efficient Inference


67. El Agente Potente: High-Throughput Agentic Atomistic Simulations


68. One Model, Two Physical Stories: Auditing Misalignment in Multi-Modal World Modeling


69. ANASSA: An Agentic AI Orchestration Framework for Spatial Intelligence


70. Crypto Accounting Bench: Evaluating Frontier and Open-Weight Models on Crypto-Asset Accounting Tasks


71. Another Blueprint In The Wall: How to Ask Frontier AI Like a Kid?


72. AI Persuasion as a Threat to Human Control


73. AcquireBound: Runtime Authorization for Resources Acquired by AI Agents


74. AppliedScientist: Automated Scientific Revision Through Iterative AI Reviewing


75. Bayesian Intelligence from the Outside


76. Moral Rebel Agents: Decision-Making Under Conflicting Obligations


77. Depth and Scale in the Sub-150M Regime: JugnuLM-53M vs JugnuLM-110M


78. Lightning Weave: Improving the Accuracy-Efficiency Frontier of Reasoning Models through Capability Composition


79. Diffusion-Based Generation of Gait Trajectories


80. DynSTEER: Dynamic Stage-wise Trajectory Evaluation and Execution-time Review for Agents


81. A note on goal-based hierarchical RL


82. AI Deployment Accountability Engineering: A Vision for Accountable AI in Safety-Critical Socio-Technical Systems


83. OptoAgent: A Trustworthy Multi-Agent Framework for Opportunistic Vision Micro-Screening in Classroom Environments


84. Beyond Scene Description: Multi-Agent Orchestration for Non-visual Access to Virtual Worlds


85. When does a scaling result justify a different allocation? A critical review of resource-allocation evidence for AI systems


86. Safety Signals to Verify NetOps Agents with Action-Level Granularity



88. Dynamic Learning Solutions: A System for Personalized Educational Video Generation


89. MOSCOPT: Mixture-of-Skills Collective Optimization for LLM Agents


90. Semantic Knowledge Technologies: what the Semantic Web lost sight of, and what it never had


91. VeriDx: Earning the Right to Diagnose with Disease-Centric Verification


92. Schizophrenia Detection from EEG Signals: A Transformer Framework with Spectrogram Representation


93. Convergent Emergence of In-Context Learning Across Modalities


94. Synthetic Data in Marketing Research: How to Evaluate and When to Trust


95. SAILOR: Solver-Assisted Interactive LLM-based Optimization Recovery


96. LoRA Fine-Tuned Models for Control Systems Course Q\&A: A Multidimensional Evaluation of Model Scale and Rank Effects


97. Map Users and Mapmakers: The Scope of Cognitive Attribution from Acquired Representations


98. ClinAgent: A ReAct-Based Agent for Conversational Access to Clinical Trial Information


99. UniCAR-RL: Seeing Better before Thinking Deeper in Visual Mathematics


100. ViperQ: Order Flow Pattern Recognition via Auction Market Theory for Reinforcement Learning Trading


101. Bypass Observation: A Conceptual Design of a Non-Intrusive Layer-Wise Semantic Extraction Architecture


102. LLM-Enhanced Multi-Agent Reinforcement Learning for Unified Electric Vehicles-Charging Station-Grid Optimization in Public Charging Systems


103. Do Not Restart: Residual Completion for Stateful Agent Handoffs


104. Partition Scores Are Not System Scores: Deployment-Fidelity Gaps in Decomposed Algorithm Selection


105. Surprising Effectiveness of Self-Demonstrations in Enhancing Schema-Ontology Mapping with LLMs


106. Homeostatic Continual Learning


107. Positioning manuscripts in the scientific landscape with agentic AI


108. How Many Thoughts Can a Vector Hold? The Capacity of Reasoning by Superposition


109. Trustworthy Agentic AI: A Comprehensive Cybersecurity and Systems Survey on Threat Landscapes, Defense Architectures, and Open Challenges


110. IBBench-Light: A Paired Evaluation of Task-Conditioned Responses to External Directives


111. MANAS-2: Constrained Reconstruction for EEG Foundation Models


112. JaxAHT: A JAX-Based Library for Ad Hoc Teamwork


113. Degraded but Not Entirely Ineffective: PE-Based Deformable Graph Neural Networks


114. Drift-Constrained Optimization: Only Direction Matters in Fine-Tuning Instruct Models


115. Windowed A-K-MDP


116. Recoverability as a System Primitive for Long-Horizon AI Agents


117. Enhancing Event Candidate Acquisition for Event Linking


118. GeoSkill:Experience-Driven Hierarchical Skill Learning with Collaborative Revision forGeospatialAgents


119. Cost Characterization of Vertically Partitioned Federated Knowledge Graphs


120. FedV-KGQA in Practice: Design Lessons and an Interactive Prototype


121. Safety as a Constraint: Fine-Tuning a LLM Recommender to Explain Itself


122. Solar Intelligence


123. Identity Is More Than Recall: A Benchmark for Persistent Identity in Deployed AI Agents


124. $τ$-Elicitation: Benchmarking multi-turn entity extraction in voice agents


125. FLoKD: Adaptive Knowledge Distillation for Federated Low-Rank LLM over Wireless Networks


126. How User-AI Mistreatment Occurs and Matters in Conversational Systems?


127. Causal multi-modal AI for personalized chemosensitivity prediction


128. Planning or Learning: Reliability and Cost in Multi-Asset Maintenance


129. A Hybrid Agentic AI Framework for Intelligent Supply Chain Analytics


130. Carbon-Aware Routing for Function Calling in Edge-Cloud LLM Systems


131. Toward a Decision-Assurance Layer for AI-Assisted Flight Planning in Air Traffic Management


132. AutoTailor: Automatic, User-Aligned Capability Selection and Adaptation for Web Agents


133. Asclepius: An Adaptive Harness for Long-Horizon Clinical Agents



135. Fraglingo: Molecular Design via Attachment-Aware Autoregressive Fragment Generation


136. Token Efficient Task Execution via Application Behavior Modeling for Web Agents


137. Grounded Adjudication of Variations across Extracted TimeLines (GAVEL): Comparing Clinical Timelines Against Their Case Reports


138. OrchSLM: Probing the Dynamics of Small Language Model Orchestration


139. Governing at Machine Speed: An Adaptive Intelligence Architecture for Real-Time AI Policy Enforcement


140. Root-Cause Attribution Is a Search Problem: Continual Search for Long-Horizon Agent Failures


141. TimeThink: Eliciting Compositional Reasoning in Timeseries Large Language Models


142. LabAgent: Customize Any Research Hubs for Scientific Discoveries Using AI Agents


143. Toward Self-Adaptive Physical AI: Can LLM Agents Manage Long-Horizon Physical Tasks?


144. Vibe Patenting: Evaluating LLM Judges for Professional Patent-Drafting Agents


145. Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement


146. Converge Then Diversify: Decoupling Convergence and Diversity in Multi-Objective Bayesian Optimisation



148. The Router Within: Eliciting Native Skill Routing from a Frozen LLM


149. Vulnerability Localization Benchmark: Measuring Agentic Security Analysis at Repository Scale


150. SlipSense: Multimodal Tactile Learning for Low-Latency and Generalized Slip Detection


151. Anatomical Grounding and Leakage-Aware Multimodal Contrastive Learning for Alzheimer’s Disease Classification from Structural MRI


152. Privacy-enhanced federated learning via asynchronous aggregation and local differential perturbation


153. Learning Multimodal One-step Flow Policy via Value-weighted Optimal Transport


154. LLM-Based Schema-Aware Split Learning for Privacy-Preserving Mental Distress Prediction Across Heterogeneous Surveys


155. K-Bench: a clinically calibrated benchmark for evaluating large language models in high-risk mental health conversations


156. Before You Poll with LLMs: A Deliberative Diagnostic Framework


157. Per-Matrix Optimality Is Not Enough: Three-Level Optimization for Low-Rank LLM Compression


158. CiteGuard-RAG: A Validation-Centered AI System for Evidence-Grounded Question Answering


159. Delegating Authorization to Misaligned Agents: Coalitional Alignment and Safe Control


160. When the World Lies: Backdoor Attacks on Latent World Models for Downstream Control


161. Transfer Learning for Socioeconomic Estimation in Forced-Displacement Settings


162. Event-Native Symbolic-Temporal Spike Encoding Framework for Heterogeneous Cyber Streams


163. Sylvas: Synergistic Learning Value based Device Scheduling in Federated Continual Learning


164. Look Before You Leap: Factual Decoding with Internal Attribution Signals


165. A Language-Guided Multimodal Foundation Model for Zero-Shot and Multi-Task Brain Signal Analysis


166. Bench2Dex: Benchmarking Visuo-Tactile Bimanual Dexterous Manipulation Across Dexterous Hands


167. More Than Just Access: Generative AI as Communication Intermediary for Blind and Low-Vision Users


168. Scalability and Performance Evaluation of Federated Learning Frameworks: A Comparative Analysis


169. Don’t Send What You Don’t Need: Question-Guided Token Pruning as a Privacy Defense for Vision-Language Models


170. Benchmarking Intra-Patient 3D Deformable Multimodal Image Registration


171. Circuit-MLLM: Topological Logic-Guided Latent-Space Visual Reasoning for Circuit Schematic Understanding


172. CIDERS: Cloud-Edge LLM Collaborative Learning via Accelerating Personalized Bilevel Optimization


173. Kaininja: Extending Native 3D Generators to the Part Level


174. Predictive Likelihood Ratios for Language Model Watermark Detection


175. ModaLens: Measuring Image Sensitivity in Report-Conditioned Medical VLMs


176. FedLTLib: A Comprehensive Benchmark for Federated Long-Tail Learning


177. Beyond AI Literacy: A Structured Review and Exploratory Meta-Analysis of Measures for Competent Generative-AI Use


178. Multi-View Molecular Representation Learning with Hierarchical Graphs and Contextualized Fingerprints


179. Through the Eyes of the Beholder: Biometric and Demographic Conditioning for Multimodal Sexism Detection


180. VideoScout: Learning Agentic Active Exploration with Adaptive Reasoning Pacing for Long Video Understanding


181. A Unified Vision-Language Model for PSMA PET/CT Report Generation, Visual Question Answering, and Lesion Segmentation


182. Self-Evolving Memory for Generative Recommendation


183. Big Brains and Changing Environments: Cause or Consequence?


184. PIVOT: Physics-Grounded Verification for AI-Generated Audio-Video Detection


185. Don’t Count the Edits, Judge by the Outcome Alone: Reward-Based Evaluation for Grammatical Error Correction


186. Specifying Reward Functions for RL Without Environment Sampling


187. The Misery of Mechanistic Interpretability: A Formal Perspective


188. Automating Attack Graph Construction for Agentic Pentesting. Towards Neuro-Symbolic Vulnerability Hunting


189. Authorship attribution and aesthetic evaluation of AI poetry: a case study with Haiku


190. How Lossless Is Lossless Speculative Decoding? The Role of Numerical Precision in Orthrus


191. Spook the Machine: Gamified Exploration of Human Imagination of Machine Fear


192. Turkish MMLU Pro: Traceable Option Augmentation and Its Validity Limits in Turkish Multiple-Choice Evaluation


193. On the role of the tokenizer in ECG transformer models


194. A Conservative OCR-Enabled Workflow for R214 Sodium Screening of South African Packaged Foods


195. CodeTS: Verifiable Text-to-Time Series Generation via Executable Code


196. IWC-Bench: Evaluating Web Application Generation from a Software Testing Perspective


197. Divide, Consult, Conquer: Capability Laundering Through Aligned LLMs


198. Robust and Efficient Communication for Multi-Agent Learning


199. End-to-End Cell Detection via Instance-aware Graph Modeling


200. Dynamic Semantic Compression for Efficient Latent-Space Inference in Large Language Models


201. Concept-Grounded Reasoning with Prompt-Driven Localization for Interpretable Structured Report Generation


202. Planning in the Backbone: DiffAdapterVLA for Native Continuous Trajectory Generation with Driving VLMs


203. Clean Scores, Buried Evidence, and Confident Wrong: A Receipt-Based Audit of Frontier Agentic QA


204. The Universe of Universes: Benefit Yield Functions, Implosion Thresholds, and Infrastructure-Aware Optimization in Multi-LLM Systems


205. When Correlations Mislead: Confounder-Aware Multi-View Urban Region Representation Learning


206. Math for AI safety: an invitation for mathematicians


207. ProtoGuide: Prototype-Driven Guidance for Class-Conditional Graph Generation


208. Pre-PEFT Probing: Weight Statistics and Perturbation Robustness for Layer Selection in VLM Vision Encoders


209. Augmenting Large Audio-Language Models with Frame-Level Grounding for Fine-Grained Temporal Perception


210. Failure-Guided Co-Evolution of Prompts and Training Data


211. TEAR: Table Extraction with Attribute Recommendation from Texts via Large Language Models


212. Interpreting hierarchical organisation of speaker embeddings


213. EMR: Self-Evolving Medical Multi-Agent System via Experience Mining and Reuse


214. PACE: Progressive Angular-to-Norm Contrastive Embedding


215. AdaVSkip: Adaptive Visual Token Skipping Across Layers For Efficient MLLMs Inference


216. MoME: Mixture-of-Memory Embeddings for Context-Aware Sparse Lookup


217. Refinement-based Flow Policy Optimization


218. DepthBenchCAD: When Does Deeper Auditing Yield More Reliable Conclusions?



220. Physics Informed Neural Network model for the dynamical study of Abdominal Aortic Aneurysm


221. ChatGPT Images 2.5 in the Wild: A Launch-Period Dataset and Detector Evaluation


222. Generate to Explore, Select to Exploit: Aligning LLM-based Headline Generation with Personalized Recommendation


223. Beyond Numerical Time Series: A Unified Benchmark for Multimodal Forecasting with Heterogeneous Context


224. Translating the Translator: Decomposing the Cost of English-Forced Inter-Agent Communication


225. Branched Optimal Transport Amortization


226. Rethinking Procedural Audio Pre-training: Source Scaling and Objective Adaptation


227. Salesforce Koa: An Enterprise Language Model for Agentic Tool Use


228. Not All Prompts Are Equal: Exploration-Guided Prompt Scaffolding for Multimodal Reinforcement Post-Training


229. Ensemble Complexity in Photovoltaic Forecasting


230. Personalizing Personal Health Interfaces: Co-Design with Generative AI


231. Mirror, Mirror on the Wall: Prompt Echoing in Small Instruct Language Models


232. SpliTEE: Improving LLM Inference on Trusted Hardware with Differentially Private GPU Outsourcing


233. Validating Hybrid-State Cache Recovery for GLM-5.3-Flash with vLLM and LMCache


234. Pick Your Poison: Learning to Select Poison Sets for Stronger LLM Backdoor Attacks


235. PIDS-Bench: Evaluating Prompt-Injection Detectors Under Over-Defense, Obfuscation, and Distribution Shift


236. Steering Generative Robot Policies with Lexicographic Preferences


237. IMPACT-VLA: Interaction-aware Multimodal Propagation Attribution via Counterfactual Trajectories for Vision-Language-Action Policies


238. ActGuard: Pre-execution Action Auditing against Indirect Prompt Injection in LLM Agents


239. LiftGCN: Efficient Energy-Preserving Graph Learning via Joukowski Spectral Lifting for Finite Element Stress Prediction


240. Online Language Adaptive Sampling for Better Distributed Cross-lingual Gains


241. CAL-MOS: Bridging Layers with Adapters for Robust MOS Prediction Across Speech Foundation Models


242. Cross-Block Conditioning in Deep Boltzmann Machines for Statistical Data Fusion


243. Neural-Network Solutions to Real-Space Charge Density and Generalization


244. Forty Shades of Blue: Quality-Diversity Alignment via Mode-Conditioned Reinforcement Learning


245. PeerPen: AI-Assisted Writing for Online Mental Health Peer Support


246. SeqMaestro: From nucleotide sequences to biological hypotheses through interpretable machine learning


247. Interpolation Is Not Invariance: Pair Count Is Not Coverage in Transformation Audits


248. Semantic Fibers and Cross-Gram Interference: A Calculus of Safety Drift in Overcomplete Representations


249. One Example Is Enough to Pass Fairness Benchmarks: Rethinking Fairness Evaluation for Aligned LLMs


250. RAIN: Region-Aware Inversion Network for Semantic Watermark Extraction


251. LLMs as Oracles: Reliance on LLMs for Subjective Personal Questions


252. A Responsive Present, a Shared Past, a Social Other: Teens’ Overreliance on Companion AI Chatbots


253. Efficiency Hallucination: Formalizing and Measuring Behavioral Calibration in LLM-Based Code Optimization


254. Enemray: Toward Capable Language Models for Hassaniya


255. Route, Don’t Fix: Regime-Dependent Decoding Correction and a Trajectory-Gated Router for Reliable Clinical LLM Answer Selection


256. A primer on evaluation methods for large language models in healthcare


257. Mind Which Bird You Favour: Parameterizing Adequacy-Fluency Balance in Meta-Evaluation of Machine Translation


258. The Stochastic Deputy: Structural Tenant Isolation for Tool-Using LLM Agents


259. How broad is that claim? Mapping Generalisation in NLP Research


260. Loop-Back Authority in LLM Agent Teams: A Paired Experiment on Flat and Hierarchical Coordination


261. TriCalRAG: A Three-Strategy, Retrieval-Augmented Benchmark for On-Premise LLM-Based Root Cause Analysis in AIOps


262. From Visual Feedback to Textual Reviews: A Multi-Agent Vision-Language Framework for Image-Grounded Review Assistance


263. Refusal Reads Only a Slice of What the Model Knows: Harm-Keyed Routing and Its Exceptions Across Model Families


264. Calibrating Interpretability Instruments Before Trusting Their Verdicts




267. OCT-FedSIR: Toward Trustworthy Federated Ophthalmic Learning under Annotation Noise


268. Carryover Drafting: Recycling Rejected States for Speculative Decoding


269. WaterKron and FlipFlop Hessian: Information-Theoretically Grounded Quantization with Kronecker-factored Hessians


270. PU classification under Non-SCAR: clustering-assisted logistic model with oversampling enhancement


271. MANE: A Multi-Path Adaptive Network for Edge Onloading of Deep Neural Networks


272. Compositional SVG Generation via VLM-Driven Hierarchical Semantic Parsing


273. Natural Language Knowledge Graph Query Execution: Leveraging Controlled Semantics in the LLM Context Window


274. Skill Composition for Legged Robot Reinforcement Learning



276. Investigating the Impacts of Generative AI on Information Seeking


277. Open-UniMo: Towards Unified Motion-Language Understanding and Generation in the Open World


278. Diagnosing Temporal Misalignment in Multichannel Time-Series Classification with Minimum Description Length


279. SENTINEL: A Multi-Pathway Architecture for Detecting Living-Off-the-Land APT Attacks on Windows Command Lines


280. AlgoRAG: Retrieval-Augmented Generation for Theoretical Computer Science Education – A Comprehensive Evaluation Framework for Algorithm Analysis and Complexity Theory


281. Disentangling Topology and Diversity in Multi-Agent LLMs for Multilingual Low-Resource Emotion Detection


282. Proving olympiad geometry theorems on a superconducting quantum processor


283. Sharing standardized image-derived data in computational pathology using DICOM


284. EdgeHAR: An Edge-Native Compact Sensor Foundation Model for Human Activity Recognition


285. Bridging the Modality Gap in Long-Form Clinical Audio: A Comparative Study of Lightweight and Heavyweight End-to-End SOAP Generation


286. Follow the Geometry, Not the Model: Cold Start Semi-Supervised Learning


287. NeuroActiSep: Detecting Factual Hallucinations from Feed-Forward Neurons in a Single Pass


288. From Visual Attribution to Clinical Reasoning: Explainable Parkinson’s Disease Screening from Hand-Drawn Patterns


289. A latent dimension of Condorcet’s jury theorem for multiple AI advisers


290. Lightweight Generalized DeepFake Face Detection with WAVIE: Wavelet Augmented Vision Intermediate Embeddings


291. Has Scientific Talent Shifted from Depth to Breadth?Evidence across Papers, Knowledge Inputs, Careers, and Teams


292. A Generative AI Integrated Multimodal Framework for Low-Latency Multi-Camera Person Re-Identification


293. Surrogate-Assisted Genetic Programming with Phenotypic Characterisation in Dynamic Multi-Mode Project Scheduling


294. A Hybrid Dependency-Aware Framework for Task Decomposition and Dynamic Agent Generation in Oracle-to-PostgreSQL Migration


295. LLaTSA: Large Language Model-Aligned General-Purpose Transient Stability Analysis


296. AURA: Unified Multimodal Framework for Conversational Music Editing


297. Biquaternionic Space with Complex-valued Attention for Temporal Knowledge Graph Completion


298. SpermYOLO: A Coordinated YOLO-Based Detector for Accurate and Efficient Sperm and Impurity Detection in Microscopic Images


299. ATTRICITE: Training an Open 4B Model for Citation Recovery toward Faithful Attribution


300. The Attribution-Compression Frontier in Retrieval-Augmented Generation


301. OpWeave: Flexible Operator Disaggregation for Heterogeneous LLM Serving


302. Assessing the Applicability of Existing Design Recommendations to AI Companion Design: A Multi-Method Study


303. Graph-Transformer Fraud Detection with Self-Supervised Pretraining and Conformal Risk Control


304. Modeling, Scaling, and Decoding: Optimizing Controllable Speech Generation with Nonverbal Vocalizations


305. Task-Specified Active Metrological Inspection with Measurement-Steered VLA Manipulation and Deterministic Evidence Gating


306. ECAS: An Edge-Controlled Agentic System for Validation-Gated Scientific Application Execution


307. Data-free On-policy Distillation


308. Entropy-Punctured Bloom Filters for Memory-Efficient Machine Learning


309. Bi-Level Routing and Sparse Spatial Attention based Multi-View BEV 3D Object Detection for Autonomous Driving


310. Towards Evolving Context Parameterization for Large Language Models


311. Signatures of Steerability in Activation Space of Language Models


312. A New Transformer-Based Approach for Audio-Based Kinship Verification and a New Uncontrolled Mandarin Kinship Speech Dataset


313. LIMBO: Lifelong Inference-Time Memory and Budget Optimization for LLM Agents


314. To do($x$) or not to do($x$): Medical Image Counterfactuals for Dataset Augmentation


315. Talking to Me or Someone Else? Rethinking Talk-to-Me Detection in Egocentric Videos


316. Real-Time Synthesis of Robust Controlled Invariant Sets for Monotone Systems


317. A Voxel-Spacing-Aware Extension of PyRadiomics for Anisotropic Texture Analysis


318. RA-CoA: Training-free Fashion Image Captioning via Retrieval-Augmented Chain-of-Attributes


319. LPA-CWM: A Learned Physical Adjudicator for Motion Reasoning with Counterfactual World Models


320. GraMRAG: Orchestrating Multi-Agent Multi-Step Reasoning via Graph Memory with Reinforcement Learning


321. AGENTQ: Quantization-Conditioned Backdoor Attacks on LLM Agents


322. Confuse the Model, Control the Flow: Understanding and Mitigating Privacy Leakage from LLM Agents with Information Flow Control


323. Rethinking the Implications of Human Feedback for Preference Learning in Human-Robot Collaboration


324. Mizan: A National Benchmark for Evaluating Large Language Models on Iraqi Arabic and the Iraqi Civic Context


325. SGWIB:Sliced Gromov-Wasserstein Information Bottleneck for Video Highlight Detection


326. Thought without systematicity? Evaluating reasoning models on rule induction tasks


327. Hardware-Aware Learned Representation Compression for Distributed In-Sensor Vision


328. CRITICS - Critical Science Without Borders: Language Models to Promote Critical Thinking in Science Education


329. Finite-Time Node Separation in Recurrent Graph Neural Networks with Persistent Gaussian Perturbations


330. DiTAR+: Dual Optimization for Robust Autoregressive Diffusion Speech Synthesis


331. Phorecaster365: A Human-Supervised Reference Architecture for Hybrid Pharmaceutical Sales Forecasting and Planning Decision Support


332. When Malicious Instructions Persist: Persistent Memory Poisoning Attack on Harness-Based Agents


333. Trustworthy, Explainable, and Sustainable Decentralized Intelligence for 6G Networks


334. Bangla Sentence Function Classification: Corpus Development, Model Benchmarking, and Interpretability


335. ReWeight: Leveraging Human Data for VLA Post-Training via Demonstration Retrieval and Sample Weighting


336. LePlanner: An Iterative Amortized Controller For World Models


337. CRAF: Cross-View Residual-Aware Fusion for Deepfake Speech Detection


338. Exploring Automated Vulnerability Identification in JavaScript Code Using Large Language Models


339. Understanding the Limits of Agentic ICD Coding


340. GEAR: From Dynamic Encoding to Dynamic Activation in Social Trajectory Prediction


341. On the Equivalence of Stochastic Control and Path Space Formulations for Schrödinger Bridges over Compact Connected Lie Groups


342. HarnessBandit: Joint Learnability-Transferability Scheduling for Multi-Harness Agentic Reinforcement Learning


343. PolicyMem: Geometric Policy Memory for LLM Governance


344. TyPatch: Transforming Patches into Typestate Rules for Kernel Bug Detection


345. Oops, Not Now: PEARL, a RAG-Based Support Agent for Gameplay and What Players Want from AI Help


346. Gap Entropy and Almost Instance-Wise Optimal Best-Arm Identification


347. Leakage-Safe and Scheduler-Aware Machine Learning for Grid Job Runtime Prediction


348. Not all Negation Cues are Equal: Affixal Negations Yield Better Negation Understanding


349. LayerRoute: Adaptive Layer-Skipping with LoRA-Preserved Quality for Efficient LLM Inference


350. FaithfulBench: Does AI Counsel Uphold or Undermine the User’s Professed Faith?


351. An Efficient and Modular Framework for Targeted Harm Mitigation in LLMS


352. Predictive audio representations for early detection and tracking of hidden dynamic objects


353. mKernel: Fast Multi-GPU, Multi-Node Fused Kernels


354. Same Patient, Different Order: Action-Level Reliability of Clinical LLM Agents Under Repeated Runs


355. Domain-Specific Jargon in Large Language Models: A Comparative Analysis between General-Purpose and Specialist Models


356. Attention Is All You Need (to Avoid Spurious Oscillations)


357. Generative Interpretability via Scalable Neuro-Symbolic Models


358. Rolling Day-Wise Mortality Prediction in Critically Ill Patients With AKI on CRRT Utilizing Machine Pressure Waveforms


359. Positive Topology and Feasible Refinement: Forcing Matrices, Positivity, and Information


360. Adaptive Phase-Switching for Communication-Efficient Federated LoRA Fine-Tuning


361. A Three-Axis Stress Test of LLM vs Classical ML for Network Intrusion Detection under Distribution Shift and Adversarial Evasion


362. One Spectrum, Two Resources: Data-Memory Scaling in Autoregressive Prediction


363. Canaries in the Bank: Auditing User-Level Privacy in Private Evolution


364. Building a Production Greek-English Speech Recognizer


365. Generative AI and Extended Reality in Collaborative Architectural Design Education: An Exploratory Studio Study


366. Edge-addition monotonicity of positive p-energy fails for every p >= 1


367. Symmetry- and Property-Aware Crystal Generation with Reinforcement Learning for Inverse Materials Design


368. Hindsight Bias in Clinical Temporal Reasoning: How Future Data Exposure Affects Large Language Model Judgment


369. Learning to Solve Hard Problems in RL for LLMs by Never Giving Up


370. Certifiably Interpretable Training of ReLU-MLPs for Boolean Tasks with Guaranteed Truth-Table Generalization


371. Chance-Constrained Belief-Space Maneuver Planning for Autonomous Collision Avoidance Under Uncertainty


372. ReCAST: Reward Credit Assignment across Timesteps for Online Diffusion Reinforcement


373. Task-Aware Federated Fine-Tuning for MoE-based Large Language Models


374. SkillAtlas: An Attack Trace Library for Agent Skills


375. ProtoCAM: Interpretable Few-Shot Mask-Guided Prototypical Learning for Breast Lesion Classification in Ultrasound Imaging


376. Bridging Thought and Action: Taming Long-Horizon Instability in Open-Source LLM Agents with a MetaTool-Enhanced ROS Framework


377. The Agentic Company OS: Substrate Inversion for Sustained Enterprise Agent Deployment


378. Pedestrian Crossing Intent Classification From Event-Based Vision Using Convolutional Spiking Neural Networks With Temporal Augmentation


379. Decoupling Error Attribution in Cloud-Native Graph-RAG: A Data Integrity Diagnostic Framework


380. Feasibility and Memory Mechanisms of Chern-Simons Context Reservoir Computation


381. Task-Based CT Protocol Optimization Using Reinforcement Learning and Virtual Imaging Trials


382. IMM-based Multiple Object Tracking using a State Prediction Neural Network


383. Adaptive Conformal Redistribution for Inter-class Transitional Uncertainty in Medical Image Classification


384. Variational Template Matching with Statistical Fusion for Anomaly Detection in Patterned Structures


385. LLaDA-UI: Bringing Block-wise Diffusion to Vision-Language GUI Agents


386. Multimodal-Multiresolution Foundation Model for Lunar Remote Sensing


387. Conflict-Predictive Variable Horizons in Multi-Drone Distributed Model Predictive Control


388. Forward-Facing Near-Infrared Adds Little to Colour for Farm-Machinery Traversability: A Site-Disjoint Evaluation of Sensor-Dependent Spatial Leakage


389. BEACON: Behavior and Appearance Control for Subject-Specific Video Generation


390. From Process Loss to Assembly Bonus: Human-Grounded Diagnosis of Multi-Agent LLM Collaboration


391. TryOnReward: Learning Foveated Consistency for Reinforcement Fine-Tuning of Virtual Try-On


392. Sampling headroom is not selection gain: a compute-value audit of test-time scaling for video world models


393. (How) Do MLLMs Report Bistable Images Like Humans?


394. An Evolutionary Computation Framework for Multi-Agent Q-Learning with Mean-Field Environmental Feedback


395. Evaluation of MLLM-Agnostic Plug-and-Play Keyframe Selection Methods for Long Video Understanding


396. ShieldVLA: Feasibility-Aware Safety Alignment for Vision-Language-Action Models


397. Chemical and geometric representation fidelity improves drug–target affinity prediction


398. Planning as Dynamics Relaxation: Hippocampal Recurrent Network Realizes Optimal Goal-Directed Navigation


399. Diagnosing Faults in Reinforcement Learning Simulators and World Models with Canonical Polynomial Invariants


400. LLMs or Naive Bayes? Old Gems or New Ways


401. Is Bash All You Need? An Empirical Study of Tool Interfaces for Enterprise Digital Worker Agents


402. From Semantic to Token Communication: The Next Paradigm for Large-Model-Driven 6G Intelligent Connectivity


403. Multilingual Agent System for Inclusive Wildfire Evacuation Guidance


404. Natural-Language to SysMLv2 Translation via Conformance-Driven Iterative Refinement


405. Physically Aware Radiomics Without Interpolation: Disentangling Voxel Geometry and Signal Modification in CT and MRI


406. Generalization Can Emerge in Tabular Foundation Models From a Single Table


407. Beyond Quacking: Deep Integration of Language Models and RAG into DuckDB


408. Towards Optimizing SQL Generation via LLM Routing