전체 AI 논문 - 2026-06-24

1. OpenThoughts-Agent: Data Recipes for Agentic Models


2. World Models in Pieces: Structural Certification for General Agents


3. Matching Tasks to Objectives: Fine-Tuning and Prompt-Tuning Strategies for Encoder-Decoder Pre-trained Language Models


4. Grading the Grader: Lessons from Evaluating an Agentic Data Analysis System


5. Accuracy and Satisfaction in Multi-Turn LLM Dialogues for NFR Assessment


6. Difference-Making without Making a Difference


7. Solving Inverse Problems of Chaotic Systems with Bidirectional Conditional Flow Matching


8. Assessing Distribution Shift in Human Activity Recognition for Domain Generalization


9. BluTrain: A C++/CUDA Framework for AI Systems


10. Can Scale Save Us From Plasticity Loss in Large Language Models?


11. Scaling Laws for Task-Specific LLM Distillation


12. Decentralised AI Training and Inference with BlockTrain


13. Cost-Optimal Decision Diagrams for Stochastic Boolean Function Evaluation


14. LaGO: Latent Action Guidance for Online Reinforcement Learning


15. CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning


16. SAFARI: Scaling Long Horizon Agentic Fault Attribution via Active Investigation


17. Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback


18. When CQs Go Wrong: Challenges in CQ Verification with OE-Assist


19. Abstractions of Queries in Ontology-Based Data Access


20. AI Tokenomics: The Economics of Tokens, Computation, and Pricing in Foundation Models


21. ScaleToT: Generalizing Structured LLM Reasoning for Billion-Scale Low-Activity User Modeling


22. Uncertainty-Aware Longitudinal Forecasting of Alzheimer’s Disease Progression Using Deep Learning


23. ASALT: Adaptive State Alignment for Lateral Transfer in Multi-agent Reinforcement Learning


24. AdversaBench: Automated LLM Red-Teaming with Multi-Judge Confirmation and Cross-Model Transferability



26. Quant Convergence: Bridging Classical Value Investing and Modern Factor Models for Systematic Equity Selection


27. GUI vs. CLI: Execution Bottlenecks in Screen-Only and Skill-Mediated Computer-Use Agents


28. Governed Shared Memory for Multi-Agent LLM Systems


29. Reinforcement Learning for Computer-Use Agents with Autonomous Evaluation


30. A specialized reasoning large language model for accelerating rare disease diagnosis: a randomized AI physician assistance trial


31. On the Smallness of the Large Language Models Scaling Exponents


32. The Latent Bridge: A Continuous Slow-Fast Channel for Real-Time Game Agents


33. CompressKV: Semantic-Retrieval-Guided KV-Cache Compression for Resource-Efficient Long-Context LLM Inference


34. Bayesian control for coding agents


35. ReM-MoA: Reasoning Memory Sustains Mixture-of-Agents Scaling


36. Can Aggregate Invariants Accelerate Continuous Subgraph Matching? Limits, Laws, and a Dynamic Spectral Index


37. Agentic AI for Bilevel Long-Term Optimization of Policy-Driven Physical Layer Systems


38. Cycle-Consistent Neural Explanation of Formal Verification Certificates


39. ATRIA: Adaptive Traceable ECG Reporting with Iterative Agents


40. Age of LLM: A Strategic 1v1 Benchmark for Reasoning, Diplomacy and Reliability of Large Language Models under Fog of War


41. PHANTOM: A Large-Scale Dataset of Multimodal Adversarial Attacks for Vision-Language Models


42. When Helpfulness Overrides Causal Caution: Context-Dependent Suppression and Recovery in LLMs


43. Accelerating Disaggregated RL for Visual Generative LLMs with Diffusion-Based Parallelism and Trainer-Assisted Generation


44. MVG-KAN: Multi-View Geo-Wind Guided KAN for PM$_{2.5}$ Forecasting


45. Prob-BBDM: a Probabilistic Brownian Bridge Diffusion Model for MRI sequence image-to-image translation


46. LemonHarness Technical Report


47. Tractable Reasoning and Conjunctive Query Answering for Defeasible DL-Lite under Rational Closure


48. Probing the Misaligned Thinking Process of Language Models


49. Towards Federated Long-Tailed Graph Learning: An Energy-Guided Dual Decoupling Approach


50. SP-Mind: An Autonomous Reasoning Agent for Spatial Proteomics Analysis


51. FlowR2A: Learning Reward-to-Action Distribution for Multimodal Driving Planning


52. Exploring the relationship between human-centric AI and firm idiosyncratic risks


53. Navigating User Behavior toward Personalized Multimodal Generation


54. Data Scale, Not Latency, Shapes Cross-Lingual Encoder Transfer in Streaming ASR


55. An Introduction to Causal Reinforcement Learning


56. The Geometry Behind Diffusion and Flow Matching: Gradient Flows and Geodesics in Wasserstein Space


57. T2D-Bench: Evidence-Gated Evaluation of LLM Outputs for Type 2 Diabetes Using a Multi-Layer Clinical-Lifestyle Knowledge Graph


58. OmniPath: A Multi-Modal Agentic Framework for Auditing Wheelchair Accessibility


59. VeryTrace: Verifying Reasoning Traces through Compilable Formalism and Structured Verification


60. ReMMD: Realistic Multilingual Multi-Image Agentic Verification for Multimodal Misinformation Detection


61. Exploring Academic Influence of Algorithms by Co-occurrence Network Based on Full-text of Academic Papers


62. Beyond Trajectory Imitation: Strategy-Guided Policy Optimization for LLM Reasoning


63. Ensemble Feature Selection and Harris Hawks Optimization for Explainable Mental Health Risk Prediction in Female Sex Workers


64. Breaking the Filter Bubble: A Semantic Pareto-DQN Framework for Multi-Objective Recommendation


65. Can Language Model Agents be Helpful Circuit Explainers in Mechanistic Interpretability?


66. Reinforcement Learning Towards Broadly and Persistently Beneficial Models


67. Safe and Generalizable Hierarchical Multi-Agent RL via Constraint Manifold Control


68. Critique of Agent Model


69. Neuro-Symbolic Drive: Rule-Grounded Faithful Reasoning for Driving VLAs


70. RIFT-Bench: Dynamic Red-teaming For Agentic AI Systems


71. InSight: Self-Guided Skill Acquisition via Steerable VLAs


72. FLUX3D: High-Fidelity 3D Gaussian Generation with Diffusion-Aligned Sparse Representation


73. It’s Complicated: On the Design and Evaluation of AI-Powered AAC Interfaces


74. IV-CoT: Implicit Visual Chain-of-Thought for Structure-Aware Text-to-Image Generation


75. Large-Language-Model Discovery of Quantum LDPC Codes through Structured Concept Evolution


76. OrbitForge: Text-to-3D Scene Generation via Reconstruction-Anchored Video Synthesis


77. EG-VQA: Benchmarking Verifiable Video Question Answering with Grounded Temporal Evidence


78. Grad Detect: Gradient-Based Hallucination Detection in LLMs


79. Paying to Know: Micro-Transaction Markets for Verified Product Information in Agentic E-Commerce


80. DeepBD: A Grounded Agentic Workflow for Variant Prioritization and Diagnosis of Genetic Birth Defects


81. Context-Aware Prediction of Student Quiz Performance with Multimodal Textbook Features


82. UniDrive: A Unified Vision-Language and Grounding Framework for Interpretable Risk Understanding in Autonomous Driving


83. Beyond U-Net: A Latent-Representation-Aligned Skip-Free Backbone for Flow-Matching Speech Enhancement


84. Task Decomposition for Efficient Annotation


85. Evaluating the Interpretability of Sparse Autoencoders with Concept Annotations


86. TACTFUL: Tactile-Driven Exploration For Object Localization and Identification in Confined Environments


87. FlowPipe: LLM-Enhanced Conditional Generative Flow Networks for Data Preparation Pipeline Construction


88. AI-PAVE-Br: Leveraging Large Language Models for Enhanced Product Attribute Value Extraction through a Golden Set Approach


89. Visualizing “We the People”: Bridging the Perception Gap through Pluralistic Data Storytelling


90. Privacy-Preserving RAG via Multi-Agent Semantic Rewriting: Achieving Confidentiality Without Compromising Contextual Fidelity


91. Infinitesimal Causality


92. Toward Self-Evolution-Ready Workflow Harnesses: A Reversible Migration Path and Convertibility Taxonomy for Expert LLM Pipelines


93. Poster: Exploring the Limits of Audio-Based Detection of Turkish Phone Call Scams


94. A Fair Evaluation of Graph Foundation Models for Node Property Prediction


95. CrossPool: Efficient Multi-LLM Serving for Cold MoE Models through KV-Cache and Weight Disaggregation


96. Red-Teaming the Agentic Red-Team


97. RetiSEM: Generalising Causal Models for Fragmented Biomedical Data


98. Adaptive Machine Learning Framework for UAV Trajectory Optimization in O-RAN


99. video-SALMONN-R$^3$: Learning to ReWatch, ReAsk, and ReAnswer for Efficient Video Understanding


100. G$^3$VLA: Geometric inductive bias for Vision-Language-Action Models


101. The African Language Tax: Quantifying the Cost, Latency, and Context Penalty of Tokenizing African Languages in Frontier LLMs


102. NoContactNoWorries: Estimating Contact through Vision and Proprioception for In-Hand Dexterous Manipulation


103. MedPCFM: Improving Medical Point Cloud Completion by Integrating Point Transformers and Flow Matching


104. Transformation Behavior of Images in Latent Space


105. Detecting AI Coding Agents in Open Source: A Validated Multi-Method Census of 180 Million Repositories


106. Entity Resolution via Batched Oracle Queries


107. Average Rankings Mask Per-Subject Optimality: A Friedman-Nemenyi Benchmark of EEG Motor-Imagery BCI Decoders


108. Female-RHINO: A Real-Time Scanner-Integrated Framework for Automated Quantitative Uterine MRI Analysis and Structured Reporting


109. On the Stability of Prompt Ranking in Large Language Model Evaluation


110. Structural Kolmogorov-Arnold Convolutions: Learnable Function on the Values or the Filter Shape as Parameter-Efficient Alternative to Per-Edge Convolutional KANs


111. What Does ODRL Mean? A Cross-Level Ontological Grounding of Permissions, Prohibitions, and Duties in UFO-L


112. ZONOS2 Technical Report


113. Real-Time Interactive Music Generation via Data-Free Streaming Consistency Distillation


114. CALIBER: Calibrating Confidence Before and After Reasoning in Language Models


115. Pigeonholing: Bad prompts hurt models to collapse and make mistakes


116. Neural Network-Based Parametric Model Reduction for Predicting Turbulent Flow for Different Vehicle Geometries


117. SURGELLM: Rethinking Multi-Task Evaluation through Task-Aware Feature Gating with Class-Balanced Normalization


118. Social Structure Matters in 3D Human-Human Interaction Generation


119. AutoSpec: Safety Rule Evolution for LLM Agents via Inductive Logic Programming


120. Inclusive Interactive Collisions for Multi-View Consistent Compositional 3D Generation


121. MMed-Bench-IR: A Heterogeneous Benchmark for Multilingual Medical Information Retrieval


122. Co-occurring associated retained concepts in Diffusion Unlearning


123. Deep Learning Approaches for 3D Medical Scene Completion: From Geometric Modeling to Generative Paradigms


124. Zero-Shot Test-Time Canonicalization using Out-of-Distribution Scoring


125. Agon: An Autonomous Large-Scale Omnidisciplinary Research System Built on Prompt Economy


126. Lightweight Transformer Models for On-Device Fault Detection: A Benchmark Study on Resource-Constrained Deployment


127. A Pāninian Foundation for Indic Language Processing


128. Breaking Shortcut Learning for Cross-Trial EEG-Guided Target Speech Extraction via Two-Stage Training


129. Metis: Bridging Text and Code Memory for Self-Evolving Agents


130. DTT-BSR+: A Generative-Regression Cascade for Music Source Restoration


131. A Benchmark for Hallucination Detection in VLMs for Gastrointestinal Endoscopy


132. DramaDirector: Geometry-Guided Short Drama Generation


133. The impact of generative artificial intelligence on academic development of Chinese students in humanities and social sciences


134. Beyond Bayer: Task-Optimal Sensor Co-Design for Robust Autonomous-Driving Segmentation


135. Predicting Poets’ Origins from Verse: A Computational Analysis of Regional Linguistic Fingerprints in the Complete Tang Poems


136. DynaWM: Dynamics-Aware Distillation with World Model and Momentum Targets for Smooth Locomotion over Continuous Stairs


137. Blockwise Policy-Drift Gating for On-Policy Distillation


138. CAVEWOMAN: How Large Language Models Behave Under Linguistic Input and Output Compression


139. PixJail: Self-Evolving Paper-to-Pipeline Reproduction for Text-to-Image Jailbreak Evaluation


140. End-to-End Radar and Communication Modulation Recognition with Neuromorphic Computing


141. Token Complexity of Certifying Stochastic-Oracle Reliability


142. Selective Capability Unlearning in End-to-End Spoken Language Understanding


143. RAVEN: A Regime-Aware Variable-context Expert Network for Financial Time Series Forecasting


144. Rapid FinFET Modelling Using an Autoencoder


145. Towards Version-aware Operations and Transaction Memories for Multi-layer MeMo


146. Fast and Slow Variational Continual Learning


147. Towards Spec Learning: Inference-Time Alignment from Preference Pairs


148. EMAgnet: Parameter-Space EMA Regularization for Policy Gradient Self-Play in Large Games


149. Learning to Trigger: Reinforcement Learning at the Large Hadron Collider


150. RASC+: Retrieval-Constrained LLM Adjudication for Clinical Value Set Authoring


151. Faithful by Construction: Claim-Anchored Attribution for Multi-Document Summarization


152. Maestro Order: A Model-Agnostic Orchestration Harness


153. Offline Reinforcement Learning for Warehouse SLAM Throughput Control


154. When Retrieval Metrics Mislead: Measuring Policy Signal in Long-Horizon Tool-Use Agents


155. Catastrophic Compositional Generation: Why Vanilla Diffusion Models Fail to Extrapolate


156. ARIA: Adaptive Region-Based Importance Allocation for Conditional Diffusion Distillation


157. The Professor: Multi-Teacher Unsupervised Prompt Distillation for Vision-Language Models


158. E-MRL: Cross-view Aligned Evidence-driven Multimodal Reinforcement Learning for Reliable 3D Tumor Analysis


159. Mind the Heads: Topological Representation Alignment for Multimodal LLMs


160. One Year Later…The Harms Persist, But So Do We!


161. Promise and challenges of heart chamber segmentation from non-contrast CT scans using contrastive unpaired image translation: a feasibility study


162. JupOtter: Cell-Level Bug Detection in Jupyter Notebooks


163. MGI: Member vs Generated Inference


164. Are Safety Guarantees in Neural Networks Safe? How to Compute Trustworthy Robustness Certifications


165. The Measurable Majority


166. Decentralized Coordination of Autonomous Traffic Through Advanced Air Mobility Corridors


167. Deciphering Fingerprints of 3D Molecular Surfaces for Accurate Epitope Prediction


168. From Spatial to Spectral: An Efficient, Frequency-Guided Feature Representation Learner for Small Object Detection


169. Ten Digits on a Train: AI-Assisted Verification of Two Eigenvalue Problems


170. From Task-Guided Conversational Graphs to Goal-Oriented Dialogue Runtimes


171. Integrated Sensing and Communications for Real-time Avatar Control in XR over 5G


172. Cryptographic certificates of validity for trustworthy AI


173. Emergent Relational Order in LLM Agent Societies: From Collective Affect to Authority Stratification


174. Listening makes Vision Clear for VLMs


175. Neuromorphic Speech Enhancement with Dual-Branch Spiking Neural Networks


176. Engineering Reliable Autonomous Systems: Challenges and Solutions


177. VeriPilot: An LLM-Powered Verilog Debugging Framework


178. Exploring Dualistic Meta-Learning to Enhance Domain Generalization in Open Set Scenarios


179. Synergizing Physically Constrained MCMC and Chemical-Informed Gaussian Processes for Reaction Network Discovery


180. JEDEL: Zero-Shot DNA-Encoded Library Design for Early-Stage Drug Discovery


181. Sol Video Inference Engine: Agent-Native Full-Stack Acceleration Framework for Efficient Video Generation


182. Low-power analogue neural networks with trainable nonlinear connections for continuous control


183. A Survey on Federated Causal Discovery and Inference


184. Weight-Space Geometry of Offline Reasoning Training


185. A Unified Framework for Runtime Verification and Model-Based Diagnosis in LOLA



187. Random coloured digraphs defined by a Markov logic network


188. Audio-visual Contrastive Alignment for Diffusion-based Visual-conditioned Speech Enhancement


189. Coordinate-Queryable Neural Field Reconstruction for EEG Spatial Super-Resolution with Unseen-Electrode Generation


190. Event-Aligned Analysis of Multi-Rater Pain Assessments Using Continuous Wearable Physiology


191. Heterogeneous 2D/1D Signal Representation Fusion for Underwater Acoustic Modulation Recognition Under Distribution Shift


192. Evaluating LLM Usage for Efficient and Explainable Numerical and Classified Implicit Sentiment Analysis of Product Desirability


193. Self-Recognition Finetuning can Prevent and Reverse Emergent Misalignment


194. FP8 is All You Need (Part 2): Efficient Ozaki-Bailey Style FFT Through Tensor-core Garner Reformulation and Kulisch Escape Route


195. SemChunk-C: Semantic Segmentation for C Code


196. Quantifying Prior Dominance in RAG Systems


197. Beyond the Autoregressive Horizon: A Comprehensive Survey of Diffusion Models, World Modelling, and State Space Models for Code


198. Reentrant value fields as delayed coupled reaction-diffusion systems on finite graphs