전체 AI 논문 - 2026-06-19

1. Toward Calibrated Mixture-of-Experts Under Distribution Shift


2. How Do Instructions Shape Speech? Cross-Attention Attribution for Style-Captioned Text-to-Speech


3. LedgerAgent: Structured State for Policy-Adherent Tool-Calling Agents


4. DeepSWIP: Quotient-WMC Counterfactuals for Neural Probabilistic Logic Programs


5. FlowEdit: Associative Memory for Lifelong Pronunciation Adaptation in Flow-Matching TTS


6. Multi-LCB: Extending LiveCodeBench to Multiple Programming Languages


7. What Do Safety-Aligned LLMs Learn From Mixed Compliance Demonstrations?


8. Context-Aware Hierarchical Bayesian Modeling of IVF Laboratory Environmental Conditions


9. Interpretable Sperm Morphology Classification via Attention-Guided Deep Learning


10. Rethinking Shrinkage Bias in LLM FP4 Pretraining: Geometric Origin, Systemic Impact, and UFP4 Recipe


11. Automating SKILL.md Generation for Computer-Using Agents via Interaction Trajectory Mining


12. SoftSkill: Behavioral Compression for Contextual Adaptation


13. Leveraging systems’ non-linearity to tackle the scarcity of data in the design of Intelligent Fault Diagnosis Systems


14. Lagrange: An Open-Vocabulary, Energy-Based Sparse Framework for Generalized End-to-End Driving


15. Confidence-Aware Automated Assessment of Student-Drawn Scientific Models


16. Navigating Unreliable Parametric and Contextual Knowledge: Explicit Knowledge Conflict Resolution for LLM Inference


17. A Multi-Agent system for Multi-Objective constrained optimization


18. Thermodynamic Measure of Intelligence


19. QMFOL: Benchmarking Large Language Model Reasoning via Quantifiable Monadic First-Order Logic Test Case Generation


20. Augmenting Game AI with Deep Reinforcement Learning


21. Beyond Accuracy: Measuring Logical Compliance of Predictive Models


22. Apparent Psychological Profiles of Large Language Models are Largely a Measurement Artifact


23. Implicit Semantic-Aware Communication Based on Hypergraph Reasoning


24. Modularity-Free Conflict-Averse Training for Generalized PINNs


25. BIM-Edit: Benchmarking Large Language Models for IFC-Based Building Information Modeling


26. RACL: Reasoning-Agent Control Layers for Continuous Metaheuristic Learning


27. Learning to Prompt: Improving Student Engagement with Adaptive LLM-based High-School Tutoring


28. ScaffoldAgent: Utility-Guided Dynamic Outline Optimization for Open-Ended Deep Research


29. Multi-Head Attention-Based Feature Extractor Integration with Soft Actor-Critic for Porosity Prediction and Process Parameter Optimization in Additive Manufacturing


30. Residual-Space Evolutionary Optimization via Flow-based Generative Models


31. Process-Verified Reinforcement Learning for Theorem Proving via Lean


32. Autonomous Event-Driven Multi-Agent Orchestration for Enterprise AI at Scale


33. Reward as An Agent for Embodied World Models


34. ENPIRE: Agentic Robot Policy Self-Improvement in the Real World


35. Advancing DialNav through Automatic Embodied Dialog Augmentation


36. PhysDrift: Bridging the Embodiment Gap in Humanoid Co-Speech Motion Generation


37. The Tao of Agency: Autotelic AI, Embedded Agency and Dissolution of the Self


38. eCNNTO: A Highly Generalizable ConvNet for Accelerating Topology Optimization


39. Multi-Agent Transactive Memory


40. MetaResearcher: Scaling Deep Research via Self-Reflective Reinforcement Learning in Adversarial Virtual Environments


41. A Systematic Evaluation of Black-Box Uncertainty Estimation Methods for Large Language Models


42. TelcoAgent: A Scalable 5G Multi-KPM Forecasting With 3GPP-Grounded Explainability



44. Think Again or Think Longer? Selective Verification for Budget-Aware Reasoning


45. CombEval: A Framework for Evaluating Combinatorial Counting in Large Language Models


46. ORAgentBench: Can LLM Agents Solve Challenging Operations Research Tasks End to End?


47. AgentFinVQA: A Deployable Multi-Agent Pipeline for Auditable Financial Chart QA


48. Beyond Entropy: Learning from Token-Level Distributional Deviations for LLM Reasoning


49. Optimal Scheduling in a Question-Answering Forum of Knowledge Workers


50. Grounded Inference: Principles for Deterministically Encapsulated Generative Models


51. Benchmarking Agentic Review Systems


52. A Comparative Study of Pretrained Transformer Models for Quranic ASR: Speech Representations, Label Formats, and Dataset Composition


53. Interpreting Neural Combinatorial Optimization via Evolving Programmatic Bottlenecks


54. GLARE: A Natural Language Interface for Querying Global Explanations


55. Beyond Static Leaderboards: Predictive Validity for the Evaluation of LLM Agents


56. Exit-and-Join Dynamics for Decentralized Coalition Formation


57. Denoising Implicit Feedback for Cold-start Recommendation


58. BrainG3N: A Dual-Purpose Tokenizer for Controllable 3D Brain MRI Generation


59. AI4SE and SE4AI Exploration: A Decade Looking Back and Forward


60. Toten: Knowledge-Based Ontological Tokenization Of Physical Quantities And Technical Notation In Brazilian Portuguese


61. Which Pairs to Compare for LLM Post-Training?


62. Configurable Clinical Information Extraction with Agentic RAG: What Works, What Breaks, and Why


63. Analyzing the Narration Gap in LLM-Solver Loops


64. Uncertainty Decomposition for Clarification Seeking in LLM Agents


65. ITNet: A Learnable Integral Transform That Subsumes Convolution, Attention, and Recurrence


66. Emergent Alignment


67. REVEAL++: Differentiable Phenotypic Grouping for Vision-Language Retinal Modeling of Alzheimer’s Disease Risk


68. LLM Doesn’t Know What It Doesn’t Know: Detecting Epistemic Blind Spots via Cross-Model Attribution Divergence on Clinical Tabular Data


69. DeXposure-Claw: An Agentic System for DeFi Risk Supervision


70. Hidden Anchors in Multi-Agent LLM Deliberation


71. Diffusion Language Models: An Experimental Analysis


72. Measuring Curriculum Alignment across Topical Coverage, Competency, and Cognitive Depth: A Longitudinal Framework Applied to CS2013 and CS2023


73. Deontic Policies for Runtime Governance of Agentic AI Systems


74. How Transparent is DiffusionGemma?


75. Structuring and Tokenizing Distributed User Interest Context for Generative Recommendation


76. SARLO-80: Worldwide Slant SAR Language Optic Dataset 80cm


77. Sovereign Execution Brokers: Enforcing Certificate-Bound Authority in Agentic Control Planes


78. Efficient and Sound Probabilistic Verification for AI Agents


79. FreeStyle: Free Control of Style-Content Dual-Reference Generation from Community LoRA Mining


80. Calibration Without Comprehension: Diagnosing the Limits of Fine-Tuning LLMs for Vulnerability Detection in Systems Software


81. Contagion Networks: Evaluator Bias Propagation in Multi-Agent LLM Systems


82. Optimal Order of Multi-Agent and General Many-Body Systems


83. UltraQuant: 4-bit KV Caching for Context-Heavy Agents


84. Analyzing Defensive Misdirection Against Model-Guided Automated Attacks on Agentic AI Systems


85. Repurposing a Speech Classifier for Guided Diffusion-Based Speech Generation


86. Multi-View Decompilation for LLM-Based Malware Classification


87. LLM agent safety, multi-turn red-teaming, jailbreak benchmarks, adversarial robustness, safety-critical systems


88. DataMagic: Transforming Tabular Data into Data Insight Video


89. CRAX: Fast Safe Reinforcement Learning Benchmarking


90. AutoPass: Evidence-Guided LLM Agents for Compiler Performance Tuning


91. Robust $Q$-learning for mean-field control under Wasserstein uncertainty in common noise


92. Boundary Embedding Shaping with Adaptive Contrastive Learning for Graph Structural Disentanglement


93. ELVA: Exploring Ranking-Driven Universal Multimodal Retrieval


94. Editorial Alignment: A Participatory Approach to Engaging Editorial Expertise in LLM-mediated Knowledge Dissemination


95. The Register Gap: A Meaning Intelligence Framework for Nigerian Public Discourse


96. Finetuning Vision-Language-Action Models Requires Fewer Layers Than You Think


97. SPOT-E: Test-Time Entropy Shaping with Visual Spotlights for Frozen VLMs


98. ScholarQuest: A Taxonomy-Guided Benchmark for Agentic Academic Paper Search in Open Literature Environments


99. Learner-based Concept Drift Detection: Analysis and Evaluation


100. FlowMaps: Modeling Long-Term Multimodal Object Dynamics with Flow Matching


101. HilDA: Hierarchical Distillation with Diffusion for Advancing Self-Supervised LiDAR Pre-trainin


102. Evaluating and Enhancing Negation Comprehension in Remote Sensing MLLMs


103. MedRLM: Recursive Multimodal Health Intelligence for Long-Context Clinical Reasoning, Sensor-Guided Screening, Evidence-Grounded Decision Support, and Community-to-Tertiary Referral Optimization


104. From Texts to Scores: Tracing the Emergence of Essay Quality Representations in Large Language Models


105. Hybrid ANN-SNN Pipeline with Local Plasticity


106. Frequency-Aware Flow Matching for Continuous and Consistent Robotic Action Generation


107. Dual-Agent Framework for Cross-Model Verified Translation of Natural-Language Protocols into Robotic Laboratory Platform


108. Sensorimotor World Models: Perception for Action via Inverse Dynamics


109. Hybrid Diffusion Transformer for Instruction-Guided Audio Editing via Rectified Flow


110. MakeupMirror: Improving Facial Attribute Preservation in Diffusion Models for Makeup Transfer


111. IHUBERT: Vector-Based Semantic Deduplication and Domain-Balanced Pretraining for Persian Resources


112. The Hidden Evolution of Disguised Visual Context inside the VLM


113. Variable-Length Tokenization via Learnable Global Merging for Diffusion Transformers


114. Evaluation of EEG Foundation Models for Event-Based Burst-Suppression Detection in ICU


115. See-and-Reach: Precise Vision-Language Navigation for UAVs within the Field of View


116. AI Economist Agent: An Agentic Framework for Model-Grounded Economic Analysis with RAG, Knowledge Graphs, and Large Language Models


117. A Neuromorphic Reinforcement Learning Framework for Efficient Pathfinding in Robotic Mobile Fulfillment Systems


118. When Lower Privileges Suffice: Investigating Over-Privileged Tool Selection in LLM Agents


119. Hierarchical Control in Multi-Agent Games: LLM-based Planning and RL Execution


120. StreamKL: Fast and Memory-Efficient KL Divergence for Boosting Attention Distillation


121. Connect the Dots: Training LLMs for Long-Lifecycle Agents with Cross-Domain Generalization Via Reinforcement Learning


122. Tri-Info: Generalizable, Interpretable Failure Prediction for VLA Models via Information Theory


123. Beyond Static Endpoints: Tool Programs as an Interface for Flexible Agentic Web Services


124. The Algorithmic-Human Manager: AI, Apps, and Workers in the Indian Gig Economy


125. ROSE: Benchmarking the Perception-to-Action Gap in Multimodal Models


126. Confidence Calibration for Multimodal LLMs: An Empirical Study through Medical VQA


127. SIMBA: ABidirectional Retrieval Forward Simulation Framework for Modeling FY-4A GIIRS Hyperspectral Infrared Radiances Toward NWP Applications


128. Triangular Consistency as a Universal Constraint for Learning Optical Flow


129. Speeding up the annotation process in semantic segmentation industrial applications


130. Spatial-Aware Reduction Framework: Towards Efficient and Faithful Visual State Space Models


131. Co-policy: Responsive Human-Robot Co-Creation for Musical Performances


132. Measuring Biological Capabilities and Risks of AI Agents


133. SL-S4Wave: Self-Supervised Learning of Physiological Waveforms with Structured State Space Models


134. FFinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming


135. PSCT-Net: Geometry-Aware Pediatric Skull CT Reconstruction via Differentiable Back-Projection and Attention-Guided Refinement


136. Large Language Models Do Not Always Need Readable Language


137. Neural Additive and Basis Models with Feature Selection and Interactions


138. When, Where, and How: Adaptive Binning for Tabular Self-Supervised Learning


139. CSWinUNETR: Segmentation of Thin Anatomical Structures in Medical Images


140. CREDENCE: Claim Reduction for Decomposition & Enhanced Credibility – Semantic Metrics and Convergence Analysis


141. Uncertainty-Aware Reward Modeling for Stable RLHF


142. ParaScale: Scale-Calibrated Camera-Motion Transfer via a Gauge-Invariant Parallax Number


143. Policy-aware Vector Search: A Vision for Fine Grained Access Control in Vector Databases


144. Improving End-to-End Speech Recognition for Dysarthric Speech through In-Domain Data Augmentation


145. Agentic Electronic Design Automation: A Handoff Perspective


146. Systematic Study of Dysarthric Speech Recognition: Spectral Features and Acoustic Models


147. Cross-Dataset, Age, and Gender Generalization: A Comprehensive Analysis of Fine-Tuning Strategies for Low-Resource Children’s ASR


148. Towards Engineering Scaling Laws with Pretraining Data Composition


149. Data Standards for Humanoid Robotics: The Missing Infrastructure for Physical AI


150. SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling


151. Temporal Self-Imitation Learning


152. Manifold Bandits: Bayesian Curriculum Learning over the Latent Geometry of Large Language Models


153. Beyond Uniform Forgetting: A Study of Sequential Direct Preference Optimization Across Preference Settings


154. QueryGaussian: Scalable and Training-Free Open-Vocabulary 3D Instance Retrieval


155. VOiLA: Vectorized Online Planning with Learned Diffusion Model for POMDP Agents


156. Bidirectional Tutoring for Developmental Motor Learning in Robots: Co-Developed Interaction Dynamics Support Stable Learning


157. NRITYAM: Language Models Meet Art and Heritage of Dance


158. Library-Aware Doubles and Iterative Repair for Large Language Model-Generated Unit Tests in OpenSIL Firmware


159. OnDeFog: Online Decision Transformer under Frame Dropping


160. AURA: Adaptive Uncertainty-aware Refinement for LLM-as-a-Judge Auditing


161. FineREX: Fine-Tuned NER-RE for Human Smuggling Knowledge Graphs


162. Efficiently Representing Algorithms With Chain-of-Thought Transformers


163. LOKI: Memory-Free Null-Space Constrained Lifelong Knowledge Editing


164. TeleMorpher: Toward Robust Simultaneous Motion-Location Editing


165. Creating Multilingual Mental Health Dialogue Datasets: Limits of Persona-Based Localization via Nationality and Language


166. Before the Labels: How Dataset Construction Shapes Suicidality Detection in Clinical Text


167. Hard or Just Unreached? Diagnosing the Sampling Blind Spot in Math-Reasoning Difficulty Estimation


168. Token Factory: Efficiently Integrating Diverse Signals into Large Recommendation Models


169. CTS-MoE: Implicit Terrain Adaptation via Mixture-of-Experts for Perceptive Locomotion


170. Formal Verification of Learned Multi-Agent Communication Policies via Decision Tree Distillation


171. RIVET: Robust Idempotent Voice Attribute Editing


172. VCG: A Multimodal Retrieval Framework for E-Commerce Video Feeds under Extreme Cold-Start Conditions


173. Before the Pull Request: Mining Multi-Agent Coordination


174. StaminaBench: Stress-Testing Coding Agents over 100 Interaction Turns


175. Latent Confounded Causal Discovery via Lie Bracket Geometry


176. FAPO: Fully Autonomous Prompt Optimization of Multi-Step LLM Pipelines


177. PrefSQA: Pairwise Preference Prediction for Speech Quality Assessment and the Critical Role of High Quality Datasets


178. IHBench: Evaluating Post-Interruption Recovery in Voice Agents with Structured Workflows


179. A BART-based approach with hierarchical strategy for Vietnamese abstractive multi-document summarization


180. FlowFake: Liquid Networks for Audio Deepfake Detection


181. Exploring Feature Extraction Technique Parameters for Acoustic Gunshot Classification


182. GDGU: A Gradient Difference-based Graph Unlearning Method for Cyberattack Localization in Electric Vehicle Charging Networks


183. Review of Machine Learning Models for Solar Energetic Particle Prediction


184. PerceptionDLM: Parallel Region Perception with Multimodal Diffusion Language Models


185. A Tool for the Synthesis of Adaptive Probabilistic Processors Based on the Ising Model


186. Techniques for Peak Memory Reduction for LoRA Fine-tuning of LLMs on Edge Devices


187. Concept Flow Models: Anchoring Concept-Based Reasoning with Hierarchical Bottlenecks


188. Can In-Context Learning Support Intrinsic Curiosity?


189. Secure Coding Drift in LLM-Assisted Post-Quantum Cryptography Development: A Gamified Fix


190. Scaling Generative Foundation Models for Chest Radiography with Rectified Flow Transformers


191. Playful Agentic Robot Learning


192. JustDiag!: A Diagnostic Justification Engine for Accountable Root Cause Analysis


193. VERITAS: Verifier-Guided Proof Search for Zero-Shot Formal Theorem Proving


194. Execution-bound advisory automation for agentic AI: a reproducible AIBOM-driven CSAF-VEX framework


195. Interpretable and Verifiable Hardware Generation with LLM-Driven Stepwise Refinement


196. Bistable by Construction: Wall-Clock-Calibrated State Monitors Have No Moment-Detection Regime at Agent Cadence


197. DynAMO:Dynamic Asset Management Orchestration via Topological Multi-Agent Scheduling


198. Improving Code-Switching ASR with Code-Mixing Guided Synthetic Speech


199. How Linear Is a Transformer Feed-Forward Block? Per-Block Linear Recoverability Is Learned, Not Architectural


200. Emyx: Fast and efficient all-atom protein generation


201. Cost-Optimal LLM Routing with Limited User Feedback under User Satisfaction Guarantees


202. Protein Representation Learning with Secondary-Structure and Energy-Filtered Hydrogen-Bond Graphs


203. cAPM: Continual AI-Assisted Pace-Mapping with Active Learning


204. ProMUSE: Progressive Multi-modal Uncertainty-guided Staged Evidential Alzheimer Disease Classification


205. Human-like autonomy emerges from self-play and a pinch of human data


206. Zero-Inflated Gaussian Distributions Enable Parameter-Space Sparsity in Estimation-of-Distribution Algorithms


207. Information Lattice Learning as Probabilistic Graphical Model Structure Learning


208. Computational Identifiability


209. Physical Atari: A Robust and Accessible Platform for Real-time Reinforcement Learning on Robots


210. Trustworthy Multi-Agent Systems: Mitigating Semantic Drift with the Argent Signaling Protocol


211. Sign-Language Datasets at Scale: A Comprehensive Survey on Resources, Benchmarks, and Annotation Standards


212. Detecting Hallucinations for Large Language Model-based Knowledge Graph Reasoning


213. Where to Place the Query? Unveiling and Mitigating Positional Bias in In-Context Learning for Diffusion LLMs via Decoding Dynamics


214. DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence


215. How LLMs Fail and Generalize in RTL Coding for Hardware Design?


216. Disentangling Linguistic Relatedness from Task Alignment in Cross-Lingual Transfer


217. Ensembles of Large Language Models for Identifying EQ-5D Studies in PubMed Based on Their Abstracts


218. Exposing the Unsaid: Visualizing Hidden LLM Bias through Stochastic Path Aggregation


219. Human-AI Agent Interaction in a Business Context


220. Human Universal Grasping