전체 AI 논문 - 2026-07-23

1. SoftReason: A Fully Differentiable Neuro-Soft-Symbolic Deductive Reasoning Architecture over High-Dimensional Perceptual Data


2. Train the Model, Not the Reader: Decodability Supervision for Verifiable Activation Explanations


3. PoTRE: Test-Time Reasoning inspired by Cognitive Heterogeneity


4. CUSUM-Shaped Inference-Time Monitoring and Targeted Re-Decoding for Quantized Small Language Model Reasoning


5. TRUST-ESD: A Risk-Calibrated and Governance-Aware AI Framework for Enterprise Strategic Decision Support Under Uncertainty


6. PRO-LONG: Programmatic Memory Enables Long-Horizon Reasoning


7. Reading and Steering Representations of Materials-Science Mechanisms in an Open-Weight Language Model


8. Global Difference Constraint Propagation for Constraint Programming


9. EvoDRC: A Self-Evolving Agentic Framework for Automated DRC Violation Repair


10. Safe Remediation as Risk-Constrained Intervention Decision in Microservice Systems


11. CLARK: Closed-loop Learning for Adaptive Reasoning over Knowledge Graphs


12. Coordinating from Memory: Graph-Structured Experience Reuse for Multi-Agent Adaptation in Dynamic Manufacturing


13. The Giant Hippocampus: From Structural Monoculture to a System of Systems


14. EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization


15. SenWorld: A Digital-Twin Simulation for Generating Context-Rich Evaluation Data


16. MOF-Sleuth: Tool-Grounded Reward Alignment for Explainable Fine-Grained MOF CIF Auditing


17. Long-Term Sequential Decision Making under Risk


18. JANUS: Foreseeing Latent Risk for Long-Horizon Agent Safety


19. DocOps: A Verifiable Benchmark for Autonomous Agents in Complex Document Operations


20. Know Your Agent: Reconnaissance-Driven Pentesting of AI Agents


21. Rewarding Better Thinking for LLM Preference Alignment


22. Silent Failures in Multimodal Agentic Search:A Diagnostic Taxonomy and Cross-Judge Evaluation


23. Symbol and Footprint Database for Electronic Components by Agentic Recognition and Generation


24. Edge Intelligence in Civil Aviation: Paradigms, Techniques, and Applications


25. Knowledge-Centric Self-Improvement


26. Sophisticated Policies from Epistemic Priors


27. The Chronos Vulnerability: A Taxonomy of Temporal Persistence and Memory-Based Deception in Agentic AI


28. FORCE-Bench: A Benchmark, Dataset, and Evaluation Harness for Agentic AI in Enterprise Finance


29. ITPEval: Benchmarking Formal Translation Across Interactive Theorem Provers


30. HyGRL: Adaptive Hybrid Graph Reasoning for Multi-Entity Questions


31. CrackedPDFs: A Controlled Benchmark for Hidden Prompt Injection in PDFs


32. Euclean: Automated Geometry Problem Formalization with Unified Verification in Lean


33. Mitigating Scaffolding Collapse in Socratic Tutors via Representation Alignment


34. Beyond Tracking or Shortcut: Composition-Bounded Predictive States in Poker Autoregressive Models


35. Spectral-LSH: Sub-Quadratic Prompt Compression via Krylov-Projected Locality-Sensitive Hashing


36. Rethinking Uncertainty Evaluation in Large Language Models


37. Geometry-Guided Constraint Learning for LLM Safety Classification


38. Logic-Guided Data Extraction with Answer Set Programming and Large Language Models


39. Statistically Grounded Sparse-Feature Interventions for Activation-Space Control in Large Language Models


40. AdaRoPE: Not All Attention Heads Should Rotate and Scale Equally


41. GraphContainer: A Unified Platform for Comparing and Debugging Graph RAG Methods


42. Lifted Representation Hypothesis in Language Models


43. Profile-Graph Memory for LLM Agents: Implicit Cross-Entity Traversal through Narrative Profiles


44. LISA: Linear-Indexed Sparse Attention for Efficient Long-Context Reasoning


45. Stochastic Primal-Dual Decoding for Multiobjective Generative Recommender Systems


46. NEXUS: Structured Runtime Safety for Tool-Using LLM Agents


47. Information Discernment in Large Language Models


48. FormulaSPIN: Self-Play Fine-Tuning for Natural Language to Spreadsheet Formula Generation


49. Benchmarking Confidential GPU Inference on NVIDIA H100 under Intel TDX


50. OpenEvoShield: Dual Non-Stationary Continual Defense for Open-World Multi-Agent System Attacks


51. Hybrid LSTM-Graph Neural Framework for Robust Financial Fraud Detection and Adversarial Resilience


52. FineServe: A Fine-Grained Dataset and Characterization of Global LLM Serving Workloads


53. Persian Pixel: A large-scale synthetic OCR dataset for Persian language


54. FMRP-LEAN: A HIPAA-Compliant AI-Augmented LIMS Architecture for End-to-End Clinical Assay Workflow Optimization


55. Generative AI floods and dilutes the market for books


56. Closing the Lab-to-Store Gap: A Data-Efficient Post-Training and Experience-Driven Learning VLA Framework for Retail Humanoids


57. Understanding Generative AI-mediated User Engagement with Academic Library Resources


58. Toward Reliable RGB-D Semantic Segmentation: Handling Missing Modalities via Condition Dropout


59. Don’t Trust the Label: License Laundering in AI Supply Chains


60. Courteous Anticipation: Improving Long-Lived Task Planning in Persistent Shared Environments


61. Sound Probabilistic Safety Bounds for Large Language Models


62. Self-supervision drives representational convergence in medical foundation models more than clinical supervision


63. The Maskability Index: Predicting Task-Objective Alignment in Pretrained Language Models


64. The Ethics of Autonomous AI Agents for Offensive Security


65. Pushing the Frontier of Full-Song Generation: Hierarchical Autoregressive Planning Meets Flow-Matching Rendering


66. On the Systematic Challenges of Culturally Loaded Machine Translation: Dream of the Red Chamber as the Cultural Lens


67. DQAOA-GPT: AI-Accelerated Distributed Quantum Optimization for Combinatorial Problems


68. Small, Free, and Effective: Orchestrating Open-Weight Small Language Models to Outperform Single LLM for Malware Analysis


69. ELSAA: Efficient Low-Rank and Sparse Attention Approximation for Training Transformers


70. The Quadrilateral Loss: Additivity as a Measurable Behavior of Dense Neural Networks


71. StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation


72. Audio-Zero: Label-Free Self-Evolution for Fine-Grained Audio Reasoning


73. Active Inference as a Convex Markov Decision Process


74. SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD


75. Formal Foundations for Known Good Reliable Die Screening in Chiplet-Based AI Systems-on-Chip


76. PRIME-SVR: Physics-infoRmed Implicit Multi-Echo Slice-to-Volume Reconstruction for Fetal T2 mapping


77. ENTRAP-VL: A Taxonomic Probe for Dual Contextual Entrainment in Vision-Language Models


78. Reinforcement Learning for Large Language Model Selective Evidence Adoption from Contaminated Retrieval Results


79. Co-Evolving LLM Evaluators and Policies via DynamicRubric


80. Language-Specific versus Cross-Lingual Knowledge Graphs for Implicit Aspect Identification in Arabic: A Comparative Study of Reasoning and Adaptation Strategies


81. Test Case Prioritization for DNNs via Neural Collapse Instability


82. A Systematic Benchmark of Intensity Normalisation Methods for 3D Knee MRI Segmentation and Cross-Domain Generalisability


83. Drift-Aware RL-based Wavelet Denoising for Network-Traffic Anomaly Detection


84. Taming the Security-Energy Paradox: A Green AI Approach to Optimized Android Malware Detection


85. Post-Training in Time Series Foundation Models: A Unifying Framework


86. Are Attributions of Consciousness to AI Chatbots Epistemically Innocent?


87. TINY_SCHILLER: A Drop-In German Drama Corpus for Small Language Models


88. Time Series Network Utilization KPI Forecasting Using Advanced AI/ML Models


89. When Shippers Become Algorithms: Candidate Exposure, Information Design, and the Concentration of LLM-Mediated Freight Markets


90. HijackKV: New Threat in Position-Independent KV Cache Reuse


91. When Does Knowledge Distillation Hurt? Reliability-Aware Distillation for Low-Resource Language Summarization


92. G-MAD: A Game-Based Data Generation Framework for Multi-View RGB-T Aerial Object Detection


93. A Framework of User Experience Principles for Human-AI Agent Interaction in the Workplace


94. OSVE: One Step Video Editing with One Step Diffusion Models


95. Defense Against LLM Backdoors using Critical Neuron Isolation Pruning


96. Overview of FinMMEval 2026 Task 2: Multilingual Financial Short-Answer Question Answering


97. PRISM-DR: Per-lesion Retinal Inference with Specialist Models for Diabetic Retinopathy


98. Memory-Augmented Multimodal Large Language Models for Small Object Understanding in Streaming Aerial Videos


99. Overview of FinMMEval 2026 Task 1: Multilingual Financial Multiple-Choice Question Answering


100. Auto-Fill: Learning to Predict Missing Values Accurately with Specialist Language Models


101. Sentence Splitter: Uncovering Latent Factual Structure for Self-Supervised Learning


102. Beyond Fail-to-Pass: Iterative Hardening of Co-Generated Bug Reproduction Tests and Fixes


103. OPIUM: Mitigating Steering Externalities and Over-Refusal via Dual Objective Latent Optimization


104. Physics-Aware Complex-Valued State Space Model with Scattering-Prior Feature Modulation for PolSAR Image Classification


105. RPPNet: Perceptually-Grouped Rhythm-Pitch Primitives for Long-Term Structure Melody Generation via Boundary-Aware Modeling


106. An Isotropy-Preserving Spectral Cap for Muon: Theory and Three Case Studies


107. Convergence-Latency-Aware Adaptive Modulation and Resource Allocation in RIS-Assisted Wireless Federated Learning


108. Learning the Arabic Dialect Continuum as a Continuous Space: A Regression Approach to Speaker Origin Prediction


109. The World Model Remembers, the Actor Forgets: Dream Rehearsal for Continual Model-Based RL


110. An Automated Framework for Extracting Reachable Attack Chains from Cyber Threat Intelligence Reports


111. Personalized Recommendation Tool Learning via Autonomous Language Agents


112. Did Alice Do Wrong? Cross-Cultural Differences in Student Perceptions of Generative AI Use in University Computing Education


113. PhenSPINE: A Standardized Benchmark for Spine Pathology Diagnosis


114. SLPO: Scaling Latent Reasoning via a Surrogate Policy


115. Reference-Free Evaluation of Reasoning in Open-Ended Question Answering


116. FedLSG: LLM-Enhanced Semantic Calibration for Federated Graph Backdoor Defense


117. PerfAgent: Profiler-Guided Iterative Refinement for Repository-Level Code Optimization


118. Anatomy of a Sound Neural Reasoner: One-Shot Amortization, First-Pass Poisoning, and Search Inertness in Clue-Rich Completion


119. Adaptive Capitulation: A Structural Failure Mode of LLM Responses in Vulnerability Contexts


120. Understanding Developer Pain Points in Federated Learning: Insights from Stack Overflow and GitHub


121. SCPP: A Unified Python Library for Soft Clustering


122. Causal dictionary learning reveals and validates transcription-factor binding features in genomic language models


123. Juxtaposition of Shallow Reservoir-Triggered Seismicity and Deep Tectonic Locking in the Qiaojia-Dongchuan Seismic Gap


124. Fine-grained Computation-Communication Overlap via Tile-level Signaling and Scheduling for Mixture-of-Experts


125. Trustworthy Privacy-Preserving Multimodal Federated Learning for Personalised Breast Cancer Prediction


126. D3VL: Understanding Driving Scenes from 3D Time Series Data and Video with Language Models


127. SynPre-FL: Synthetic data-driven pretraining integrated Federated Learning training framework


128. Hybrid LLM-Guided Search for Quantum Reservoir Architecture Design


129. Integrity of peer-to-peer distributed LLM inference under malicious nodes


130. ModPack: An Extensible Teleoperation Interface for Bimanual Mobile Manipulation


131. MoA-Structured Decode Attention DNF Derivation, KV-Cache Accumulation, GQA/MQA, and OpenACC Kernel


132. Predictive Extrema, Unprofitable Policies: An AI-Assisted Audit of Candle-Based Binance Spot Timing Models


133. REGEN: Replay-recycling for Expert-to-Generalist distillation with Offline Reinforcement Learning


134. Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents


135. Unlearning as Distribution Restoration: A Controlled Counterfactual Study, a Validated Selective Screen, and the Limits of Oracle-Free Certification


136. BaseRT: Advancing Best-in-Class LLM Inference with Apple M5 Neural Accelerators


137. Building Trust in Autonomous Commerce: A Verifiable Global Event Timeline and AI-Ready Fraud Intelligence Layer


138. ChainWatch: A Kill Chain-Aligned Sequential Detection Framework for Multi-Step Attacks in MCP-Based AI Agent Systems


139. BRIM: Workload-Balanced Dual-Sided Bit-Serial Sparse Inference Accelerator


140. ChannelGuard: Safe Models Do Not Compose into Safe Multi-Agent Systems


141. Making Single-Cell Data Distillation Auditable: Traceable Real-Cell Coresets via Discrete Min-Max Selection


142. JailMeter: An Evidence-Based Evaluation Framework for Jailbreak Attacks on Large Language Models


143. Opto-ViT-v2: Noise-Resilient On-Chip Fine-Tuning for Photonic Near-Sensor Vision Transformer Accelerators


144. Auditing Retrieval-Augmented LLM Hypotheses for Longitudinal Cell Painting Morphology


145. Structured Latent Space Modeling over Multi-Scale Temporal Patches for Multivariate Time Series Forecasting


146. Recovering Clinical Utility Under Differential Privacy: Empirical Validation of Adaptive Federated Aggregation on Heterogeneous Cardiovascular Datasets


147. Leveraging Offline Supervision for Efficient and Generalizable Reinforcement Learning in Large-Scale Vision-Language-Action Models


148. From Trajectories to Prefixes: Reusing Teacher Trajectories via Replayed Prefixes and Online Continuation


149. Cross-Subject Semantic Decoding with Shared-Space Alignment for Generalized Neural Representation Learning


150. Decodable but Not Detectable: A Leakage Fingerprint for Near-OOD Benchmarks


151. LAARA: Layer-Aware Adaptive Rank Allocation for Parameter-Efficient Fine-Tuning


152. Simulating Eutopia: Revisiting Long-term Fairness with Outcomes, Performativity, and Dynamics


153. Scale-Aware Learning of Chaotic Dynamics on Unstructured Meshes via Binned Spectral Losses


154. Challenges of Explainability in Continual Learning for Time Series Forecasting


155. Economic Evaluations of Language Models


156. Stateful Guardrails for Multi-Turn LLM Systems: A Conversational Risk Accumulation Framework