전체 AI 논문 - 2026-07-20

1. CRAFT: Clustering Rubrics to Diagnose Weak LLM Capabilities and Generate Targeted Fine-Tuning Data


2. Harmonizing AI Safety Thresholds


3. SciForge: An AI-Native, Multimodal Workbench for Scientific Discovery


4. Closing the AI Trust Gap: The Case for Independent Certification for Trustworthy AI


5. A Formally Grounded ODRL Evaluator: Implementation and Comparison


6. DSWorld: A Data Science World Model for Efficient Autonomous Agents


7. Knowledge-Centric Agents for Workflow Generation


8. AgentFAIR: A Multi-Agent Collaborative Framework for FAIRness Evaluation of Geospatial Datasets


9. NeurOWL: An LLM-Based Neural-symbolic Framework for Incomplete OWL Ontology Reasoning


10. Behavioral Controllability of Agentic Models for Information Extraction: From Fixed Workflows to Reflective Agents


11. S1-Omni: A Unified Multimodal Reasoning Model for Scientific Understanding, Prediction, and Generation


12. ToolVerse: Unlocking Massive Environments and Long-Horizon Tasks for Agentic Reinforcement Learning


13. Neuro-Symbolic AI for LEED compliance: Document-Centric Benchmarking, Deterministic Numeric Checking, and When Multimodal Hurts


14. MGDT: MLLM-Guided Diffusion Transformer with Relation-Adaptive Mixture-of-Experts for Multimodal Knowledge Graph Completion


15. SeerGuard: A Safety Framework for Mobile GUI Agents via World Model Prediction


16. Logic, Optimization, and Artificial Intelligence


17. A Critical Analysis of Trustworthy AI Tools, Mark Frameworks, and the Implementation Chasms


18. From Black Box to Executable Logic: Explainable Reinforcement Learning through Prolog Expert Systems


19. Beyond a Joke: Multi-Angle Reasoning for Detecting and Explaining Harmful Humor in Memes


20. Do Coding Agents Need Executable World Models, Simplification, and Verification to Solve ARC-AGI-3?


21. DrawingVQA: A Real-World Benchmark for Multi-Depth Visual-Textual Reasoning on Construction Drawings


22. Precise but Uncoupled: Reviewer Precision Does Not Guarantee Critique Uptake in Multi-Agent Math Reasoning


23. AnovaX: A Local, Multi-Agent Voice Assistant with LLM Planning, Typed Executors, and Adaptive Recovery


24. Cura 1T: Specialized Model for Agentic Healthcare


25. Causal-Audit: Explicit and Auditable Graph-based Reasoning via Target-Aware Causal Chain Construction


26. GraphDx: A Cost-Aware Knowledge-Enhanced Multi-Agent Framework for Sequential Diagnosis


27. Evaluating Open-Weight LLMs for Generating Structured Threat Information for Autonomous Vehicle Vulnerabilities


28. When Does Muon Help Agentic Reinforcement Learning?


29. An Exam for Active Observers


30. When Do Multi-Agent Systems Help? An Information Bottleneck Perspective


31. ToolSciVer: Multimodal Scientific Claim Verification with Visual Tool Augmented Reinforcement Learning


32. A Methodology for Auditable Trustworthiness Levels in AI Lifecycle Governance


33. Understanding Reasoning from Pretraining to Post-Training


34. DADiff: Diffusion-Driven Cross-Domain Policy Adaptation for Reinforcement Learning


35. HCIG: A Hierarchical Cross-Modal Incongruity Graph Network for Multimodal Sarcasm and Cyberbullying Detection


36. JoyNexus: Service-Oriented Multi-Tenant Post-Training for VLA Models


37. LLM-Powered Agentic AI for 5G/6G Networks: A Tutorial and Survey on Architectures, Protocols, and Standardization


38. Spatial Normalization for Cross-Domain Retinal Layer Segmentation in Optical Coherence Tomography


39. When Model Merging Rivals Joint Multi-Task Reinforcement Learning: A Task-Vector Geometry Analysis


40. Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning


41. Loop the Loopies!


42. Revisiting data-driven dynamic security assessment with a tabular foundation model


43. Rethinking Quantum Continual Learning with Quantum Fisher Information


44. Candidate Attended Dialogue State Tracking Using BERT


45. DPNeXt: A Lightweight Multi-Scale Feature Fusion Framework for Efficient ViT-Based Multi-Task Dense Prediction


46. Robustness of Reinforcement Learning-Based Congestion Management in Low-Voltage Grids


47. Sociocultural Influences on Opinion Formation: Word of Mouth Dynamics, Mass Media and Behavioural Development


48. When Not to Automate: A Formal Protocol for Human Preservation in AI-Optimized Organizations


49. On the Failure of Boundary-Seeking Distillation in Bottlenecked Generative Architectures


50. Orbis 2: A Hierarchical World Model for Driving


51. Induction in Both Directions: A Mechanistic Analysis of In-Context Learning in Masked Diffusion Language Models


52. Perceived AGI: Believability as Dimensional Completeness, Not Capability


53. DECODEM: Data Extraction from Corporate Organizational Documents via Enhanced Methods


54. EgoExoMoCap: Distributed Ego-Exo Human Motion Capture


55. Conditional Reliability of Toxicity Signals for Multilingual and Code-Mixed Abuse Detection


56. Agentic Synthesis against Counterexample-Supplemented Sketches


57. Test-Time Noise Guided Adaptation for Realistic Autoregressive Video Generation


58. CAMMAR: Culture-Aware Matryoshka for Metaphorical Arabic Representations


59. RTL-Sequencer: Towards Scalable RTL Timing Prediction with the Sequence-based Paradigm


60. In-context learning of closed form solution to simple linear regression task using transformer with linear self-attention


61. Knowledge-Assisted Multi-Graph Dependency Learning for Multivariate Time Series Anomaly Detection in Multi-Stage Industrial Processes


62. On the Geometry of Learned Representations in Event-Based Multi-Modal Egomotion Estimation


63. Modularized Dynamic-Granularity Video LLM for Multi-Event Long Video Understanding


64. AquaAugmentor: A Novel Feature Augmentation Algorithm for Water Potability Prediction


65. Scaling Time Series Classification via XAI-Driven Data Reduction


66. GeoChrono: Benchmarking and Rethinking Long-Term Temporal Understanding in Remote Sensing


67. AuEmoChat: Authentic Emotion Understanding and Rendering for Conversational Speech Synthesis


68. Debiasing Text-to-Image Evaluation via Implicit Cultural Alignment Reward Modeling


69. Map as a Prompt: Learning Multi-Modal Spatial-Signal Foundation Models for Cross-scenario Wireless Localization


70. Efficient Difficulty-Aware Dynamic Routing for Diffusion-Based Real-World Image Super-Resolution


71. Toward a mechanistic understanding of inference in visual cortex and diffusion models


72. On the Structure of Address in Multi-Party Dialogue: From Discrete Labels to Continuous Levels


73. IMBench: A Benchmark for Intuitive Robotic Manipulation


74. A cubical formalisation of topos causal models: intervention, sheaf gluing, and the intuitionistic do-calculus


75. Think at 5 Hz, Act at 20 Hz: Asynchronous Fast-Slow Vision-Language-Action Inference for Closed-Loop Driving


76. AEGIS: Assay-Aware Protocol Validation and Runtime Monitoring for Open-Source Liquid Handling Robots


77. Process Reward Informed Tree Rollout for Effective Multi-Turn RL


78. Scalable LLM Agent Tool Access in the Cloud


79. Field-Aware RankMixer with Dual-Stream Bilinear Fusion for the Tencent UNI-REC Challenge


80. MemoGuard: An Adaptive Runtime for Guarding Against Memory Traps in Communication-Limited Robot Navigation


81. Information-Directed Sampling for Causal Bandits


82. Ask Twice, Look Twice: Prompt Echoing Resolves the Question-First Paradox in Vision-Language Models


83. Hard Rules, Soft Preferences: Bridging Reasoning, Learning, and Optimization for Personalized Packing Checklist Generation


84. Evolutionary Algorithm-Guided LLMs for Physics-Informed Neural Network Design


85. From Feasibility to Desirability: Plan, Learn, Adapt (PLA) Framework for Personalized On-Device Itinerary Generation


86. CoWeaver: A Bi-directional, Learnable and Explainable Matching Engine for Mixed Human-Agent Science Collaboration


87. Kolmogorov–Arnold Networks for Small Language Models


88. Recursive Harness Self-Improvement


89. SLAPBench: Benchmarking Multimodal Large Language Models for Four-Finger SLAP Fingerprint Verification


90. Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching


91. An Auto-Scaling Approach for Serverless Environments Based on a Multi-Expert Consensus Mechanism


92. LLM-Driven AutoML for Cross-Lingual Handwritten OCR: Closed-Loop Neural Architecture Search with GPT-5, GPT-4o, and Claude Sonnet 4


93. Verbalizable Representations Form a Global Workspace in Language Models


94. FLINT: Fingerprinting Federated Learning Architectures from 5G PHY-Layer Side Channels


95. Design-Based Supervised Learning with Noisy Human Labels


96. Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation


97. AI Trading: Evaluating Large Language Models for Technical Market Analysis


98. Partial Information Decomposition as a Multi-Contrast 3D MRI Selection Strategy for Resource-Constrained Deep Neural Network Training in Brain Tumor Segmentation


99. Large Language Models as Unified Multimodal Learners for Clinical Prediction


100. Lazy Arithmetic using Systolic Arrays for Closing the Verification Gap on Embedded Systems


101. Data-driven Video Codec with Implicit Neural Representations


102. AV-JEPA: Extending LeJEPA to Audio-Visual Self-Supervised Learning


103. Structure of the Circular-Dyadic Convolution Error


104. How Does Empowering Users with Greater System Control Affect News Filter Bubbles?


105. Empathy as Predictive Misalignment Tolerance: A Co-Regulation Framework and the Regime Structure of Dialogue Repair