LLM 관련 주요 논문 - 2026-07-27

1. The Regression Tax: Decomposing Why Skills Help and Hurt LLM Agents


2. TRACE-ROUTER: Task-Consistent and Adaptive Online Routing for Agentic AI


3. Dynamic Capability Scoping for Enterprise AI Agents: A Synthetic Dataset and Three-Source Permission Architecture


4. SceneActBench: Can Agents Act on the 3D Scenes They See?


5. Agentic Root Cause Analysis through Evidence-Grounded Reasoning


6. IDEAgent: Agentic Quality-Diversity Search for Research Idea Generation


7. Deconstructing Off-Policy Ratios: Entropy-Scaled Trust Regions for Asynchronous Reinforcement Learning


8. Industrial Tokenization for LLM-Based Health Intelligence: A Federated Architecture for Industrial Evidence Integration


9. Zero-Shot Mission-Level Evaluation for Aerial MLLM Agents


10. Learning as Reasoning Unfolds: Progressive Rollout Allocation for Efficient Reinforcement Learning


11. Semiotic logical hexagon theory for LLM logical reasoning


12. DAGForge: Auditable Causal DAG Authoring with Biomedical Literature


13. Persistent Computational State: A Session-Centric Runtime for Generative World Models


14. Do Modules Stay in Their Lane? Role Drift in Compound LLM Systems


15. Discrete Action Space as a Prerequisite for GRPO Convergence in Small-Model Continuous Control


16. Trajectory-Aware Retrieval Agents for Temporal Decision- Making


17. FBLayout: Optimizing Memory Layout for Efficient LLM Finetuning on Mobile GPUs


18. From Profiles to Steering Vectors: Global Sparse Priors and Local Semantic Calibration for Personalized Text Generation


19. Do VLMs Read or Rewrite? On Transcription Faithfulness in Vision-Language Models


20. Lost in Context: Addressing Context Anxiety in Large Language Models


21. Household Movement Detection in Mixed-Format Occupancy Data Using LLM-Based Entity Resolution


22. The Hard Decision Layer: Evidence for Committed Inference in Transformers


23. Procedural Knowledge Is Not Low-Rank: Why LoRA Fails to Internalize Multi-Step Procedures


24. SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text


25. Coupled Hierarchical Search over Topology and Execution for Agentic Workflow Synthesis


26. AgentKVShift: Efficient KV Cache Reuse for Agentic Memory Systems


27. Transferable Latency Prediction for Fast LLM Screening on Heterogeneous Edge Devices


28. Securing Multimodal AI through Internal Information Decomposition


29. Risk Is Not the Target: A Monotonic Framework for Evaluating Wildfire Operational Risk Signals


30. FlowEvo: Self-Evolving Agents through the Co-Evolution of Workflows and Executable Skills


31. Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science


32. CausalForge: A Formally Grounded, Self-Improving Agentic Framework for Automated Research in Causal Inference


33. MineValiCoder: Reliable Code Generation with Test Case Quality Mining and Bipartite Graph-Based Mutual Validation


34. Beyond Perspectives: A Trio-Ethnography of Interpretation Evolution in LLM-Supported Programming Education


35. A Self-Calibrating Agentic AI Framework for Autonomous Edge Resource Allocation


36. HiKV: Hierarchical Importance-Aware KV Cache with Hardware Acceleration for LLM Decoding


37. Teachy Mini: Development and Preliminary Evaluation of a Knowledge-Based Generative Social Robot for Higher Education


38. Cross-Tokenizer On-Policy Distillation via Byte-Prefix Marginalization


39. Towards Trustworthy and Cost-Efficient Data Integration: From Naïve RAG to Agentic RAG


40. IFCLoRA: Topology-Aware Rank Allocation for Parameter-Efficient Fine-Tuning


41. Why Large Language Models and Humans Converge and Diverge in Evaluating Creativity


42. Filling Before Advancing: Capability-Gap-Driven Post-Training for Scenario-Specialized Remote Sensing MLLMs


43. From Isolated Tasks to Structured Capabilities: A Multilayer Taxonomy for Large Language Models


44. DBA-Bench: A Production-Fidelity Benchmark for LLM-Based Database Operations Agents


45. dRAE: Representation Autoencoder with Hyper-Spherical Codes


46. Benchmarking Text-to-SQL under Role-Based Access Control


47. MEUSLI: a Multilingual Projector for LLM-based ASR and Beyond


48. Benchmarking Fine-tuning and Retrieval Strategies for a Multimodal Language Model on the NRC Reactor Operator Licensing Examination


49. EVL-MCoT: Enhanced Vision-Language Multi-CoT for Harmful Meme Detection


50. Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning


51. Unified Static-Dynamic Pruning for Efficient LLM Inference


52. Data Quality over Capacity: Internalizing Documents into LoRA Adapters for Closed-Book QA


53. SCALE: Self-Supervised Constraint-Aware Layout GEneration for Local P&R DRV Fixing at Advanced Nodes


54. ToolGuardian: Declarative Security for AI Agent-Tool Interactions


55. Graph-Theoretic Neural Network Fragmentation with Covariant Direct Molecular Force Learning: Enabling Coupled-Cluster Accuracy AIMD for Fluxional Systems


56. Probing Latent Colombian Identity Inferences in Qwen2.5-7B with Natural Language Autoencoders


57. Every Model Cheats: Prompt-Level Mitigation of Cheating on Offensive Cyber Tasks


58. Co-design of LLM-based preference agents: participation may drive overtrust


59. A Defense of the Quadratic Model


60. Enhancing SLMs for Sustainable Code Optimization in Radio-Astronomy


61. Pixels for Programs? A Cross-Provider Case Study of Input-Token Accounting for Source Code as Text and Images


62. Ordered Action Tokens for Visuomotor Policy Learning


63. Cross-Model LLM Code Review: Should you use Claude to review Codex or vice versa?


64. Tool-Guided Retrieval-Augmented Repair for Securing LLM-Generated C Code


65. Adversarial Style Optimization: Enhancing VLM Jailbreaks by GRPO-based Stylistic Triggers Optimization


66. From Obligation to Specification: A Survey on Validating EU AI Act Requirements in RE


67. Decoupled Attention Fusion: Accelerating RAG with Efficient KV Cache Reuse


68. Control panels to clarify user intent with Large Language Models