03 / ARTICLE LEDGER
关联文章
按发布时间倒序
- RoboPocket: Improve Robot Policies Instantly with Your Phone 2026-03-06 · ARXIV
- Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought 2026-03-06 · ARXIV
- POET-X: Memory-efficient LLM Training by Scaling Orthogonal Transformation 2026-03-06 · ARXIV
- Leveraging LLM Parametric Knowledge for Fact Checking without Retrieval 2026-03-06 · ARXIV
- Cheap Thrills: Effective Amortized Optimization Using Inexpensive Labels 2026-03-06 · ARXIV
- Censored LLMs as a Natural Testbed for Secret Knowledge Elicitation 2026-03-06 · ARXIV
- TaxonRL: Reinforcement Learning with Intermediate Rewards for Interpretable Fine-Grained Visual Reasoning 2026-03-05 · ARXIV
- SimpliHuMoN: Simplifying Human Motion Prediction 2026-03-05 · ARXIV
- Robustness of Agentic AI Systems via Adversarially-Aligned Jacobian Regularization 2026-03-05 · ARXIV
- AgentIR: Reasoning-Aware Retrival for Deep Research Agents 2026-03-05 · ARXIV
- A Dual-Helix Governance Approach Towards Reliable Agentic AI for WebGIS Development 2026-03-05 · ARXIV
- $τ$-Knowledge: Evaluating Conversational Agents over Unstructured Knowledge 2026-03-05 · ARXIV
- ZipMap: Linear-Time Stateful 3D Reconstruction with Test-Time Training 2026-03-05 · ARXIV
- Tuning Just Enough: Lightweight Backdoor Attacks on Multi-Encoder Diffusion Models 2026-03-05 · ARXIV
- Monitoring Emergent Reward Hacking During Generation via Internal Activations 2026-03-05 · ARXIV
- Hindsight Quality Prediction Experiments in Multi-Candidate Human-Post-Edited Machine Translation 2026-03-05 · ARXIV
- FedCova: Robust Federated Covariance Learning Against Noisy Labels 2026-03-05 · ARXIV
- Speculative Speculative Decoding 2026-03-04 · ARXIV
- Valet: A Standardized Testbed of Traditional Imperfect-Information Card Games 2026-03-04 · ARXIV
- LoGeR: Long-Context Geometric Reconstruction with Hybrid Memory 2026-03-04 · ARXIV
- Inherited Goal Drift: Contextual Pressure Can Undermine Agentic Goals 2026-03-04 · ARXIV
- CFG-Ctrl: Control-Based Classifier-Free Diffusion Guidance 2026-03-04 · ARXIV
- Tether: Autonomous Functional Play with Correspondence-Driven Trajectory Warping 2026-03-04 · ARXIV
- Why Adam Can Beat SGD: Second-Moment Normalization Yields Sharper Tails 2026-03-04 · ARXIV
- Odin: Multi-Signal Graph Intelligence for Autonomous Discovery in Knowledge Graphs 2026-03-04 · ARXIV
- Multi-Scale Adaptive Neighborhood Awareness Transformer For Graph Fraud Detection 2026-03-04 · ARXIV
- MoECLIP: Patch-Specialized Experts for Zero-shot Anomaly Detection 2026-03-04 · ARXIV
- Interpreting Speaker Characteristics in the Dimensions of Self-Supervised Speech Features 2026-03-04 · ARXIV
- From Complex Dynamics to DynFormer: Rethinking Transformers for PDEs 2026-03-04 · ARXIV
- Evaluating Performance Drift from Model Switching in Multi-Turn LLM Systems 2026-03-04 · ARXIV
- Compact Prompting in Instruction-tuned LLMs for Joint Argumentative Component Detection 2026-03-04 · ARXIV
- Beyond Task Completion: Revealing Corrupt Success in LLM Agents through Procedure-Aware Evaluation 2026-03-04 · ARXIV
- Tool Verification for Test-Time Reinforcement Learning 2026-03-03 · ARXIV
- Symbol-Equivariant Recurrent Reasoning Models 2026-03-03 · ARXIV
- Sketch2Colab: Sketch-Conditioned Multi-Human Animation via Controllable Flow Distillation 2026-03-03 · ARXIV
- Reasoning Core: A Scalable Procedural Data Generation Suite for Symbolic Pre-training and Post-Training 2026-03-03 · ARXIV
- Partial Causal Structure Learning for Valid Selective Conformal Inference under Interventions 2026-03-03 · ARXIV
- Multi-Head Low-Rank Attention 2026-03-03 · ARXIV
- Frontier Models Can Take Actions at Low Probabilities 2026-03-03 · ARXIV
- Conformal Policy Control 2026-03-03 · ARXIV
- Adaptive Confidence Regularization for Multimodal Failure Detection 2026-03-03 · ARXIV
- Probabilistic Learning and Generation in Deep Sequence Models 2026-03-03 · ARXIV
- Principled Fast and Meta Knowledge Learners for Continual Reinforcement Learning 2026-03-03 · ARXIV
- MC-Search: Evaluating and Enhancing Multimodal Agentic Search with Structured Long Reasoning Chains 2026-03-03 · ARXIV
- KVSlimmer: Theoretical Insights and Practical Optimizations for Asymmetric KV Merging 2026-03-03 · ARXIV
- Knowledge without Wisdom: Measuring Misalignment between LLMs and Intended Impact 2026-03-03 · ARXIV
- Evaluating AI Grading on Real-World Handwritten College Mathematics: A Large-Scale Study Toward a Benchmark 2026-03-03 · ARXIV
- Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization 2026-03-03 · ARXIV
- CHIMERA: Compact Synthetic Data for Generalizable LLM Reasoning 2026-03-03 · ARXIV
- BioProAgent: Neuro-Symbolic Grounding for Constrained Scientific Planning 2026-03-03 · ARXIV