03 / ARTICLE LEDGER

关联文章

按发布时间倒序

  1. Imitating What Works: Simulation-Filtered Modular Policy Learning from Human Videos 2026-02-16 · ARXIV
  2. CoPE-VideoLM: Codec Primitives For Efficient Video Language Models 2026-02-16 · ARXIV
  3. Asynchronous Verified Semantic Caching for Tiered LLM Architectures 2026-02-16 · ARXIV
  4. Creative Ownership in the Age of AI 2026-02-14 · ARXIV
  5. UniT: Unified Multimodal Chain-of-Thought Test-time Scaling 2026-02-13 · ARXIV
  6. Scaling Verification Can Be More Effective than Scaling Policy Learning for Vision-Language-Action Alignment 2026-02-13 · ARXIV
  7. On-Policy Context Distillation for Language Models 2026-02-13 · ARXIV
  8. MonarchRT: Efficient Attention for Real-Time Video Generation 2026-02-13 · ARXIV
  9. CM2: Reinforcement Learning with Checklist Rewards for Multi-Turn and Multi-Step Agentic Tool Use 2026-02-13 · ARXIV
  10. AttentionRetriever: Attention Layers are Secretly Long Document Retrievers 2026-02-13 · ARXIV
  11. Agentic Test-Time Scaling for WebAgents 2026-02-13 · ARXIV
  12. The Pensieve Paradigm: Stateful Language Models Mastering Their Own Context 2026-02-13 · ARXIV
  13. Stop Unnecessary Reflection: Training LRMs for Efficient Reasoning with Adaptive Reflection and Length Coordinated Penalty 2026-02-13 · ARXIV
  14. P-GenRM: Personalized Generative Reward Model with Test-time User-based Scaling 2026-02-13 · ARXIV
  15. On the Complexity of Offline Reinforcement Learning with $Q^\star$-Approximation and Partial Coverage 2026-02-13 · ARXIV
  16. Meta-Sel: Efficient Demonstration Selection for In-Context Learning via Supervised Meta-Learning 2026-02-13 · ARXIV
  17. Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation 2026-02-13 · ARXIV
  18. KAN-FIF: Spline-Parameterized Lightweight Physics-based Tropical Cyclone Estimation on Meteorological Satellite 2026-02-13 · ARXIV
  19. Few-Shot Design Optimization by Exploiting Auxiliary Information 2026-02-13 · ARXIV
  20. Capability-Oriented Training Induced Alignment Risk 2026-02-13 · ARXIV
  21. Weight Decay Improves Language Model Plasticity 2026-02-12 · ARXIV
  22. TabICLv2: A better, faster, scalable, and open tabular foundation model 2026-02-12 · ARXIV
  23. SCRAPL: Scattering Transform with Random Paths for Machine Learning 2026-02-12 · ARXIV
  24. GENIUS: Generative Fluid Intelligence Evaluation Suite 2026-02-12 · ARXIV
  25. Diffusion-Pretrained Dense and Contextual Embeddings 2026-02-12 · ARXIV
  26. Data-Efficient Hierarchical Goal-Conditioned Reinforcement Learning via Normalizing Flows 2026-02-12 · ARXIV
  27. Data Repetition Beats Data Scaling in Long-CoT Supervised Fine-Tuning 2026-02-12 · ARXIV
  28. Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling 2026-02-12 · ARXIV
  29. ROCKET: Rapid Optimization via Calibration-guided Knapsack Enhanced Truncation for Efficient Model Compression 2026-02-12 · ARXIV
  30. LoRA-Squeeze: Simple and Effective Post-Tuning and In-Tuning Compression of LoRA Modules 2026-02-12 · ARXIV
  31. From Buffers to Registers: Unlocking Fine-Grained FlashAttention with Hybrid-Bonded 3D NPU Co-Design 2026-02-12 · ARXIV
  32. Enhancing Predictability of Multi-Tenant DNN Inference for Autonomous Vehicles' Perception 2026-02-12 · ARXIV
  33. CLI-Gym: Scalable CLI Task Generation via Agentic Environment Inversion 2026-02-12 · ARXIV
  34. Towards Explainable Federated Learning: Understanding the Impact of Differential Privacy 2026-02-11 · ARXIV
  35. Quantum-Audit: Evaluating the Reasoning Limits of LLMs on Quantum Computing 2026-02-11 · ARXIV
  36. Learning on the Manifold: Unlocking Standard Diffusion Transformers with Representation Encoders 2026-02-11 · ARXIV
  37. CODE-SHARP: Continuous Open-ended Discovery and Evolution of Skills as Hierarchical Reward Programs 2026-02-11 · ARXIV
  38. Causality in Video Diffusers is Separable from Denoising 2026-02-11 · ARXIV
  39. Biases in the Blind Spot: Detecting What LLMs Fail to Mention 2026-02-11 · ARXIV
  40. Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning 2026-02-11 · ARXIV
  41. Step-resolved data attribution for looped transformers 2026-02-11 · ARXIV
  42. Olaf-World: Orienting Latent Actions for Video World Modeling 2026-02-11 · ARXIV
  43. Anagent For Enhancing Scientific Table & Figure Analysis 2026-02-11 · ARXIV
  44. The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies 2026-02-11 · ARXIV
  45. Statistical benchmarking of transformer models in low signal-to-noise time-series forecasting 2026-02-11 · ARXIV
  46. Hybrid Responsible AI-Stochastic Approach for SLA Compliance in Multivendor 6G Networks 2026-02-11 · ARXIV
  47. CoFEH: LLM-driven Feature Engineering Empowered by Collaborative Bayesian Hyperparameter Optimization 2026-02-11 · ARXIV
  48. Code2World: A GUI World Model via Renderable Code Generation 2026-02-11 · ARXIV
  49. Robustness Is a Function, Not a Number: A Factorized Comprehensive Study of OOD Robustness in Vision-Based Driving 2026-02-10 · ARXIV
  50. Data Science and Technology Towards AGI Part I: Tiered Data Management 2026-02-10 · ARXIV