03 / ARTICLE LEDGER

关联文章

按发布时间倒序

  1. How Transparent is DiffusionGemma? 2026-06-19 · ARXIV
  2. Native Active Perception as Reasoning for Omni-Modal Understanding 2026-06-18 · ARXIV
  3. Visual Verification Enables Inference-time Steering and Autonomous Policy Improvement 2026-06-17 · ARXIV
  4. The Value Axis: Language Models Encode Whether They're on the Right Track 2026-06-16 · ARXIV
  5. Gaze Heads: How VLMs Look at What They Describe 2026-06-15 · ARXIV
  6. EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments 2026-06-12 · ARXIV
  7. Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models 2026-06-11 · ARXIV
  8. When to Align, When to Predict: A Phase Diagram for Multimodal Learning 2026-06-10 · ARXIV
  9. OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics 2026-06-09 · ARXIV
  10. How reliable are LLMs when it comes to playing dice? 2026-06-08 · ARXIV
  11. TailLoR: Protecting Principal Components in Parameter-Efficient Continual Learning 2026-06-05 · ARXIV
  12. STRIDE: Training Data Attribution via Sparse Recovery from Subset Perturbations 2026-06-04 · ARXIV
  13. Neuron Populations Exhibit Divergent Selectivity with Scale 2026-06-03 · ARXIV
  14. Mitigating Perceptual Judgment Bias in Multimodal LLM-as-a-Judge via Perceptual Perturbation and Reward Modeling 2026-06-02 · ARXIV
  15. Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models 2026-06-01 · ARXIV
  16. Physics Is All You Need? A Case Study in Physicist-Supervised AI Development of Scientific Software 2026-05-29 · ARXIV
  17. PEFT-Arena: Understanding Parameter-Efficient Finetuning from a Stability-Plasticity Perspective 2026-05-28 · ARXIV
  18. Algorithmic Monocultures in Hiring 2026-05-27 · ARXIV
  19. MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research 2026-05-26 · ARXIV
  20. SkillOpt: Executive Strategy for Self-Evolving Agent Skills 2026-05-25 · ARXIV
  21. Tokenisation via Convex Relaxations 2026-05-22 · ARXIV
  22. Variance Reduction for Expectations with Diffusion Teachers 2026-05-21 · ARXIV
  23. Atoms of Thought: Universal EEG Representation Learning with Microstates 2026-05-20 · ARXIV
  24. DashAttention: Differentiable and Adaptive Sparse Hierarchical Attention 2026-05-19 · ARXIV
  25. IVGT: Implicit Visual Geometry Transformer for Neural Scene Representation 2026-05-18 · ARXIV
  26. EntityBench: Towards Entity-Consistent Long-Range Multi-Shot Video Generation 2026-05-15 · ARXIV
  27. WARDEN: Endangered Indigenous Language Transcription and Translation with 6 Hours of Training Data 2026-05-14 · ARXIV
  28. AlphaGRPO: Unlocking Self-Reflective Multimodal Generation in UMMs via Decompositional Verifiable Reward 2026-05-13 · ARXIV
  29. ELF: Embedded Language Flows 2026-05-12 · ARXIV
  30. LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling 2026-05-11 · ARXIV
  31. ActCam: Zero-Shot Joint Camera and 3D Motion Control for Video Generation 2026-05-08 · ARXIV
  32. Taming Outlier Tokens in Diffusion Transformers 2026-05-07 · ARXIV
  33. A Closed-Form Adaptive-Landmark Kernel for Certified Point-Cloud and Graph Classification 2026-05-06 · ARXIV
  34. SpecKV: Adaptive Speculative Decoding with Compression-Aware Gamma Selection 2026-05-05 · ARXIV
  35. Enhancing RL Generalizability in Robotics through SHAP Analysis of Algorithms and Hyperparameters 2026-05-05 · ARXIV
  36. HyCOP: Hybrid Composition Operators for Interpretable Learning of PDEs 2026-05-04 · ARXIV
  37. Turning the TIDE: Cross-Architecture Distillation for Diffusion Large Language Models 2026-04-30 · ARXIV
  38. Recursive Multi-Agent Systems 2026-04-29 · ARXIV
  39. Personalized Worked Example Generation from Student Code Submissions using Pattern-based Knowledge Components 2026-04-28 · ARXIV
  40. The Last Human-Written Paper: Agent-Native Research Artifacts 2026-04-28 · ARXIV
  41. Spend Less, Fit Better: Budget-Efficient Scaling Law Fitting via Active Experiment Selection 2026-04-27 · ARXIV
  42. Seeing Fast and Slow: Learning the Flow of Time in Videos 2026-04-24 · ARXIV
  43. SpeechParaling-Bench: A Comprehensive Benchmark for Paralinguistic-Aware Speech Generation 2026-04-23 · ARXIV
  44. Generalization at the Edge of Stability 2026-04-22 · ARXIV
  45. MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval 2026-04-21 · ARXIV
  46. Towards Real-Time ECG and EMG Modeling on $μ$ NPUs 2026-04-21 · ARXIV
  47. ASMR-Bench: Auditing for Sabotage in ML Research 2026-04-20 · ARXIV
  48. MM-WebAgent: A Hierarchical Multimodal Web Agent for Webpage Generation 2026-04-17 · ARXIV
  49. SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments 2026-04-16 · ARXIV
  50. SceneCritic: A Symbolic Evaluator for 3D Indoor Scene Synthesis 2026-04-15 · ARXIV