03 / ARTICLE LEDGER

关联文章

按发布时间倒序

  1. Beyond Trial-and-Error: Agentic Optimization for Image-to-Video Adherence 2026-08-13 · ARXIV
  2. Class Activation Mapping in Explainable Computer Vision: A Method-Centered Review of CNN, Transformer, and Foundation-Model-Era Visual Explanations 2026-08-13 · ARXIV
  3. Constructing Dynamic Master Logic Models as Knowledge Graphs for Complex System Diagnostics Using Retrieval-Augmented Large Language Models 2026-08-13 · ARXIV
  4. DreamFly: Causal Memory and Receding-Horizon Diffusion Planning for Aerial Vision-Language Navigation 2026-08-13 · ARXIV
  5. AVA-Encoder: Towards Agent-Native Video Representation Learning 2026-08-13 · ARXIV
  6. Two-stage Odd Residual Flows for Mean-Preserving Probabilistic Time Series Forecasting 2026-08-13 · ARXIV
  7. AlbumentationsX: One Augmentation Pipeline for Images and Related Annotations 2026-08-13 · ARXIV
  8. sLTN: Structural Logic Tensor Networks 2026-08-13 · ARXIV
  9. The Illusion of Cross-Lingual Safety in Low-Resource Languages 2026-08-13 · ARXIV
  10. Scheduling Mixed RL Rollouts Beyond Prefix Locality 2026-08-13 · ARXIV
  11. MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment 2026-08-13 · ARXIV
  12. A Quantum Roadmap for Softmax Attention: Exact Born-Rule Analogs for Softmax Attention on the Probability Simplex 2026-08-13 · ARXIV
  13. How to Verify Consistency of Probabilistic Claims 2026-08-12 · ARXIV
  14. Test-Time Self-Evolving GUI Visual Grounding via Reflection-Guided On-Policy Self-Distillation 2026-08-12 · ARXIV
  15. Long-Horizon AI Research for Grothendieck Constant: A Case Study in Human-AI Mathematical Collaboration 2026-08-12 · ARXIV
  16. Beyond a Bag of Features: Set-Level Instability in Sparse Autoencoders 2026-08-12 · ARXIV
  17. ConVAWG: A Retrieval-Grounded Framework for Controlled Synthetic Dialogue Generation in Violence Against Women and Girls 2026-08-12 · ARXIV
  18. Stealing Reasoning Traces from Proprietary LLM APIs 2026-08-12 · ARXIV
  19. Sci-VBench: Evaluating Knowledge- and Reasoning-Intensive Video Generation in Science Domains 2026-08-12 · ARXIV
  20. ArchAgent v2: A Case Study with the Data Prefetching Championship 2026-08-12 · ARXIV
  21. Energy-Structured Latent World Models with Neural Time Fields for Physically Constistent Open-World Motion Planning 2026-08-12 · ARXIV
  22. Financial Numerical Prediction and Allocation as Token Generation 2026-08-12 · ARXIV
  23. SHE: Trajectory-driven Safety Harness Evolution for LLM Agents 2026-08-12 · ARXIV
  24. Space-Creating versus Dead Possession: An Off-Ball Possession-Quality Index for Broadcast Football 2026-08-12 · ARXIV
  25. Consilience for Verifier-Free Test-Time Scaling 2026-08-11 · ARXIV
  26. Decoding-Level Taboo: A Diagnostic Stress Test for LLM Robustness 2026-08-11 · ARXIV
  27. DSLE: A Learning Environment for Dark Souls Boss Encounters 2026-08-11 · ARXIV
  28. GENCO - A Unified Neural Solver Embedded in a Development Framework for Steady-State Grid Analysis 2026-08-11 · ARXIV
  29. From Values to Benchmarks: Evaluating Large Language Models for Governmental Use in Dutch 2026-08-11 · ARXIV
  30. Multimodal Model Diffing for Feature Discovery and Control 2026-08-11 · ARXIV
  31. Beyond Naturalness: Probing Automated Text-To-Speech Evaluators on Linguistically Grounded Dimensions 2026-08-11 · ARXIV
  32. KGCaRe: Explainable Complex Conditional Question Answering using Automatic Knowledge Graph Construction and Context Retrieval with LLMs 2026-08-11 · ARXIV
  33. A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs While Improving Accuracy 2026-08-11 · ARXIV
  34. TEPA: Revoking Stale Memories for Conflict-Robust Language Agents 2026-08-11 · ARXIV
  35. Diffusion LLMs as Targets and Adversaries: Mechanistic Safety Exploits 2026-08-11 · ARXIV
  36. SABRE: Scalable and Automated Benchmarking of VLMs under Stress 2026-08-11 · ARXIV
  37. Post-Grokking Collapse at the Representation-Readout Interface in Muon-Trained Transformers 2026-08-11 · ARXIV
  38. Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing 2026-08-11 · ARXIV
  39. PsychoAgent: An Affect-Sensitive Cognitive Architecture for Conflict-Aware Memory in LLM Agents 2026-08-11 · ARXIV
  40. An Exploratory Evaluation of LLM-Assisted Rewriting of Moderate-Complexity Financial Sentences for DisCoCat-Based Sentiment Analysis 2026-08-10 · ARXIV
  41. Blast Radius 2026-08-10 · ARXIV
  42. Taxonomy-Driven Analysis of Open-Source AI Risk Mitigation Tools 2026-08-10 · ARXIV
  43. SkillProx: Self-Evolving Agent Skills via Proximal Textual Gradient Descent 2026-08-10 · ARXIV
  44. Interaction Creates Dynamical AI Behavior Absent in Isolation 2026-08-10 · ARXIV
  45. CoinRAG: Contextualized Information Nugget KV Cache Reuse for Long-Context RAG 2026-08-10 · ARXIV
  46. CreativeInstruct: Scalably Teaching LLMs to Balance Quality, Creativity, and Diversity 2026-08-10 · ARXIV
  47. MirrorWorld: Taming Video Diffusion Models for Mirror Reflection Generation 2026-08-10 · ARXIV
  48. Beyond Top-K: Replacing Black-Box Retrieval with Interpretable Agentic Operations 2026-08-08 · ARXIV
  49. RRC: Unlocking Generative Reward Models in LLM Reinforcement Learning via Ranking-Based Reward Construction 2026-08-08 · ARXIV
  50. Does FLAIR super-resolution erase or hallucinate small white-matter lesions? 2026-08-08 · ARXIV