03 / ARTICLE LEDGER

关联文章

按发布时间倒序

  1. Iberdrola enhances IT operations using Amazon Bedrock AgentCore | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
  2. Data Science and Technology Towards AGI Part I: Tiered Data Management 2026-02-10 · ARXIV
  3. Building real-time voice assistants with Amazon Nova Sonic compared to cascading architectures | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
  4. Automated Reasoning checks rewriting chatbot reference implementation | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
  5. ARO: A New Lens On Matrix Optimization For Large Models 2026-02-10 · ARXIV
  6. LangChain 进阶实战:当 Memory 遇上 OutputParser,打造有记忆的结构化助手 2026-02-10 · JUEJIN
  7. ANCRe: Adaptive Neural Connection Reassignment for Efficient Depth Scaling 2026-02-10 · ARXIV
  8. Rethinking Latency Denial-of-Service: Attacking the LLM Serving Framework, Not the Model 2026-02-10 · ARXIV
  9. MARTI-MARS$^2$: Scaling Multi-Agent Self-Search via Reinforcement Learning for Code Generation 2026-02-10 · ARXIV
  10. Emergent Misalignment is Easy, Narrow Misalignment is Hard 2026-02-10 · ARXIV
  11. Learning a Generative Meta-Model of LLM Activations 2026-02-09 · ARXIV
  12. Endogenous Resistance to Activation Steering in Language Models 2026-02-09 · ARXIV
  13. DAWN: Dependency-Aware Fast Inference for Diffusion LLMs 2026-02-09 · ARXIV
  14. LLMs as the new high level language 2026-02-08 · HACKER_NEWS
  15. Experts Have World Models. LLMs Have Word Models 2026-02-07 · HACKER_NEWS
  16. Evaluating and mitigating the growing risk of LLM-discovered 0-days 2026-02-07 · HACKER_NEWS
  17. PhysicsAgentABM: Physics-Guided Generative Agent-Based Modeling 2026-02-06 · ARXIV
  18. Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory 2026-02-06 · ARXIV
  19. DyTopo: Dynamic Topology Routing for Multi-Agent Reasoning via Semantic Matching 2026-02-06 · ARXIV
  20. DFlash: Block Diffusion for Flash Speculative Decoding 2026-02-06 · ARXIV
  21. CommCP: Efficient Multi-Agent Coordination via LLM-Based Communication with Conformal Prediction 2026-02-06 · ARXIV
  22. Show HN: BioTradingArena – Benchmark for LLMs to predict biotech stock movements 2026-02-06 · HACKER_NEWS
  23. LLMs could be, but shouldn't be compilers 2026-02-06 · HACKER_NEWS
  24. RRAttention: Dynamic Block Sparse Attention via Per-Head Round-Robin Shifts for Long-Context Inference 2026-02-06 · ARXIV
  25. DLM-Scope: Mechanistic Interpretability of Diffusion Language Models via Sparse Autoencoders 2026-02-06 · ARXIV
  26. Subliminal Effects in Your Data: A General Mechanism via Log-Linearity 2026-02-05 · ARXIV
  27. Rethinking the Trust Region in LLM Reinforcement Learning 2026-02-05 · ARXIV
  28. Reinforced Attention Learning 2026-02-05 · ARXIV
  29. Multi-Head LatentMoE and Head Parallel: Communication-Efficient and Deterministic MoE Parallelism 2026-02-05 · ARXIV
  30. CoT is Not the Chain of Truth: An Empirical Internal Analysis of Reasoning LLMs for Fake News Generation 2026-02-05 · ARXIV
  31. AstrBotDevs/AstrBot 2026-02-05 · GITHUB_TRENDING
  32. Understanding and Exploiting Weight Update Sparsity for Communication-Efficient Distributed RL 2026-02-04 · ARXIV
  33. Accelerating Scientific Research with Gemini: Case Studies and Common Techniques 2026-02-04 · ARXIV
  34. Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability 2026-02-03 · ARXIV
  35. RLAnything: Forge Environment, Policy, and Reward Model in Completely Dynamic RL System 2026-02-03 · ARXIV
  36. Reward-free Alignment for Conflicting Objectives 2026-02-03 · ARXIV
  37. RE-TRAC: REcursive TRAjectory Compression for Deep Search Agents 2026-02-03 · ARXIV
  38. MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents 2026-02-03 · ARXIV
  39. Expanding the Capabilities of Reinforcement Learning via Text Feedback 2026-02-03 · ARXIV
  40. AgentRx: Diagnosing AI Agent Failures from Execution Trajectories 2026-02-03 · ARXIV
  41. Show HN: I built "AI Wattpad" to eval LLMs on fiction 2026-02-03 · HACKER_NEWS
  42. Sparse Reward Subsystem in Large Language Models 2026-02-03 · ARXIV
  43. Reliable Use of Lemmas via Eligibility Reasoning and Section$-$Aware Reinforcement Learning 2026-02-03 · ARXIV
  44. Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning 2026-02-03 · ARXIV
  45. How RLHF Amplifies Sycophancy 2026-02-03 · ARXIV
  46. Error Taxonomy-Guided Prompt Optimization 2026-02-03 · ARXIV
  47. UPA: Unsupervised Prompt Agent via Tree-Based Search and Selection 2026-02-02 · ARXIV
  48. TEON: Tensorized Orthonormalization Beyond Layer-Wise Muon for Large Language Model Pre-Training 2026-02-02 · ARXIV
  49. FOCUS: DLLMs Know How to Tame Their Compute Bound 2026-02-02 · ARXIV
  50. My iPhone 16 Pro Max produces garbage output when running MLX LLMs 2026-02-02 · HACKER_NEWS