03 / ARTICLE LEDGER

关联文章

按发布时间倒序

  1. Colored Petri Nets, LLMs, and distributed applications 2026-02-14 · HACKER_NEWS
  2. 架构图不再手画:用 LikeC4 + AI,让架构“活”起来 2026-02-14 · JUEJIN
  3. CM2: Reinforcement Learning with Checklist Rewards for Multi-Turn and Multi-Step Agentic Tool Use 2026-02-13 · ARXIV
  4. AttentionRetriever: Attention Layers are Secretly Long Document Retrievers 2026-02-13 · ARXIV
  5. Agentic Test-Time Scaling for WebAgents 2026-02-13 · ARXIV
  6. Evaluating Multilingual, Context-Aware Guardrails: A Humanitarian LLM Use Case 2026-02-13 · HACKER_NEWS
  7. The Pensieve Paradigm: Stateful Language Models Mastering Their Own Context 2026-02-13 · ARXIV
  8. P-GenRM: Personalized Generative Reward Model with Test-time User-based Scaling 2026-02-13 · ARXIV
  9. Meta-Sel: Efficient Demonstration Selection for In-Context Learning via Supervised Meta-Learning 2026-02-13 · ARXIV
  10. Weight Decay Improves Language Model Plasticity 2026-02-12 · ARXIV
  11. Data Repetition Beats Data Scaling in Long-CoT Supervised Fine-Tuning 2026-02-12 · ARXIV
  12. Improving 15 LLMs at Coding in One Afternoon. Only the Harness Changed 2026-02-12 · HACKER_NEWS
  13. 万字长文深入解析Skill/MCP/RAG/Agent/OpenClaw底层逻辑 2026-02-12 · JUEJIN
  14. The Problem with LLMs 2026-02-12 · HACKER_NEWS
  15. Show HN: CodeRLM – Tree-sitter-backed code indexing for LLM agents 2026-02-12 · HACKER_NEWS
  16. Quantum-Audit: Evaluating the Reasoning Limits of LLMs on Quantum Computing 2026-02-11 · ARXIV
  17. Biases in the Blind Spot: Detecting What LLMs Fail to Mention 2026-02-11 · ARXIV
  18. Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning 2026-02-11 · ARXIV
  19. The Death of Traditional Testing: Agentic Development Broke a 50-Year-Old Field, JiTTesting Can Revive It 2026-02-11 · BLOGS_PODCASTS
  20. How LinqAlpha assesses investment theses using Devil’s Advocate on Amazon Bedrock | Amazon Web Services 2026-02-11 · BLOGS_PODCASTS
  21. AI 名词解释 2026-02-11 · JUEJIN
  22. AI Design-to-Code 的两个根本问题,和我的解法 2026-02-11 · JUEJIN
  23. The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies 2026-02-11 · ARXIV
  24. CoFEH: LLM-driven Feature Engineering Empowered by Collaborative Bayesian Hyperparameter Optimization 2026-02-11 · ARXIV
  25. Scale LLM fine-tuning with Hugging Face and Amazon SageMaker AI | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
  26. Iberdrola enhances IT operations using Amazon Bedrock AgentCore | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
  27. Data Science and Technology Towards AGI Part I: Tiered Data Management 2026-02-10 · ARXIV
  28. Building real-time voice assistants with Amazon Nova Sonic compared to cascading architectures | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
  29. Automated Reasoning checks rewriting chatbot reference implementation | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
  30. ARO: A New Lens On Matrix Optimization For Large Models 2026-02-10 · ARXIV
  31. LangChain 进阶实战:当 Memory 遇上 OutputParser,打造有记忆的结构化助手 2026-02-10 · JUEJIN
  32. ANCRe: Adaptive Neural Connection Reassignment for Efficient Depth Scaling 2026-02-10 · ARXIV
  33. Rethinking Latency Denial-of-Service: Attacking the LLM Serving Framework, Not the Model 2026-02-10 · ARXIV
  34. MARTI-MARS$^2$: Scaling Multi-Agent Self-Search via Reinforcement Learning for Code Generation 2026-02-10 · ARXIV
  35. Emergent Misalignment is Easy, Narrow Misalignment is Hard 2026-02-10 · ARXIV
  36. Learning a Generative Meta-Model of LLM Activations 2026-02-09 · ARXIV
  37. Endogenous Resistance to Activation Steering in Language Models 2026-02-09 · ARXIV
  38. DAWN: Dependency-Aware Fast Inference for Diffusion LLMs 2026-02-09 · ARXIV
  39. LLMs as the new high level language 2026-02-08 · HACKER_NEWS
  40. Experts Have World Models. LLMs Have Word Models 2026-02-07 · HACKER_NEWS
  41. Evaluating and mitigating the growing risk of LLM-discovered 0-days 2026-02-07 · HACKER_NEWS
  42. PhysicsAgentABM: Physics-Guided Generative Agent-Based Modeling 2026-02-06 · ARXIV
  43. Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory 2026-02-06 · ARXIV
  44. DyTopo: Dynamic Topology Routing for Multi-Agent Reasoning via Semantic Matching 2026-02-06 · ARXIV
  45. DFlash: Block Diffusion for Flash Speculative Decoding 2026-02-06 · ARXIV
  46. CommCP: Efficient Multi-Agent Coordination via LLM-Based Communication with Conformal Prediction 2026-02-06 · ARXIV
  47. Show HN: BioTradingArena – Benchmark for LLMs to predict biotech stock movements 2026-02-06 · HACKER_NEWS
  48. LLMs could be, but shouldn't be compilers 2026-02-06 · HACKER_NEWS
  49. RRAttention: Dynamic Block Sparse Attention via Per-Head Round-Robin Shifts for Long-Context Inference 2026-02-06 · ARXIV
  50. DLM-Scope: Mechanistic Interpretability of Diffusion Language Models via Sparse Autoencoders 2026-02-06 · ARXIV