03 / ARTICLE LEDGER
关联文章
按发布时间倒序
- Colored Petri Nets, LLMs, and distributed applications 2026-02-14 · HACKER_NEWS
- 架构图不再手画:用 LikeC4 + AI,让架构“活”起来 2026-02-14 · JUEJIN
- CM2: Reinforcement Learning with Checklist Rewards for Multi-Turn and Multi-Step Agentic Tool Use 2026-02-13 · ARXIV
- AttentionRetriever: Attention Layers are Secretly Long Document Retrievers 2026-02-13 · ARXIV
- Agentic Test-Time Scaling for WebAgents 2026-02-13 · ARXIV
- Evaluating Multilingual, Context-Aware Guardrails: A Humanitarian LLM Use Case 2026-02-13 · HACKER_NEWS
- The Pensieve Paradigm: Stateful Language Models Mastering Their Own Context 2026-02-13 · ARXIV
- P-GenRM: Personalized Generative Reward Model with Test-time User-based Scaling 2026-02-13 · ARXIV
- Meta-Sel: Efficient Demonstration Selection for In-Context Learning via Supervised Meta-Learning 2026-02-13 · ARXIV
- Weight Decay Improves Language Model Plasticity 2026-02-12 · ARXIV
- Data Repetition Beats Data Scaling in Long-CoT Supervised Fine-Tuning 2026-02-12 · ARXIV
- Improving 15 LLMs at Coding in One Afternoon. Only the Harness Changed 2026-02-12 · HACKER_NEWS
- 万字长文深入解析Skill/MCP/RAG/Agent/OpenClaw底层逻辑 2026-02-12 · JUEJIN
- The Problem with LLMs 2026-02-12 · HACKER_NEWS
- Show HN: CodeRLM – Tree-sitter-backed code indexing for LLM agents 2026-02-12 · HACKER_NEWS
- Quantum-Audit: Evaluating the Reasoning Limits of LLMs on Quantum Computing 2026-02-11 · ARXIV
- Biases in the Blind Spot: Detecting What LLMs Fail to Mention 2026-02-11 · ARXIV
- Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning 2026-02-11 · ARXIV
- The Death of Traditional Testing: Agentic Development Broke a 50-Year-Old Field, JiTTesting Can Revive It 2026-02-11 · BLOGS_PODCASTS
- How LinqAlpha assesses investment theses using Devil’s Advocate on Amazon Bedrock | Amazon Web Services 2026-02-11 · BLOGS_PODCASTS
- AI 名词解释 2026-02-11 · JUEJIN
- AI Design-to-Code 的两个根本问题,和我的解法 2026-02-11 · JUEJIN
- The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies 2026-02-11 · ARXIV
- CoFEH: LLM-driven Feature Engineering Empowered by Collaborative Bayesian Hyperparameter Optimization 2026-02-11 · ARXIV
- Scale LLM fine-tuning with Hugging Face and Amazon SageMaker AI | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
- Iberdrola enhances IT operations using Amazon Bedrock AgentCore | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
- Data Science and Technology Towards AGI Part I: Tiered Data Management 2026-02-10 · ARXIV
- Building real-time voice assistants with Amazon Nova Sonic compared to cascading architectures | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
- Automated Reasoning checks rewriting chatbot reference implementation | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
- ARO: A New Lens On Matrix Optimization For Large Models 2026-02-10 · ARXIV
- LangChain 进阶实战:当 Memory 遇上 OutputParser,打造有记忆的结构化助手 2026-02-10 · JUEJIN
- ANCRe: Adaptive Neural Connection Reassignment for Efficient Depth Scaling 2026-02-10 · ARXIV
- Rethinking Latency Denial-of-Service: Attacking the LLM Serving Framework, Not the Model 2026-02-10 · ARXIV
- MARTI-MARS$^2$: Scaling Multi-Agent Self-Search via Reinforcement Learning for Code Generation 2026-02-10 · ARXIV
- Emergent Misalignment is Easy, Narrow Misalignment is Hard 2026-02-10 · ARXIV
- Learning a Generative Meta-Model of LLM Activations 2026-02-09 · ARXIV
- Endogenous Resistance to Activation Steering in Language Models 2026-02-09 · ARXIV
- DAWN: Dependency-Aware Fast Inference for Diffusion LLMs 2026-02-09 · ARXIV
- LLMs as the new high level language 2026-02-08 · HACKER_NEWS
- Experts Have World Models. LLMs Have Word Models 2026-02-07 · HACKER_NEWS
- Evaluating and mitigating the growing risk of LLM-discovered 0-days 2026-02-07 · HACKER_NEWS
- PhysicsAgentABM: Physics-Guided Generative Agent-Based Modeling 2026-02-06 · ARXIV
- Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory 2026-02-06 · ARXIV
- DyTopo: Dynamic Topology Routing for Multi-Agent Reasoning via Semantic Matching 2026-02-06 · ARXIV
- DFlash: Block Diffusion for Flash Speculative Decoding 2026-02-06 · ARXIV
- CommCP: Efficient Multi-Agent Coordination via LLM-Based Communication with Conformal Prediction 2026-02-06 · ARXIV
- Show HN: BioTradingArena – Benchmark for LLMs to predict biotech stock movements 2026-02-06 · HACKER_NEWS
- LLMs could be, but shouldn't be compilers 2026-02-06 · HACKER_NEWS
- RRAttention: Dynamic Block Sparse Attention via Per-Head Round-Robin Shifts for Long-Context Inference 2026-02-06 · ARXIV
- DLM-Scope: Mechanistic Interpretability of Diffusion Language Models via Sparse Autoencoders 2026-02-06 · ARXIV