03 / ARTICLE LEDGER
关联文章
按发布时间倒序
- Iberdrola enhances IT operations using Amazon Bedrock AgentCore | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
- Data Science and Technology Towards AGI Part I: Tiered Data Management 2026-02-10 · ARXIV
- Building real-time voice assistants with Amazon Nova Sonic compared to cascading architectures | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
- Automated Reasoning checks rewriting chatbot reference implementation | Amazon Web Services 2026-02-10 · BLOGS_PODCASTS
- ARO: A New Lens On Matrix Optimization For Large Models 2026-02-10 · ARXIV
- LangChain 进阶实战:当 Memory 遇上 OutputParser,打造有记忆的结构化助手 2026-02-10 · JUEJIN
- ANCRe: Adaptive Neural Connection Reassignment for Efficient Depth Scaling 2026-02-10 · ARXIV
- Rethinking Latency Denial-of-Service: Attacking the LLM Serving Framework, Not the Model 2026-02-10 · ARXIV
- MARTI-MARS$^2$: Scaling Multi-Agent Self-Search via Reinforcement Learning for Code Generation 2026-02-10 · ARXIV
- Emergent Misalignment is Easy, Narrow Misalignment is Hard 2026-02-10 · ARXIV
- Learning a Generative Meta-Model of LLM Activations 2026-02-09 · ARXIV
- Endogenous Resistance to Activation Steering in Language Models 2026-02-09 · ARXIV
- DAWN: Dependency-Aware Fast Inference for Diffusion LLMs 2026-02-09 · ARXIV
- LLMs as the new high level language 2026-02-08 · HACKER_NEWS
- Experts Have World Models. LLMs Have Word Models 2026-02-07 · HACKER_NEWS
- Evaluating and mitigating the growing risk of LLM-discovered 0-days 2026-02-07 · HACKER_NEWS
- PhysicsAgentABM: Physics-Guided Generative Agent-Based Modeling 2026-02-06 · ARXIV
- Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory 2026-02-06 · ARXIV
- DyTopo: Dynamic Topology Routing for Multi-Agent Reasoning via Semantic Matching 2026-02-06 · ARXIV
- DFlash: Block Diffusion for Flash Speculative Decoding 2026-02-06 · ARXIV
- CommCP: Efficient Multi-Agent Coordination via LLM-Based Communication with Conformal Prediction 2026-02-06 · ARXIV
- Show HN: BioTradingArena – Benchmark for LLMs to predict biotech stock movements 2026-02-06 · HACKER_NEWS
- LLMs could be, but shouldn't be compilers 2026-02-06 · HACKER_NEWS
- RRAttention: Dynamic Block Sparse Attention via Per-Head Round-Robin Shifts for Long-Context Inference 2026-02-06 · ARXIV
- DLM-Scope: Mechanistic Interpretability of Diffusion Language Models via Sparse Autoencoders 2026-02-06 · ARXIV
- Subliminal Effects in Your Data: A General Mechanism via Log-Linearity 2026-02-05 · ARXIV
- Rethinking the Trust Region in LLM Reinforcement Learning 2026-02-05 · ARXIV
- Reinforced Attention Learning 2026-02-05 · ARXIV
- Multi-Head LatentMoE and Head Parallel: Communication-Efficient and Deterministic MoE Parallelism 2026-02-05 · ARXIV
- CoT is Not the Chain of Truth: An Empirical Internal Analysis of Reasoning LLMs for Fake News Generation 2026-02-05 · ARXIV
- AstrBotDevs/AstrBot 2026-02-05 · GITHUB_TRENDING
- Understanding and Exploiting Weight Update Sparsity for Communication-Efficient Distributed RL 2026-02-04 · ARXIV
- Accelerating Scientific Research with Gemini: Case Studies and Common Techniques 2026-02-04 · ARXIV
- Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability 2026-02-03 · ARXIV
- RLAnything: Forge Environment, Policy, and Reward Model in Completely Dynamic RL System 2026-02-03 · ARXIV
- Reward-free Alignment for Conflicting Objectives 2026-02-03 · ARXIV
- RE-TRAC: REcursive TRAjectory Compression for Deep Search Agents 2026-02-03 · ARXIV
- MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents 2026-02-03 · ARXIV
- Expanding the Capabilities of Reinforcement Learning via Text Feedback 2026-02-03 · ARXIV
- AgentRx: Diagnosing AI Agent Failures from Execution Trajectories 2026-02-03 · ARXIV
- Show HN: I built "AI Wattpad" to eval LLMs on fiction 2026-02-03 · HACKER_NEWS
- Sparse Reward Subsystem in Large Language Models 2026-02-03 · ARXIV
- Reliable Use of Lemmas via Eligibility Reasoning and Section$-$Aware Reinforcement Learning 2026-02-03 · ARXIV
- Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning 2026-02-03 · ARXIV
- How RLHF Amplifies Sycophancy 2026-02-03 · ARXIV
- Error Taxonomy-Guided Prompt Optimization 2026-02-03 · ARXIV
- UPA: Unsupervised Prompt Agent via Tree-Based Search and Selection 2026-02-02 · ARXIV
- TEON: Tensorized Orthonormalization Beyond Layer-Wise Muon for Large Language Model Pre-Training 2026-02-02 · ARXIV
- FOCUS: DLLMs Know How to Tame Their Compute Bound 2026-02-02 · ARXIV
- My iPhone 16 Pro Max produces garbage output when running MLX LLMs 2026-02-02 · HACKER_NEWS