03 / ARTICLE LEDGER
关联文章
按发布时间倒序
- AV-AIVAT: 74x Cheaper Agent Evaluation with Certified Anytime-Valid Stopping in Imperfect-Information Games 2026-08-07 · ARXIV
- Tracing the Heart: An Evidence-Linked Pipeline for Heart-Failure Feature Engineering 2026-08-07 · ARXIV
- [AINews] AMD buys Taalas 2026-08-07 · BLOGS_PODCASTS
- The Bitter Lesson of Tool Calling 2026-08-07 · ARXIV
- Control agent behaviors and cost beyond a single action: new capabilities in Amazon Bedrock AgentCore 2026-08-07 · BLOGS_PODCASTS
- Configure rate limits for AI traffic on AgentCore gateway 2026-08-07 · BLOGS_PODCASTS
- Securing AI agents with temporal policies in Amazon Bedrock AgentCore 2026-08-07 · BLOGS_PODCASTS
- Inside vLLM: Anatomy of a High-Throughput LLM Inference System (2025) 2026-08-07 · HACKER_NEWS
- Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning 2026-08-06 · ARXIV
- Born Against, or why hobby programming communities are against LLM usage 2026-08-06 · HACKER_NEWS
- [AINews] Jeff, Sanjay, Oriol, and Quoc depart DeepMind; Demis to Chair; Koray to SVP — what is going on at GDM??? 2026-08-06 · BLOGS_PODCASTS
- Run production AI agents in n8n with Amazon Bedrock AgentCore harness 2026-08-06 · BLOGS_PODCASTS
- LLMs won't break symmetric crypto 2026-08-06 · HACKER_NEWS
- Argus: A General-Purpose Agentic Runtime for Long-Horizon Reasoning 2026-08-06 · ARXIV
- How we built an MCP bridge to give our AgentCore-hosted AI agent access to local MCP tools 2026-08-06 · BLOGS_PODCASTS
- How Mobileye transformed support operations using Amazon Bedrock AgentCore 2026-08-06 · BLOGS_PODCASTS
- Can Large Language Models Recover Semantic Optimization Opportunities That Compilers Miss? 2026-08-06 · ARXIV
- From User Sequences to Scaling Laws: A Multi-Stage Architecture for Meta’s Ads Ranking 2026-08-06 · BLOGS_PODCASTS
- Beating GPT-5.6 Sol on retrieval with 100x cheaper open models 2026-08-06 · HACKER_NEWS
- How LendingTree built a multi-agent mortgage assistant on Amazon Bedrock 2026-08-06 · BLOGS_PODCASTS
- Agogic: Performance-Timed Music Tokens for LLM-Native Text-to-Symbolic-Music Generation 2026-08-06 · ARXIV
- Test-Time Scaling in Reasoning LLMs: Inference Regimes, Evaluation, and Reproducibility 2026-08-06 · ARXIV
- How we built a realtime system for responsive voice AI in six months 2026-08-05 · BLOGS_PODCASTS
- TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning 2026-08-05 · ARXIV
- Position: LLMs Can't Jump 2026-08-05 · HACKER_NEWS
- Inside our 353,000-person vibe coding course 2026-08-05 · BLOGS_PODCASTS
- WorldCup Arena: Prospective, Leakage-Free Evaluation of Frontier LLMs on a Live Tournament 2026-08-05 · ARXIV
- Zero-Mem: Zero-Token Memory Operations for LLM Agents 2026-08-05 · HACKER_NEWS
- Unpacking ChatGPT Work: the Agent for a Billion Users 2026-08-05 · BLOGS_PODCASTS
- SocietyBench: Forecasting Counterfactual Social-World Evolution 2026-08-05 · ARXIV
- ParVL: Parallel Scaling and Expandable Compute Allocation for Multimodal LLMs 2026-08-05 · ARXIV
- [AINews] Megakernels are so dead and so back 2026-08-05 · BLOGS_PODCASTS
- Structured Memory for Edge Language Models: Persistent Context and Corpus Retrieval via O(1) SSM State Injection 2026-08-05 · ARXIV
- Introducing Web Search on Amazon Bedrock for foundation model grounding 2026-08-05 · BLOGS_PODCASTS
- Mistral's Shieldstral: 3B open-weights model for multimodal moderation 2026-08-05 · HACKER_NEWS
- AtumAI: A Principled Framework for Agentic Generation of Datacenter Control-Plane Policies 2026-08-05 · ARXIV
- Automated web insight extraction with Amazon Bedrock AgentCore 2026-08-05 · BLOGS_PODCASTS
- The latest AI news we announced in July 2026 2026-08-04 · BLOGS_PODCASTS
- Agent skills that bring team coding standards to Claude Code and Codex 2026-08-04 · HACKER_NEWS
- GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning 2026-08-04 · ARXIV
- The Condition-Number Barrier in Sparse Least Squares 2026-08-04 · ARXIV
- The benefits of medical AI assistance vary based on user expertise 2026-08-04 · BLOGS_PODCASTS
- [AINews] Qwen 3.8 Max(2.4T) and 27B, new open weights models for Coding and Cowork 2026-08-04 · BLOGS_PODCASTS
- From weeks to minutes: How Formula 1® uses agentic AI on AWS to accelerate data operations 2026-08-04 · BLOGS_PODCASTS
- The Inference Engineering Masterclass — Philip Kiely & Ali Taha, Baseten 2026-08-04 · BLOGS_PODCASTS
- LLMs reward expertise 2026-08-04 · HACKER_NEWS
- AgentHPOBench: A Benchmark For Evaluating LLM Agents as Sequential Hyperparameter Optimizers 2026-08-04 · ARXIV
- GEM Training: How Meta Doubled the Efficiency of Its LLM-Scale Ads Foundation Model 2026-08-04 · BLOGS_PODCASTS
- ExtractBench: A Benchmark for Schema-Guided Enterprise Document Extraction 2026-08-03 · ARXIV
- TokTier: Exact Stateful Tokenization for Agentic LLM Serving 2026-08-03 · ARXIV