03 / ARTICLE LEDGER
关联文章
按发布时间倒序
- Challenges in Evaluating Explanation Methods for Static and Evolving Data 2026-08-07 · ARXIV
- Kitesurf: Agent-first browser that runs in V8 isolates 2026-08-07 · HACKER_NEWS
- CalibForge: Adversarial Solver Calibration for Scaling Learnable Terminal Tasks 2026-08-07 · ARXIV
- Resourced Authority A Mechanism-Design Model for Participatory Governance of Deployed AI Agents 2026-08-07 · ARXIV
- The Low Frequency Trap: Video Language Models Fail at Simple Event Bookkeeping 2026-08-07 · ARXIV
- Interviewing Engineers in the AI Era: Lessons from a Year of Rebuilding 2026-08-07 · HACKER_NEWS
- How HSP GRUPPE builds AI capabilities for tax advisory 2026-08-07 · BLOGS_PODCASTS
- AV-AIVAT: 74x Cheaper Agent Evaluation with Certified Anytime-Valid Stopping in Imperfect-Information Games 2026-08-07 · ARXIV
- Investigating Artificial Intelligence Digital Sovereignty in Mobile Shopping Apps: A Case Study of Nigeria 2026-08-07 · ARXIV
- Tracing the Heart: An Evidence-Linked Pipeline for Heart-Failure Feature Engineering 2026-08-07 · ARXIV
- Software development with AI is starting to feel like cooking steak 2026-08-07 · HACKER_NEWS
- [AINews] AMD buys Taalas 2026-08-07 · BLOGS_PODCASTS
- The Bitter Lesson of Tool Calling 2026-08-07 · ARXIV
- jina-reranker-v3.5:通过混合 Attention 和自蒸馏实现更快的 Listwise Reranking 2026-08-07 · JUEJIN
- Control agent behaviors and cost beyond a single action: new capabilities in Amazon Bedrock AgentCore 2026-08-07 · BLOGS_PODCASTS
- Humans missed 1 in 3 threats approving AI agent commands across 40k game runs 2026-08-07 · HACKER_NEWS
- Configure rate limits for AI traffic on AgentCore gateway 2026-08-07 · BLOGS_PODCASTS
- Securing AI agents with temporal policies in Amazon Bedrock AgentCore 2026-08-07 · BLOGS_PODCASTS
- Predicting Brain Morphometry with MT-GNN: Mesh Evolution in Continuous Time with Graph-Based Metric Tensor Embeddings 2026-08-07 · ARXIV
- Inside vLLM: Anatomy of a High-Throughput LLM Inference System (2025) 2026-08-07 · HACKER_NEWS
- Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation for Modern Greek across Specialist Domains 2026-08-06 · ARXIV
- Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning 2026-08-06 · ARXIV
- OctoLong: Mid-Training On Cross-Repository Code Contexts Enhances Long-Context Modeling 2026-08-06 · ARXIV
- Born Against, or why hobby programming communities are against LLM usage 2026-08-06 · HACKER_NEWS
- [AINews] Jeff, Sanjay, Oriol, and Quoc depart DeepMind; Demis to Chair; Koray to SVP — what is going on at GDM??? 2026-08-06 · BLOGS_PODCASTS
- LLMs won't break symmetric crypto 2026-08-06 · HACKER_NEWS
- Argus: A General-Purpose Agentic Runtime for Long-Horizon Reasoning 2026-08-06 · ARXIV
- How we built an MCP bridge to give our AgentCore-hosted AI agent access to local MCP tools 2026-08-06 · BLOGS_PODCASTS
- Prime Agent: A self-improving RLM agent 2026-08-06 · HACKER_NEWS
- Can Large Language Models Recover Semantic Optimization Opportunities That Compilers Miss? 2026-08-06 · ARXIV
- From User Sequences to Scaling Laws: A Multi-Stage Architecture for Meta’s Ads Ranking 2026-08-06 · BLOGS_PODCASTS
- Beating GPT-5.6 Sol on retrieval with 100x cheaper open models 2026-08-06 · HACKER_NEWS
- Assessment of Conditional Diffusion Model for Synthetic Histopathology Image Generation 2026-08-06 · ARXIV
- I'm Scared a Stranger Will Call My Novel AI, So I Built GitHub for Words 2026-08-06 · HACKER_NEWS
- Agogic: Performance-Timed Music Tokens for LLM-Native Text-to-Symbolic-Music Generation 2026-08-06 · ARXIV
- Test-Time Scaling in Reasoning LLMs: Inference Regimes, Evaluation, and Reproducibility 2026-08-06 · ARXIV
- Launch HN: HyperProbe (YC S26) – Agents that do read-only debugging in prod 2026-08-06 · HACKER_NEWS
- Circles powers telco personalization with OpenAI technology 2026-08-06 · BLOGS_PODCASTS
- PAST-Bench: Benchmarking the Foundations of Recursive Self-Improvement in Personal Agents 2026-08-05 · ARXIV
- How we built a realtime system for responsive voice AI in six months 2026-08-05 · BLOGS_PODCASTS
- Cloudflare OS: an open platform for agents, apps, and work 2026-08-05 · HACKER_NEWS
- TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning 2026-08-05 · ARXIV
- Position: LLMs Can't Jump 2026-08-05 · HACKER_NEWS
- Inside our 353,000-person vibe coding course 2026-08-05 · BLOGS_PODCASTS
- WorldCup Arena: Prospective, Leakage-Free Evaluation of Frontier LLMs on a Live Tournament 2026-08-05 · ARXIV
- Teaching AI to speak the language of pathology 2026-08-05 · BLOGS_PODCASTS
- Zero-Mem: Zero-Token Memory Operations for LLM Agents 2026-08-05 · HACKER_NEWS
- Unpacking ChatGPT Work: the Agent for a Billion Users 2026-08-05 · BLOGS_PODCASTS
- SocietyBench: Forecasting Counterfactual Social-World Evolution 2026-08-05 · ARXIV
- ParVL: Parallel Scaling and Expandable Compute Allocation for Multimodal LLMs 2026-08-05 · ARXIV