03 / ARTICLE LEDGER

关联文章

按发布时间倒序

  1. Challenges in Evaluating Explanation Methods for Static and Evolving Data 2026-08-07 · ARXIV
  2. Kitesurf: Agent-first browser that runs in V8 isolates 2026-08-07 · HACKER_NEWS
  3. CalibForge: Adversarial Solver Calibration for Scaling Learnable Terminal Tasks 2026-08-07 · ARXIV
  4. Resourced Authority A Mechanism-Design Model for Participatory Governance of Deployed AI Agents 2026-08-07 · ARXIV
  5. The Low Frequency Trap: Video Language Models Fail at Simple Event Bookkeeping 2026-08-07 · ARXIV
  6. Interviewing Engineers in the AI Era: Lessons from a Year of Rebuilding 2026-08-07 · HACKER_NEWS
  7. How HSP GRUPPE builds AI capabilities for tax advisory 2026-08-07 · BLOGS_PODCASTS
  8. AV-AIVAT: 74x Cheaper Agent Evaluation with Certified Anytime-Valid Stopping in Imperfect-Information Games 2026-08-07 · ARXIV
  9. Investigating Artificial Intelligence Digital Sovereignty in Mobile Shopping Apps: A Case Study of Nigeria 2026-08-07 · ARXIV
  10. Tracing the Heart: An Evidence-Linked Pipeline for Heart-Failure Feature Engineering 2026-08-07 · ARXIV
  11. Software development with AI is starting to feel like cooking steak 2026-08-07 · HACKER_NEWS
  12. [AINews] AMD buys Taalas 2026-08-07 · BLOGS_PODCASTS
  13. The Bitter Lesson of Tool Calling 2026-08-07 · ARXIV
  14. jina-reranker-v3.5:通过混合 Attention 和自蒸馏实现更快的 Listwise Reranking 2026-08-07 · JUEJIN
  15. Control agent behaviors and cost beyond a single action: new capabilities in Amazon Bedrock AgentCore 2026-08-07 · BLOGS_PODCASTS
  16. Humans missed 1 in 3 threats approving AI agent commands across 40k game runs 2026-08-07 · HACKER_NEWS
  17. Configure rate limits for AI traffic on AgentCore gateway 2026-08-07 · BLOGS_PODCASTS
  18. Securing AI agents with temporal policies in Amazon Bedrock AgentCore 2026-08-07 · BLOGS_PODCASTS
  19. Predicting Brain Morphometry with MT-GNN: Mesh Evolution in Continuous Time with Graph-Based Metric Tensor Embeddings 2026-08-07 · ARXIV
  20. Inside vLLM: Anatomy of a High-Throughput LLM Inference System (2025) 2026-08-07 · HACKER_NEWS
  21. Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation for Modern Greek across Specialist Domains 2026-08-06 · ARXIV
  22. Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning 2026-08-06 · ARXIV
  23. OctoLong: Mid-Training On Cross-Repository Code Contexts Enhances Long-Context Modeling 2026-08-06 · ARXIV
  24. Born Against, or why hobby programming communities are against LLM usage 2026-08-06 · HACKER_NEWS
  25. [AINews] Jeff, Sanjay, Oriol, and Quoc depart DeepMind; Demis to Chair; Koray to SVP — what is going on at GDM??? 2026-08-06 · BLOGS_PODCASTS
  26. LLMs won't break symmetric crypto 2026-08-06 · HACKER_NEWS
  27. Argus: A General-Purpose Agentic Runtime for Long-Horizon Reasoning 2026-08-06 · ARXIV
  28. How we built an MCP bridge to give our AgentCore-hosted AI agent access to local MCP tools 2026-08-06 · BLOGS_PODCASTS
  29. Prime Agent: A self-improving RLM agent 2026-08-06 · HACKER_NEWS
  30. Can Large Language Models Recover Semantic Optimization Opportunities That Compilers Miss? 2026-08-06 · ARXIV
  31. From User Sequences to Scaling Laws: A Multi-Stage Architecture for Meta’s Ads Ranking 2026-08-06 · BLOGS_PODCASTS
  32. Beating GPT-5.6 Sol on retrieval with 100x cheaper open models 2026-08-06 · HACKER_NEWS
  33. Assessment of Conditional Diffusion Model for Synthetic Histopathology Image Generation 2026-08-06 · ARXIV
  34. I'm Scared a Stranger Will Call My Novel AI, So I Built GitHub for Words 2026-08-06 · HACKER_NEWS
  35. Agogic: Performance-Timed Music Tokens for LLM-Native Text-to-Symbolic-Music Generation 2026-08-06 · ARXIV
  36. Test-Time Scaling in Reasoning LLMs: Inference Regimes, Evaluation, and Reproducibility 2026-08-06 · ARXIV
  37. Launch HN: HyperProbe (YC S26) – Agents that do read-only debugging in prod 2026-08-06 · HACKER_NEWS
  38. Circles powers telco personalization with OpenAI technology 2026-08-06 · BLOGS_PODCASTS
  39. PAST-Bench: Benchmarking the Foundations of Recursive Self-Improvement in Personal Agents 2026-08-05 · ARXIV
  40. How we built a realtime system for responsive voice AI in six months 2026-08-05 · BLOGS_PODCASTS
  41. Cloudflare OS: an open platform for agents, apps, and work 2026-08-05 · HACKER_NEWS
  42. TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning 2026-08-05 · ARXIV
  43. Position: LLMs Can't Jump 2026-08-05 · HACKER_NEWS
  44. Inside our 353,000-person vibe coding course 2026-08-05 · BLOGS_PODCASTS
  45. WorldCup Arena: Prospective, Leakage-Free Evaluation of Frontier LLMs on a Live Tournament 2026-08-05 · ARXIV
  46. Teaching AI to speak the language of pathology 2026-08-05 · BLOGS_PODCASTS
  47. Zero-Mem: Zero-Token Memory Operations for LLM Agents 2026-08-05 · HACKER_NEWS
  48. Unpacking ChatGPT Work: the Agent for a Billion Users 2026-08-05 · BLOGS_PODCASTS
  49. SocietyBench: Forecasting Counterfactual Social-World Evolution 2026-08-05 · ARXIV
  50. ParVL: Parallel Scaling and Expandable Compute Allocation for Multimodal LLMs 2026-08-05 · ARXIV