03 / ARTICLE LEDGER

关联文章

按发布时间倒序

  1. SPO++: Stream-Aligned Policy Optimization for Asynchronous Agentic RL 2026-08-26 · ARXIV
  2. Recursive Experiential-Working Memory Evolution for Long-Horizon Agent Harnesses 2026-08-26 · ARXIV
  3. Mitigating Reasoning-Induced Misalignment via Safety-Direction Penalty 2026-08-26 · ARXIV
  4. Advancing price-performance for developers with GPT‑5.6 in Kiro 2026-08-26 · BLOGS_PODCASTS
  5. When Names Cross Scripts: A Source-Grounded Benchmark for Historical Entity Reconciliation in the Mongol World 2026-08-26 · ARXIV
  6. Introducing the Admin plugin for ChatGPT Work and Codex 2026-08-26 · BLOGS_PODCASTS
  7. The Measurement Revolution? Credible Measurement and Inference in the Age of AI 2026-08-26 · ARXIV
  8. Show HN: I made a Raspberry with Qwen my local car AI 2026-08-26 · HACKER_NEWS
  9. Agentic observability with Amazon OpenSearch Service MCP Apps 2026-08-26 · BLOGS_PODCASTS
  10. OpenAI Jalapeño: Better than Nvidia Blackwell 2026-08-26 · HACKER_NEWS
  11. EarthVerse: Benchmarking Scientific Agents Across Dynamic Earth Systems and Natural Hazards 2026-08-26 · ARXIV
  12. 5 ways to upgrade your home decor with Google Search 2026-08-26 · BLOGS_PODCASTS
  13. Correcting a learned physical invariant improves world-model rollouts 2026-08-26 · ARXIV
  14. AI is hitting entry-level jobs hardest, Stanford study finds 2026-08-26 · HACKER_NEWS
  15. The full stack behind abundant intelligence 2026-08-26 · BLOGS_PODCASTS
  16. Predicting Multiple Clinical Outcomes Related to Functional Recovery and Social Isolation Among Older Adults After Lower-Limb Fracture or Hip Replacement 2026-08-26 · ARXIV
  17. How much of HN is AI? 2026-08-26 · HACKER_NEWS
  18. Jalapeño’s first results show industry-leading speed and efficiency in AI inference 2026-08-25 · BLOGS_PODCASTS
  19. Interpretable AI with Local Distillation 2026-08-25 · ARXIV
  20. Anthropic tells staff to work from home due to possible security team strike 2026-08-25 · HACKER_NEWS
  21. The Interaction Tax: When Communication Erases Diversity in Multi-Agent Teams 2026-08-25 · ARXIV
  22. OpenAI restores 5-hour Codex and Work limits for ChatGPT Plus users 2026-08-25 · HACKER_NEWS
  23. How AI Assistance Affects Human Skill Development: A Study of Learning with Logic Puzzles 2026-08-25 · ARXIV
  24. Building a restaurant telephony AI host with Amazon Connect 2026-08-25 · BLOGS_PODCASTS
  25. Apple introduces M6 and M5 Ultra for a big leap in performance and AI compute 2026-08-25 · HACKER_NEWS
  26. ConvergeFlow: Language Flow with Provable Convergence to Token Embeddings 2026-08-25 · ARXIV
  27. Agentic Resource Discovery (ARD): An open specification for agent discovery 2026-08-25 · BLOGS_PODCASTS
  28. Prime Agent: A Self-Improving RLM Harness 2026-08-25 · ARXIV
  29. MTIA 300: Meta’s First Training Chip with Built-in NICs and Communication-Offloading Engines 2026-08-25 · BLOGS_PODCASTS
  30. Provably adaptive sampling with uniform and remasking discrete diffusion models 2026-08-25 · ARXIV
  31. Physics-Constrained Deep Learning Model for Contactless Blood Pressure Monitoring from Triaxial Bodyseismography 2026-08-25 · ARXIV
  32. Disrupting a new covert influence campaign from Russia 2026-08-25 · BLOGS_PODCASTS
  33. Training AI to Paint with Code 2026-08-25 · HACKER_NEWS
  34. MetaRoCE: A New RDMA Transport Built for AI-Scale Ethernet 2026-08-25 · BLOGS_PODCASTS
  35. EG-ARSA: An Expert-Grounded Open Model for Visual Road Safety Auditing in Low-Resource Settings 2026-08-25 · ARXIV
  36. SWE Refactor Bench: Can Coding Agents Complete a Long-Horizon, Whole-Repository Stack Migration? 2026-08-25 · ARXIV
  37. ReWorld: An Interactive World Model with Long-Horizon Memory 2026-08-25 · ARXIV
  38. Headlong: A Microharness for Persistent Agents 2026-08-25 · HACKER_NEWS
  39. LLMs could control their host machines by exploiting inference engines 2026-08-25 · HACKER_NEWS
  40. How to Train a Critic Stably and Efficiently 2026-08-25 · ARXIV
  41. OmniScientist: An Omni-Modal Omni-Discipline AI Scientist 2026-08-14 · ARXIV
  42. OpenAI appoints Dali Rajic as Chief Revenue Officer 2026-08-14 · BLOGS_PODCASTS
  43. AutoDesign: Meta-Harness Optimization for Long-Horizon Agentic Design 2026-08-14 · ARXIV
  44. Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed 2026-08-14 · BLOGS_PODCASTS
  45. One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL 2026-08-14 · ARXIV
  46. The builder’s guide to GPT‑5.6 2026-08-14 · BLOGS_PODCASTS
  47. Diagram-MMU: A Multi-Modal Benchmark for Scientific Diagrams 2026-08-14 · ARXIV
  48. AI At Home Part 1: A Box Of Scraps 2026-08-14 · HACKER_NEWS
  49. How Organizations Use AI: Evidence from ChatGPT [pdf] 2026-08-14 · HACKER_NEWS
  50. A Cascaded Unsupervised-Supervised NLP Pipeline for Detecting Accusatory Language in Public Procurement 2026-08-14 · ARXIV