03 / ARTICLE LEDGER

关联文章

按发布时间倒序

  1. Stealing Reasoning Traces from Proprietary LLM APIs 2026-08-12 · ARXIV
  2. Daybreak models are now available on AWS 2026-08-12 · BLOGS_PODCASTS
  3. Suzanne: AI tool for designing and manufacturing physical products 2026-08-12 · HACKER_NEWS
  4. Sci-VBench: Evaluating Knowledge- and Reasoning-Intensive Video Generation in Science Domains 2026-08-12 · ARXIV
  5. 🔬The BioAI Phase Shift - Matthew McPartlon & Neil Patil, Chai Discovery 2026-08-12 · BLOGS_PODCASTS
  6. Accelerate cyber defense with OpenAI and AWS: Daybreak Red & Daybreak Blue now available to eligible customers on Amazon Bedrock 2026-08-12 · BLOGS_PODCASTS
  7. ArchAgent v2: A Case Study with the Data Prefetching Championship 2026-08-12 · ARXIV
  8. OpenAI’s head of ethics leaves less than a year after joining 2026-08-12 · HACKER_NEWS
  9. First Orion accelerates QA automation using Amazon Nova Act 2026-08-12 · BLOGS_PODCASTS
  10. Energy-Structured Latent World Models with Neural Time Fields for Physically Constistent Open-World Motion Planning 2026-08-12 · ARXIV
  11. Why Go Is an Ideal Language for AI-Assisted Software Engineering 2026-08-12 · HACKER_NEWS
  12. Financial Numerical Prediction and Allocation as Token Generation 2026-08-12 · ARXIV
  13. What I learned by putting GitHub Copilot behind a MitM proxy 2026-08-12 · HACKER_NEWS
  14. SHE: Trajectory-driven Safety Harness Evolution for LLM Agents 2026-08-12 · ARXIV
  15. AMIE, our research medical AI system, demonstrates real-time clinical video consultation capabilities in a first-of-its-kind study. 2026-08-12 · BLOGS_PODCASTS
  16. 番外篇二:《不到十行代码,我用ReactAgen搭了个会自己调工具的 Agent》 2026-08-12 · JUEJIN
  17. Stealing Reasoning Traces from Proprietary LLM APIs 2026-08-12 · HACKER_NEWS
  18. Space-Creating versus Dead Possession: An Off-Ball Possession-Quality Index for Broadcast Football 2026-08-12 · ARXIV
  19. How Pixieset achieved 35% AI feature adoption by solving the right problem with Amazon Bedrock 2026-08-12 · BLOGS_PODCASTS
  20. Consilience for Verifier-Free Test-Time Scaling 2026-08-11 · ARXIV
  21. Apple Silicon and macOS VMs: 11–16× Faster LLM Inference with Llama.cpp 2026-08-11 · HACKER_NEWS
  22. 从“即兴创作”到“规格先行”,华为云码道(CodeArts)代码智能体持续深耕企业级规范驱动开发能力 2026-08-11 · JUEJIN
  23. How to organize Claude Code for product work 2026-08-11 · HACKER_NEWS
  24. Decoding-Level Taboo: A Diagnostic Stress Test for LLM Robustness 2026-08-11 · ARXIV
  25. DSLE: A Learning Environment for Dark Souls Boss Encounters 2026-08-11 · ARXIV
  26. GENCO - A Unified Neural Solver Embedded in a Development Framework for Steady-State Grid Analysis 2026-08-11 · ARXIV
  27. The Water Footprint of AI 2026-08-11 · HACKER_NEWS
  28. From Values to Benchmarks: Evaluating Large Language Models for Governmental Use in Dutch 2026-08-11 · ARXIV
  29. How Claude marks AI-generated content 2026-08-11 · HACKER_NEWS
  30. Multimodal Model Diffing for Feature Discovery and Control 2026-08-11 · ARXIV
  31. As AI eats the web, the internet’s collective memory is disappearing 2026-08-11 · HACKER_NEWS
  32. 让不同大模型共享一个 Agent:Pi 如何统一 Provider 与 Context Handoff 2026-08-11 · JUEJIN
  33. DeepSeek: Reverse Engineering an AI Assistant by Interviewing Itself 2026-08-11 · HACKER_NEWS
  34. [AINews] Muse Glimmer and Spark: Open Weights return Personal Superintelligence promise 2026-08-11 · BLOGS_PODCASTS
  35. 光本位联合文心快码打造面向光电芯片研发的全栈AI Agent Lightmate 2026-08-11 · JUEJIN
  36. Beyond Naturalness: Probing Automated Text-To-Speech Evaluators on Linguistically Grounded Dimensions 2026-08-11 · ARXIV
  37. KGCaRe: Explainable Complex Conditional Question Answering using Automatic Knowledge Graph Construction and Context Retrieval with LLMs 2026-08-11 · ARXIV
  38. What's the best programming language for coding agents? 2026-08-11 · HACKER_NEWS
  39. A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs While Improving Accuracy 2026-08-11 · ARXIV
  40. TEPA: Revoking Stale Memories for Conflict-Robust Language Agents 2026-08-11 · ARXIV
  41. Humanising LLM Outputs Is Dumb 2026-08-11 · HACKER_NEWS
  42. Show HN: Needle2: 14MB agentic LLM for phones, wearables, smart home and robots 2026-08-11 · HACKER_NEWS
  43. Diffusion LLMs as Targets and Adversaries: Mechanistic Safety Exploits 2026-08-11 · ARXIV
  44. SABRE: Scalable and Automated Benchmarking of VLMs under Stress 2026-08-11 · ARXIV
  45. Mark Zuckerberg attacks 'closed' AI rivals as Meta returns to open models 2026-08-11 · HACKER_NEWS
  46. Post-Grokking Collapse at the Representation-Readout Interface in Muon-Trained Transformers 2026-08-11 · ARXIV
  47. Launch HN: Stoa Markets (YC S26) – A Marketplace for GPUs and AI Servers 2026-08-11 · HACKER_NEWS
  48. What building an AI-native finance function taught me 2026-08-11 · BLOGS_PODCASTS
  49. Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows 2026-08-11 · HACKER_NEWS
  50. Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing 2026-08-11 · ARXIV