03 / ARTICLE LEDGER
关联文章
按发布时间倒序
- Stealing Reasoning Traces from Proprietary LLM APIs 2026-08-12 · ARXIV
- Daybreak models are now available on AWS 2026-08-12 · BLOGS_PODCASTS
- Suzanne: AI tool for designing and manufacturing physical products 2026-08-12 · HACKER_NEWS
- Sci-VBench: Evaluating Knowledge- and Reasoning-Intensive Video Generation in Science Domains 2026-08-12 · ARXIV
- 🔬The BioAI Phase Shift - Matthew McPartlon & Neil Patil, Chai Discovery 2026-08-12 · BLOGS_PODCASTS
- Accelerate cyber defense with OpenAI and AWS: Daybreak Red & Daybreak Blue now available to eligible customers on Amazon Bedrock 2026-08-12 · BLOGS_PODCASTS
- ArchAgent v2: A Case Study with the Data Prefetching Championship 2026-08-12 · ARXIV
- OpenAI’s head of ethics leaves less than a year after joining 2026-08-12 · HACKER_NEWS
- First Orion accelerates QA automation using Amazon Nova Act 2026-08-12 · BLOGS_PODCASTS
- Energy-Structured Latent World Models with Neural Time Fields for Physically Constistent Open-World Motion Planning 2026-08-12 · ARXIV
- Why Go Is an Ideal Language for AI-Assisted Software Engineering 2026-08-12 · HACKER_NEWS
- Financial Numerical Prediction and Allocation as Token Generation 2026-08-12 · ARXIV
- What I learned by putting GitHub Copilot behind a MitM proxy 2026-08-12 · HACKER_NEWS
- SHE: Trajectory-driven Safety Harness Evolution for LLM Agents 2026-08-12 · ARXIV
- AMIE, our research medical AI system, demonstrates real-time clinical video consultation capabilities in a first-of-its-kind study. 2026-08-12 · BLOGS_PODCASTS
- 番外篇二:《不到十行代码,我用ReactAgen搭了个会自己调工具的 Agent》 2026-08-12 · JUEJIN
- Stealing Reasoning Traces from Proprietary LLM APIs 2026-08-12 · HACKER_NEWS
- Space-Creating versus Dead Possession: An Off-Ball Possession-Quality Index for Broadcast Football 2026-08-12 · ARXIV
- How Pixieset achieved 35% AI feature adoption by solving the right problem with Amazon Bedrock 2026-08-12 · BLOGS_PODCASTS
- Consilience for Verifier-Free Test-Time Scaling 2026-08-11 · ARXIV
- Apple Silicon and macOS VMs: 11–16× Faster LLM Inference with Llama.cpp 2026-08-11 · HACKER_NEWS
- 从“即兴创作”到“规格先行”,华为云码道(CodeArts)代码智能体持续深耕企业级规范驱动开发能力 2026-08-11 · JUEJIN
- How to organize Claude Code for product work 2026-08-11 · HACKER_NEWS
- Decoding-Level Taboo: A Diagnostic Stress Test for LLM Robustness 2026-08-11 · ARXIV
- DSLE: A Learning Environment for Dark Souls Boss Encounters 2026-08-11 · ARXIV
- GENCO - A Unified Neural Solver Embedded in a Development Framework for Steady-State Grid Analysis 2026-08-11 · ARXIV
- The Water Footprint of AI 2026-08-11 · HACKER_NEWS
- From Values to Benchmarks: Evaluating Large Language Models for Governmental Use in Dutch 2026-08-11 · ARXIV
- How Claude marks AI-generated content 2026-08-11 · HACKER_NEWS
- Multimodal Model Diffing for Feature Discovery and Control 2026-08-11 · ARXIV
- As AI eats the web, the internet’s collective memory is disappearing 2026-08-11 · HACKER_NEWS
- 让不同大模型共享一个 Agent:Pi 如何统一 Provider 与 Context Handoff 2026-08-11 · JUEJIN
- DeepSeek: Reverse Engineering an AI Assistant by Interviewing Itself 2026-08-11 · HACKER_NEWS
- [AINews] Muse Glimmer and Spark: Open Weights return Personal Superintelligence promise 2026-08-11 · BLOGS_PODCASTS
- 光本位联合文心快码打造面向光电芯片研发的全栈AI Agent Lightmate 2026-08-11 · JUEJIN
- Beyond Naturalness: Probing Automated Text-To-Speech Evaluators on Linguistically Grounded Dimensions 2026-08-11 · ARXIV
- KGCaRe: Explainable Complex Conditional Question Answering using Automatic Knowledge Graph Construction and Context Retrieval with LLMs 2026-08-11 · ARXIV
- What's the best programming language for coding agents? 2026-08-11 · HACKER_NEWS
- A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs While Improving Accuracy 2026-08-11 · ARXIV
- TEPA: Revoking Stale Memories for Conflict-Robust Language Agents 2026-08-11 · ARXIV
- Humanising LLM Outputs Is Dumb 2026-08-11 · HACKER_NEWS
- Show HN: Needle2: 14MB agentic LLM for phones, wearables, smart home and robots 2026-08-11 · HACKER_NEWS
- Diffusion LLMs as Targets and Adversaries: Mechanistic Safety Exploits 2026-08-11 · ARXIV
- SABRE: Scalable and Automated Benchmarking of VLMs under Stress 2026-08-11 · ARXIV
- Mark Zuckerberg attacks 'closed' AI rivals as Meta returns to open models 2026-08-11 · HACKER_NEWS
- Post-Grokking Collapse at the Representation-Readout Interface in Muon-Trained Transformers 2026-08-11 · ARXIV
- Launch HN: Stoa Markets (YC S26) – A Marketplace for GPUs and AI Servers 2026-08-11 · HACKER_NEWS
- What building an AI-native finance function taught me 2026-08-11 · BLOGS_PODCASTS
- Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows 2026-08-11 · HACKER_NEWS
- Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing 2026-08-11 · ARXIV