03 / ARTICLE LEDGER

关联文章

按发布时间倒序

  1. BioProAgent: Neuro-Symbolic Grounding for Constrained Scientific Planning 2026-03-03 · ARXIV
  2. Taming Momentum: Rethinking Optimizer States Through Low-Rank Approximation 2026-03-02 · ARXIV
  3. Do LLMs Benefit From Their Own Words? 2026-03-02 · ARXIV
  4. DARE-bench: Evaluating Modeling and Instruction Fidelity of LLMs in Data Science 2026-03-02 · ARXIV
  5. CUDA Agent: Large-Scale Agentic RL for High-Performance CUDA Kernel Generation 2026-03-02 · ARXIV
  6. Building specialized AI without sacrificing intelligence: Nova Forge data mixing in action | Amazon Web Services 2026-03-02 · BLOGS_PODCASTS
  7. Build a serverless conversational AI agent using Claude with LangGraph and managed MLflow on Amazon SageMaker AI | Amazon Web Services 2026-03-02 · BLOGS_PODCASTS
  8. 从入门到辞职|从ChatGPT到OpenClaw,跟上智能时代的进化 2026-03-02 · JUEJIN
  9. Right-sizes LLM models to your system's RAM, CPU, and GPU 2026-03-02 · HACKER_NEWS
  10. Spring AI 实战:从零构建类 OpenClaw 的自主 Agent 2026-03-02 · JUEJIN
  11. SafeGen-LLM: Enhancing Safety Generalization in Task Planning for Robotic Systems 2026-03-02 · ARXIV
  12. Ollama 入门指南:本地大模型实践 2026-03-02 · JUEJIN
  13. 推理平台全景 2026-03-01 · JUEJIN
  14. 轻松接入大语言模型API -04 2026-03-01 · JUEJIN
  15. The Science of Detecting LLM-Generated Text (2024) 2026-03-01 · HACKER_NEWS
  16. Deterministic Programming with LLMs 2026-03-01 · HACKER_NEWS
  17. 一天一个开源项目(第36篇):EverMemOS - 跨 LLM 与平台的长时记忆 OS,让 Agent 会记忆更会推理 2026-02-28 · JUEJIN
  18. OpenClaw:开源AI助手平台的革命之路 2026-02-28 · JUEJIN
  19. 基于 Go 语言与 DeepSeek-V3 构建企业级自动化代码审计系统深度解析 2026-02-28 · JUEJIN
  20. 一文搞懂激活函数! 2026-02-27 · JUEJIN
  21. Understanding Usage and Engagement in AI-Powered Scientific Research Tools: The Asta Interaction Dataset 2026-02-27 · ARXIV
  22. We gave terabytes of CI logs to an LLM 2026-02-27 · HACKER_NEWS
  23. Show HN: Badge that shows how well your codebase fits in an LLM's context window 2026-02-27 · HACKER_NEWS
  24. AI Agent框架探秘:拆解 OpenHands(9)--- AgentController 2026-02-27 · JUEJIN
  25. Dify 构建 FE 工作流:前端团队可复用 AI 工作流实战 2026-02-27 · JUEJIN
  26. 如何用 AI 做业务级 Code Review 2026-02-27 · JUEJIN
  27. Scaling Search Relevance: Augmenting App Store Ranking with LLM-Generated Judgments 2026-02-27 · ARXIV
  28. Mitigating Legibility Tax with Decoupled Prover-Verifier Games 2026-02-27 · ARXIV
  29. Discourse-Aware Dual-Track Streaming Response for Low-Latency Spoken Dialogue Systems 2026-02-27 · ARXIV
  30. Agency and Architectural Limits: Why Optimization-Based Systems Cannot Be Norm-Responsive 2026-02-27 · ARXIV
  31. 灵感启发:日产文章 100 篇,打造“实时热点洞察”引擎 2026-02-26 · JUEJIN
  32. Recovered in Translation: Efficient Pipeline for Automated Translation of Benchmarks and Datasets 2026-02-26 · ARXIV
  33. Large model inference container – latest capabilities and performance enhancements | Amazon Web Services 2026-02-26 · BLOGS_PODCASTS
  34. New method could increase LLM training efficiency 2026-02-26 · BLOGS_PODCASTS
  35. crawl4ai:AI时代的数据采集利器——从入门到实战 2026-02-26 · JUEJIN
  36. Show HN: ZSE – Open-source LLM inference engine with 3.9s cold starts 2026-02-26 · HACKER_NEWS
  37. Enhancing LLM-Based Test Generation by Eliminating Covered Code 2026-02-26 · ARXIV
  38. CxMP: A Linguistic Minimal-Pair Benchmark for Evaluating Constructional Understanding in Language Models 2026-02-26 · ARXIV
  39. Why Pass@k Optimization Can Degrade Pass@1: Prompt Interference in LLM Post-training 2026-02-25 · ARXIV
  40. On Data Engineering for Scaling LLM Terminal Capabilities 2026-02-25 · ARXIV
  41. Learning from Trials and Errors: Reflective Test-Time Planning for Embodied LLMs 2026-02-25 · ARXIV
  42. Large-Scale Online Deanonymization with LLMs 2026-02-25 · HACKER_NEWS
  43. AI Agent框架探秘:拆解 OpenHands(8)--- CodeActAgent 2026-02-25 · JUEJIN
  44. VAUQ: Vision-Aware Uncertainty Quantification for LVLM Self-Evaluation 2026-02-25 · ARXIV
  45. Tool Building as a Path to "Superintelligence" 2026-02-25 · ARXIV
  46. Beyond the Star Rating: A Scalable Framework for Aspect-Based Sentiment Analysis Using LLMs and Text Classification 2026-02-25 · ARXIV
  47. An Expert Schema for Evaluating Large Language Model Errors in Scholarly Question-Answering Systems 2026-02-25 · ARXIV
  48. Mercury 2: Fast reasoning LLM powered by diffusion 2026-02-25 · HACKER_NEWS
  49. Skill-Inject: Measuring Agent Vulnerability to Skill File Attacks 2026-02-24 · ARXIV
  50. RCCLX: Innovating GPU Communications on AMD Platforms 2026-02-24 · BLOGS_PODCASTS