03 / ARTICLE LEDGER
关联文章
按发布时间倒序
- BioProAgent: Neuro-Symbolic Grounding for Constrained Scientific Planning 2026-03-03 · ARXIV
- Taming Momentum: Rethinking Optimizer States Through Low-Rank Approximation 2026-03-02 · ARXIV
- Do LLMs Benefit From Their Own Words? 2026-03-02 · ARXIV
- DARE-bench: Evaluating Modeling and Instruction Fidelity of LLMs in Data Science 2026-03-02 · ARXIV
- CUDA Agent: Large-Scale Agentic RL for High-Performance CUDA Kernel Generation 2026-03-02 · ARXIV
- Building specialized AI without sacrificing intelligence: Nova Forge data mixing in action | Amazon Web Services 2026-03-02 · BLOGS_PODCASTS
- Build a serverless conversational AI agent using Claude with LangGraph and managed MLflow on Amazon SageMaker AI | Amazon Web Services 2026-03-02 · BLOGS_PODCASTS
- 从入门到辞职|从ChatGPT到OpenClaw,跟上智能时代的进化 2026-03-02 · JUEJIN
- Right-sizes LLM models to your system's RAM, CPU, and GPU 2026-03-02 · HACKER_NEWS
- Spring AI 实战:从零构建类 OpenClaw 的自主 Agent 2026-03-02 · JUEJIN
- SafeGen-LLM: Enhancing Safety Generalization in Task Planning for Robotic Systems 2026-03-02 · ARXIV
- Ollama 入门指南:本地大模型实践 2026-03-02 · JUEJIN
- 推理平台全景 2026-03-01 · JUEJIN
- 轻松接入大语言模型API -04 2026-03-01 · JUEJIN
- The Science of Detecting LLM-Generated Text (2024) 2026-03-01 · HACKER_NEWS
- Deterministic Programming with LLMs 2026-03-01 · HACKER_NEWS
- 一天一个开源项目(第36篇):EverMemOS - 跨 LLM 与平台的长时记忆 OS,让 Agent 会记忆更会推理 2026-02-28 · JUEJIN
- OpenClaw:开源AI助手平台的革命之路 2026-02-28 · JUEJIN
- 基于 Go 语言与 DeepSeek-V3 构建企业级自动化代码审计系统深度解析 2026-02-28 · JUEJIN
- 一文搞懂激活函数! 2026-02-27 · JUEJIN
- Understanding Usage and Engagement in AI-Powered Scientific Research Tools: The Asta Interaction Dataset 2026-02-27 · ARXIV
- We gave terabytes of CI logs to an LLM 2026-02-27 · HACKER_NEWS
- Show HN: Badge that shows how well your codebase fits in an LLM's context window 2026-02-27 · HACKER_NEWS
- AI Agent框架探秘:拆解 OpenHands(9)--- AgentController 2026-02-27 · JUEJIN
- Dify 构建 FE 工作流:前端团队可复用 AI 工作流实战 2026-02-27 · JUEJIN
- 如何用 AI 做业务级 Code Review 2026-02-27 · JUEJIN
- Scaling Search Relevance: Augmenting App Store Ranking with LLM-Generated Judgments 2026-02-27 · ARXIV
- Mitigating Legibility Tax with Decoupled Prover-Verifier Games 2026-02-27 · ARXIV
- Discourse-Aware Dual-Track Streaming Response for Low-Latency Spoken Dialogue Systems 2026-02-27 · ARXIV
- Agency and Architectural Limits: Why Optimization-Based Systems Cannot Be Norm-Responsive 2026-02-27 · ARXIV
- 灵感启发:日产文章 100 篇,打造“实时热点洞察”引擎 2026-02-26 · JUEJIN
- Recovered in Translation: Efficient Pipeline for Automated Translation of Benchmarks and Datasets 2026-02-26 · ARXIV
- Large model inference container – latest capabilities and performance enhancements | Amazon Web Services 2026-02-26 · BLOGS_PODCASTS
- New method could increase LLM training efficiency 2026-02-26 · BLOGS_PODCASTS
- crawl4ai:AI时代的数据采集利器——从入门到实战 2026-02-26 · JUEJIN
- Show HN: ZSE – Open-source LLM inference engine with 3.9s cold starts 2026-02-26 · HACKER_NEWS
- Enhancing LLM-Based Test Generation by Eliminating Covered Code 2026-02-26 · ARXIV
- CxMP: A Linguistic Minimal-Pair Benchmark for Evaluating Constructional Understanding in Language Models 2026-02-26 · ARXIV
- Why Pass@k Optimization Can Degrade Pass@1: Prompt Interference in LLM Post-training 2026-02-25 · ARXIV
- On Data Engineering for Scaling LLM Terminal Capabilities 2026-02-25 · ARXIV
- Learning from Trials and Errors: Reflective Test-Time Planning for Embodied LLMs 2026-02-25 · ARXIV
- Large-Scale Online Deanonymization with LLMs 2026-02-25 · HACKER_NEWS
- AI Agent框架探秘:拆解 OpenHands(8)--- CodeActAgent 2026-02-25 · JUEJIN
- VAUQ: Vision-Aware Uncertainty Quantification for LVLM Self-Evaluation 2026-02-25 · ARXIV
- Tool Building as a Path to "Superintelligence" 2026-02-25 · ARXIV
- Beyond the Star Rating: A Scalable Framework for Aspect-Based Sentiment Analysis Using LLMs and Text Classification 2026-02-25 · ARXIV
- An Expert Schema for Evaluating Large Language Model Errors in Scholarly Question-Answering Systems 2026-02-25 · ARXIV
- Mercury 2: Fast reasoning LLM powered by diffusion 2026-02-25 · HACKER_NEWS
- Skill-Inject: Measuring Agent Vulnerability to Skill File Attacks 2026-02-24 · ARXIV
- RCCLX: Innovating GPU Communications on AMD Platforms 2026-02-24 · BLOGS_PODCASTS