02Chain-of-thought obfuscation learned from output supervision can generalise to unseen tasksARXIV ArXiv AI Agent阅读文章 arrow_forward
01#490 – State of AI in 2026: LLMs, Coding, Scaling Laws, China, Agents, GPUs, AGI | Lex Fridman PodcastBLOGS_PODCASTS 博客与播客 AI Agent阅读文章 arrow_forward
01Towards a science of scaling agent systems: When and why agent systems workHACKER_NEWS Hacker News AI Agent阅读文章 arrow_forward
01Show HN: Zuckerman – minimalist personal AI agent that self-edits its own codeHACKER_NEWS Hacker News AI Agent阅读文章 arrow_forward
01What I learned building an opinionated and minimal coding agentHACKER_NEWS Hacker News AI Agent阅读文章 arrow_forward
01Generative AI and Wikipedia editing: What we learned in 2025HACKER_NEWS Hacker News 生成式 AI阅读文章 arrow_forward
31Autonomous cars, drones cheerfully obey prompt injection by road signHACKER_NEWS Hacker News AI 安全阅读文章 arrow_forward
31A Step Behind the Bleeding Edge: A Philosophy on AI in DevHACKER_NEWS Hacker News阅读文章 arrow_forward
31Show HN: I trained a 9M speech model to fix my Mandarin tonesHACKER_NEWS Hacker News阅读文章 arrow_forward
31Show HN: Pinecone Explorer – Desktop GUI for the Pinecone vector databaseHACKER_NEWS Hacker News 数据库阅读文章 arrow_forward
31Show HN: I built an AI conversation partner to practice speaking languagesHACKER_NEWS Hacker News阅读文章 arrow_forward
30Reasoning While Asking: Transforming Reasoning Large Language Models from Passive Solvers to Proactive InquirersARXIV ArXiv 大语言模型阅读文章 arrow_forward
30Late Breaking Results: Conversion of Neural Networks into Logic Flows for Edge ComputingARXIV ArXiv阅读文章 arrow_forward
30Hybrid Linear Attention Done Right: Efficient Distillation and Effective Architectures for Extremely Long ContextsARXIV ArXiv阅读文章 arrow_forward
30FineInstructions: Scaling Synthetic Instructions to Pre-Training ScaleARXIV ArXiv 大语言模型阅读文章 arrow_forward
30Mamdani to kill the NYC AI chatbot caught telling businesses to break the lawHACKER_NEWS Hacker News阅读文章 arrow_forward
30Show HN: Amla Sandbox – WASM bash shell sandbox for AI agentsHACKER_NEWS Hacker News AI Agent阅读文章 arrow_forward
30Taisei Corporation shapes the next generation of talent with AIBLOGS_PODCASTS 博客与播客 生成式 AI阅读文章 arrow_forward
30Moltworker: a self-hosted personal AI agent, minus the minisHACKER_NEWS Hacker News AI Agent阅读文章 arrow_forward
30Why Adam Works Better with $β_1 = β_2$: The Missing Gradient Scale Invariance PrincipleARXIV ArXiv阅读文章 arrow_forward
30Language-based Trial and Error Falls Behind in the Era of ExperienceARXIV ArXiv AI Agent阅读文章 arrow_forward
30EWSJF: An Adaptive Scheduler with Hybrid Partitioning for Mixed-Workload LLM InferenceARXIV ArXiv 大语言模型阅读文章 arrow_forward
30Epistemic Context Learning: Building Trust the Right Way in LLM-Based Multi-Agent SystemsARXIV ArXiv AI Agent阅读文章 arrow_forward
29When Flores Bloomz Wrong: Cross-Direction Contamination in Machine Translation EvaluationARXIV ArXiv 大语言模型阅读文章 arrow_forward
29SokoBench: Evaluating Long-Horizon Planning and Reasoning in Large Language ModelsARXIV ArXiv 大语言模型阅读文章 arrow_forward
29Project Genie: Experimenting with infinite, interactive worldsHACKER_NEWS Hacker News阅读文章 arrow_forward
29Post-Training Fairness Control: A Single-Train Framework for Dynamic Fairness in RecommendationARXIV ArXiv阅读文章 arrow_forward
29PatchFormer: A Patch-Based Time Series Foundation Model with Hierarchical Masked Reconstruction and Cross-Domain Transfer Learning for Zero-Shot Multi-Horizon ForecastingARXIV ArXiv阅读文章 arrow_forward
29Exploring Transformer Placement in Variational Autoencoders for Tabular Data GenerationARXIV ArXiv阅读文章 arrow_forward
29Deep Researcher with Sequential Plan Reflection and Candidates Crossover (Deep Researcher Reflect Evolve)ARXIV ArXiv AI Agent阅读文章 arrow_forward