AI FRONTIER
Signals worth understanding before they become obvious.
Advancing Private AI Compute with secure, server-side memory
Introducing private, server-side memory to Private AI Compute for personal AI.
This matters because a capability that was previously difficult or impractical is moving closer to real product use.
Cartograph: Federated Tool Discovery with Operator-Attested Retrieval for AI Agents
论文摘要:Cartograph:提出或使用新的基准评测,用来衡量模型在特定任务上的表现。
A different system architecture is emerging here, not just another model-size or benchmark update.
Auditing and Repairing LLM-as-Judge Failures in a Production Text-to-SQL Pipeline
论文摘要:Auditing and Repairing LLM-as-Judge Failures in a Production Text-to-SQL Pipeline:围绕大语言模型的能力、评测或对齐问题展开,适合关注方法与实验结果。
This looks mature enough to test in a real workflow rather than only read about.
Stealth Apart, Harm Together: Skill Cascading Attacks on Skill-Based Agent Systems
论文摘要:A skill is a modular package of natural-language instructions, executable scripts, and reference resources that an agent can load at runtime to extend its capabilities for a s...
A different system architecture is emerging here, not just another model-size or benchmark update.