AI论文速递 2026年08月12日(HuggingFace Daily Papers)¶
数据来源:https://huggingface.co/papers 采集时间:2026-08-12
📌 重点关注¶
- SPIEval: Evaluating Large Language Models as Mobile Assistants over Scattered Personal Information | arXiv — 【重点关注】 Large language models (LLMs) are increasingly deployed as mobile assistants, ... 💡 LLM手机助手评估基准,移动端+AI转型的核心方向
- A^2E : An End-to-End Agent Auditing Engine | arXiv — 【重点关注】 With the rapid advancement of large language models (LLMs), harnesses have be... 💡 端到端Agent审计引擎,是Agent可靠落地的关键基础设施
- Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection | arXiv — 【重点关注】 The malicious use of generative artificial intelligence to create highly real... 💡 多Agent协作做深度伪造检测,多Agent架构实战参考
📋 其他值得关注¶
- Motif 3: Technical Report | arXiv — We introduce Motif 3, a decoder-only Mixture-of-Experts language model with 3...
- Stealing Reasoning Traces from Proprietary LLM APIs | arXiv — Leading large language model providers now conceal their models' step-by-step...
- Agent Memory Distillation: Empowering Small LLM Agents with Hierarchical Teacher Memory | arXiv — Memory systems have shown promise for improving agent performance, but their ...
- RoMeRL: Balancing Feedback Coverage and the Memory-Reward Trap in Self-Evolving Agent Memory via Reduced-Order Utility States | arXiv — Learning-based memory systems for self-evolving LLM agents face two tightly c...
- On-Policy Self-Distillation without Any Supervision | arXiv — On-policy (Self-)Distillation (OPD / OPSD) has shown strong potential for pos...
- The Next Screenshot Knows: Gated Hindsight Distillation for Mobile GUI Agents | arXiv — GUI agents are commonly trained offline from successful interaction trajector...
- OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse Prefetching | arXiv — Large language model (LLM) inference serving is increasingly constrained by m...