📰 科技脉搏
2026-09-10
🏢 公司追踪
Anthropic 最新研究:An alignment assessment of recent cybersecurity incidents…
📅 09-09📎 Anthropic Research
🔥 AI & 科技
研究探索急诊科复诊质量审查中的人类决策过程,并评估AI辅助筛查的潜力。通过分析临床医生的判断模式,揭示人机协作在识别高风险复诊病例中的价值,为优化急诊质量改进流程提供实证依据。
📅 09-09📎 arXiv AI Papers
提出MOONWALK系统,用于动画/VFX前期制作的初级—主管评审流程,通过“意图—证据—行动”对齐来中介协作。它将反馈目标、依据和修改动作结构化关联,提升审查透明度、可追溯性与效率,减少理解偏差和返工。
📅 09-09📎 arXiv AI Papers
Foundation vision-language models (VLMs) exhibit broad intelligence about the world, yet translating this intelligence into robot control remains challenging. We present Show-Harness, an Embodied Harness that enables VLMs to “play” robots through a compact semantic interface linking intent to action
📅 09-09📎 arXiv AI Papers
Enterprises deploy systems, not checkpoints. Usable capability depends jointly on weights, serving route, precision, output contract, and harness, yet all 18 audited benchmarks score advertised model identifiers. We treat this as measurement error and give a protocol that makes it reportable. It has
📅 09-09📎 arXiv AI Papers
Joint-Embedding Predictive Architecture (JEPA) world models learn a compact latent representation of the world that supports prediction and planning, but their capability to learn physics and generate physically realistic dynamics remains hitherto untested. In this work, we introduce SemiGroup-JEPA
📅 09-09📎 arXiv AI Papers
提出JarvisGUI,面向跨设备GUI智能体的动态任务组合框架。它可将复杂用户目标自动拆解并组合为跨设备操作序列,提升异构终端上的自动化执行与泛化能力,为跨设备智能助理提供新思路。
📅 09-09📎 arXiv AI Papers
提出ConvMem卷积记忆机制,用于增强大模型长上下文推理。通过卷积结构压缩与检索历史信息,缓解长序列中的遗忘与计算开销,在长文档理解、多轮推理等任务上提升性能与效率。
📅 09-09📎 arXiv AI Papers
提出面向鲁棒大模型的层选择性遗忘方法,仅对关键层进行参数更新,避免全模型微调带来的能力损伤。该方法能精准移除目标知识,同时保留通用性能,实现更高效、可控的LLM遗忘。
📅 09-09📎 arXiv AI Papers
提出基于本体的记忆生命周期管理框架,解决大语言模型长期对话中的连贯性问题。通过结构化记忆组织与动态更新机制,实现持久上下文保持,提升模型在多轮交互中的一致性与信息延续能力。
📅 09-09📎 arXiv AI Papers
系统评估基础模型在在线内容审核中的表现,对比指令驱动与示例驱动两种策略操作化方式。研究发现示例驱动方法在政策执行一致性和边界案例处理上更具优势,为自动化内容治理提供实践指导。
📅 09-09📎 arXiv AI Papers
Google Gemini 产品更新:4 ways Gemini makes administrative chores quick and easy…
📎 Gemini Updates
💰 财经 & 投资
Ben Thompson 科技战略分析:OpenAI Does Math, Reward-Hacking, Meta Launches Personal Agent…
📎 Stratechery
由 Hermes 自动生成 | 查看历史摘要