扶摇AI知识笔记AI 前沿知识库
首页 / 全部文章

全部文章

共 200 篇内容,来自 15 个平台 · 每日同步更新

200 篇内容
InfoQ 中文
大模型基础InfoQ 中文

神秘模型 Union Alpha 突袭!上线首日跑掉20亿Token,部分网友实测称性能直逼 Astra

点击查看原文>

token
InfoQ 中文
智能体应用InfoQ 中文

百度智能云首发产业智能体操作系统,要实现 AI 的商业和技术飞轮

点击查看原文>

智能体智能体操作系统
InfoQ 中文
产业观察InfoQ 中文

9 月 20 日杭州|FDE 在现场:从 AI 组织进化到客户现场,企业 AI 到底卡在哪里?

点击查看原文>

极客公园
智能体应用极客公园

马斯克再暗示合并特斯拉和 SpaceX;传 iPhone 18 Pro 系列卖爆;大疆 Pocket 4P「珠光白」3799 开售|极客早知道

全球首款 AI 智能体手机努比亚 NaviX Ultra 上市:5999 元起,首销销售额「一秒破亿」 9 月 16 日,中兴旗下努比亚今日下午正式发布 NaviX Ultra,官方称其为「…

智能体
InfoQ 中文
大模型基础InfoQ 中文

阶跃发布全新语音大模型 StepAudio 3系列:覆盖语音识别、生成、实时交互与音乐创作

点击查看原文>

大模型
极客公园
物理 AI极客公园

追觅四大赛道 IFA 首秀:一套技术,四个出口

追觅在 IFA 上展出的不是四条产品线,是一套技术的四个出口。 作者|张勇毅 编辑|郑玄 去年一年,全球扫地机器人的出货量约 2412 万台。同一年,全球人形机器人的出货量约 1.8 万台。…

人形机器人具身具身智能
InfoQ 中文
产业观察InfoQ 中文

边说话边推理、边聊天边调用工具,谷歌 Gemini 3.8 Live 要攻克语音 Agent 的沉默时刻

点击查看原文>

雷科技
产业观察雷科技

质量没有捷径!雷军谈AI造车:仿真再强,也要落地真实道路测试

雷军再谈AI造车。

雷锋网
产业观察雷锋网

豆包座舱助手发布,首款合作车型即将开启预售

9月17日,豆包与火山引擎联合打造的AI原生车载助手——豆包座舱助手正式发布。 豆包座舱助手致力于成为聪明懂你、交互自然有温度、安全可靠的座舱 AI 伙伴,带来更加方便、愉悦的驾驶出行体验。…

InfoQ 中文
产业观察InfoQ 中文

让仿真走到设计前端,第叁范式发布国内首款 GPU 原生跨尺度系统级电磁仿真软件

点击查看原文>

gpu
掘金
智能体应用掘金

让 AI 用上你的资料,RAG 是怎么做到的?

下周去上海出差,你想订一家每晚 650 元的酒店,于是问 AI:“这个价格,公司能报销吗?” AI 可能知道出差报销通常要提供哪些凭证。但要回答你这笔住宿费能不能报销,它需要看到你公司的差旅…

rag
雷科技
产业观察雷科技

警惕!大型AI企业正在垄断就业赛道,大学千万别盲目跟风

不然的话,未来的年轻人只会越来越依赖AI。

就业
雷锋网
智能体应用雷锋网

蚂蚁集团成为上海第48届世界技能大赛国家战略合作伙伴

9月22日至27日,第48届世界技能大赛将于上海举行。作为全球规模最大、影响力最广的职业技能赛事,本届大赛共设64个竞赛项目,汇聚了来自全球七十余个国家的顶尖技能人才同台竞技。 作为本届大赛…

就业战略智能体
雷锋网
产业观察雷锋网

蚂蚁集团全员接入千问办公,打造大型企业办公标杆

9月17日,记者获悉,蚂蚁集团正式接入千问办公,作为全公司的智能办公Agent底座,供全员使用,打造大型企业办公的标杆。 图示 蚂蚁集团和千问办公达成合作 蚂蚁集团是全球领先的数字支付与科技…

钛媒体
产业观察钛媒体

聊天即干活,Claude不用再切换入口了

入口合并,额度也被一起合并了。

钛媒体
产业观察钛媒体

端侧AI之战正式打响

端侧AI商业化加速,企业聚焦轻量模型与硬件开发,智能硬件迎来升级浪潮。

商业化
爱范儿
视觉模型爱范儿

用 Seedance 2.5 拍《沙丘 3》,我们替你把 AI 短片的坑全都踩过了

 看到这条视频,你会不会以为《沙丘 3》提前泄露了。 这其实是一部由我们团队的两位同事,花 12 天时间,结 […] #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩…

seedance
雷锋网
产业观察雷锋网

独家解读丨国产GPU四份半年报出炉,赚钱后才发现二级市场「难哄」

早前,壁仞科技、天数智芯、沐曦股份和摩尔线程陆续交出上市后的首份或最新半年报。9月,燧原科技也正式启动申购,国产GPU正在以前所未有的密度进入资本市场。 但当越来越多国产GPU公司开始接受公…

gpu估值商业化
雷锋网
大模型基础雷锋网

独家丨前腾讯混元预训练负责人姚星丞,加入Thinking Machines Lab

雷峰网独家获悉,腾讯混元大语言模型预训练负责人姚星丞,近日已加入北美创业公司Thinking Machines Lab。 据知情人士透露,Thinking Machines Lab此次给姚星…

大语言模型预训练
雷锋网
大模型基础雷锋网

独家丨BAT 天价挖 Gemini 大牛,可惜都挖错了

据雷峰网独家获悉,一位 Google Gemini 的预训练专家,于近期已加入国内某互联网大厂。 该研究员在 Google 的职级或是 L7- L8,接近总监级别,长期在英国工作,入职后主要…

大模型模型训练算力
雷科技
产业观察雷科技

深度思维联合创始人发声:AI发展再快,也绝不能凌驾于安全之上

AI卷速度,也要顾安全。

少数派
产业观察少数派

派早报:佳能发布 EOS R8 Mark II、GPT-5.5 即将下线等

惠普发布 ZBook Ultra G3a 16 移动工作站、影石发布 Mic Pro 腾讯会议版 AI 录音领夹麦等。 查看全文

InfoQ 中文
产业观察InfoQ 中文

汪涛详解华为AI战略:算力为核心,昇腾960提前登场,PB级 KV Cache把基础设施推入新阶段

点击查看原文>

战略算力
钛媒体
物理 AI钛媒体

机器人开始制造机器人了?万台产能背后残酷大考才刚开始

10 分钟下线一台人形机器人,可是量产不等于商业化通关。

人形机器人商业化机器人
钛媒体
物理 AI钛媒体

机器人企业还没拿到“大结果”

一边失血,一边前行。

机器人
雷锋网
大模型基础雷锋网

折叠屏还在比大小,努比亚已经在抢另一个入口

今年9月,手机行业最显眼的变化之一仍然是折叠。 华为继续做三折叠,苹果也终于拿出了自己的折叠屏手机。厂商还在研究一个熟悉的问题:怎样在有限的机身里,装进更大的屏幕。 另一种变化,发生在屏幕里…

大模型
掘金
产业观察掘金

我让 AI 当面试官面了我一轮:第 3 个追问我就卡住了(附 10 道追问清单)

把简历亮点交给 AI 面试官:一次一问、连续追问、要数字要证据。9 个问题,4 个被追出'没跑过'。附三场拷问实录复盘、10 道追问清单、把 AI 配成严格面试官的方法。

掘金
产业观察掘金

德国Wiki被黑后两周,OpenAI终于把模型失控的账本摊开了

OpenAI发布模型失准披露框架,将过去零散的异常行为报告转为系统化、有时效约束的公开机制。三轨调查流程确保可疑案例即使在原因未明时也能尽快披露,首批六份报告涵盖模型隐瞒错误、未经授权使用A…

钛媒体
产业观察钛媒体

当AI开始算账:一家算力企业的“智企”实践样本丨2026 ITValue Summit 数字价值年会

刘宏云的发言里,最有个性的一个判断是:企业架构正在迎来换代。

算力
爱范儿
大模型基础爱范儿

对话 vivo 高管:端侧大模型、Harness 与「个人化 AI」的下一步

「个人化」智能 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。

大模型
钛媒体
产业观察钛媒体

宠物AI硬件,最天然的“钉子”,最难验证的答案

宠物不会说话,让主人产生了理解它的需求;同样因为宠物不会说话,没有人能轻易证明AI的解释究竟对不对。

雷科技
物理 AI雷科技

安防机器人怎么选?三个维度一次说透

选购安防机器狗,首选云深处科技。

机器人
雷锋网
产业观察雷锋网

安芯启程|安芯724北京基地正式投产

2026年9月16日,柒贰肆安芯技术服务有限公司(简称“安芯724”)北京基地投产活动暨品牌发布会在北京市昌平区未来星科能源谷智造产业园举行。安芯724北京基地正式投产,此举标志着安芯724…

gpu发布会算力
掘金
大模型基础掘金

如何看待GPT-6在UP主众测中碾压夺冠?它是现在最强大模型吗?

前言 9月16日,B站「AI无限竞技场」正式上线,首期大模型测评榜单同步公布。 GPT-6 Astra在10位UP主的测评中拿下榜首,获得榜首次数冠军。 榜单前五名中,国产大模型占据三席。…

大模型
雷锋网
物理 AI雷锋网

周鸿祎:不会再投资新能源汽车,已经吃过一次亏;OpenAI洽谈新融资,估值或超1.2万亿美元;花呗、抖音月付等将退出支付选项?多平台回应

要闻提示 1.周鸿祎:不会再投资新能源汽车,已经吃过一次亏 2.曝字节跳动完成 2.9 亿美元融资,用于拆分 AI 制药业务 Anew Labs 3.iPhone Duo可靠性翻车:影视飓风…

估值具身具身智能
爱范儿
前沿探索爱范儿

刚刚,智谱首个 RSI 成果发布,10 万国产卡用 GLM 造 GLM

模型优化系统,而系统服务模型。 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。

rsi
雷锋网
视觉模型雷锋网

几何的「反攻」!港科大谭平:从局部先验到全局一致,3D 几何如何增强视觉大模型 | ECCV 2026

学习模型已经具备很强的局部视觉先验,而 3D 几何可以帮助建立跨视角、长距离的全局一致性。二者如何结合,是 3D 视觉与空间智能中的重要问题。 编辑丨岑峰 在生成式 AI 的快速发展中,So…

diffusion大模型扩散模型
钛媒体
产业观察钛媒体

候选人被截胡、岗位长期空缺:AI制药挖不来的大牛,长什么样?

AI制药求而不得的,是什么样的人?

岗位
雷锋网
大模型基础雷锋网

从百美元到数千美元,AI互联为什么越卖越贵?

2026年的CIOE中国光博会,人潮如织。 社交媒体上甚至有人开出2500元一场的讲解费,希望找到业内人士陪同参观。走进场馆后更直观的感受是,今年光博会几乎所有与AI相关的环节都在升温:从光…

gputoken大语言模型
InfoQ 中文
视觉模型InfoQ 中文

从概率生成到稳定交付:AIGC专业内容生产的工程化挑战

点击查看原文>

aigc
钛媒体
产业观察钛媒体

从130亿美元到1050亿美元,张一鸣的钱为什么越来越多

张一鸣卸任经营一线,账面身家却大幅攀升。财富增长根植于字节跳动的业务跃迁与AI估值重估,未上市企业的账面财富背后,同样暗藏预期波动的现实考验。

估值
极客公园
物理 AI极客公园

京东押注物理 AI,冲在前面的是一群 95 后

去年,AI 行业最热闹的是模型榜单、推理能力和多模态生成。所有人都在问:这个模型有多强? 但到了 2026 年,问题变了。人们开始追问:它到底能干什么? 这个问题,京东在今年的 JDD 大会…

具身具身智能多模态
InfoQ 中文
产业观察InfoQ 中文

亚马逊 CTO 要来中国了,说实话,这比泛滥的 AI 发布会值得关注

点击查看原文>

发布会
雷锋网
物理 AI雷锋网

九识建成首个L4万卡集群,无人驾驶进入多模态大模型新范式

一家做无人车的公司,为什么需要一万多张GPU? 9月10日,九识CEO孔旗在广州宣布战略升级为“城市级物理AI”。一周后的9月17日,CTO庄立首次披露支撑这次升级的底层能力:九识已建成业内…

gpurobot多模态
掘金
大模型基础掘金

为什么你的 AI 越聊越“笨”,还越来越慢?

你有没有发现,Agent 刚开始还挺聪明,聊久了却越来越慢、消耗的 token 越来越多,甚至开始忘记前面说过的要求? 比如,国庆假期快到了,你让 Agent 帮你规划一次旅行。目的地、日期…

token
雷科技
智能体应用雷科技

个体网支付宝小程序及智能体亮相,为个体工商户提供政策查询等服务

个体网支付宝小程序及个体网智能体首次公开亮相并开放使用。

政策智能体
钛媒体
视觉模型钛媒体

世界模型入口战:字节向左,京东向右

世界模型带来的不仅是技术变革,更是人机交互方式的根本转变。

世界模型
雷锋网
物理 AI雷锋网

专访零跑周洪涛:2026年冲刺智驾头部,零跑的进度条比想象中快

朱江明很少会在午休时间离开办公室。 但在体验智驾新版本后的一个中午,他又独自一人,开着搭载智驾 4.0的D19在杭州市区痛快跑了一个多小时。在公司内部群,他以少有的振奋接连赞扬了智驾团队。…

世界模型战略智驾
钛媒体
产业观察钛媒体

万亿市场背后,AI数据基础设施的竞赛已经开始

模型能力快速逼近应用需求的同时,数据正在成为企业AI规模化落地的新瓶颈。

掘金
智能体应用掘金

一天一个开源项目(第220篇):WeKnora —— 腾讯开源的企业级知识框架,从 RAG 问答到 Wiki 自进化

WeKnora 是腾讯开源的 LLM 知识框架,融合 RAG 快速问答、ReAct 智能体多步推理和 Wiki 模式自动生成互联知识库三大能力,支持 20+ LLM 提供商、10+ 文档格式…

llmrag 开源
InfoQ 中文
产业观察InfoQ 中文

vivo 把 Agent 做进操作系统:6000 多项原子技能开放调用,AgentOS 预览版亮相

点击查看原文>

掘金
智能体应用掘金

[Agent Memory / 强化学习] MemPO源码学习笔记 --- (3)--- Rollout思路

[Agent Memory / 强化学习] MemPO源码学习笔记 --- (3)--- Rollout思路 0x00 概要 0x01 Rollout 主要内容 1.1 时间线 1.2 纯…

agent memorymemory
InfoQ 中文
智能体应用InfoQ 中文

Snowflake World Tour 上海站 Keynote——Snowflake 企业级智能体

点击查看原文>

智能体
InfoQ 中文
产业观察InfoQ 中文

Linux Foundation CEO:AI史上最大投资潮背后,真正托底的是开源

点击查看原文>

开源
掘金
产业观察掘金

ERP排程总在纸上谈兵?JVS-APS如何用真实产能约束打通计划与执行闭环

问题背景:为什么ERP排程常成‘理想模型’? ERP原生排程模块普遍采用无限产能假设——仅依据BOM结构、标准工时和工序逻辑推算理论完工时间,未将设备可用性、人员排班、模具占用、换型耗时等真…

掘金
大模型基础掘金

DeepSeek Harness 系列(09):可观测性——怎么知道 Agent 在干什么

Token 消耗多少?哪步最慢?出错了在哪里?这篇文章讲 dsh 的可观测性机制:Session 事件日志、Token 计量、遥测 Seam,OpenTelemetry

token
掘金
产业观察掘金

DeepSeek Harness 的插件,到底该怎么用?

DeepSeek Harness 发布不到两天,GitHub Star 就逼近了 10 万。 一个 Agent Harness 能获得这么高的关注度,我自然也想看看它到底有什么不同。 它有一…

爱范儿
产业观察爱范儿

ColorOS 17 发布,OPPO 想让 AI 往前一步,主动一些聪明一些

手机操作系统里的 AI,正在从被动走向主动。 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。

钛媒体
产业观察钛媒体

Altman 对话 Benioff:开源模型会不会失控

当开源模型的能力逼近闭源,而监管还没准备好,下一次逃逸,谁负责?

开源监管
InfoQ 中文
大模型基础InfoQ 中文

Agent 的经济账,不能只算 Token——阿里用 Qoder Cloud Agents 给出答案

点击查看原文>

token
钛媒体
产业观察钛媒体

AI赢得了芯片,输不起一次停电

算力的尽头是电力。

算力芯片
掘金
大模型基础掘金

AI用量v0.1.11更新发布 新增 jusage doctor 诊断指令 托盘展示token 和余额 新增 AutoClaw 支持

🚀 本次更新 🖥️ Desktop ✨ macOS 菜单栏托盘默认同时显示今日 Token 和金额,可在「应用设置」中选择仅 Token、仅金额或关闭。#157 @Xxcool ✨ 新增 A…

token
雷锋网
智能体应用雷锋网

AI办公之外,百度智能云为何盯上「产业智能体操作系统」?

“整个下半年,就看AI办公圈谁能杀出重围了”。一位大模型从业者对雷峰网表示。 从五月开始,各个互联网大厂争相出场,写字楼电梯间、机场候机室,到处都是Workbuddy、千问办公、豆包工作、百…

大模型智能体智能体操作系统
掘金
产业观察掘金

AI 技能地图

AI 工具更新得太快了。 很多时候,不是我们不愿意学,而是不知道该先学什么。 最近,吴恩达老师分享了一张 AI 工程技能地图,刚好回答了这个问题。 他的团队分析了 10,000 多个招聘职位…

雷科技
物理 AI雷科技

4亿美元到账!地瓜C轮融资成功,机器人行业的“高通”来了

机器人基础设施开始“标准化”。

机器人融资
雷科技
产业观察雷科技

魅族前高管新动态!PANDAER不甘只做配件,AI硬件才是星辰大海

PANDAER不再只是配件厂。

雷锋网
物理 AI雷锋网

被AI放大的光学「戏份」,让「配角」舜宇的生意越做越大

作者 | 陈悦琳 编辑 | 郭思 几乎同时迎来“量产元年”的人形机器人和智能眼镜,正在放大光学的“戏份”。 据Counterpoint统计,今年上半年全球人形机器人出货量突破2.2万台,同比…

人形机器人战略机器人
雷锋网
大模型基础雷锋网

芯片从业者拆解OpenAI造芯,还能再快3个月?

量产一代、研发一代、预研一代,这一芯片公司习以为常的战略节奏,如今正被大模型公司压得更紧。 按照OpenAI今年8月公布的进展,尽管第一代自研芯片Jalapeño(下以“小辣椒”代称)计划于…

大模型战略芯片
极客公园
产业观察极客公园

腾讯、字节、阿里「会战」AI 办公之后:Agent 领域格局已变

「这仗打错了,最多是白烧几十亿;但如果不打,可能是生死问题。」 这是创新工场联合首席执行官、管理合伙人汪华对大厂重金押注 AI 办公的判断。 2026 年的这个夏天,中国科技巨头发起了 Ag…

雷锋网
产业观察雷锋网

碧桂园服务与支付宝全面深化合作,共筑“智慧物业”数字化新标杆

2026年9月15日,碧桂园服务与支付宝正式签署合作协议。双方宣布将围绕物业服务升级、年终缴费大促及智慧社区建设等领域开展深度合作。此次合作标志着双方合作全面升级,通过数字化技术为业主打造更…

极客公园
物理 AI极客公园

特斯拉、Figure 还在攻克量产,小鹏机器人已经走下产线

9 月 8 日,一台高阶通用人形机器人小鹏 IRON 完成自动化总装,从产线上走了下来。 如果只看画面,这好像是又一段科技公司放出来的机器人视频。在这个被大模型和具身智能席卷的时代,科技圈每…

robot人形机器人具身
极客公园
物理 AI极客公园

没有方向盘、没有踏板、没有后视镜:特斯拉最疯狂的车来了

9 月 3 日,特斯拉「悄悄」办了一场发布会。 没有直播,没有全球媒体轰炸,只请了少数受邀者参加。活动开始前,特斯拉的网红粉丝们发布了照片和视频。特斯拉还在网站上设置了倒计时,为该活动预热。…

估值发布会自动驾驶
爱范儿
产业观察爱范儿

早报|雷军同日到访宇树与B站/罗永浩差评带来流量,野人先生单日涨粉近3万/鸿蒙智行确认问界合作调整,赛力斯主导

· Google 向全体工程师开放 Claude,Gemini 仍是默认模型 · 鸿蒙智行确认问界合作模式调整,赛力斯主导五项业务 · Waymo 计划 2027 年在东京推出全无人出租车…

爱范儿
产业观察爱范儿

我和我的 AI 手机吃了 3 顿饭

把 AI 计算放在最适合计算 AI 的地方去。 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。

极客公园
智能体应用极客公园

当智能体开始替人花钱,如何证明「它是谁」?

在蚂蚁集团,负责 AI 支付安全的陈树鹏是个深度的智能体使用者。日常写代码、梳理汇报、提炼会议纪要,AI 几乎接管了他工作流里的大部分环节。 这种高度依赖背后,却存在着一个清晰的分水岭。一旦…

智能体
爱范儿
物理 AI爱范儿

对话蚂蚁灵波 CEO 朱兴:机器人还吃不了「粗粮」

具身智能会有自己的「ChatGPT 时刻」吗 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。

具身具身智能机器人
雷锋网
大模型基础雷锋网

基元律动与无问芯穹达成战略合作,推进高质量Token供给与应用

2026年 9 月 15 日,AI基础设施公司基元律动(TokenRhythm)与无问芯穹(Infinigence AI)签署战略合作协议。双方将结合无问芯穹的Agentic Infra产品…

agentictoken开源
极客公园
智能体应用极客公园

在飞书的上下文底座上,豆包开工了

今年,不少公司的工作群里,将会多一些「新同事」。 它们是 AI,但可以像普通员工一样被拉进群聊,阅读此前的讨论和文档,参与会议、承接任务,并调用不同的业务工具完成工作。你甚至可以看到 Age…

智能体
雷锋网
产业观察雷锋网

华为坤灵升级“4+10+N”场景化方案,发布“经纬计划”和50个样板点

[中国,上海,2026年9月16日] 今日,华为坤灵新品发布会2026在上海举行。华为常务董事、ICT BG CEO杨超 斌宣布,华为坤灵全面升级 “4+10+N” 一站式场景化方案,发布2…

发布会算力
极客公园
产业观察极客公园

助听器躺赚三十年暴利,被 AI 打破了

一个塞在耳朵里的小东西,卖到十万一副,丢失甚至需要动用全城搜索? 助听器,是一门奇特的生意。 整个全球市场,近九成的助听器市场份额被五家外资公司把持:索诺瓦(峰力)、Demant(奥迪康)、…

市场份额芯片
雷科技
产业观察雷科技

助力中小企业AI化!华为坤灵双管齐下:一站式场景方案+高密度服务网

华为坤灵让智能世界真·触手可及。

arXiv cs.AI
物理 AIarXiv cs.AI

rMuscle: Robotic Muscle Memory for Efficient Vision-Language-Action Model Inference

Factory work is a promising early scenario for embodied AI: assigning repetitive manual jobs to…

embodiedmemoryrobot
arXiv cs.CV
产业观察arXiv cs.CV

Zero-Shot Cross-Lingual Recognition of Sign Language Handshapes

Sign language processing advances rapidly for high-resource languages such as American Sign Lan…

arXiv cs.LG
产业观察arXiv cs.LG

WaveTLM: Reliable Time-Series Language Modeling through Task Compilation

Time-series language models provide a shared natural-language interface across temporal tasks,…

arXiv cs.CV
产业观察arXiv cs.CV

Video-Based Markerless Motion Capture for Clinical and Rehabilitation Biomechanics: A PRISMA-ScR Scoping Review of Validated Architectures, Clinical Readiness, and Emerging Methods

Background.. Video-based markerless motion capture promises movement analysis without the cost,…

arXiv cs.CV
视觉模型arXiv cs.CV

VibeAvatar: Aligning Phonetic Kinematics and Human Aesthetics for High-Fidelity Talking Avatar Synthesis

Multi-modal talking avatar synthesis aims to generate realistic talking videos from a reference…

diffusion
arXiv cs.CV
物理 AIarXiv cs.CV

Track, Articulate, Act: Generating Articulation from Casual Human Videos

Human videos contain rich causal evidence for robot manipulation: they reveal how hand motion i…

embodiedrobot
arXiv cs.RO
智能体应用arXiv cs.RO

Towards Interaction Regulation from Human Feedback via Free Energy Minimization

A central challenge across control and learning is the design of mechanisms regulating the inte…

autonomous agent
arXiv cs.CV
大模型基础arXiv cs.CV

Toward Markerless Video-based Tremor Analysis: Objective Quantification of Pathological Tremor in Mouse Preclinical Models

Tremor is a movement disorder characterized by involuntary, rhythmic oscillations of body parts…

llm
极客公园
物理 AI极客公园

Token 之后,谁来组织 AI 计算?Arm 寻找下一代计算的答案

如果说今年 3 月,Arm 推出首款自研 CPU 芯片,是这家公司向市场释放的一个信号,它开始尝试突破过去「只提供底层架构授权」的角色边界。那么在上海举办的 Arm Everywhere C…

token物理 ai物理世界
arXiv cs.CL
产业观察arXiv cs.CL

TeleAntiFraud 2.0: A Refreshable, Profile-Grounded, and Audio-Based Benchmark for Telecom Fraud Detection

Telecom fraud scripts evolve rapidly and are often designed to resemble routine service convers…

arXiv cs.AI
智能体应用arXiv cs.AI

Taming the Agentic RAN: Stability-Guaranteed Arbitration of Autonomous AI Agents in O-RAN

The O-RAN control plane is becoming agentic: autonomous AI agents, deployed as rApps by differe…

agenticrsi
arXiv cs.CL
产业观察arXiv cs.CL

Structured Claim-Level Discourse Representations for Dense Health Narratives

Health discourse in social media videos often contains densely entangled claims spanning multip…

arXiv cs.AI
智能体应用arXiv cs.AI

Social Laws for Multi-agent Coordination in Stochastic Environments

In multi-agent environments, coordinating agents to prevent interference and ensure robust indi…

multi-agent
arXiv cs.AI
产业观察arXiv cs.AI

Securing quantum error correction against misleading advice from AI agents

Can an attacker turn influence over an artificial intelligence (AI) adviser into a harmful quan…

arXiv cs.CL
产业观察arXiv cs.CL

ScienceIDE: Turning World's Scientific Codebase into Agent Learnable Environments

Scientific code repositories encode decades of human knowledge in executable models, methods, a…

arXiv cs.RO
产业观察arXiv cs.RO

SOL-SLAM: Inverse Compositional Gauss-Newton Direct Registration for Fast Sonar-Only Local SLAM

Autonomous underwater navigation typically relies on complex and expensive multi-modal sensor s…

arXiv cs.RO
产业观察arXiv cs.RO

SEAM: Submap-Anchored Evidence for Lifelong LiDAR Mapping under Trajectory Deformation

We propose SEAM, a LiDAR-based lifelong mapping framework. Instead of relying on a single ancho…

arXiv cs.AI
产业观察arXiv cs.AI

Reporting Practice Matters: The Impact of Reference Choice on Chest X-ray Report Evaluation

Radiologists follow heterogeneous reporting practices. Two radiologists examining the same imag…

arXiv cs.CV
视觉模型arXiv cs.CV

ReFigBench: Benchmarking Scientific Figure Reconstruction as Editable PowerPoint Artifacts

Multimodal coding agents are expected to turn visual inputs into usable artifacts, and they act…

multimodal
arXiv cs.CV
视觉模型arXiv cs.CV

RankGround: Efficient High-Resolution GUI Grounding via Lightweight Reranker-Guided Crop Selection

Graphical User Interface (GUI) grounding is a fundamental perception task for multimodal agents…

multimodalvlm
arXiv cs.AI
产业观察arXiv cs.AI

RLLBC-Lib: An Educational Code Library for Reinforcement Learning and Learning-Based Control

Reinforcement learning (RL) is an exciting concept as well as a remarkable success story worth…

arXiv cs.AI
产业观察arXiv cs.AI

Probabilistic Linear Explanations

Formal explainability provides mathematically grounded justifications for individual prediction…

arXiv cs.LG
大模型基础arXiv cs.LG

Preventing Model Collapse: A Fisher-Rao Perspective on the Dynamics of Training with Synthetic Data

Large Language Models (LLMs) are now routinely trained using synthetic data, since high-quality…

llmrsi
arXiv cs.AI
前沿探索arXiv cs.AI

Prepared Or Unprepared? Evaluating Healthcare Workforce Readiness for Clinical Adoption of Artificial Intelligence in Nigeria

Artificial intelligence (AI) is increasingly integrated into healthcare systems worldwide, yet…

rsi
arXiv cs.CV
物理 AIarXiv cs.CV

PointZero: 3D Point Track Completion for Learning Transferable 3D Dynamics

World models endow perceptual systems with the ability to predict how scenes evolve under inter…

robot
arXiv cs.CL
产业观察arXiv cs.CL

Playing log(N)-Questions over Wikipedia Abstracts: Communication Efficiency Between Paired Frontier Models

We evaluate six frontier language models on the two-agent $\log(N)$-Questions game. A questione…

arXiv cs.LG
产业观察arXiv cs.LG

Physics-based prediction, uncertainty quantification and decision-making for IN718 crystallographic texture intensity across LPBF defocus regimes

Reliable prediction of crystallographic texture in laser powder bed fusion is critical for link…

arXiv cs.CV
物理 AIarXiv cs.CV

PhysVGGT: Feed-Forward Dense Physical Property Estimation from A Single Image

Physical properties, such as friction, hardness, stiffness, and density, govern how robots shou…

robot
arXiv cs.CL
产业观察arXiv cs.CL

PersonaPath: Towards Knowledge-Centric Personalized Learning Path Planning

Adaptive learning systems commonly formulate learning path planning as Exercise-Centric (EC) re…

arXiv cs.CV
智能体应用arXiv cs.CV

PULSE: Unlocking Practical Image Compression on Single-Thread CPU

Despite recent progress in learned image compression, existing methods remain computationally e…

agentic
arXiv cs.CV
视觉模型arXiv cs.CV

PANORAMA: Panoptic Grounded Captioning via Mask Proposal Selection

Intelligent systems that act in the world require image understanding that is both comprehensiv…

vlm
极客公园
产业观察极客公园

OpenAI、Anthropic、谷歌联手研究 AI 安全;微信员工辟谣 AI 助手小微「偷隐私」;美国已在太空部署武器

因加大 AI 投入 报道称字节跳动 2026 上半年净利润下降 据报道,由于加大了在人工智能领域的投资,字节跳动 2026 年上半年净利润同比出现下降。同时,受 TikTok 等海外业务推动…

营收隐私
arXiv cs.AI
大模型基础arXiv cs.AI

Objective vs. Search: Decomposing What Makes a Good Tokeniser

Two dominant tokenisation algorithms are used by modern language models: byte-pair encoding (BP…

token
arXiv cs.CV
视觉模型arXiv cs.CV

NormLift: From Lifted Features To Semantic Reliability In 3D Gaussian Splatting

Training-free weighted aggregation is widely used to lift 2D semantic features onto 3D Gaussian…

gaussian splatting
arXiv cs.CL
大模型基础arXiv cs.CL

Monitoring and Discovering Reward Hacking with Internal Representations during LLM Evaluations

As models scale, reward hacking becomes more frequent, more sophisticated, and more consequenti…

llm
arXiv cs.CV
产业观察arXiv cs.CV

Mask IPL: Noise-Free Intrinsic Position Learning via Computation Graph Clipping for Event-Based Spike-Driven Tracking

Spiking Neural Networks (SNNs) match the event-driven nature of event cameras and naturally ext…

arXiv cs.AI
产业观察arXiv cs.AI

MUSE: Benchmarking Large Vision-Language Models on Multi-Modal Understanding in Situated Education

Large vision-language models have achieved remarkable progress in multi-modal understanding, ye…

arXiv cs.CL
智能体应用arXiv cs.CL

Long-Lived Characters, Local Inference: Incremental Memory Maintenance for Game NPCs

A game character should not have to reread its entire life before every conversation. For local…

memory
arXiv cs.CV
大模型基础arXiv cs.CV

Learning Where to Focus: Self-Supervised Multi-Scale ViTs for Histopathology

Pathologists diagnose diseases by first locating suspicious tissue and then examining it at hig…

tokentransformer
arXiv cs.LG
产业观察arXiv cs.LG

Learning Lyapunov Operators for Nonlinear Systems

Constructing Lyapunov functions for nonlinear dynamical systems is a central problem in stabili…

arXiv cs.RO
物理 AIarXiv cs.RO

Learning Holistic Whole-Body Loco-Manipulation with a Bipedal Mobile Manipulator

Bipedal loco-manipulation enables robots to interact with objects beyond the nominal workspace…

robot
arXiv cs.RO
物理 AIarXiv cs.RO

KINO: A Keyframe Interface for VLM Planning and Whole-Body Control in Humanoid Loco-Manipulation

Humanoid loco-manipulation requires robots to interpret task instructions and scene semantics w…

humanoidrobotvlm
arXiv cs.LG
产业观察arXiv cs.LG

Interpretable Multi-Instance Learning Enables Early Prediction of Key Molecular Alterations from Routine Flow Cytometry in Acute Myeloid Leukemia

Background: Molecular testing for NPM1 and FLT3-ITD mutations guides critical early treatment d…

arXiv cs.CV
物理 AIarXiv cs.CV

In-Context Robot Learning with VLM Agents

Enabling robots to adapt to unfamiliar environments as readily as humans remains a moonshot goa…

agenticembodiedrobot
arXiv cs.CL
产业观察arXiv cs.CL

How Much is a Human Right Worth? ECtHR-NPD: A Benchmark for Predicting Non-Pecuniary Damage Awards

Existing legal benchmarks cover diverse tasks, while continuous monetary remedies remain compar…

arXiv cs.LG
前沿探索arXiv cs.LG

How Model Growth, Recursion, and Boundary Operators Influence Scaling Exponents

Scaling laws predict how loss decreases with increases in computation. We show, contrary to con…

rsitransformer
arXiv cs.AI
大模型基础arXiv cs.AI

Higher-order pruning of experts in mixture-of-experts language models

Mixture-of-Experts (MoE) language models suffer from large parameter counts, which create a sig…

memorymoe
arXiv cs.CV
视觉模型arXiv cs.CV

Geometry beneath the Waves: Dense Priors for Sparse-View Underwater 3D Gaussian Splatting

Underwater 3D reconstruction supports applications ranging from marine ecosystem monitoring and…

gaussian splattingrsi
arXiv cs.AI
产业观察arXiv cs.AI

Flag Game: A Toy Model for Mechanistic Swarm Interpretability

Emergent coordinated behaviors of AI agents are starting to present critical safety risks. A ke…

arXiv cs.LG
产业观察arXiv cs.LG

Fast Learning Rates for Physics-Informed Kernel Methods

In physics-informed machine learning, a target function $u^*$ is learned from noisy value obser…

arXiv cs.CL
大模型基础arXiv cs.CL

FRAUDSkill: Structured Frozen-Weight Skill Optimization for Audio Anti-Fraud Detection

Large audio-language models have shown promise for anti-fraud detection by directly processing…

fine-tun
arXiv cs.CV
智能体应用arXiv cs.CV

FIVE-VLA: Fast and EffectIVE Autonomous Driving with Recurrent Action Memory

State-of-the-art vision-language-action models (VLA) for autonomous driving face critical limit…

memorytoken
arXiv cs.LG
智能体应用arXiv cs.LG

Exponential Hardness of Off-Policy Evaluation under History-Dependent Logging

Can a logged dataset visit every hidden state frequently and still be exponentially uninformati…

memory
arXiv cs.RO
物理 AIarXiv cs.RO

Examining the Difference in Human Behavior Between Virtual and Real-World Human-Robot Teaming

Prototyping and evaluating human-robot teaming (HRT) scenarios in the real-world is costly. Vir…

robot
arXiv cs.LG
智能体应用arXiv cs.LG

Evidence-Grounded Agentic Formulation Development in an Autonomous Laboratory

Self-emulsifying drug delivery systems (SEDDS) can improve the oral bioavailability of poorly s…

agentic
arXiv cs.CL
大模型基础arXiv cs.CL

EviGen: Predictive Evidence Scaffolding for Verifiable Clinical Rationale Generation

Longitudinal electronic health records (EHRs) capture years of patient history across notes, co…

llm
arXiv cs.RO
物理 AIarXiv cs.RO

ElastiQP: An Always-Feasible QP Solver for Constrained Robot Control

As robot capabilities increase, quadratic programming (QP)-based controllers must account for a…

robot
arXiv cs.AI
前沿探索arXiv cs.AI

Dreaming the Sound of Contact: Leveraging Video and Audio Generation for Zero-Shot Force-Aware Manipulation and Data Generation

Recent advances in video generation allow robots to learn manipulation trajectories from genera…

agirobot
arXiv cs.AI
产业观察arXiv cs.AI

Double descent is the principle of least action

The test error of a model plotted against its number of parameters $d$ falls, peaks when the mo…

arXiv cs.AI
产业观察arXiv cs.AI

Decodable but Misrouted: Sparse Features Uncover a Readout Gap in Vision-Language Models for Harmful Meme Detection

When a large vision-language model misclassifies a harmful meme, the failure may reflect missin…

arXiv cs.CV
前沿探索arXiv cs.CV

DISTA-Net++: Rethinking Infrared Small Target Unmixing Beyond Sub-Pixel Separation

Long-range infrared imaging frequently confronts dense target clusters whose diffraction-limite…

agi
arXiv cs.CV
产业观察arXiv cs.CV

Copy What Is Seen, Generate What Is Not: Training-Free Anomaly-Aware Video Restoration

A surveillance system that detects an anomaly often has to repair the footage as well, yet the…

arXiv cs.LG
视觉模型arXiv cs.LG

Comprehensive reconstruction of collider events with hypergraph representation learning and graph-conditioned diffusion

In particle collider experiments, event reconstruction is the task of inferring the kinematics…

diffusion
arXiv cs.AI
智能体应用arXiv cs.AI

Cognitive Extensions for Dual-Process Language Agents: Memory and Self-Reflection in Interactive Environments

Language agents remain brittle in interactive environments, where success requires long-horizon…

memory
arXiv cs.RO
物理 AIarXiv cs.RO

CaSCo: Cascade-Aware Soft-Collision Motion Planning

Conventional motion planning treats collision as a binary constraint, although contact with dif…

robot
arXiv cs.RO
产业观察arXiv cs.RO

Body-Motion Control of a Simulated Aerial Swarm from a First-Person View

First-person-view (FPV) teleoperation of aerial swarms requires an operator to coordinate colle…

arXiv cs.CL
产业观察arXiv cs.CL

Beyond frequency measures: Can contextual embeddings capture meaning change in scientific texts?

Identifying technological trends is a core scientometric task, yet traditional frequency-based…

arXiv cs.AI
智能体应用arXiv cs.AI

Beyond Outcomes: Dual-View Relational Learning for Efficient Agent Benchmarking

Agent benchmarks are substantially more costly to evaluate than conventional LLM benchmarks. Be…

agenticllm
arXiv cs.RO
物理 AIarXiv cs.RO

Asymptotically Optimal Multi-Robot Task and Motion Planning

Multi-robot task and motion planning (MR-TAMP) requires jointly reasoning about discrete task d…

robot
arXiv cs.AI
产业观察arXiv cs.AI

Affora: A Design System for Agent-Friendly Interfaces

Computer-use agents increasingly operate software designed for people, but interfaces often lea…

arXiv cs.CV
前沿探索arXiv cs.CV

Adaptive Convolutional Sparse Coding via Information Bottleneck for Robust Visual Signal Representation

Visual signals require compact yet sufficient representations for robust downstream prediction.…

rsi
arXiv cs.RO
视觉模型arXiv cs.RO

AdaGeoVLN: Selective Geometry Across Representation Depth and Navigation Time for Vision-Language Navigation

Vision-language navigation requires aligning language with visual observations while maintainin…

vlm
arXiv cs.AI
大模型基础arXiv cs.AI

ASLEval: Measuring Privacy Exposure Displacement in LLM Agent Sessions

Privacy evaluations of tool-using LLM agents often inspect a designated action, final response,…

llm
雷科技
产业观察雷科技

AI交易分化,全球AI基建龙头联想集团锚定未来确定性

全球AI基建龙头联想集团的“确定性”答案。

arXiv cs.MA
产业观察arXiv cs.MA

ABM-SIRTEM: A Hybrid Agent-Based and Epidemiological Model for Pandemic Response

The COVID-19 pandemic has had profound impacts on global health, social structures, and economi…

arXiv cs.AI
大模型基础arXiv cs.AI

A Zeroth-Order Paradigm for LLM Preference Alignment

Direct preference alignment methods are widely used to align large language models (LLMs) with…

llmmemory
arXiv cs.LG
产业观察arXiv cs.LG

A General Kernel Framework for Non-CND Distance Measures Using |D|-Dimensional Sparse Landmark Embeddings

Kernel methods, and Gaussian Processes (GPs) in particular, require a Hilbertian distance measu…

雷锋网
智能体应用雷锋网

2nm天玑9600 Pro,把旗舰SoC竞争推向「融合计算」

智能体让旗舰SoC从拼核心,转向拼系统。 过去数年,旗舰手机芯片的升级路径非常清晰,制程更先进、CPU主频更高、GPU、ISP更强,再加上算力不断增长的NPU。 这套方法在生成式AI进入手机…

gpu发布会智能体
arXiv cs.RO
产业观察arXiv cs.RO

"What's going to happen after I'm gone?": Parent Perspectives on Technology in Supporting Independent Living for Adults with Intellectual Disabilities

Adults with intellectual and developmental disabilities (IDD) are increasingly transitioning fr…

极客公园
物理 AI极客公园

造物 100 #06|自动驾驶上轮椅了,口袋相机学会飞行,AI 教练上了雪场

硬件创新力正在消失。 9 月 10 日,苹果发布了筹备多年的首款折叠屏 iPhone Duo,又是一款集大成之作,困扰多年的折叠屏折痕有了「苹果」解法。但我们明显发现硬件创新的品类迭代速度开…

无人机自动驾驶
爱范儿
产业观察爱范儿

豆包工作和飞书,把中国第一个团队 Agent 拉进了工作群

为团队而生的办公 Agent #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。

极客公园
智能体应用极客公园

给 AI 发外设,戴森派摄像头进嘴、绿联给充电宝开了扇窗

AI 正在把所有硬件重新做一遍,这句话说了三年,这周轮到了一把牙刷。 今天,IFA 柏林开展,一千九百个品牌都在不同角度展示自己的创新力,连戴森都掏出了一款牙刷来告诉行业,我也能用 AI 玩…

智能体生态
爱范儿
产业观察爱范儿

早报|iOS27正式推送/比亚迪高管:燃油车在中国没有未来/iPhone 18 Pro渠道降价900元,苹果称不干预

· 广汽一汽签署重组意向协议,一汽拟成第二大股东 · 中国地震台网中心:正与苹果沟通 iOS 地震预警 · 马斯克旗下公司撤回对苹果的反垄断诉讼,继续起诉 OpenAI #欢迎关注爱范儿官方…

极客公园
产业观察极客公园

对话小宇宙 kyth:播客的护城河是真实,AI 无法取代的是人的立场

头图来源:小宇宙 过去一年,播客被推到了内容行业的聚光灯下。 B 站拿出 10 亿元流量扶持视频播客,罗永浩、鲁豫等名人带着数小时的长对谈进入市场。小红书也在持续补齐视频与音频播客能力,抖音…

商业化生态
爱范儿
前沿探索爱范儿

号称人类造的最后一个 AI 要来了,刷屏全网的 RSI 是什么

AI 开始造 AI,但距离「智能爆炸」还有多远? #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。

rsi
极客公园
前沿探索极客公园

iOS 27、MacOS 27 正式发布;豆包手机助手消费者版亮相;李想:「大车」趋势一定会结束

特朗普抨击 Anthropic CEO:AI 发展不能踩刹车,美国有「高智商总统」就足够 特朗普周一在社交媒体发文,明确表示其不认可 AI 数据中心引发的民众反弹及外界对前沿模型的忧虑。他写…

前沿数据中心监管
爱范儿
产业观察爱范儿

iOS 27 正式版体验:Siri AI 终于开窍了,老 iPhone 升级也有新东西

值得升级。 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。

arXiv cs.MA
智能体应用arXiv cs.MA

ToMAS: A Pilot Failure-Grounded Theory-of-Mind Benchmark from Multi-Agent LLM Failures

LLM-based multi-agent systems can fail even when communication succeeds because agents do not c…

llmmulti-agentrsi
arXiv cs.MA
智能体应用arXiv cs.MA

Skill-based Agentic Evaluation for Real-time Data Science Tasks

We present a framework for evaluating data-science agents on live, continuously updated data us…

agenticllm
arXiv cs.MA
产业观察arXiv cs.MA

Set-membership localization of intermittent RF sources using a fleet of collaborating UAVs

This paper proposes a set-membership approach (SMA) to localize radio frequency (RF) sources ob…

arXiv cs.MA
产业观察arXiv cs.MA

PaperDoctor: Evidence-Grounded and Actionable Feedback for Scientific Papers in Progress

Autoresearch agents are reshaping the research ecosystem, but they can also let flawed claims e…

arXiv cs.MA
智能体应用arXiv cs.MA

Multi-Agent Learning with Cooperation-Driven Optimization Dynamics

Multilayer Artificial Neural Networks trained via backpropagation are the basic blocks of many,…

multi-agent
arXiv cs.MA
智能体应用arXiv cs.MA

Mo' Models, Mo' Problems: How to best select model pools when designing Multi-Agent Systems

Multi-agent Systems (MAS) combine multiple model outputs to solve complex reasoning tasks. Howe…

llmmulti-agentrsi
arXiv cs.MA
产业观察arXiv cs.MA

Intervention problems in the Linear Threshold Model: A general formulation and new results

We study an optimal intervention problem for linear threshold models. This is a popular class o…

arXiv cs.MA
物理 AIarXiv cs.MA

Exact Fusion and Coordinated Exploration in Multi-Robot Active Inference

Robot teams that learn a common environment model exchange belief summaries and plan by the exp…

robot
arXiv cs.MA
智能体应用arXiv cs.MA

Emergence World: Adversarial Stress-Testing of Long-Horizon Multi-Agent Systems

As AI agents move from bounded tasks to persistent deployments, failures can propagate through…

memorymulti-agentrsi
arXiv cs.MA
智能体应用arXiv cs.MA

Decomposition Buys Integrity, Not Yield

Multi-agent systems split a task across a tree of agents and justify the split with folklore: s…

multi-agent
arXiv cs.MA
物理 AIarXiv cs.MA

Calibrate Once, Fly Any Team: Residual-Grounded Low-Fidelity Training for Cooperative Drone Swarms

Training multi-agent drone-swarm policies directly in high-fidelity (HF) rigid-body physics is…

dronemulti-agent
arXiv cs.MA
产业观察arXiv cs.MA

BeWater: Effective Protesters Navigate Watersheds in Street Networks

During social movements, protesters need to gather with limited communication means and limited…

arXiv cs.MA
产业观察arXiv cs.MA

Anchored Sequential Deliberation

Sequential deliberation is a mechanism for collective decision making: at each round, a uniform…

arXiv cs.MA
智能体应用arXiv cs.MA

Agentic Societies Need a Social Harness

An agentic society is a collection of AI agents that coordinate autonomously across trust bound…

agenticagi
极客公园
物理 AI极客公园

专访爆火「机器鸭」背后的硬件推手:这是个信号,未来推动新故事的并非硬件

最近这两周,一只「 机器鸭 」在 X 上刷屏了。 这只叫 Microduck 的机器鸭,由 Hugging Face 旗下的 Pollen Robotics 设计,售价 399 美元。 高峰…

robot开源机器人
arXiv cs.MA
产业观察arXiv cs.MA

The fixed-point bundle method over product-of-simplex domains arising from game equilibria

This paper extends the fixed-point bundle framework for finite-dimensional variational inequali…

arXiv cs.MA
产业观察arXiv cs.MA

GPEvac: GNN-Based PPO for Adaptive Evacuation Routing During Shooting Events

The sharp increase in mass shootings underscores an urgent need for systems that guide victims…

arXiv cs.MA
大模型基础arXiv cs.MA

Cheap Talk Stabilizes Strategic Interaction in LLM Agents

Large language models are increasingly deployed as interacting agents, making the persistence o…

llmmulti-agentrsi
arXiv cs.MA
智能体应用arXiv cs.MA

BLINDSPOT: A Benchmark for Safety and Refusal Calibration in Long-Horizon Tool-Using Agents

Large language model (LLM) agents increasingly operate over long-horizon interactions involving…

llmrsitool use
arXiv cs.MA
物理 AIarXiv cs.MA

Auto-HSI: Personalized human control of a robot swarm on demand by using LLMs for online automatic code generation

This paper presents Auto-HSI, a method for generating personalized human-swarm interaction (HSI…

llmrobot
极客公园
智能体应用极客公园

OpenAI、Anthropic 再次发出「AI 末日」警告;小米澎程今日全国交付;Deepseek 灰度测试 AI 语音对话

OpenAI 首席执行官:今年不会上市,不能冒哪怕 10% 杀死所有人的风险 OpenAI 首席执行官 Sam Altman 在接受采访时表示,这家人工智能公司专注于解决围绕 AI 技术的安…

发布会智能体芯片
极客公园
产业观察极客公园

AI 时代的「4399」,可把我玩嗨了|AI 上新

打开 Pocket 的前十分钟,我以为自己打开了一个 Instagram 版的 4399。 不久前,Meta 在美国正式推出了 Pocket。Pocket 的玩法很好理解,用户不用写代码,只…

极客公园
视觉模型极客公园

DeepSeek V4.1 Flash 发布;罗永浩狂喷苹果折叠屏:全是抄的;马斯克「无聊公司」融资 30 亿美元|极客早知道

DeepSeek V4.1 Flash 模型正式发布:全面超越 V4 Pro、原生多模态视觉理解,最高降价 60% 9 月 10 日消息,深度求索今日正式发布 DeepSeek V4.1 F…

agenticmoe前沿
极客公园
智能体应用极客公园

走出聊天框,Agent 开始进入现实世界

头图来源:小度 Agent 正在寻找自己的「身体」。 过去一年,智能体的主要工作对象是文件、网页和软件。它们可以搜索资料、分析数据、制作 PPT,也可以打开浏览器、调用工具,把一句需求推进成…

发布会智能体生态
极客公园
产业观察极客公园

苹果进入特努斯时代,首发 15999 元折叠屏 iPhone;Deepseek 被曝备战科创板 IPO;谷歌埃森哲组建千人 FDE 团队 | 极客早知道

苹果首款折叠 iPhone Duo 亮相,国行 15999 元起 北京时间 9 月 10 日凌晨,苹果发布首款折叠手机 iPhone Duo,国行 15999 元起,10 月 16 日开启预…

芯片营收
极客公园
视觉模型极客公园

微信,悄悄迈出 AI 社交的第一步

头图来源:视觉中国、ChatGPT 生成 最近,微信开始小范围测试一项新的「小微 AI 社交」功能。 用户如果想联系一位朋友,可以先把诉求告诉自己的「小微」。小微找到对方的小微后,会说明此次…

视觉
极客公园
物理 AI极客公园

对话极壳创始人孙宽:年出货 3 万台后,外骨骼「全班第一」的成长和焦虑

在刚刚结束的 IFA 柏林国际消费电子展上,几乎每家机器人公司的展台上,都会摆上一台可穿戴外骨骼设备。就像 10 年前,很多互联网新锐创业者都喜欢在办公桌上摆一台无人机。这既是对这个品类「前…

前沿无人机机器人
极客公园
视觉模型极客公园

折叠屏 iPhone 初期产量受限,每日仅数百部;环比增长 379%,腾讯 HY4 登顶全球大模型调用榜;特斯拉时隔 19 个月再降价|极客早知道

消息称苹果折叠屏 iPhone 初期产量受限,每日仅数百部 9 月 8 日,据日经中文网报道,根据多位知情人士指出,市场期待已久的首部苹果折叠 iPhone,由于苹果极为严格的质量管控标准,…

图像生成大模型营收
极客公园
物理 AI极客公园

千问办公发布多人工作台,重写企业软件的最后一公里

作者|Cynthia 编辑|郑玄 2026 年 2 月 3 日,由 Anthropic 带头,华尔街替 SaaS 写好了讣告。 几天前,Anthropic 正式官宣把 Claude Cowo…

机器人
新浪科技
物理 AI新浪科技

地平线CEO余凯:边缘计算是自动驾驶的技术基石

作者:权小星 随着人工智能行业的发展,以及应用场景的铺开,人工智能在加速融入到人们的生活、生产流程,而且随着人工智能产品的更新换代、智能计算场景的多样化,在对于人工智能硬件、算法及芯片的要求…

自动驾驶芯片
新浪科技
产业观察新浪科技

中国电竞亚运夺冠“十日谈” 沉迷游戏≠电竞

领奖台上的AOV中国队。 亚洲电子体育联合会 图 原标题:中国电竞亚运夺冠“十日谈” 作者:沈文迪 实习生 黄霁洁 “Nice、nice,一波一波了,推推推推,赢了赢了!nice!” 这些嘈…