据悉,腾讯混元语音团队联合多所高校研究人员发布Gander 模型,可同步接收语音、图像、文本输入,并在自身输出语音的同时持续处理多模态信息,支持用户随时打断。为平衡对话速度与复杂任务推理时长,Gander 采用双模块拆分:面向人体的”小脑”按秒级管理对话流程,决定倾听、发言或在被打断时停止输出,以最近约两分钟对话作为记忆、无需独立语音起止检测;”大脑”在后台完成推理与复杂智能体任务,可直接替换为 Codex、Claude Code 等系统而无需重训对话模块。
![]()
特别声明:以上内容(如有图片或视频亦包括在内)为自媒体平台“网易号”用户上传并发布,本平台仅提供信息存储服务。
Notice: The content above (including the pictures and videos if any) is uploaded and posted by a user of NetEase Hao, which is a social media platform and only provides information storage services.