Tencent Gander: AI คุยเสียงพร้อมทำงานเบื้องหลัง แบบแยกสมอง
โดย Nokka (นก-กา) | 23 กันยายน 2569 Gander จาก Tencent: โมเดลเปิดที่คุยเสียงพร้อมทำงานทับซ้อน บทความนี้เขียนโดย AI (โมเดล glm-5.3 ของผู้ให้บริการ ollama-cloud) ผ่าน Hermes Agent จาก Nous Research ตรวจสอบและเรียบเรียงโดย Nokka ลองนึกภาพ: คุณกำลังคุยกับผู้ช่วย AI ด้วยเสียง กลางคันคุณขอให้มันไปหาข้อมูลเรื่องหนึ่ง แล้วคุยเรื่องอื่นต่อได้เลย ไม่ต้องรอ ไม่ต้องวางสาย…
Tencent unveiled project Gander in early November, showcasing an AI that can converse in real-time, understand speech, recognize images, and work in the background simultaneously without interrupting the conversation. This innovative approach is based on a dual-brain system inspired by the anatomy of the human cerebellum and brain, designed to handle speech interactions and background tasks separately.
The cerebellum, or "small brain," manages real-time voice conversations, visual input from sent photos, and rapid response generation. In contrast, the larger brain focuses on more demanding tasks like data retrieval, coding, or invoking external tools. These two components are connected via a queue manager, allowing the small brain to continue interacting with the user while the large brain completes heavy tasks.
Currently, Tencent offers no public online service for Gander; users must download and set up the system independently, requiring three graphics cards and significant memory. Despite the high memory requirements, the project remains research-focused rather than commercial, with models available for download under the permissive Apache 2.0 license.
While the system demonstrates promising voice interaction capabilities, it still lags behind commercial AI services like GPT-Realtime in performance. The main advantage lies in the open-source nature, enabling developers to customize and run their own instances at a lower cost compared to proprietary solutions. However, the current performance and setup complexity make it more suitable for researchers and developers interested in exploring the mechanics of dual-brain AI rather than everyday users seeking a ready-to-use voice assistant.
Written by urgent.news from Dev.to's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.