Signals 4 · free daily AI digest

NemotronLabs VoiceChat: An Open Full-duplex Speech-to-Speech Model with Tool Calling Capabilities

Paper recorded by Signals 4 on 2026-09-18 in cs.AI. Abstract reproduced from arXiv; link to the original below.

Published 2026-09-18 on arXiv · recorded by Signals 4 on 2026-09-21

Category: cs.AI · 人工智能 · first seen 2026-09-21

Abstract

We introduce NemotronLabs VoiceChat, an open full-duplex speech-to-speech model with native tool-calling capabilities. NemotronLabs VoiceChat combines a streaming speech encoder and decoder-only language model with parallel specialized output streams for agent text and structured function calls, an auxiliary RNN-T branch for incremental user transcription, and a streaming TTS decoder. This design

Read on arXiv →

#8 most recent of 320 cs.AI papers we have recorded · ↑ newer: A Lie Detector Test for Language Models: Reading Knowledge a Model Won · ↓ older: Learning Cardiac Features: ECG Biometrics Across Time and~Exercise
Cite this page: NemotronLabs VoiceChat: An Open Full-duplex Speech-to-Speech Model with Tool Calling Capabilities: the #8 most recent of 320 cs.AI papers we have recorded (as of 2026-09-18). Source: Signals 4 (Signals API) — https://data.jiangzhang.ca/signals4/t/papers/nemotronlabs-voicechat-an-open-full-duplex-speech-to-speech-model-with-tool-call.html
Free to quote with attribution to “Signals 4 (Signals API)”. Machine-readable: papers.json
Related: More cs.AI papers · arXiv signals · All papers · Today in AI
Get 4 AI signals a day by email — free.
Subscribe free → See all plans →
Get 4 AI signals a day by email — free
All models · All repos · By company · Daily editions