Active Speaker Detection

NVIDIA audio-visual neural active speaker detection microservice.

Updated 8/30/20260 code examples

About the Model

为什么要选择 NVIDIA Active Speaker Detection?

音视频联合多模态检测模型,可在多人会议与复杂视频场景中毫秒级定位正在发言的目标人物。

How to Access for Free (via NVIDIA NIM)

No detailed access description available.

Try it in your browser

Pick a model, paste your own free key, and run. Your key is sent once to call the provider and never stored on our servers.

Pick a model, paste your own free key, and run. Your key is sent once to call the provider and never stored on our servers.

Code Examples

No examples provided yet.

Community Feedback

Comments are tied to your GitHub account — sign in to join the discussion.