videoaudiodetectagegender_mapper

July 17, 2026 · View on GitHub

Detect age and gender (male, female, child) from video audio signals using a pretrained wav2vec2 model. This operator processes videos tagged as containing speech and classifies the speaker's age and gender from the audio stream. It must be operated after video_tagging_from_audio_mapper.

使用预训练的 wav2vec2 模型从视频音频信号中检测年龄和性别(男性、女性、儿童)。此算子处理被标记为包含语音的视频,并从音频流中分类说话者的年龄和性别。它必须在 video_tagging_from_audio_mapper 之后运行。

Type 算子类型: mapper

Tags 标签: gpu, hf, video

🔧 Parameter Configuration 参数配置

name 参数名type 类型default 默认值desc 说明
hf_audio_mapper<class 'str'>'audeering/wav2vec2-large-robust-24-ft-age-gender'HuggingFace model for age/gender classification.
tag_field_name<class 'str'>'audio_speech_attribute'field name to store the age/gender results.
args''extra args
kwargs''extra args