whisper_large_v3_encoder_fp16
UpdateMay 28, 2026 17:26PublishedMay 28, 2026 17:26
141
0
0
whisper_large_v3_encoder_fp16 - 1
Avatar
User_jh05wj
Type
Other
Basic Model
Other
Published time
May 28, 2026 17:26
File Signature
bc4fb6a9bcba671a2ea09a0fe1e9369b18ee78330abfa91b358b99382dc801d5
whisper_large_v3_encoder_fp16.safetensors
1.58 GB

Whisper Large v3 Audio Encoder Model.

Used to extract speech features, perform lip-syncing, and analyze emotional rhythm, enhancing lip accuracy and audio alignment in digital human video broadcasting.

Suitable for LongCat Avatar audio-driven workflows.

Gallery

No creation yet