What it may solve
Give text-only DeepSeek Harness agents video understanding: scene-aware frame sampling + VLM + optional ASR transcript fused into timeline evidence. / 给纯文本模型的视频理解Plugin(场景感知抽帧 + VLM + 可选语音转录)
Imported third-party catalog description; not a Registry verification conclusion.