@chain_note_tw目標 55 秒4 個片段成本 $0.15負責人 柯昱翔更新 2026/07/29 05:26Higgsfield 任務
半自動任務流程:系統負責把 Prompt、素材與檢查清單準備到可以直接照做, 人只做 Higgsfield 上真正需要人的那一步。
這一版刻意不做瀏覽器機器人
Higgsfield 的生成品質需要人眼判斷,替它寫一支會壞掉的自動化腳本只是把風險藏起來。 我們把可以自動化的部分做到底:Prompt 自動帶入正確的主持人、背景、構圖與禁止事項, 素材與檔名一次備妥,上傳強制綁定 segment_id。人負責的只剩「看一眼、按生成、丟回來」, 而且每一步都留在流程裡可追蹤、可重做。
批次進度
需要動畫的段落
3
音檔就緒
3
已下載任務包
3
已上傳成果
0
待處理
3
蘇雨桐
@chain_note_tw
ctv_zh_tw_female_teacher_02
Higgsfield Prompt14.8sproj_03_seg_0301_v1.mp4
[HOST] A friendly Taiwanese female educator in her late 20s, shoulder-length hair, cream knit sweater, standing in front of a clean pastel studio wall. Warm even lighting, gentle hand gestures while explaining. Vertical 9:16 framing, waist-up composition, no on-screen text. [SHOT] Vertical 9:16, 1080x1920, 30fps. Full-frame host shot, chest-up composition, eyes on the lens, head centred with breathing room above. [BACKGROUND] Locked to the registered studio backdrop supplied as background.png. Do not redress, relight or replace the set — every segment of this project must cut together. [LIGHTING] Soft key from camera-left, gentle rim separation from the backdrop, constant colour temperature and exposure across all segments. [ACTION] Opening hook: already speaking at frame 1 — no dead air, no intro beat. Energy starts high and settles by the second sentence. Delivery follows the host's registered register (親切、耐心的教學口吻;先用生活比喻,再帶名詞解釋,語速偏慢). [LIP-SYNC] Drive the mouth from the supplied audio.wav (14.8s of Mandarin narration). Mouth stops the moment the audio stops. [DURATION] 14.8s, one continuous take, no cuts. [IDENTITY] Match host_reference.png exactly — same person, hair, wardrobe and framing as every other segment of project proj_03. This project is locked to 蘇雨桐 (@chain_note_tw).
由主持人的 base_prompt 加上本段語境(版型、開場/中段/CTA、對嘴長度、專案一致性)組出, 不需要人工再編輯。
Negative Prompt畫面不得生成字幕spec 第 6 節:字幕由 Ditto-Vid 在最終合成時另外加,Higgsfield 畫面本身不得生成字幕。
任務包內容已備妥
task.json
任務參數:prompt、negative_prompt、target_duration、輸出檔名
project_id=proj_03|segment_id=seg_0301|host_id=host_02
audio.wav
該段 Cartesia 音檔,作為 Higgsfield 的對嘴音軌
14.8 秒|voice_id=ctv_zh_tw_female_teacher_02
host_reference.png
主持人參考圖,確保十位主持人不會互相串味
蘇雨桐(host_02)的固定形象,Demo 以漸層色塊代替真實圖檔
background.png
該主持人的固定背景,讓所有片段能剪在一起
同一專案所有片段共用,不得在 Higgsfield 端替換
instructions.txt
操作步驟、輸出檔名與 QA 檢查清單
輸出檔名 proj_03_seg_0301_v1.mp4|目標 14.8 秒
蘇雨桐
@chain_note_tw
ctv_zh_tw_female_teacher_02
Higgsfield Prompt14.6sproj_03_seg_0303_v1.mp4
[HOST] A friendly Taiwanese female educator in her late 20s, shoulder-length hair, cream knit sweater, standing in front of a clean pastel studio wall. Warm even lighting, gentle hand gestures while explaining. Vertical 9:16 framing, waist-up composition, no on-screen text. [SHOT] Vertical 9:16, 1080x1920, 30fps. Tight picture-in-picture crop — the host is composited at 28% frame width in the bottom-right corner with a 4% margin, so keep the head centred in the upper third of the plate and leave clean space around it. [BACKGROUND] Locked to the registered studio backdrop supplied as background.png. Do not redress, relight or replace the set — every segment of this project must cut together. [LIGHTING] Soft key from camera-left, gentle rim separation from the backdrop, constant colour temperature and exposure across all segments. [ACTION] Mid-roll explanation: measured delivery, small purposeful gestures on key numbers, no large body movement that would break the composite. Delivery follows the host's registered register (親切、耐心的教學口吻;先用生活比喻,再帶名詞解釋,語速偏慢). [LIP-SYNC] Drive the mouth from the supplied audio.wav (14.6s of Mandarin narration). Mouth stops the moment the audio stops. [DURATION] 14.6s, one continuous take, no cuts. [IDENTITY] Match host_reference.png exactly — same person, hair, wardrobe and framing as every other segment of project proj_03. This project is locked to 蘇雨桐 (@chain_note_tw).
由主持人的 base_prompt 加上本段語境(版型、開場/中段/CTA、對嘴長度、專案一致性)組出, 不需要人工再編輯。
Negative Prompt畫面不得生成字幕spec 第 6 節:字幕由 Ditto-Vid 在最終合成時另外加,Higgsfield 畫面本身不得生成字幕。
任務包內容已備妥
task.json
任務參數:prompt、negative_prompt、target_duration、輸出檔名
project_id=proj_03|segment_id=seg_0303|host_id=host_02
audio.wav
該段 Cartesia 音檔,作為 Higgsfield 的對嘴音軌
14.6 秒|voice_id=ctv_zh_tw_female_teacher_02
host_reference.png
主持人參考圖,確保十位主持人不會互相串味
蘇雨桐(host_02)的固定形象,Demo 以漸層色塊代替真實圖檔
background.png
該主持人的固定背景,讓所有片段能剪在一起
同一專案所有片段共用,不得在 Higgsfield 端替換
instructions.txt
操作步驟、輸出檔名與 QA 檢查清單
輸出檔名 proj_03_seg_0303_v1.mp4|目標 14.6 秒
蘇雨桐
@chain_note_tw
ctv_zh_tw_female_teacher_02
Higgsfield Prompt14.6sproj_03_seg_0304_v1.mp4
[HOST] A friendly Taiwanese female educator in her late 20s, shoulder-length hair, cream knit sweater, standing in front of a clean pastel studio wall. Warm even lighting, gentle hand gestures while explaining. Vertical 9:16 framing, waist-up composition, no on-screen text. [SHOT] Vertical 9:16, 1080x1920, 30fps. Full-frame host shot, chest-up composition, eyes on the lens, head centred with breathing room above. [BACKGROUND] Locked to the registered studio backdrop supplied as background.png. Do not redress, relight or replace the set — every segment of this project must cut together. [LIGHTING] Soft key from camera-left, gentle rim separation from the backdrop, constant colour temperature and exposure across all segments. [ACTION] Closing call-to-action: settle the pace, lean in slightly on the final sentence, hold a calm confident expression for the last 0.5s without moving out of frame. Delivery follows the host's registered register (親切、耐心的教學口吻;先用生活比喻,再帶名詞解釋,語速偏慢). [LIP-SYNC] Drive the mouth from the supplied audio.wav (14.6s of Mandarin narration). Mouth stops the moment the audio stops. [DURATION] 14.6s, one continuous take, no cuts. [IDENTITY] Match host_reference.png exactly — same person, hair, wardrobe and framing as every other segment of project proj_03. This project is locked to 蘇雨桐 (@chain_note_tw).
由主持人的 base_prompt 加上本段語境(版型、開場/中段/CTA、對嘴長度、專案一致性)組出, 不需要人工再編輯。
Negative Prompt畫面不得生成字幕spec 第 6 節:字幕由 Ditto-Vid 在最終合成時另外加,Higgsfield 畫面本身不得生成字幕。
任務包內容已備妥
task.json
任務參數:prompt、negative_prompt、target_duration、輸出檔名
project_id=proj_03|segment_id=seg_0304|host_id=host_02
audio.wav
該段 Cartesia 音檔,作為 Higgsfield 的對嘴音軌
14.6 秒|voice_id=ctv_zh_tw_female_teacher_02
host_reference.png
主持人參考圖,確保十位主持人不會互相串味
蘇雨桐(host_02)的固定形象,Demo 以漸層色塊代替真實圖檔
background.png
該主持人的固定背景,讓所有片段能剪在一起
同一專案所有片段共用,不得在 Higgsfield 端替換
instructions.txt
操作步驟、輸出檔名與 QA 檢查清單
輸出檔名 proj_03_seg_0304_v1.mp4|目標 14.6 秒
不需要主持人動畫的段落
原片滿版段落直接使用 YouTube 原素材與原音,不進 Higgsfield 流程。
- 段 2seg_0302原片滿版已驗證Higgsfield 不適用