Back Office · office.temerarii.xyz
One published post, granular — the Social for W34 Tue, its copy, its output-policy format, and the video master it derives from. Part of the day's full output set.
postW34-Tue-social-5kindSocialweekW34dayTuedate2026-08-25campaignlongform-youtubecadence5/day floor × 9 channels

Copy the published post text

The Stack
copy ready · render pending

Output-policy spec format · dimensions (asset_specs.output_policy)

formatnative cut · 9:16 · 1:1 · 16:9 dims1080×1920 · 1080×1080 · 1920×1080 cadence5/day floor × 9 channels

Channels 9 destinations

TikTokInstagramLinkedInX/TwitterFacebookThreadsPinterestBlueskyYouTube

This post a distinct social asset — its own angle, storyboard, and cuts

thread-the-stack-W34-Tue
5-distinct-social/day · 9:16 master → 1:1 / 16:9 cuts per channel

Channel-cuts this asset → 9 native captions (one master · per-platform aspect+copy)

ChannelNative caption
TiktokThe tool: Whisper, for uncanny transcription. The limit: on a long file it drifts and the timestamps wander off. The fix: chunk the audio, then run WhisperX to force word-level alignment back onto the waveform. Captions snap to each word. (AI-assisted)
InstagramThe Stack: Whisper. The limit: it drifts on long audio. The fix: Chunk the audio. Run WhisperX for word-level alignment. Captions snap to each spoken word. #whisper #transcription #captions #ai #buildinpublic
LinkedinThe tool is Whisper, and the transcription is uncanny. The limit nobody mentions: on a long file it drifts, and the timestamps wander off the actual audio. The fix: chunk the audio first, then run WhisperX to force word-level alignment back onto the real waveform. Now captions snap to each spoken word, so auto-generated subtitles look hand-tuned and your edits land on the syllable. Take one long recording, chunk it, run WhisperX, and compare the timestamps. The fix sells itself.
XTool: Whisper. Uncanny transcription. Limit: on long audio it drifts and timestamps wander. Fix: chunk the audio, run WhisperX for word-level alignment. Captions snap to each word.
FacebookThe tool is Whisper, and the transcription is uncanny. The limit is that on a long file it drifts and the timestamps wander off the audio. The fix: chunk the audio first, then run WhisperX to force word-level alignment back onto the real waveform. Now captions snap to each spoken word and look hand-tuned. Try it on one long recording and compare the timestamps.
ThreadsThe tool is Whisper, uncanny transcription. The limit: on a long file it drifts and the timestamps wander off. The fix: chunk the audio, then run WhisperX to force word-level alignment onto the real waveform. Captions snap to each spoken word. Try it on one clip and compare.
PinterestFixing Whisper transcription drift on long audio with WhisperX. The method: chunk the audio first, then run WhisperX to force word-level alignment back onto the real waveform so captions snap to each spoken word. A plain guide to accurate auto-generated subtitles and word-level timestamps.
BlueskyTool: Whisper. Uncanny transcription. Limit: on long audio it drifts and timestamps wander. Fix: chunk the audio, run WhisperX for word-level alignment. Captions snap to each word.
YoutubeTitle: The Stack: Fix Whisper's Drift on Long Audio with WhisperX The tool is Whisper, and the transcription is uncanny. The limit nobody mentions: on a long file it drifts and the timestamps wander off the actual audio. The fix: chunk the audio first, then run WhisperX to force word-level alignment back onto the real waveform. Now captions snap to each spoken word, so auto-generated subtitles look hand-tuned and your edits land on the syllable. Take one long recording, chunk it, run WhisperX, and compare the timestamps.

Composition layer × scene 6 scenes · this post's OWN storyboard (distinct per asset)

#Layer (comp_id · still · tier)BeatTimecodeMotionLogoAudioVO / on-screen / caption
1s1
matches intent
shared field
signature-3d
hook0–6.9skinetic-buildicon·wireframe♪ node_lock
The tool is Whisper, uncanny transcription. The limit: on a long file it drifts and the timestamps wander off.
on-screen: Whisper drifts on long audio
expected on screen: black ground · torus hero in the shared Signal Field · Nuntius leads · mark · kinetic-build · icon·wireframe logo · caption bottom-left
spec (the prompt): comp_id shared Signal Fieldvisual markshape torusground blacktreatment wireframemotion kinetic-buildpower laser-lockinstrument laser-trace→Whisper drifts on long audio
2s2
matches intent
ChecklistCard
template
teach6.9–13.600000000000001skinetic-buildicon·wireframe♪ node_lock
So we chunk the audio first, then run WhisperX to force word-level alignment back onto the real waveform.
on-screen: Chunk + WhisperX align
expected on screen: black ground · a ChecklistCard panel over a dimmed Signal Field · Nuntius leads · node-graph · kinetic-build · icon·wireframe logo · caption bottom-left
spec (the prompt): comp_id ChecklistCardvisual node-graphshape torusground blacktreatment wireframemotion kinetic-buildpower morphinstrument morph+laser→Chunk + WhisperX aligncurate items, nodes
3s3
matches intent
DiffCard
template
build13.6–20.9stype-onicon·wireframe♪ node_lock
Now captions snap to each spoken word, so auto-generated subtitles look hand-tuned and your edits land on the syllable.
on-screen: Caption that snaps to the word
expected on screen: black ground · a DiffCard panel over a dimmed Signal Field · Nuntius leads · code · type-on · icon·wireframe logo · caption bottom-left
spec (the prompt): comp_id DiffCardvisual codeshape torusground blacktreatment wireframemotion type-onpower summoninstrument draw-on→Caption that snaps to the wordcurate codeLines, lines
4s4
matches intent
LegacyCard
template
proof20.9–27.7sreceipts-counticon·wireframe♪ node_lock
So take one long recording, chunk it, and run WhisperX. Compare the timestamps. The fix sells itself.
on-screen: Run WhisperX on a clip
expected on screen: black ground · a LegacyCard panel over a dimmed Signal Field · Nuntius leads · receipts · receipts-count · icon·wireframe logo · caption bottom-left
spec (the prompt): comp_id LegacyCardvisual receiptsshape torusground blacktreatment wireframemotion receipts-countpower receiptsinstrument spotlight→Run WhisperX on a clipcurate statsLabels
5s5
first render · fix pending
shared field
signature-3ddead_airgeneric_scene
futurist27.7–32.7sspatial-parallaxicon·wireframe♪ bed
The Stack
expected on screen: black ground · torus hero in the shared Signal Field · Nuntius leads · spatial-parallax · icon·wireframe logo · caption bottom-left
spec (the prompt): comp_id shared Signal Fieldshape torusground blacktreatment wireframemotion spatial-parallaxpower spatial-parallaxinstrument laser-fire→The Stack
6s6
first render · fix pending
shared field
signature-3ddead_airgeneric_scene
resolve32.7–37.7scoalescenceicon·wireframe♪ bed_out
The Stack
expected on screen: black ground · torus hero in the shared Signal Field · Nuntius leads · coalescence · coalescence · icon·wireframe logo · caption bottom-left
spec (the prompt): comp_id shared Signal Fieldvisual coalescenceshape torusground blacktreatment wireframemotion coalescencepower coalescenceinstrument coalescence→The Stack

Cross-links this post in the day's output set