Korean caption fonts on Sume: why weight-shift needs Pretendard
Sume's weight-shift and korean-ad caption styles animate a font weight axis that only Pretendard has. Other faces keep color and scale emphasis only.

On Sume, the weight-shift and korean-ad caption styles animate the wght axis, and only Pretendard has that axis. If you set font to any other Hangul face, the style keeps its color and scale emphasis but loses the weight travel. If you omit font, korean-ad and weight-shift use Pretendard. The details are on Video captions, read 2026-10-06.
Which style uses which face?
The docs give a default face per style, and font changes only the face.
| Style | Default face | Weight travel |
|---|---|---|
korean-ad | Pretendard | Yes |
weight-shift | Pretendard | Yes |
highlight, pill-karaoke, editorial-emphasis | Pretendard | Not the headline feature |
black-outline, clip-wipe | Do Hyeon | No |
What breaks if I pick the wrong font?
Two guards protect you. Korean text on slam, punch or tiktok-green returns 400 caption_hangul_text_latin_style, because those Latin display faces have no Hangul glyphs. A Hangul font on a Latin style returns 400 caption_font_requires_hangul_style. In both cases you pay nothing, because the error comes at request time and not after a render.
- Send
language: "ko"as a speech-to-text hint; it never selects a style or font. - Send
style: "korean-ad"explicitly if you want the ad karaoke look; with nostyle, Korean text resolves toblack-outline. editorial-emphasisalways uses Black Han Sans for its emphasis line.
Does the font list change per request?
font is optional and the API rejects names outside its list instead of falling back, so a caption never silently uses a face you did not select.
Sources
Related posts
More in Media tools
- Captions for a voiceover: reuse TTS word timings or run STT?
If you generated the voice on Sume, ask for timestamps.words and skip a 1-cent-a-minute STT pass. If the audio is someone else's, run STT. Here is the split.
- Captions language field: a speech hint, not a style or font
On Sume video captions, language only hints speech-to-text. It never picks the style or font, so Korean needs a Hangul style such as black-outline.
- Check a Short has sound and picture before adding it to a YouTube show
YouTube asks shows for high-quality video and sound and says no-visuals videos likely will not count. Probe each Short with Sume video inspect first.
- compose_video_has_no_audio: a silent clip under a still on Sume
Timeline compose warns compose_video_has_no_audio and still completes. The output is silent. How to add a voice-over afterward with a Timeline render.
Written by Sume