긴 영상 → 자막 넣은 Shorts

AI에 긴 영상과 편집 방향을 전달하면, 선별하고 화면 비율을 맞추고 자막까지 넣은 소셜 클립 묶음을 받을 수 있습니다.

  • 비디오
  • 파이프라인 5개
  • ≥ 2 크레딧

AI 에이전트

단말기

처음 오셨나요? 빠른 시작
01
Homebrew brew install pipe2-ai/tap/pipe2
↳
Go go install github.com/pipe2-ai/pipe2-cli/cmd/pipe2@latest
02
API 토큰 발급 ↗ echo "$PAT" | pipe2 auth login --token -

결과 변경

--input필수
asset_url기본값—Source video: a YouTube or social URL, direct media URL, local file, or existing Pipe2 asset.
--instructions
string기본값Find the five strongest self-contained moments. Prefer clear ideas, surprising claims, useful explanations, or complete stories. Each clip should make sense without the rest of the video.Editorial brief describing which moments to select and what makes a good clip.
--max-clips
int기본값5Maximum number of clips Video Trim may return.
--max-seconds
int기본값45Maximum duration of each selected clip, in seconds.
--aspect
enum기본값9:16Output aspect ratio for every clip.
9:16 · 1:1 · 4:5 +1
9:161:14:516:9
--preset
enum기본값serif-editorialCaption style used for every clip.
tiktok-bold-yellow · minimal-white · subtle-drop +3
tiktok-bold-yellowminimal-whitesubtle-dropkaraoke-gradientbig-serifserif-editorial
--corrections
string기본값Optional comma-separated word corrections in the form from=to,from=to.
--lang
string기본값autoISO 639-1 transcription language, or auto.
--parallel
int기본값4Maximum number of selected clips processed in parallel.
--position
enum기본값autoCaption position. Auto keeps text away from the main subject when possible.
auto · top · middle +1
autotopmiddlebottom
작동 방식 파이프라인 5개
  1. Video Trim이 선택한 각 아이디어를 끊기지 않게 유지하는 데 쓰는 타임스탬프 발화 지도를 만듭니다.

    입력 소스: OpenClaw Creator: Why 80% Of Apps Will Disappear
    텍스트
    Today, I'm sitting down with Peter Steinberger, the creator of OpenClaw, the open source personal AI agent that has completely taken over the internet.
    The GitHub repo exploded to over 160,000 stars practically overnight.
    The community has built countless projects like Maltbook, where bots talk among themselves,
    and now the bots are even renting humans to do tasks in the real world.
    In our conversation, we discuss his aha moment, his contrarian development philosophies, and what this means for builders in 2026.
    Let's dive in.
    [upbeat music] So good to see you, man.
    Hey, what's up?
    Um, so you've made something people want.
    …
  2. 원본 전체를 검토하고 편집 방향에 따라 가장 강한 순간을 골라 별도의 클립으로 반환합니다.

    비디오
  3. 선택된 각 클립을 지정한 소셜용 화면 비율로 다시 구성하고, 화자와 중요한 내용이 화면에 보이도록 유지합니다.

    비디오
  4. 다시 구성한 각 클립을 전사해 자막이 최종 편집과 타이밍에 맞도록 합니다.

    텍스트
    Are apps just gonna go away?
    Uh, I think 80% of them are going away.
    Why do I need MyFitnessPal?
    Like, my agent already knows that I'm making bad decisions.
    I'm at, I don't know,
    uh, Smashburger something,
    and it will already assume that I eat what I like to eat.
    If I don't make a comment, it will just, like, automatically track it, or I make a picture and it'll just store it somewhere.
    I don't even need to care, right?
  5. 해당 전사본을 선택한 스타일과 위치로 각 클립에 자막으로 입힙니다.

    비디오
효과적인 이유와 변경할 수 있는 항목

이 방식이 효과적인 이유

레시피는 먼저 원본의 타임스탬프 전사본을 만듭니다. 그다음 video-trim이 전체 녹화본을 검토하고, 후보 순간을 지시 사항과 비교해 그 자체로 완결된 클립 여러 개를 반환합니다. 별도의 하이라이트 목록을 맞춰 둘 필요도, 타임스탬프 파일을 준비할 필요도 없습니다.

이후 선택된 각 클립은 개별 마무리 과정을 거칩니다. 화면 비율 조정이 전사보다 먼저 이루어지므로 전사본과 입힌 자막이 최종 컷과 정확히 일치합니다. 각 분기는 병렬로 실행됩니다. 클립을 더 많이 요청하면 작업량은 늘지만, 레시피가 하나의 긴 순차 대기열이 되지는 않습니다.

중요한 설정

  • instructions: 주제와 편집 기준을 함께 설명하세요. “그 자체로 완결된 가장 강력한 주장 다섯 가지”가 “흥미로운 순간”보다 낫습니다.
  • max_clips: 클립 수의 상한입니다. 조건을 충족하는 순간이 적을 때 약한 클립으로 수를 채우도록 강요하지 않습니다.
  • max_seconds: 각 클립의 최대 길이입니다. 훅에는 짧게, 완결된 설명에는 길게 설정하세요.
  • aspect: Shorts, Reels, TikTok에는 9:16을 사용하고, 게시할 곳에 따라 1:1, 4:5, 16:9로 바꾸세요.

레시피 옵션에서 자막 스타일, 언어, 교정, 병렬 실행도 설정할 수 있습니다.

참고 사항

원본이 좋아도 요청한 수만큼 적합한 순간이 있다는 보장은 없습니다. 에이전트는 같은 아이디어를 반복하거나 약한 순간을 잘라내기보다 더 적은 수의 클립을 반환할 수 있습니다. 게시하기 전에 각 클립을 원래 맥락에서 검토하세요. 따로 떼어 낸 인용은 원래 대화에 있던 중요한 단서를 잃을 수 있습니다.

샘플 갤러리는 출처를 밝힌 OpenClaw 인터뷰에서, 중단해도 이어서 실행할 수 있는 한 번의 실행으로 만들었습니다. 출력 형식을 보여 주는 예시이며, 새로 실행하면 AI가 다른 순간을 고를 수 있습니다.