Pick the source
Drop a file, paste captions, or reference a link for attribution. Nothing starts until you confirm you have the rights to it.
Video2Blog reads what is said and what is shown — slides, product screens, numbers — and writes a structured article where every paragraph links to its exact moment. You review the evidence, not a wall of AI text.
Free · 20 minutes of video · no card · runs on Cloudflare Workers AI
The team compared three plans side by side before changing anything. ▶ 4:17
After the switch, the free plan converted at 40% once customers paid per processed minute. ▶ 10:57
inferenceThe free tier now works as a trial for heavy users. ▶ 11:20
The heavy lifting on the video happens in your browser. The model work happens on Cloudflare. Every step leaves evidence you can check.
Drop a file, paste captions, or reference a link for attribution. Nothing starts until you confirm you have the rights to it.
The file is decoded on your device. It is never uploaded or stored — only what the next steps need leaves your machine.
Audio becomes 30-second parts. Key frames are picked where the screen changes, and their text is read by OCR — slides, UI labels, numbers.
Each audio part is transcribed once with timestamps and language detection. A part is never transcribed — or paid for — twice.
Speech and screens are merged into ~45-second evidence windows, then written into an article in one structured call. Every paragraph must cite its windows.
A deterministic pass checks citations, verbatim quotes and numbers against the source, and points at the exact paragraph to review.
From the reviewed article: LinkedIn post, X thread, newsletter section, YouTube description and quotes — same citations, one call.
Click a transcript line or a key moment: the video, the timeline and the article stay in sync. This is sample data — your own video gives you the same views.
What one SaaS team learned by dropping per-seat plans
The team put three plans side by side before changing anything; the cheapest one mostly served as an anchor.
“On the free plan we now see a forty percent conversion rate” — once customers paid per processed minute instead of per seat.
inferenceThe free tier now behaves like a trial for heavy users.
Speech, key frames and on-screen text aligned in ~45-second windows. The article, the checks and the search all read from it.
Picked where the screen changes — a duplicate only if three signals agree.
Slides, UI labels and numbers, with a confidence score. Low-confidence reads are flagged.
Citations, verbatim quotes, numbers absent from the source — computed on every save.
Review the evidence next to the claim instead of re-watching an hour.
Find what was said or shown across every source you processed.
Derived formats keep their citations. Identical requests reuse what was already generated — no second bill.
Transcript-only tools summarise what they heard and fill the gaps. Video2Blog writes only what the source says or shows — and proves it.
Usage-based pricing is a proven strategy used by most successful SaaS companies. The team saw conversions roughly double after switching.
After the switch, the free plan converted at 40% once customers paid per processed minute. ▶ 10:52
inferenceThe free tier now works as a trial for heavy users. ▶ 14:50
No black box: here is where each step runs and what is kept.
No credits to decode. Re-writing, AI edits, language versions and exports of a processed source don't count.
Try it on one video.
For a weekly video or podcast.
For teams publishing from demos and webinars.
For agencies and content teams.
Prices in USD. Paid plans are opening in stages — during the preview, contact us to upgrade. Unused hours don't roll over.
No. When you add a file, your browser decodes it locally. Only 30-second audio parts (for transcription) and up to 60 selected screenshots are sent. The audio parts are discarded once transcribed; the transcript, screenshots and on-screen text are kept so you can edit and re-generate.
Cloudflare Workers AI: Whisper large-v3-turbo for speech and Gemma for writing. There is no third-party AI provider and no API key to bring. Your content is not used to train models.
The writer only sees the evidence timeline and must cite it. A separate, non-AI check then verifies that citations exist, quotes appear word-for-word in the transcript and numbers appear in the speech or on screen. Anything that fails is flagged at the exact paragraph.
You can paste it to reference the video: we read its public title, channel, thumbnail and chapters for attribution and timestamp links. YouTube's terms don't allow third-party services to download videos, so to build the article we need the file (if it is yours or you have permission) or its captions — both export from YouTube Studio.
No. Neither platform offers an approved way for a service to list another account's videos, and we don't scrape them. Upload the videos you have rights to; paste the post link to keep it as the credited source.
Video and audio files (MP4, MOV, WebM, MP3, M4A, WAV) prepared in your browser, and captions or timestamped transcripts (.vtt, .srt, .txt). Direct media links and podcast feeds are fetched server-side on deployments where the media engine is enabled.
Drop in a demo, a webinar or a tutorial. Review the evidence, not a wall of AI text.
Free · 20 minutes of video · no card