What is Video Editor?
Video Editor is a browser-based, non-destructive video editing workspace. Upload one video and cut it into a sequence, or add any number of additional video/image overlay tracks and compose them into split-screen, picture-in-picture, or multi-participant video-call layouts — without uploading your footage to a server or account.
What can Video Editor do?
- Trim — drag either edge of a clip on the timeline, or set precise numeric in/out points. Dragging snaps to the playhead, other clips' edges, and markers for a clean cut, with a small on-screen highlight when it snaps — hold Alt to drag without snapping.
- Split — cut a clip in two at the playhead.
- Multi-select — Ctrl/Cmd-click to add clips to a selection, Shift-click to select a range, then delete or duplicate all of them in one action (one undo step, original order preserved). Click empty timeline space to clear the selection.
- Delete — remove an unwanted segment, or every selected clip at once.
- Join — merge adjacent clips back together.
- Free-form positioning — drag any clip left or right to place it wherever you want on the timeline (gaps allowed, overlaps rejected), with snapping to the playhead, other clips' edges, and markers; drag any clip in a multi-selection and the whole group moves together.
- Duplicate — one click to repeat a clip, or duplicate a whole multi-clip selection as one block.
- Freeze frame — turn the current frame into a still image dropped into the sequence.
- Undo/redo — step backward and forward through your edit history.
- Timeline thumbnails and waveform — see a filmstrip and audio waveform on every clip, not just a plain colored bar.
- Timeline zoom — zoom in for frame-accurate trims on a long clip, zoom back out for a quick overview; the main track and every overlay track scroll together.
- Text and titles — headings, subtitles, lower thirds, quotes, callouts, and watermark text, each with its own font size, color, background (with adjustable opacity), outline, drop shadow, character spacing, line spacing, multi-line text, an entrance animation (fade/slide/pop), and an optional custom uploaded font — one click to duplicate any text layer, or drag it directly on the preview to reposition it.
- Logo / watermark image — upload an image, position it in a corner, and control its size and opacity.
- Rotate & flip — rotate any clip 90°/180°/270° and flip it horizontally or vertically, useful for footage shot the wrong way or mirrored webcam captures.
- Speed — 0.25× to 4× per clip.
- Fade in/out — per clip, video and audio together.
- Transitions — cut, fade, dip to black, dip to white, crossfade, slide, wipe, zoom, or blur between clips.
- Color filters — brightness, contrast, saturation, grayscale, and warm/cool/vintage/cinematic presets.
- Manual crop & pan — drag a focus pad to choose which part of an oversized frame stays visible, plus a zoom slider to crop in tighter than a straight fill.
- Aspect-ratio-safe guides — an optional on-screen overlay showing action-safe/title-safe zones while you compose, never present in the export.
- Silence removal — scan a clip for silent stretches and review them before removing.
- Audio normalization — one click to analyze and level out a clip's volume, or set it manually.
- Reverse clip — plays a clip's video and audio backwards in the export.
- Master audio — a project-wide volume control, mute-all, and fade in/out that apply on top of every clip's own volume and fades, plus a live level meter while playing.
- Timeline markers — press M (or the flag button) to drop a labeled marker at the playhead, for planning cuts, noting music changes, or leaving yourself notes — markers are for your own navigation and aren't burned into the export.
- Multi-track overlays — add any number of video or image overlay tracks, not just one.
- Track mute / solo / lock / hide — silence one overlay track's audio, solo it to hear it alone, lock a track to protect it from accidental edits, or hide it from the picture without deleting it.
- Split-screen — two videos side by side or stacked, with an adjustable divider.
- Picture-in-picture / video call — a main video with one or more repositionable, resizable overlay tiles, each independently positioned.
- Person cutout — automatically remove a picture-in-picture overlay's background (no green screen needed) using on-device machine learning, with edge softness, an optional drop shadow, and an optional colored outline glow — nothing uploaded, works live in the preview and the export.
- Video-call templates — Bottom strip, Side strip, or Corners one-click layouts that arrange every overlay tile at once for a multi-participant call.
- Shapes — rectangle, circle, line, and arrow annotations, each with its own color, fill, stroke width, position, and timing.
- Per-clip audio — keep, mute, replace, or mix in different audio per clip.
- Audio ducking — automatically lowers mixed-in background audio while a clip's own audio has signal, so voice and music don't fight for attention.
- Reframe for social platforms — Landscape (16:9), Square (1:1), or Vertical (9:16), with one-click YouTube/TikTok/Instagram/LinkedIn presets.
- Reframe background — when fitting the whole frame into a new shape leaves empty space, fill it with a solid color, a color gradient, a blurred/zoomed copy of your own video, or a custom uploaded image.
- Export controls — choose resolution (480p/720p/1080p), quality (small/balanced/high), and frame rate (original, or a fixed 24/30/60fps), with an estimated output size shown before you export and a Cancel button once it's running. Every export carries a small "convertam.app" mark in the bottom-right corner.
- Auto Captions — transcribe your edited timeline directly (no export needed first), edit the transcript, download SRT/VTT/TXT, or burn captions straight into your final exported video.
- Burn Subtitles — already have an .srt or .vtt file? Upload it, style it (font size, position, color, outline/background), preview it synced to playback, and burn it permanently into your exported video — no transcription, no AI, nothing uploaded.
- Clean audio — reduce hum, rumble, and hiss on a clip's audio right from its audio controls, with optional voice clarity and normalization, before it's mixed into your final export.
- Keyboard shortcuts — Space to play/pause, S to split, D to duplicate, Delete to remove, M to drop a marker, Ctrl/Cmd+Z to undo.
Non-destructive editing, explained
Every edit in Video Editor works by changing which part of your original file plays and in what order — your uploaded file itself is never modified or re-encoded until you export. That means you can trim, split, delete, and reorder freely, undo any of it, and only pay the cost of real rendering once, at export time.
Multi-track composition: split-screen, picture-in-picture, and video calls
Add a video or image overlay to unlock composition — and add as many more as you need, each becoming its own overlay track with its own settings. Split screen places the main video and one overlay side by side or one above the other, with a divider you can drag to adjust the balance (only available with exactly one overlay track). Picture-in-picture keeps your main video full-size and places each overlay as its own smaller, independently positioned and resized tile — with two or more overlays, that becomes a genuine multi-participant video call layout. Both modes crop instead of stretching mismatched aspect ratios, so no video is ever distorted. Each overlay track's header has its own mute, solo, hide, and lock buttons — mute or solo control that track's audio in the mix, hide removes it from the picture without deleting it, and lock protects its clips from accidental trims or deletes until you unlock it again.
Video-call templates
With two or more overlay tracks, the Composition panel offers one-click video-call templates — Bottom strip and Side strip line every overlay tile up along an edge at matching size; Corners places one tile in each corner. Applying a template switches every overlay to picture-in-picture mode and arranges them automatically; you can still drag any tile afterward to fine-tune its position, or switch back to per-track controls to position each one by hand.
Need to record something first?
Video Editor is exclusively for editing video you already have. To create a new recording, use the separate Screen Recorder tool — it captures your screen, a window, or a browser tab, with an optional microphone and an optional webcam picture-in-picture, entirely in your browser. Once you stop recording there, one click on "Open in Video Editor" hands it straight to your timeline here, ready to trim, compose, and export.
Exporting for TikTok, Instagram, and YouTube
Pick a Frame in the Composition panel to choose the shape your export is cropped to: Landscape (16:9) for YouTube, Square (1:1) for an Instagram feed post, or Vertical (9:16) for TikTok, Reels, and Shorts. This works with a single video too — no second clip or overlay required, the whole export is simply reframed. Use Crop to fill to fill the new frame edge to edge (cropping the sides or top/bottom as needed), or Fit whole frame to keep the entire original picture visible with space left over on the sides or top/bottom — choose what fills that space: a solid Color, a Gradient between two colors, a Blurred zoomed-in copy of your own footage (the common look for turning a 16:9 video into a clean 9:16 one), or a Custom image you upload.
Text, titles, speed, and effects
The Text & titles panel adds always-on-top text layers — pick a Heading, Subtitle, Lower third, Simple text, Watermark, Quote, or Callout starting point, then set its own font size, color, background (and its opacity), outline, drop shadow, character spacing, line spacing (multi-line text is supported), bold/italic, alignment, position, an entrance animation, and start/end timing, and duplicate any layer in one click. Upload your own font file if the built-in ones aren't enough. The Shapes panel works the same way for rectangle, circle, line, and arrow annotations — each with its own color, filled/outline style, stroke width, position or endpoints, and timing. Logo/watermark follows the same pattern for an image instead of text. Per clip, the Clip panel adds Speed (0.25× to 4×), Fade in/out (video and audio together), Rotate (90° steps) and Flip horizontal/vertical, Volume (with a one-click Normalize audio button that analyzes the clip and levels it to a consistent loudness), Reverse to play the clip backwards, Duck background to automatically lower mixed-in background audio while the clip's own audio has signal, Filters (brightness, contrast, saturation, grayscale, or a one-click preset), and Crop / pan — a draggable focus pad plus a zoom slider for choosing exactly which part of the frame survives a crop. A separate Master audio control applies a project-wide volume, mute-all, fade in/out, and a live level meter on top of every clip's own settings. Choosing a Transition — Fade, Dip to black, Dip to white, Crossfade, Slide, Wipe, Zoom, or Blur — sets that clip's fade-out (and, for Fade/Dip, the next clip's fade-in) to match automatically; every option past Dip genuinely blends both clips' video together during the transition (a push, a hard-edged reveal, a scale-in, or a soft blur dissolve, respectively) rather than fading through a color.
Common use cases
- Trimming footage — cut a long recording down to the part you need.
- Removing mistakes — split around an unwanted section and delete it.
- Reordering clips — rearrange multiple takes into the right sequence.
- Reaction and commentary videos — picture-in-picture your face over gameplay or footage.
- Interview and comparison videos — split screen two speakers or two angles side by side.
- Multi-participant video calls or panels — arrange three or four speakers with a video-call template.
- Tutorials and product demos — bring in a Screen Recorder recording, then add shape annotations, text, and captions to point things out.
Supported formats
MP4, WebM, MOV, and other video formats your browser can natively play and decode.
Privacy and temporary processing
Editing, composition, and export all happen locally in your browser — your video is never uploaded for those steps. Auto Captions is the one exception: it first renders just your edited timeline's audio locally (never the video), then sends a compressed copy of that rendered audio to our transcription provider, processed for that single request and not stored afterward. If you don't use Auto Captions, nothing about your video — or its audio — ever reaches Convertam's servers.
Limitations
- Overlay tiles, not an equal grid: with picture-in-picture and video-call layouts, the main track stays full-size underneath and overlays are placed as smaller tiles on top — there's no mode that shrinks every participant, main track included, into equal-size grid cells.
- Rendering speed: export runs in your browser, so it takes real processing time proportional to your video's length, composition, effects, and your device's power — using speed, fades, filters, text, shapes, extra overlay tracks, or a watermark takes longer to export than a straight trim.
- Crossfade's incoming clip: during the blend, the next clip is held on its own first frame (not actively playing) while it dissolves in — this avoids any jump or repeat once its own official slot begins, at the cost of that clip not visibly moving until the crossfade finishes.
- Reversed clip preview: browsers can't natively play video backwards, so a reversed clip previews as a best-effort silent scrub rather than smooth playback — the exported video plays it backwards properly, with its audio correctly reversed too.
- Auto Captions rendering time: generating the timeline's audio for transcription happens in real time (roughly as long as your edited timeline's own length), since it plays the mix back to capture it — shorter than a full video export, but not instant for a long timeline.
- Multi-select scope: a selection can only include clips from one track at a time — it's for shared delete/duplicate actions on a sequence of clips, not for moving several clips freely across the timeline at once.
- Person cutout accuracy: background removal uses an on-device machine learning model tuned for a single person facing the camera in reasonable lighting — it isn't pixel-perfect chroma-key quality, works best with one clearly visible person, and the first time you enable it there's a brief pause while the (self-hosted, never uploaded-to) model loads.
