AI video moderation, with the evidence attached

Timestamped evidence for every flag.

Echosaw checks each upload against a fixed harm policy and returns timestamped evidence.

Moderate a video free

Works with your own files — not just YouTube links.

Encrypted in transit (TLS 1.2+) and at rest (AES-256)Cited, timestamped answersFiles up to 5GB
Video, audio, images, documentsTimestamped evidenceFlagged media stays privateWorks with private files

What Echosaw moderates

Echosaw is multimodal AI video moderation software that checks each upload against a fixed harm policy — nudity, sexual content, violence, hate speech, and profanity — across video frames, the audio transcript, images, and documents. Echosaw is one engine: it ingests the media, analyzes it, scores it against the harm policy or rubric you enter, and summarizes what it found. It is built for teams that publish or review uploaded media and need a record of what was flagged and why.

What you upload

Upload video (MP4, MOV, AVI, MKV, WebM), audio (MP3, WAV, M4A, FLAC, OGG, AAC), or images (PNG, JPG, WebP, GIF, HEIC) up to 5 GB each, or documents (PDF, PPTX, TXT, Markdown) up to 25 MB, or paste a public YouTube, Vimeo, or Rumble link.

See It In Action

Real meeting. Real analysis. Real insights.

Explore a full Echosaw analysis of a recorded meeting: timestamped transcript, key moments, AI insights, generated outputs, and a chat grounded in the video.

  1. 1. Upload video, audio, images, or documents
  2. 2. Analyze transcripts, timelines, and key moments
  3. 3. Ask questions and get cited answers

What comes back

A report with a visual sensitivity rating (none, possible, or likely), an allow or review action, content warnings, and — on Growth and above — the flagged categories with confidence scores and the time ranges where they occur. Flagged media cannot be made public. Results are available in the web app and through the REST API and MCP server, so you can push findings into your CRM, ticketing, or VMS.

How Echosaw fits your workflow

Echosaw flags each issue, attaches timestamped evidence, and gates publishing: flagged media cannot be made public in the Echosaw Public Library. Your team reviews the evidence and decides what happens next, in the web app or in your own systems through the API. Echosaw analyzes media as it arrives — record in the browser, upload, or send files through the API — and returns findings as close to real time as recorded analysis gets. Need continuous feed ingestion? Contact us.

What Echosaw moderates, by modality

One analysis covers every modality in the file. Categories are fixed; they are not user-configurable.

ModalityWhat is checkedCategories flaggedEvidence returned
Video framesFrames across the whole videoNudity, sexual activity, suggestive content, violence, graphic violence or gore, visually disturbing contentCategory, confidence, start and end time of each flagged span
Audio transcriptThe speech-to-text transcript of the audio trackProfanity, hate speech, insult, sexual, violence or threat, graphicCategory and confidence; profanity instances with their timestamps
ImagesThe image itself, plus text visible in itNudity, sexual content, suggestive content, violence, disturbing contentModeration labels with confidence scores; visible text in the Screen Text tab
DocumentsText extracted from PDF, PPTX, TXT, and MarkdownToxicity categories including hate speech, sexual content, harassment or abuse, profanity, insult, threatsClean, flagged, or blocked status with the categories found

Video content moderation across picture and sound

Most harmful video is harmful in one channel, not both. A clip can look ordinary and contain a slur in the dialogue, or be silent and show graphic violence. Echosaw treats video moderation as two checks on one file: the frames are checked for visual categories, and the transcript is checked for spoken ones. Both land in the same report.

Visual flags carry the time range where they occur. In the report, flagged moments show on the Timeline next to the scene they belong to, and content warnings appear in the Insights tab, so a reviewer jumps to the second that matters instead of watching the whole file. Click any finding and the clip plays from that moment.

How AI moderation works in Echosaw

Echosaw is one engine: it ingests the media, analyzes it, scores it against the policy or rubric, and summarizes what it found, with timestamped findings on a dashboard you act from. For AI moderation, the policy is a fixed harm policy. Machine learning models score the frames and the transcript, and Echosaw rolls those scores up into a visual sensitivity rating and a recommended action: allow, or review.

Documents are handled the same way. The extracted text is checked before the report is generated; a document with pervasive severe content is blocked rather than analyzed.

Publish gating: flagged media stays private

When you try to make media public in the Echosaw Public Library, Echosaw checks its moderation flags first. Media flagged for profanity, nudity, sexual content, hate speech, or violence cannot be published, and the response lists the flags that blocked it. You can contact support to request an appeal.

Video moderation API and MCP

The video moderation API is the same REST API that runs every analysis. Send a file or URL to POST /v1/analyze, poll for status, and read the safety section of the results. AI agents can do the same through the Echosaw MCP server, where echosaw_set_media_visibility returns the moderation flags when media cannot be published. See the API reference.

AI content moderation services vs. software

AI content moderation services usually mean one of two things: a vendor that supplies human moderators, or software that scores content for you. Echosaw is multimodal AI moderation software: it gives your team the flags, the timestamps, and the evidence; your team makes the call.

If you need to check content against your own policy rather than a fixed harm list, load that policy into rubric scoring — the rubric is your policy, and the evidence is the media — and every finding comes back with a timestamp in the source. The same review model, applied to camera footage, is Echosaw for surveillance.

Pricing

Moderation is part of every analysis, not an add-on. Plans are Starter $9, Growth $19, Pro $29, and Agency $49 per month, plus usage pricing per minute of media. Agency has the lowest rates: $0.43 per minute for audio plus video and $0.22 per minute for audio only, for videos up to 210 minutes. Example: on Agency, a 10-minute clip costs 10 × $0.43 = $4.30 in usage on top of the $49 monthly plan. Starter shows the sensitivity rating, the action, and content warnings; Growth and above add moderation evidence details. Your first three uploads are free. See pricing.

Video moderation FAQ

What is video moderation?
Video moderation is checking a video for harmful content before it is shown to others. Echosaw does video content moderation on the picture and the sound: frames are checked for nudity, sexual content, violence, and disturbing imagery, and the transcript is checked for profanity, hate speech, insults, threats, and sexual or graphic language.
Does Echosaw have a video moderation API?
Yes. Submit media with POST /v1/analyze, poll GET /v1/analysis/status/{mediaId}, and read the safety section of GET /v1/analysis/results/{mediaId}. The same results are available to AI agents through the Echosaw MCP server.
Which AI moderation tools for video content are included?
Every video, audio, image, and document analysis includes moderation. Video gets visual moderation of frames plus moderation of the transcript; audio gets transcript moderation; images get visual moderation; documents get moderation of the extracted text.
Does Echosaw remove content automatically?
Echosaw flags content, attaches timestamped evidence, and gates publishing: media with moderation flags cannot be made public in the Echosaw Public Library. Your team decides what happens next, in the web app or in your own systems through the REST API and MCP server.
Can I moderate against my own rules?
The moderation categories are fixed. To check a video against your own written policy, use rubric scoring: load the policy as a rubric and each finding comes back with a timestamp in the source.
How much does AI moderation cost?
Moderation is included in every analysis. Plans are $9, $19, $29, and $49 per month plus usage pricing. Agency, at $49 per month, has the lowest rates: $0.43 per minute for audio plus video and $0.22 per minute for audio only, for videos up to 210 minutes. Example: on Agency, a 10-minute clip costs 10 × $0.43 = $4.30 in usage on top of the $49 monthly plan. Your first three uploads are free. Detailed moderation evidence starts on the $19 Growth plan.

Last updated: September 28, 2026

Ready to bring powerful multimodal AI to your media operations?

Trusted at scale to extract semantic insights, build intelligent timelines, deliver accurate transcripts, analyze audio and visual content, and generate synthetic media — with full control and security. Start with our Starter plan for $9/month — usage-based pricing so you only pay for what you analyze.