AI video moderation, with the evidence attached
Timestamped evidence for every flag.
Echosaw checks each upload against a fixed harm policy and returns timestamped evidence.
Works with your own files — not just YouTube links.
What Echosaw moderates
Echosaw is multimodal AI video moderation software that checks each upload against a fixed harm policy — nudity, sexual content, violence, hate speech, and profanity — across video frames, the audio transcript, images, and documents. Echosaw is one engine: it ingests the media, analyzes it, scores it against the harm policy or rubric you enter, and summarizes what it found. It is built for teams that publish or review uploaded media and need a record of what was flagged and why.
What you upload
Upload video (MP4, MOV, AVI, MKV, WebM), audio (MP3, WAV, M4A, FLAC, OGG, AAC), or images (PNG, JPG, WebP, GIF, HEIC) up to 5 GB each, or documents (PDF, PPTX, TXT, Markdown) up to 25 MB, or paste a public YouTube, Vimeo, or Rumble link.
Real meeting. Real analysis. Real insights.
Explore a full Echosaw analysis of a recorded meeting: timestamped transcript, key moments, AI insights, generated outputs, and a chat grounded in the video.
- 1. Upload video, audio, images, or documents
- 2. Analyze transcripts, timelines, and key moments
- 3. Ask questions and get cited answers
What comes back
A report with a visual sensitivity rating (none, possible, or likely), an allow or review action, content warnings, and — on Growth and above — the flagged categories with confidence scores and the time ranges where they occur. Flagged media cannot be made public. Results are available in the web app and through the REST API and MCP server, so you can push findings into your CRM, ticketing, or VMS.
How Echosaw fits your workflow
Echosaw flags each issue, attaches timestamped evidence, and gates publishing: flagged media cannot be made public in the Echosaw Public Library. Your team reviews the evidence and decides what happens next, in the web app or in your own systems through the API. Echosaw analyzes media as it arrives — record in the browser, upload, or send files through the API — and returns findings as close to real time as recorded analysis gets. Need continuous feed ingestion? Contact us.
What Echosaw moderates, by modality
One analysis covers every modality in the file. Categories are fixed; they are not user-configurable.
| Modality | What is checked | Categories flagged | Evidence returned |
|---|---|---|---|
| Video frames | Frames across the whole video | Nudity, sexual activity, suggestive content, violence, graphic violence or gore, visually disturbing content | Category, confidence, start and end time of each flagged span |
| Audio transcript | The speech-to-text transcript of the audio track | Profanity, hate speech, insult, sexual, violence or threat, graphic | Category and confidence; profanity instances with their timestamps |
| Images | The image itself, plus text visible in it | Nudity, sexual content, suggestive content, violence, disturbing content | Moderation labels with confidence scores; visible text in the Screen Text tab |
| Documents | Text extracted from PDF, PPTX, TXT, and Markdown | Toxicity categories including hate speech, sexual content, harassment or abuse, profanity, insult, threats | Clean, flagged, or blocked status with the categories found |
Video content moderation across picture and sound
Most harmful video is harmful in one channel, not both. A clip can look ordinary and contain a slur in the dialogue, or be silent and show graphic violence. Echosaw treats video moderation as two checks on one file: the frames are checked for visual categories, and the transcript is checked for spoken ones. Both land in the same report.
Visual flags carry the time range where they occur. In the report, flagged moments show on the Timeline next to the scene they belong to, and content warnings appear in the Insights tab, so a reviewer jumps to the second that matters instead of watching the whole file. Click any finding and the clip plays from that moment.
How AI moderation works in Echosaw
Echosaw is one engine: it ingests the media, analyzes it, scores it against the policy or rubric, and summarizes what it found, with timestamped findings on a dashboard you act from. For AI moderation, the policy is a fixed harm policy. Machine learning models score the frames and the transcript, and Echosaw rolls those scores up into a visual sensitivity rating and a recommended action: allow, or review.
Documents are handled the same way. The extracted text is checked before the report is generated; a document with pervasive severe content is blocked rather than analyzed.
Publish gating: flagged media stays private
When you try to make media public in the Echosaw Public Library, Echosaw checks its moderation flags first. Media flagged for profanity, nudity, sexual content, hate speech, or violence cannot be published, and the response lists the flags that blocked it. You can contact support to request an appeal.
Video moderation API and MCP
The video moderation API is the same REST API that runs every analysis. Send a file or URL to POST /v1/analyze, poll for status, and read the safety section of the results. AI agents can do the same through the Echosaw MCP server, where echosaw_set_media_visibility returns the moderation flags when media cannot be published. See the API reference.
AI content moderation services vs. software
AI content moderation services usually mean one of two things: a vendor that supplies human moderators, or software that scores content for you. Echosaw is multimodal AI moderation software: it gives your team the flags, the timestamps, and the evidence; your team makes the call.
If you need to check content against your own policy rather than a fixed harm list, load that policy into rubric scoring — the rubric is your policy, and the evidence is the media — and every finding comes back with a timestamp in the source. The same review model, applied to camera footage, is Echosaw for surveillance.
Pricing
Moderation is part of every analysis, not an add-on. Plans are Starter $9, Growth $19, Pro $29, and Agency $49 per month, plus usage pricing per minute of media. Agency has the lowest rates: $0.43 per minute for audio plus video and $0.22 per minute for audio only, for videos up to 210 minutes. Example: on Agency, a 10-minute clip costs 10 × $0.43 = $4.30 in usage on top of the $49 monthly plan. Starter shows the sensitivity rating, the action, and content warnings; Growth and above add moderation evidence details. Your first three uploads are free. See pricing.
Moderation guides
- Content moderation toolsWhat content moderation software should do, and how Echosaw does it.
- AI image moderationPhoto and picture moderation with labels, confidence, and publish gating.
- SurveillanceThe same timestamped review, applied to camera footage.
- Rubric scoringScore a recording against your own policy or rubric.
Video moderation FAQ
- What is video moderation?
- Video moderation is checking a video for harmful content before it is shown to others. Echosaw does video content moderation on the picture and the sound: frames are checked for nudity, sexual content, violence, and disturbing imagery, and the transcript is checked for profanity, hate speech, insults, threats, and sexual or graphic language.
- Does Echosaw have a video moderation API?
- Yes. Submit media with POST /v1/analyze, poll GET /v1/analysis/status/{mediaId}, and read the safety section of GET /v1/analysis/results/{mediaId}. The same results are available to AI agents through the Echosaw MCP server.
- Which AI moderation tools for video content are included?
- Every video, audio, image, and document analysis includes moderation. Video gets visual moderation of frames plus moderation of the transcript; audio gets transcript moderation; images get visual moderation; documents get moderation of the extracted text.
- Does Echosaw remove content automatically?
- Echosaw flags content, attaches timestamped evidence, and gates publishing: media with moderation flags cannot be made public in the Echosaw Public Library. Your team decides what happens next, in the web app or in your own systems through the REST API and MCP server.
- Can I moderate against my own rules?
- The moderation categories are fixed. To check a video against your own written policy, use rubric scoring: load the policy as a rubric and each finding comes back with a timestamp in the source.
- How much does AI moderation cost?
- Moderation is included in every analysis. Plans are $9, $19, $29, and $49 per month plus usage pricing. Agency, at $49 per month, has the lowest rates: $0.43 per minute for audio plus video and $0.22 per minute for audio only, for videos up to 210 minutes. Example: on Agency, a 10-minute clip costs 10 × $0.43 = $4.30 in usage on top of the $49 monthly plan. Your first three uploads are free. Detailed moderation evidence starts on the $19 Growth plan.
Last updated: September 28, 2026
Ready to bring powerful multimodal AI to your media operations?
Trusted at scale to extract semantic insights, build intelligent timelines, deliver accurate transcripts, analyze audio and visual content, and generate synthetic media — with full control and security. Start with our Starter plan for $9/month — usage-based pricing so you only pay for what you analyze.