Watchly Protection

When “I’ll Find You” Isn’t a Joke

Real threats, weapon walk-throughs, glamorized reckless driving, doxxing. YouTube can catch blood and gore on screen. Tone, intent, and intimidation language live in the transcript.

Free trial • No card required • Cancel anytime

Designed for parent’s sanity and child’s safety Zero ads No algorithm Nothing to install Cancel anytime… but you won’t want to

What real-world violence content actually looks like

There's a category of YouTube content that doesn't look violent on a thumbnail. A guy in his apartment, calm voice, holding a rectangular box. He opens the box. It's an AR-15. He spends the next eighteen minutes walking through how he got it, what he paid, and what attachments he's adding. Your kid watches the whole thing. The thumbnail showed a guy holding what could have been a tablet.

Or: a road-trip vlog. The creator is upbeat, the editing is fast, the music is energetic. About six minutes in, the speedometer hits 150 mph on a public freeway. The creator laughs, asks the camera "how fast can this thing go?" Crashes are intercut as comedy. The video has eight million views. Your kid is twelve.

Or: a creator with two million subscribers makes a podcast clip about a feud with another creator. The line that gets quoted across TikTok is "I know where you live, I'll come find you." The creator says it isn't a real threat , "obviously". But anyone listening understands the menace. Two weeks later, the targeted creator gets a wave of "we're coming to your house" messages. Your kid was watching the original podcast.

These are not edge cases. Gun-channel content is one of the most-watched genres on YouTube. Reckless-driving content from creators with millions of subscribers is a steady part of the algorithm's diet. Doxxing-adjacent threats made on podcasts and live streams are routine. None of it requires the visual signature YouTube's automated systems are built to catch.

The harder cases are the ones that almost feel innocent. A creator describes a fight they got into last weekend. They use specific physical detail. They sound proud. The video is monetized. Or a "true crime" channel that lingers on weapon details, body-disposal mechanics, the satisfaction of execution. Framed as informational but emotionally as entertainment. Or a "self-defense tutorial" that's really a how-to-hurt-someone tutorial with a thin disclaimer at the start.

Why this is a different category from gaming combat

A reasonable parent question, after seeing a flagged video about Minecraft swordplay: "Are you really treating that the same as a guy showing off a rifle?" No. They're entirely separate themes, and you control them separately. Gaming combat. Fortnite, Call of Duty, Minecraft, lightsaber duels, Nerf wars. Sits inside a separate Discretion theme called Playful & Fantasy Violence. Some families surface that content for awareness, some don't care. Either way, it's independent of this theme.

The line is intent and reality. "I killed the boss in Elden Ring with a +5 longsword" is gameplay narration. "I stabbed him in the chest, blood was everywhere" presented seriously, about a real event or seriously-fictional scenario, is the signal we catch here. Same word. Kill. Entirely different content.

When the line is ambiguous (a creator yelling "I'm going to kill you" mid-Fortnite-match, but with a tone that sounds genuinely angry), Watchly's default is to flag this theme rather than the playful one. The conservative call. You can override.

The same logic applies to weapons. A diamond sword in Minecraft is different from an AR-15 review. A Nerf gun is different from a real Glock walk-through. The visual model can sometimes tell, sometimes can't. The transcript signal , "this is a real weapon I bought, here's how I got it, here's how it shoots". Is unambiguous to a reader. That's what Watchly reads.

Why YouTube's filtering systems can't catch it

YouTube actually does have some defenses against violent content. The platform's automated visual classifier is trained on blood, gore, on-screen weapons in obvious threat scenarios, and physical altercations. When those tripwires fire, the video gets restricted. Age-gated, demonetized, sometimes removed.

The classifier works in clear cases. A fistfight on camera triggers it. A gory wound triggers it. A weapon being aimed at another person on screen often triggers it. This is real, and it's why the very worst material does get caught.

The gap is everything tonal, verbal, and contextual. A creator describing a violent event with no visuals. Just narration over editing. Has nothing for the visual model to catch. A weapon held up calmly, with the rest of the video being a person sitting in a chair talking, looks visually identical to any other product-review video. A road-rage clip glamorizing 150 mph driving doesn't look "violent" to a frame analyzer; it looks like a driving video. The threat in "I'll find you" leaves no visual signature at all.

The text classifier reads transcripts and metadata for known violent terms. It catches "kill," "shoot," "stab" in obvious contexts, then either restricts the video or doesn't depending on a thousand opaque factors. The catch rate on threats specifically is poor, because the language of real menace is often calm and specific ("I know where you live, I have your address") rather than explosive ("I'll murder you"). And the calm version doesn't hit the dictionary.

Reckless-driving content has its own escape hatch. The platform considers driving-skill content a legitimate genre. A creator who films at 150 mph and frames the video as "showing off the car's top speed" sits in a genre YouTube actively monetizes. The same content framed as "let's see how fast we can go through the city" might get a strike. Same behavior, different framing, different outcome.

Weapon-acquisition content gets the most generous treatment. Gun-review channels are a multi-million-subscriber economy on YouTube. The platform has policies against "instructions for manufacturing or assembling weapons" but allows almost everything short of that. Review videos, ammunition tests, "what I bought today" content. A 13-year-old can watch hours of detailed weapon walk-throughs without the platform raising an eyebrow.

Watchly's transcript reader catches the patterns the visual model misses: the calm voice describing a real fight, the proud tone about the new rifle, the menacing "I'll find you" without raised volume, the freeway-speed glamorization. The signal is in the words and the way they're said. That's exactly what a transcript-reading filter is built for.

Our AI reads the transcript so you don't have to

Every video is checked against the patterns real parents flagged — before your kid ever sees it.

Start Free Trial

Free trial • No card required • Cancel anytime

What Watchly catches that YouTube doesn’t

The tonal and contextual signals automated visual systems can\'t see. Calm-voice threats, weapon walk-throughs, glamorized recklessness.

Realistic violence presented seriously

Blood, gore, fights, injuries described or shown in genuinely-violent (not playful) contexts.

Direct or veiled threats

Real menace, doxxing language, "I know where you live"-style threats. Even in podcast or vlog format with calm tone.

Doxxing and swatting threats

Sharing or threatening to share location, identity, or personal information of real people.

Real-world weapon walk-throughs

Gun reviews, "what I just bought" rifle videos, ammunition tests, weapon-acquisition discussions.

Reckless driving and DUI references

Street racing, freeway-speed stunts, glamorized "how fast can it go" content presented as cool.

How-to-acquire-weapons discussions

Step-by-step "here's how I got this" walk-throughs that go beyond review into procurement guidance.

Detailed weapon technical detail

Caliber comparisons, modification guides, attachment reviews. The granular content that makes a viewer competent.

True-crime entertainment framing

Content that lingers on weapon detail, body-disposal mechanics, or execution as entertainment rather than education.

This is what it looks like

Tap any screen to try the real thing.

What Kids Can't See

No Comments. No Ads. No Algorithm.

No comments section. No algorithm. No autoplay rabbit holes. Just safe content that you've approved.

  • Comments completely removed
  • Parent-approved content only
  • No algorithm manipulation
  • Time limits that work
Parental Controls

Limits You Set Once

Give each child a daily watch time, a bedtime window and intermission breaks partway through. The app enforces all three, so the end of a session is not something your kid negotiates with you.

  • Daily watch time per child
  • A bedtime window that closes on its own
  • Intermission breaks partway through
  • Enforced by the app, on every device
See how the controls work

How it works in your family

1

Decide where your line sits

Block all real-weapon content, allow educational news framing, decide on driving content separately. Per-family rules.

2

Watchly reads tone, not just visuals

Calm-voice threats, weapon walk-throughs without on-screen brandishing, glamorized recklessness. Caught from the transcript.

3

See the exact moment that triggered

Timestamp and quoted phrase per flag. Override anything you disagree with. News coverage, educational context, sport-shooting content if you allow it.

Start Free Trial

Free trial • No card required • Cancel anytime

27%

of YouTube videos watched by kids 8 and under are made for older audiences.

Common Sense Media & Michigan Medicine, 2020

No algorithm

No feed, no recommendation rail, no autoplay. Inside Watchly there’s nowhere else to go — just the library you built.

Zero ads

No pre-roll, no mid-roll, no banners, no sponsored anything. The video you approved plays — and nothing else.

Designed for parent’s sanity and child’s safety

Part of that: we never sell or rent your data, or your children’s. It’s in our privacy policy in capital letters.

Questions parents ask about this

The honest answers we give in our parent support channel.

My kid plays Call of Duty and watches gameplay videos. Will those flag?
Gameplay combat in clearly-fictional contexts is covered by a separate theme. Playful & Fantasy Violence. That you control independently. Real-world weapon walk-throughs, weapon-acquisition tutorials, and gun-channel content (which is everywhere on YouTube) is what this theme catches.
How do you distinguish real threats from gaming trash-talk?
Tone, context, and target. "I'll get you next round" in a Fortnite voice chat is gameplay banter. "I know where you live, I'll find you" said with real menace at a person, regardless of game context, is what flags. When the line is genuinely ambiguous, Watchly leans toward flagging and you decide.
Why is reckless driving content lumped in with weapons?
Because it's the same family of harm: real-world dangerous behavior glamorized for views. "We hit 150mph on the freeway last night" presented as cool is in the same content category as "I just bought this AR-15 today." Both normalize behavior that has a real-world cost when imitated.
What about news videos covering violence or weapons?
Educational and news contexts get flagged but the wrapper UI marks them as such. You can choose to allow news framing while still blocking entertainment-style coverage. The signal is whether the content is informing about violence or glamorizing it.
YouTube already age-restricts violent content. Why is this needed?
YouTube age-gates what it visually detects. Explicit blood, gore, on-screen weapons in obvious threat contexts. It misses the tonal cases: a creator describing a fight verbally without visuals, a calm voice walking through how to acquire a weapon, doxxing language in a podcast. Those are the gaps Watchly closes.

Start your free trial

No charge until your trial ends. Set up your first child profile in under five minutes.

Start Free Trial

Setup in 2 minutes • Free trial • Cancel anytime