←Back to blog
May 10, 202610 min read

How to Hide Negative Instagram Comments Automatically (and in Bulk)

Hide negative, toxic and spam comments on Instagram without turning comments off — Instagram's free Hidden Words filter, bulk-managing comments, and AI auto-hide for brands getting hundreds a day.

A Reel goes viral. Two-hundred-and-some comments. The vast majority are positive — "Love this!" "Where can I get this?" "🔥🔥🔥". Then there's that 5-10%:

Every viewer who scrolls down to read comments — and IG's algorithm rewards posts where viewers do read comments — sees that 5-10% mixed in with the love. Your conversion drops. Your brand vibes deteriorate. The post that should have been a win turns into a recruitment ad for your competitor.

The short answer: turn on Instagram's free Hidden Words filter first, use Manage on a post's comments to clear many at once, and if you're getting hundreds of comments a day, add an AI auto-hide tool that judges each comment's meaning rather than a word list. The rest of this guide walks through each, including when not to hide.

Step 1: Turn on Instagram's built-in Hidden Words filter (free)

Before you pay for anything, switch on what Instagram already gives you:

  1. Go to your profile → ☰ menu → Hidden Words (on some app versions it's under Privacy → Hidden Words).
  2. Under Offensive words and phrases, turn on Hide comments. This hides comments Instagram's own filter considers offensive.
  3. Turn on Advanced comment filtering if you see it, which catches more borderline comments.
  4. Under Custom words and phrases, tap Manage custom words and phrases and add words you keep seeing: competitor handles, scam phrases ("DM me to earn"), slurs specific to your niche. Then turn on Hide comments in that section.

Hidden comments aren't deleted. You can still see and unhide them under View hidden comments on the post.

Hidden Words is genuinely useful, but it's a keyword list. It can't tell "this is fire 🔥" (praise) from "you should be fired", and it misses negativity that doesn't use a listed word: "worst customer service ever, never buying again" sails straight through.

Step 2: Bulk-manage comments on a post

When a post has already collected a pile of bad comments, you don't have to swipe them one by one:

  1. Open the post's comments.
  2. Tap the ⋯ (or Manage) button at the top of the comment list.
  3. Select the comments you want to act on (Instagram allows up to 25 at a time).
  4. Choose Delete, or tap More to Restrict or Block those accounts.

Restricting is the quiet option: the person's future comments are only visible to them unless you approve them. It's a good fit for repeat trolls who escalate when they notice they've been blocked.

This works for a cleanup after the fact. It doesn't help during the first hour after a post takes off, which is exactly when negative comments do the most damage.

Step 3: Auto-hide by meaning when you get hundreds a day

If you're getting hundreds of comments a day, neither a word list nor manual cleanup keeps up. You have three ways to handle that volume. Two are bad. The third works.

Option 1: Turn off comments (don't do this)

You can disable comments per-post or for your whole account. The downsides are immediate:

This is the equivalent of solving spam by deleting your inbox. Don't do this.

Option 2: Manually moderate (you'll burn out)

You can hide individual comments by tapping the three dots → Hide. IG even has a Comment Moderation panel inside the app. Realistically:

A manual workflow is fine for a 1-comment-an-hour pace. It does not scale to a viral moment.

Option 3: Auto-hide the actual negativity (this works)

Modern AI sentiment classification is really good at this specific task. Small models like OpenAI's gpt-4o-mini reliably label a comment as positive, neutral or negative in English and most major languages, in about 150 milliseconds and for a fraction of a cent per call.

If you can classify a comment in ~150 ms and the result is negative, you can hide that comment via Meta's official Comment Moderation API in another ~200 ms. End to end, less than half a second between "the comment is posted" and "the comment is hidden from public view".

That's faster than the troll can refresh to confirm their burn landed.

The "still DM them" piece is critical

Hiding without responding sends the wrong signal. The commenter sees their comment vanish and one of two things happens:

  1. They post an angrier comment ("WHY DID YOU HIDE MY COMMENT, censorship!")
  2. They post on a different one of your Reels, making the problem multiply

The fix is counterintuitive but consistent: hide the public comment, but still DM the commenter privately. The DM acknowledges them as a human, takes the conversation off the public stage, and gives them a place to vent or be helped. You're not censoring them — you're moving the conversation to a more appropriate channel.

ReplyAtlas's auto-shield does exactly this. The classifier marks the comment NEGATIVE → we hide via Meta's Comment Moderation API → and the same automation's regular DM still fires, going to the commenter directly. They feel heard; the public post stays clean.

How the classification actually works

We use OpenAI's gpt-4o-mini as a classifier (not for replies — the classifier is the only OpenAI call we make on the auto-shield path). The prompt is roughly:

Classify the sentiment of this Instagram comment as POSITIVE, NEUTRAL, or NEGATIVE. Comment: "[user's comment text]"

That's it. No fine-tuning, no human-in-the-loop. The model returns one word.

Why this works as well as it does:

What it costs

The underlying model call costs a fraction of a cent per comment, so even a viral post with tens of thousands of comments costs a few dollars to classify. In ReplyAtlas, classification runs on your plan's included AI credits, and most accounts never get near their limit. See pricing for what each plan includes.

When to NOT use auto-shield

A few scenarios where this isn't the right tool:

Setting it up

Auto-shield is enabled per-automation (not per-account). The toggle lives on Step 3 of the New Automation modal in ReplyAtlas. You set up your normal Comment automation — keyword, DM template, etc. — and check one box.

More detail on how it behaves: Comment Auto-Shield.

A few minutes of setup; results visible from the very next viral comment storm.

FAQ

How can I automatically hide negative comments on my brand's Instagram?

Start with Instagram's free Hidden Words filter (☰ menu → Hidden Words → turn on Hide comments, plus your own custom word list). For brands getting a lot of comments, add an AI auto-hide tool such as ReplyAtlas's Comment Auto-Shield, which hides comments based on meaning rather than a keyword list, within about a second of posting.

How do I bulk hide or delete comments on Instagram?

Open the post's comments, tap ⋯ / Manage, select up to 25 comments, then choose Delete, or More → Restrict / Block. To hide comments automatically as they arrive instead of cleaning up afterwards, use Hidden Words or an AI auto-hide tool.

Can I hide comments without turning comments off?

Yes. Hidden Words, Restrict and AI auto-hide all keep comments open for everyone else. Turning comments off entirely hurts reach, because comments are one of the engagement signals Instagram uses to distribute a post.

Does this work for Live comments too?

Yes. The same toggle covers both COMMENT and LIVE_COMMENT triggers. Story-replies and postback triggers don't have a public comment to hide, so auto-shield is silently ignored on those.

Can I set a confidence threshold?

Not in v1. The classifier returns a discrete POSITIVE / NEUTRAL / NEGATIVE label, and only NEGATIVE triggers the hide. We don't expose the underlying probability. If you want a softer threshold, route NEGATIVE comments to your /inbox via sentiment routing instead of auto-hiding.

What if the classifier is wrong?

The hide is best-effort — it doesn't block the DM from sending. So even on a mis-classification, the commenter gets a regular DM. They can also manually un-hide on the IG side: hidden comments aren't deleted, they're just marked invisible to other viewers.

Does this trigger any Meta penalty?

No. The Comment Moderation API is an officially-supported endpoint. Meta provides it precisely so creators can moderate at scale. We're not doing anything platform-side that would put your account at risk.

Does it work on competitors' posts where my account is mentioned?

No. You can only moderate comments on posts owned by your connected IG account. Comments on someone else's post — even ones that tag you — are theirs to moderate.

What languages does the classifier support?

Reliably: English, Spanish, Portuguese, French, German, Hindi, Tamil, Arabic, Italian, Dutch, Filipino, Indonesian, Japanese, Korean, Chinese (Simplified + Traditional). Less reliably: less-common languages and heavy slang. Mixed-language comments (Hinglish, Spanglish) generally classify well.


The shield doesn't make your comment section perfect. But it removes the spike of negativity that disproportionately damages your post's reach during the critical first-hour-after-posting window. That alone is usually enough to justify the upgrade.

Try ReplyAtlas free and toggle the shield on any Comment automation. Comment Auto-Shield is included on paid plans; see pricing for current plans.

Ready to try it on your own Instagram?

Free Starter plan · 1,000 DMs included · No credit card · Setup in 60 seconds.

Get started — free