Skip to main content
Live
Main content

Instagram's AI Content label is flagging real photos and missing generated ones

Meta's detection system is tagging Canva background-removed photos as AI while fully generated Gemini images slip through untagged.

Jaeden Schafer
Editor in Chief · · 5 min read
Meta logo

Instagram's AI Content label is misfiring in both directions. Meta's system is stamping the tag on original photos that used nothing more than a background-removal tool, while fully AI-generated images posted with C2PA and SynthID watermarks are moving through the feed unlabeled. This is the second time in roughly two years the detection system has broken this way — Meta ran into an almost identical problem with its 'Made by AI' tag in February 2024.

The pattern users are reporting centers on Canva. Multiple Threads users say the AI Content label appears on any photo touched by Canva's Background Remover, even when the underlying image is a straightforward photograph. One user reported that removing a single speckle from an image was enough to get the entire photo flagged as AI on Instagram.

Canva has since told content strategist Jess Bruno that some of its assistive AI tools 'were being tagged as generative,' and says the tagging is now corrected. Canva's background removal help page states the tool 'doesn't add Canva's AI-generated content metadata to your design.' Threads users report mixed results after the fix — some Canva-edited images are still being tagged.

every time there is bg remover involved
Threads user, Instagram user report

Key facts

  • 01Instagram's AI Content label is being auto-applied to photos edited with Canva's Background Remover, a machine-learning tool, not a generative one.
  • 02Fully AI-generated images made in Google Gemini with C2PA and SynthID metadata embedded were not tagged by Instagram in tests.
  • 03The only reliable trigger for Instagram's AI Content label in testing was content created in Meta's own AI app.
  • 04Meta ran into a nearly identical labeling problem in February 2024 with its 'Made by AI' tag on Adobe-edited photos.
  • 05Test images have been live on a public Instagram account for almost two weeks with no AI tag applied, despite Meta's crackdown on AI accounts announced last week.

The distinction Meta appears to be missing is the one between assistive machine learning and generative AI. Background removal and object selection have used machine learning inside tools like Photoshop for over a decade. Those features do not synthesize new pixels from a prompt, and they are not what the AI Content label was built to flag.

Meta has been vague about its detection method since launch. In 2024 the company said it would scan for IPTC and C2PA metadata to identify AI-generated or AI-manipulated content, and promised to tune the system to better reflect 'the amount of AI used in an image.' It has not published updated criteria, and did not respond to a request for clarification on the current system.

The false positives are hitting brands as well as individual users. About Face, the cosmetics company founded by singer Halsey, had a recent Instagram post auto-tagged with the AI Content label. The brand's social manager said no generative AI was used and pointed to standard editing inside Apple's Photos app on an iPhone.

This photo was taken on my iPhone and then slightly edited by myself in the photo app. Our brand uses real artist[s] and real people to create everything you see across all our platforms
About Face social manager, Cosmetics brand founded by Halsey

Only a narrow set of iPhone editing features should trigger a generative label — Spatial Reframing, Extend, and the updated Clean Up added in iOS 27. Those Apple Intelligence tools embed Google's SynthID watermark. According to Google's Gemini-based verification, the About Face image did not contain SynthID, which makes Meta's basis for the tag unclear.

The false negatives are the more damaging half of the failure. The Verge's Jess Weatherbed posted images fully generated in Google's Nano Banana model inside Gemini — depicting scenes and clothing she had never worn — and both images carried C2PA and SynthID signals. Instagram tagged neither. In the same test batch, images edited with Canva's Background Remover, Photoshop's background eraser, and Adobe Firefly's generative features also went untagged.

Related · from this week
Meta launches Content Seal AI watermarking, three years after promising labels
Jaeden Schafer · 5 min read →

The only content that reliably triggered the AI Content label was material created inside Meta's own AI app. That result held even on a brand-new Instagram account posting AI images in quick succession — behavior that should look like an AI content farm to any moderation system. The account was not flagged, and this came after Meta announced a crackdown on AI-generated accounts last week. The test images have been live and public for almost two weeks.

There is a defensible reason for Meta to keep its detection heuristics private — publishing them helps bad actors evade the label. That defense weakens when the system tags a phone photo touched by a background eraser and misses a fully synthetic portrait carrying industry-standard watermarks. The signal-to-noise ratio of the label is now low enough that a viewer cannot use it to decide whether an image is real.

The commercial stakes for Meta are the credibility of the tag itself. Instagram sold the AI Content label as a trust primitive at a moment when generative image tools are cheap and abundant, and the label only works if it maps to something reliable. Right now it maps to Meta AI usage plus noise, which puts the company in the awkward position of having built a labeling system whose most consistent function is marking its own product. Until Meta names the metadata it actually scans for and treats C2PA and SynthID as first-class signals from other vendors, the label will keep training users to ignore it — the exact opposite of its purpose.

ShareXLinkedInEmail
AI Box

Every AI model. One chat.

The latest models from ChatGPT, Claude, Gemini, Sora, ElevenLabs — 80+ models in a single chat. Compare answers side by side. Pick the best one every time.

  • ChatGPT, Claude, Gemini, Grok, DeepSeek — in one chat
  • Generate images & video with Sora, Veo, Ideogram
  • Compare any two models side by side
  • From $8.99/mo · 80+ models, all included
Try AI Boxaibox.ai
Trusted by 3,000+ teams
Got a tip?

Working on something we should cover, or seeing a story we missed? Send leads, documents, or feedback to hello@aichatdaily.com. For sensitive tips, see our secure tips page for Signal and PGP options.

Spotted an error? Email hello@aichatdaily.com with the URL and the issue, or read our full corrections policy.

AI Box Daily briefingFree · Daily · No fluff

Stay ahead of everyone in AI.

The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.

Loved by 10,000+ AI professionals
Free forever. Unsubscribe with one click.

The briefing read inside teams at

Keep reading

More from Security

Meta logo
Security

Meta launches Content Seal AI watermarking, three years after promising labels

The new invisible watermark only covers images from Meta's newest Muse model, leaving three years of AI output undetectable by its own tool.

Jaeden Schafer5 min read
Meta logo
Security

Meta rolls back Instagram AI tagging in 3 days after opt-out backlash

The feature let users generate images of public Instagram accounts by default. It survived 72 hours before Meta pulled it.

Jaeden Schafer5 min read
Meta logo
Security

Meta contractors posed as minors to probe ChatGPT, Gemini, and Character.AI

A project called Cannes ran 45,000 prompts through rival chatbots in August 2025 alone, using dummy under-18 accounts.

Jaeden Schafer5 min read