AIToday
AI Safety & AlignmentAI Business & IndustryThe Verge AIPublished: Sep 4, 2026, 22:01 JST2 min read

Instagram's AI labels misfire again, flagging real photos

Instagram's AI labels misfire again, flagging real photos

Key takeaway

  • Instagram's AI content labels are wrongly tagging real photos. Users report flags after simple edits like background removal.

  • The system also misses some actual AI images.

  • Meta hasn't clarified how it decides what gets labeled.

3 Key Points

  1. What happened

    Users report Instagram's "AI Content" label is wrongly appearing on images they edited with tools like Canva's Background Remover, and even on photos taken with an iPhone and lightly edited. Meanwhile, some AI-generated images go untagged.

  2. Why it matters

    The visible labels are meant to help people trust what they see, but the misfires are leaving users confused about what is real. This echoes a 2024 problem when Meta's "Made by AI" label swept up photos with minor generative edits.

  3. What to watch

    Meta hasn't explained what signals trigger the label. One journalist's tests suggest only images edited or fully generated with Meta's own AI app reliably get the tag, while other AI tools' outputs went unlabeled.

Ask the AI about this article →

Context & Analysis

This story follows a similar episode in 2024, when Instagram's "Made by AI" label wrongly flagged photos with minor generative edits. Meta then said it would adjust its approach to reflect "the amount of AI used in an image," but it has not revealed what specific indicators it scans for. Now the same kind of confusion is resurfacing, and Meta has not responded to requests for clarification.

The problem appears to stem from vague detection methods, as even simple assistive editing tools—like background removal, which uses machine learning but not text-to-picture generation—can trigger the label. Canva says some of its tools were wrongly categorized, but reports persist. Meanwhile, actual AI-generated content sometimes slips through, as one journalist found that only Meta's own AI outputs reliably received the tag in tests.

For ordinary users, the result is a trust deficit: if real photos get flagged and some fake ones don't, the labels lose their purpose. The photographer and brands like Halsey's cosmetics company have publicly clarified that no AI was used, yet the tags remain. Meta's opaque process is leaving users to guess what will be labeled, which undermines the feature's stated goal of helping people spot synthetic content at a glance.

FAQ

What triggers the 'AI Content' label on Instagram?
According to one journalist's tests, only images edited or fully generated with Meta's own AI app reliably got tagged. Other tools like Canva's Background Remover, Photoshop, and Adobe Firefly did not trigger it in their tests.
Is Canva's Background Remover causing the false labels?
Canva acknowledged that some of its assistive AI tools were being tagged as generative and said the issue is fixed. However, some Threads users still report being tagged after using the tool, so the situation remains unclear.

Get the latest AI Safety & Alignment news every morning

For example, today's edition would include:

  • Infosys and CrowdStrike partner on AI-discovered vulnerabilitiesSiliconANGLE AI · 35m ago
  • OpenAI's Astra model thinks beyond human oversightSemafor Tech · 35m ago
  • OpenAI President Greg Brockman Discusses Astra and Alignment in Stratechery InterviewStratechery (Ben Thompson) · 3h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleNvidia exec: AI reshaping chip design