AIToday
AI Safety & AlignmentFortune AIPublished: Oct 2, 2026, 22:00 JST

IRC tests AI in crisis zones, finds one answer in four still fails

IRC tests AI in crisis zones, finds one answer in four still fails

3 Key Points

  1. What happened

    The International Rescue Committee's research lead says the aid sector should be 'all in on AI' after testing tools in refugee camps. Its Signpost service reached 20 million people, and an AI mentor called aprendIA is on track to reach nearly a million children in Nigeria by the end of 2026.

  2. Why it matters

    That discipline, it argues, is what lets aid groups reach more people faster and for less.

  3. What to watch

    The IRC says people in crisis should never be anyone's test case, so the real test is whether publishing failures — rather than hiding them — becomes standard practice across the sector. Its satellite-imagery pilot launches in Somalia next quarter.

WHO IT HITSAid-agency program and research staff deciding whether to deploy AI tools in crisis settings, and humanitarian donors weighing funding for them, are the audience this argument is aimed at.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

The IRC's argument lands in a sector under acute strain. The commentary notes there are more conflicts today than at any time since the Second World War, with almost 260 million people needing assistance and over 118 million forcibly displaced, while aid has contracted 25% since 2024 and clinics and schools are closing. Against that backdrop, the IRC frames AI not as a novelty but as a way to keep serving people when resources shrink — Signpost has reached 20 million people, and aprendIA is on track to reach nearly a million children in Nigeria by the end of 2026.

What distinguishes the IRC's position is the order of operations it describes. Before Signpost's AI assistant went near a client, it was run for six months across Greece, Italy, El Salvador and Kenya while moderators and protection officers scored answers on whether they were safe, client-centered and trauma-informed. The pass rate rose from 52% to 77%, but the IRC says it still failed roughly one answer in four — so the tool was not deployed on its own, every response went through a human, and the most sensitive questions never reached the model. The IRC says it publishes those failures because a sector that learns its lessons privately keeps repeating them.

The stakes the IRC describes hinge on whether that kind of discipline travels. The organization argues that because it works where the margin for error is smallest, applications like its own could help set guardrails for the whole industry, and that staying away would leave people in war, displacement and climate disaster as the last to benefit from tools everyone else enjoys. Whether other aid groups adopt the same publish-your-failures posture, or treat AI as something to avoid, is likely to shape how quickly these tools reach the people the IRC says need them most.

FAQ
Did the IRC deploy its AI assistant on its own after testing?
No. It still failed roughly one answer in four, so every response went through a human and the most sensitive questions never reached the model at all.
How much did the IRC's AI assistant improve during testing?
Over six months across Greece, Italy, El Salvador and Kenya, its pass rate climbed from 52% to 77%, and staff reported working about 70% faster.
What else is the IRC testing AI for?
It is piloting AI and satellite imagery to find children missing from census records, with pilots launching in Somalia next quarter, and testing photo-to-digital capture for malnutrition treatment data.

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleAI博覧会 in 博多: 5 sessions, one shared problem