Zürcher Nachrichten - AI's blind spot: tools fail to detect their own fakes

EUR -
AED 4.135003
AFN 74.312152
ALL 91.998807
AMD 407.29472
ANG 2.015836
AOA 1032.483179
ARS 1713.371989
AUD 1.624961
AWG 2.029499
AZN 1.91856
BAM 1.955775
BBD 2.262771
BDT 138.128209
BGN 1.895446
BHD 0.423495
BIF 3389.456062
BMD 1.125935
BND 1.437881
BOB 13.498825
BRL 5.867029
BSD 1.123485
BTN 108.843589
BWP 15.531799
BYN 3.383956
BYR 22068.333923
BZD 2.259471
CAD 1.604402
CDF 2600.911191
CHF 0.932682
CLF 0.028256
CLP 1115.712356
CNY 7.548891
CNH 7.557396
COP 3725.451706
CRC 514.793327
CUC 1.125935
CUP 26.96265
CVE 110.263571
CZK 24.454533
DJF 200.057407
DKK 7.479143
DOP 67.07913
DZD 150.570048
EGP 58.776238
ERN 16.889031
ETB 183.500221
FJD 2.530259
FKP 0.851099
GBP 0.850212
GEL 2.933107
GGP 0.851099
GHS 13.194829
GIP 0.851099
GMD 82.760588
GNF 9882.871886
GTQ 8.585889
GYD 235.006954
HKD 8.83567
HNL 30.162609
HRK 7.540056
HTG 147.118093
HUF 368.676725
IDR 20141.295479
ILS 3.436727
IMP 0.851099
INR 108.456295
IQD 1471.780921
IRR 1966490.662268
ISK 137.285738
JEP 0.851099
JMD 177.857694
JOD 0.798333
JPY 177.718077
KES 145.82811
KGS 98.463484
KHR 4556.940927
KMF 493.160101
KPW 1013.342222
KRW 1512.705907
KWD 0.347723
KYD 0.936188
KZT 504.773457
LAK 25235.672864
LBP 100605.895823
LKR 371.535184
LRD 192.10751
LSL 18.784656
LTL 3.324595
LVL 0.681068
LYD 7.194907
MAD 11.161855
MDL 20.07674
MGA 4963.935651
MKD 61.574202
MMK 2363.962729
MNT 4050.351228
MOP 9.080782
MRU 44.894418
MUR 54.214216
MVR 17.407387
MWK 1948.074747
MXN 20.447892
MYR 4.599338
MZN 71.951471
NAD 18.784656
NGN 1498.384001
NIO 41.339464
NOK 10.829927
NPR 174.149742
NZD 2.004693
OMR 0.432034
PAB 1.123485
PEN 3.89325
PGK 5.087934
PHP 70.464983
PKR 311.170466
PLN 4.386363
PYG 6572.914794
QAR 4.095147
RON 5.339077
RSD 117.478477
RUB 94.165779
RWF 1663.878431
SAR 4.214702
SBD 9.095387
SCR 15.468899
SDG 677.254284
SEK 11.304508
SGD 1.439625
SHP 0.850115
SLE 27.702139
SLL 23610.293318
SOS 642.091676
SRD 42.573311
STD 23304.589613
STN 24.499682
SVC 9.829873
SYP 14639.412412
SZL 18.780757
THB 37.722942
TJS 10.341466
TMT 3.940774
TND 3.347757
TOP 2.710982
TRY 55.314396
TTD 7.617901
TWD 35.829296
TZS 2960.361624
UAH 50.551445
UGX 4482.941887
USD 1.125935
UYU 45.279413
UZS 13250.828227
VES 974.472491
VND 29257.994447
VUV 134.924899
WST 3.139874
XAF 655.957
XAG 0.018651
XAU 0.000271910232
XCD 3.042897
XCG 2.024774
XDR 0.796095
XOF 655.957
XPF 119.331742
YER 266.062577
ZAR 18.752511
ZMK 10134.773796
ZMW 22.075714
ZWL 362.550741
SSP 6432.009114
MXV 2.313642
  • NGG

    0.7900

    76.12

    +1.04%

  • RIO

    1.3400

    94.21

    +1.42%

  • BTI

    0.2849

    52.64

    +0.54%

  • GSK

    -0.1500

    47.03

    -0.32%

  • BCC

    0.2900

    74.99

    +0.39%

  • BCE

    -0.1900

    19.73

    -0.96%

  • AZN

    -0.8000

    156.9

    -0.51%

  • CMSC

    0.0800

    20.24

    +0.4%

  • CMSD

    0.0900

    20.43

    +0.44%

  • RYCEF

    0.1200

    19.42

    +0.62%

  • BP

    0.2900

    44.79

    +0.65%

  • RBGPF

    0.0900

    65.09

    +0.14%

  • VOD

    0.4400

    16.82

    +2.62%

  • JRI

    0.1300

    10.9

    +1.19%

  • RELX

    -0.0700

    33.42

    -0.21%

AI's blind spot: tools fail to detect their own fakes
AI's blind spot: tools fail to detect their own fakes / Photo: Chris Delmas - AFP

AI's blind spot: tools fail to detect their own fakes

When outraged Filipinos turned to an AI-powered chatbot to verify a viral photograph of a lawmaker embroiled in a corruption scandal, the tool failed to detect it was fabricated -- even though it had generated the image itself.

Text size:

Internet users are increasingly turning to chatbots to verify images in real time, but the tools often fail, raising questions about their visual debunking capabilities at a time when major tech platforms are scaling back human fact-checking.

In many cases, the tools wrongly identify images as real even when they are generated using the same generative models, further muddying an online information landscape awash with AI-generated fakes.

Among them is a fabricated image circulating on social media of Elizaldy Co, a former Philippine lawmaker charged by prosecutors in a multibillion-dollar flood-control corruption scam that sparked massive protests in the disaster-prone country.

The image of Co, whose whereabouts has been unknown since the official probe began, appeared to show him in Portugal.

When online sleuths tracking him asked Google's new AI mode whether the image was real, it incorrectly said it was authentic.

AFP's fact-checkers tracked down its creator and determined that the image was generated using Google AI.

"These models are trained primarily on language patterns and lack the specialized visual understanding needed to accurately identify AI-generated or manipulated imagery," Alon Yamin, chief executive of AI content detection platform Copyleaks, told AFP.

"With AI chatbots, even when an image originates from a similar generative model, the chatbot often provides inconsistent or overly generalized assessments, making them unreliable for tasks like fact-checking or verifying authenticity."

Google did not respond to AFP’s request for comment.

- 'Distinguishable from reality' -

AFP found similar examples of AI tools failing to verify their own creations.

During last month's deadly protests over lucrative benefits for senior officials in Pakistan-administered Kashmir, social media users shared a fabricated image purportedly showing men marching with flags and torches.

An AFP analysis found it was created using Google's Gemini AI model.

But Gemini and Microsoft's Copilot falsely identified it as a genuine image of the protest.

"This inability to correctly identify AI images stems from the fact that they (AI models) are programmed only to mimic well," Rossine Fallorina, from the nonprofit Sigla Research Center, told AFP.

"In a sense, they can only generate things to resemble. They cannot ascertain whether the resemblance is actually distinguishable from reality."

Earlier this year, Columbia University's Tow Center for Digital Journalism tested the ability of seven AI chatbots -- including ChatGPT, Perplexity, Grok, and Gemini -- to verify 10 images from photojournalists of news events.

All seven models failed to correctly identify the provenance of the photos, the study said.

- 'Shocked' -

AFP tracked down the source of Co's photo that garnered over a million views across social media -- a middle-aged web developer in the Philippines, who said he created it "for fun" using Nano Banana, Gemini's AI image generator.

"Sadly, a lot of people believed it," he told AFP, requesting anonymity to avoid a backlash.

"I edited my post -- and added 'AI generated' to stop the spread -- because I was shocked at how many shares it got."

Such cases show how AI-generated photos flooding social platforms can look virtually identical to real imagery.

The trend has fueled concerns as surveys show online users are increasingly shifting from traditional search engines to AI tools for information gathering and verifying information.

The shift comes as Meta announced earlier this year it was ending its third-party fact-checking program in the United States, turning over the task of debunking falsehoods to ordinary users under a model known as "Community Notes."

Human fact-checking has long been a flashpoint in hyperpolarized societies, where conservative advocates accuse professional fact-checkers of liberal bias, a charge they reject.

AFP currently works in 26 languages with Meta's fact-checking program, including in Asia, Latin America, and the European Union.

Researchers say AI models can be useful to professional fact-checkers, helping to quickly geolocate images and spot visual clues to establish authenticity. But they caution that they cannot replace the work of trained human fact-checkers.

"We can't rely on AI tools to combat AI in the long run," Fallorina said.

burs-ac/sla/sms

A.Ferraro--NZN