Berliner Boersenzeitung - AI's blind spot: tools fail to detect their own fakes

EUR -
AED 4.252233
AFN 76.408927
ALL 92.707989
AMD 423.133445
ANG 2.07206
AOA 1062.783383
ARS 1722.544132
AUD 1.629599
AWG 2.085336
AZN 1.967476
BAM 1.953886
BBD 2.332075
BDT 142.07899
BGN 1.96391
BHD 0.43667
BIF 3456.869129
BMD 1.157716
BND 1.477933
BOB 13.448652
BRL 6.021166
BSD 1.157851
BTN 110.641147
BWP 15.587664
BYN 3.521905
BYR 22691.233421
BZD 2.328789
CAD 1.605688
CDF 2631.48879
CHF 0.939359
CLF 0.026911
CLP 1059.159755
CNY 7.806712
CNH 7.806733
COP 3629.740638
CRC 520.083828
CUC 1.157716
CUP 30.679474
CVE 110.155673
CZK 24.204388
DJF 205.749615
DKK 7.476107
DOP 67.752757
DZD 153.955413
EGP 58.098353
ERN 17.36574
ETB 187.309699
FJD 2.552937
FKP 0.855548
GBP 0.854759
GEL 3.022383
GGP 0.855548
GHS 12.792985
GIP 0.855548
GMD 85.096825
GNF 10171.904595
GTQ 8.835117
GYD 242.246437
HKD 9.082103
HNL 31.040193
HRK 7.533607
HTG 151.50963
HUF 364.136987
IDR 20631.656673
ILS 3.434098
IMP 0.855548
INR 110.793247
IQD 1516.894514
IRR 1591381.962892
ISK 142.202466
JEP 0.855548
JMD 183.430326
JOD 0.820868
JPY 184.664383
KES 149.808188
KGS 101.241917
KHR 4688.749507
KMF 494.345281
KPW 1041.944732
KRW 1639.245284
KWD 0.357491
KYD 0.964951
KZT 533.734889
LAK 26120.076366
LBP 103677.400969
LKR 384.193674
LRD 210.161419
LSL 18.717024
LTL 3.418434
LVL 0.700291
LYD 7.356794
MAD 10.733886
MDL 19.933389
MGA 4987.050456
MKD 61.459003
MMK 2430.957301
MNT 4163.321731
MOP 9.356473
MRU 46.42053
MUR 54.331358
MVR 17.887029
MWK 2007.781305
MXN 19.725623
MYR 4.701252
MZN 73.989667
NAD 18.717024
NGN 1568.51947
NIO 42.607713
NOK 10.9069
NPR 177.024307
NZD 1.962114
OMR 0.445147
PAB 1.157851
PEN 3.897332
PGK 5.126912
PHP 71.422999
PKR 321.35063
PLN 4.315687
PYG 6974.078372
QAR 4.232599
RON 5.241095
RSD 117.390112
RUB 98.347085
RWF 1699.228229
SAR 4.337367
SBD 9.317973
SCR 15.979876
SDG 695.210715
SEK 11.017578
SGD 1.479439
SHP 0.857712
SLE 28.360339
SLL 24276.724575
SOS 661.743234
SRD 43.743725
STD 23962.383592
STN 24.475708
SVC 10.131944
SYP 15052.623205
SZL 18.721177
THB 38.261936
TJS 10.68146
TMT 4.063583
TND 3.383542
TOP 2.787502
TRY 55.4561
TTD 7.846213
TWD 36.909023
TZS 3056.373722
UAH 51.795794
UGX 4307.724673
USD 1.157716
UYU 46.393955
UZS 13729.551592
VES 893.266858
VND 30337.947541
VUV 135.833049
WST 3.165764
XAF 655.300967
XAG 0.017572
XAU 0.000262
XCD 3.128785
XCG 2.086847
XDR 0.818565
XOF 655.306622
XPF 119.331742
YER 274.60945
ZAR 18.801188
ZMK 10420.835634
ZMW 21.813351
ZWL 372.784077
  • BCC

    -1.2600

    81.98

    -1.54%

  • NGG

    0.2500

    81.3

    +0.31%

  • GSK

    0.7600

    50.28

    +1.51%

  • BCE

    -0.1050

    23.365

    -0.45%

  • CMSC

    -0.1100

    21.34

    -0.52%

  • RIO

    1.5500

    97.23

    +1.59%

  • BTI

    -1.3300

    55.73

    -2.39%

  • JRI

    -0.1300

    12.48

    -1.04%

  • RYCEF

    0.4600

    21.25

    +2.16%

  • CMSD

    -0.0328

    21.18

    -0.15%

  • RBGPF

    -3.5100

    68.65

    -5.11%

  • VOD

    -0.2200

    16.2

    -1.36%

  • AZN

    0.4600

    156.91

    +0.29%

  • RELX

    -0.8600

    33.57

    -2.56%

  • BP

    0.3250

    42.855

    +0.76%

AI's blind spot: tools fail to detect their own fakes
AI's blind spot: tools fail to detect their own fakes / Photo: Chris Delmas - AFP

AI's blind spot: tools fail to detect their own fakes

When outraged Filipinos turned to an AI-powered chatbot to verify a viral photograph of a lawmaker embroiled in a corruption scandal, the tool failed to detect it was fabricated -- even though it had generated the image itself.

Text size:

Internet users are increasingly turning to chatbots to verify images in real time, but the tools often fail, raising questions about their visual debunking capabilities at a time when major tech platforms are scaling back human fact-checking.

In many cases, the tools wrongly identify images as real even when they are generated using the same generative models, further muddying an online information landscape awash with AI-generated fakes.

Among them is a fabricated image circulating on social media of Elizaldy Co, a former Philippine lawmaker charged by prosecutors in a multibillion-dollar flood-control corruption scam that sparked massive protests in the disaster-prone country.

The image of Co, whose whereabouts has been unknown since the official probe began, appeared to show him in Portugal.

When online sleuths tracking him asked Google's new AI mode whether the image was real, it incorrectly said it was authentic.

AFP's fact-checkers tracked down its creator and determined that the image was generated using Google AI.

"These models are trained primarily on language patterns and lack the specialized visual understanding needed to accurately identify AI-generated or manipulated imagery," Alon Yamin, chief executive of AI content detection platform Copyleaks, told AFP.

"With AI chatbots, even when an image originates from a similar generative model, the chatbot often provides inconsistent or overly generalized assessments, making them unreliable for tasks like fact-checking or verifying authenticity."

Google did not respond to AFP’s request for comment.

- 'Distinguishable from reality' -

AFP found similar examples of AI tools failing to verify their own creations.

During last month's deadly protests over lucrative benefits for senior officials in Pakistan-administered Kashmir, social media users shared a fabricated image purportedly showing men marching with flags and torches.

An AFP analysis found it was created using Google's Gemini AI model.

But Gemini and Microsoft's Copilot falsely identified it as a genuine image of the protest.

"This inability to correctly identify AI images stems from the fact that they (AI models) are programmed only to mimic well," Rossine Fallorina, from the nonprofit Sigla Research Center, told AFP.

"In a sense, they can only generate things to resemble. They cannot ascertain whether the resemblance is actually distinguishable from reality."

Earlier this year, Columbia University's Tow Center for Digital Journalism tested the ability of seven AI chatbots -- including ChatGPT, Perplexity, Grok, and Gemini -- to verify 10 images from photojournalists of news events.

All seven models failed to correctly identify the provenance of the photos, the study said.

- 'Shocked' -

AFP tracked down the source of Co's photo that garnered over a million views across social media -- a middle-aged web developer in the Philippines, who said he created it "for fun" using Nano Banana, Gemini's AI image generator.

"Sadly, a lot of people believed it," he told AFP, requesting anonymity to avoid a backlash.

"I edited my post -- and added 'AI generated' to stop the spread -- because I was shocked at how many shares it got."

Such cases show how AI-generated photos flooding social platforms can look virtually identical to real imagery.

The trend has fueled concerns as surveys show online users are increasingly shifting from traditional search engines to AI tools for information gathering and verifying information.

The shift comes as Meta announced earlier this year it was ending its third-party fact-checking program in the United States, turning over the task of debunking falsehoods to ordinary users under a model known as "Community Notes."

Human fact-checking has long been a flashpoint in hyperpolarized societies, where conservative advocates accuse professional fact-checkers of liberal bias, a charge they reject.

AFP currently works in 26 languages with Meta's fact-checking program, including in Asia, Latin America, and the European Union.

Researchers say AI models can be useful to professional fact-checkers, helping to quickly geolocate images and spot visual clues to establish authenticity. But they caution that they cannot replace the work of trained human fact-checkers.

"We can't rely on AI tools to combat AI in the long run," Fallorina said.

burs-ac/sla/sms

(K.Lüdke--BBZ)