Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
Can LLM-Generated Misinformation Be Detected?
2024-01-16
· ICLR 2024 poster ·
anchor
Findings
IC-1175
ChatGPT can be prompted to generate misinformation with near-perfect success for implicit methods but is largely resistant to explicit misinformation requests
IC-1176
ChatGPT-generated misinformation is harder for humans to detect than human-written misinformation with the same semantics
IC-1177
LLM-generated misinformation is harder for LLM detectors to detect than human-written misinformation with the same semantics