Persuasion Tools: AI takeover without AGI or agency?
Abstract
[epistemic status: speculation] > I'm envisioning that in the future there will also be systems where you can input any conclusion that you want to argue (including moral conclusions) and the target audience, and the system will give you the most convincing arguments for it. At that point people won't be able to participate in any online (or offline for that matter) discussions without risking their object-level values being hijacked. --Wei Dai > What if most people already live in that world? A world in which taking arguments at face value is not a capacity-enhancing tool, but a security vulnerability? Without trusted filters, would they not dismiss highfalutin arguments out of hand, and focus on whether the person making the argument seems friendly, or unfriendly, using hard to fake group-affiliation signals? --Benquo > 1. AI-powered memetic warfare makes all humans effectively insane. --Wei Dai, listing nonstandard AI doom scenarios This post speculates about persuasion tools—how likely they are to get better in the future relative to countermeasures, what the effects of this might be, and what implications there are for what we should do now. To avert eye-rolls, let me say up front that I don’t think the world is likely to be driven insane by AI-powered memetic warfare. I think progress in persuasion tools will probably be gradual and slow, and defenses will improve too, resulting in an overall shift in the balance that isn’t huge: a deterioration of collective epistemology, but not a massive one.
Cite this
@online{Kokotajlo2020b,
title = {Persuasion Tools: AI takeover without AGI or agency?},
author = {Daniel Kokotajlo},
url = {https://www.lesswrong.com/s/dZMDxPBZgHzorNDTt/p/qKvn7rxP2mzJbKfcA},
year = {2020},
date = {2020-11-20},
howpublished = {LessWrong},
keywords = {},
pubstate = {published},
tppubtype = {online}
}