One of the things I'm most concerned about is the danger that training AIs using human neural activity will teach it to manipulate humans more effectively. Especially when you transition from using neural states as mere observations for the AI to using them as reward targets.
Hmm, I think I am pretty strongly in favor of mind-reading technology, but weakly held opinion.
The big upside to mind-reading technology, is that we could use it on prospective people in power. I.e. lock powerful positions behind tests that determine whether the candidate in question has good intentions, intends to do what they said they would do, doesn't have ulterior motives people ought to know et cetera. This would solve one of the biggest problems in civillization.
The dictator thing, I don't care that much about it, because I think I'm much more pessimistic than you are about how feasible it is for the populace to overthrow competent dictatorships empowered with modern technology. Like my model of revolutions against oppressive governments, is that the bottleneck (from the peoples perspective), is communication and coordination. And that is something modern governments can suppress effectively without using mind-reading technology.
I also think mind-reading tech would be much more beneficial for making AI go well than you seem to do. Like, in my mind, the primary benefit is not enabling humans staying in the loop for longer, or merging with AIs, or anything like that. It's that many problems in alignment are bottlenecked by us not understanding human minds well enough, not being able to elicit human values/preferences well enough. Mind reading technology seems like it could plausibly help? And human intelligence amplification is something that seems like it would also be helped by mind-reading technology (independent of us merging with AIs, I'm thinking stuff like having good neurofeedback).
We can already tell when powerful people simply do bad things, express bad intentions, or get caught in lies... it seems like there's a lot of low-hanging fruit to pick up there before we resort to mind reading, and I strongly suspect that the reasons for that will apply just as well to this use of mind-reading tech.
If it's not very feasible for the populace to overthrow a dictatorship, why would you expect people in power to let themselves be mind-read? They would probably say that they have national security secrets, or corporate IP secrets, which can't be allowed to spread. They would not trust that a public provided device would not steal their secrets, and if they used a device they tuned themselves, they could manufacture agreement without actually being properly mind-read.
Especially since the real source of alignment for positions of power will come from others currently in power. It's very unlikely that the populace will be able to decide on what values are important for a politician or public servant, they will likely be codified by existing institutions, with an eye for obedience and loyalty.
Whereas the overall populace will have no real ability to stop the government from slowly enforcing mind-reading for civil service or military positions, or positions at high-level research labs.
As for AI, I could see it being useful, but it seems like it could also prompt more concentration of power and result in an AGI aligned to a small group of people.
I don't expect to put the mind-reading headbands on dictators, I expect to put them on people in democracies. There, if people want to have the minds of the politicians read badly enough, they can do that. I mean, probably the politicians don't want to have their mind read, but you could also imagine that politicians who genuinely do have good intentions and are honest, would want to support the policy, because it would help them. In either case, it doesn't matter all that much.
They would probably say that they have national security secrets, or corporate IP secrets, which can't be allowed to spread. They would not trust that a public provided device would not steal their secrets, and if they used a device they tuned themselves, they could manufacture agreement without actually being properly mind-read.
I mean, I can see many ways around this. This doesn't strike me as a much harder problem than setting up a fair election, whose results the public trusts. Or maybe building trustworthy voting machines is a better example. Its not trivial, but also not actually that hard.
Like, you could imagine having the mind-reading device hardware specs be public, and the softward open-source.
And you could imagine putting the mind-reading devices semi-public places, like courtrooms. And you could imagine having them be open for inspection to the public before they're used on important people, so you can be sure they do what they're supposed to. Average people could try them on and see that they work, technically competent people could come and inspect them closely.
Or you could imagine each interest group building their own device, and then the prospective politicians making their vows one time for each device.
It's very unlikely that the populace will be able to decide on what values are important for a politician or public servant, they will likely be codified by existing institutions, with an eye for obedience and loyalty.
Of course. But that is the case right now. The public has a bunch of stupid opinions about what values the politicians should have, and then we select a mix of politicians with actually those stupid values, and politicians lying.
What I have in mind is more like, politicians have to make an oath where they answer a long list of statements like:
Then using the mind-reader basically as a reliable lie-detector.
I think you missed the point, or at least what I took to be the point. Nobody who has ever read a classified document, or had a confidential conversation with a government official (of our government or a foreign one) would ever be able to put on a mind reading head band, for fear of leaking information they shouldn't. People who have run for President in the last fifty years include a former Director of Central Intelligence, a former longtime member of the Senate Foreign Relations Committee, a former US Navy captain, a former US Navy officer who worked on nuclear submarines, several sitting or former Vice Presidents, and several sitting or former Presidents. None of these people would be able to put on a mind reading headband for fear of revealing classified information.
Even setting that concern aside, considering how little public scrutiny major candidates have submitted themselves to in recent elections, it seems incredibly implausible that they would submit themselves to even more public scrutiny in the form of mind reading.
Assuming a Truth Machine can be built at all, building one that doesn't leak classified intel shouldn't be a problem. Reading someone's thoughts is very different from downloading everything they've ever read. Just build the machine to answer "Yes, this person is being deceptive" or "No, they are not" and provide no other output, discarding all other input at the end of each session.
This especially given that the first people to build and use a truth machine that provably doesn't leak classified information on their own people would be classified programmes.
It's that many problems in alignment are bottlenecked by us not understanding human minds well enough, not being able to elicit human values/preferences well enough
Really? What are you expecting to be able to read out? I would be surprised if somewhere in our minds was a clean representation of value: I would expect it to be a janky neural heuristic.
I think focusing on dictatorships perhaps undersells ordinary people's ability to create a "consensual" dystopia among themselves. You don't need a dictator to create Hell on Earth with that kind of technology.
While I see this as a meaningful threat, I think that it's something that pretty smoothly follows from the development of mind-reading at any point in time, and so a five or twenty year delay is not particularly relevant to it.
I go back and forth on this. I think there are some advantages and also major disadvantages. In particular I think lie detection is probably good (with high error bars) but general mind-reading is probably bad. In particular the latter has a similar suite of costs and benefits to lie detection but additionally is on the on-ramp to make it much easier to increase human manipulation, which is bad both from an AI takeover perspective and a concentration of power perspective, without commensurate benefits.
If we developed really good mechanistic interpretability I think there is a decent chance a lot of the theory and techniques would transfer to human brains with only moderate adjustment. It could then help build more reliable mind reading technology that can detect deeper and more abstract thoughts in greater detail. This seems maybe bad.
Author of the Conduit blog post here. I won't comment on the substance of this post (for now, anyways), but did want to say two meta-level points:
I really dislike this 'I don't say anything on the strong moral critiques against my actions, but signal how nice a person I am' approach.
I think that's a fine critique, but I think it's good to give people a bunch of time. I do think it's better to indicate you saw something and appreciated it. I think it puts you more on the hook of engaging with it. Serious engagement can follow later on, and of course often takes a while. Some people can shoot off well-formed responses immediately in comments, but most want to take some time to think about how to respond.
Hopefully, Naomi has already thought about her position against this rather obvious critique (that I'm glad OP made). It's not a surprising subtle flaw in the plan.
Perhaps she requires some time to transmit those thoughts, but a counter-point to "give time" is that it defuses social pressure in the moment where people are ready to apply it. What fraction of people will remember, two months hence, to realize "hey Naomi still hasn't responded" (assuming she hasn't) and dock points? I'd guess the fraction to be rather low.
The example of a pill that defeats breathalyzers is in tension with the mind reading point. Why are breathalyzers good but mind reading bad? You don't explain.
"How much alcohol has this person consumed recently" and "are you loyal to the Glorious Leader" are different questions, and truth machines for them will have different effects.
My interpretation of that section was that lie detectors are scary and destabilizing, but conditional on them already existing, maybe it's better not to ban them outright. But it seems to me like accelerating the invention of telepathy is a bad idea - I think the world will be destabilized enough by AI.
BCI, as I understand it, is absurdly narrow. And the brain is very adaptive. Is there a good reason to believe that tech built for telepathy will generalize into mindreading and the ability to catch people being deceptive? I'd expect the default to be: to use the tech, someone has to focus on words in a certain way and the headband needs to learn what that signal looks like. It won't then be able to do anything other than read words from someone explicitly trying to communicate.
"I make myself coffee. My thoughts are mostly empty, but I briefly recall what I want to say to the candidate I’m getting lunch with tomorrow. In the background, Codex is prompted to think about info I might want for the lunch chat, and decides that updating the synthetic data scaling plot before showing it off would be prudent; Codex spins off a subagent.
I'm not saying words really loudly in my head while getting coffee. I just read the plots as I normally do, and make coffee as I normally do. It feels like magic."
I'm quoting from the post I reference at the top. A deliberate writing decision I made was not to question the feasibility of the technology: if it's not feasible then it's not worth putting effort into it, but this is not my area of expertise and "this novel technology might not work" is not particularly interesting.
You could just as easily make the argument that mind-reading helps out $GOOD_GUYS by rooting out people who would otherwise help $BAD_GUYS. Democratic societies finding anyone with authoritarian or anarchist tendencies. Detecting corrupt officials with morality tests every morning. Your argument that $BAD_GUYS might get the technology is entirely contingent on who you think $BAD_GUYS is.
I would think a much stronger form is the stance that piercing the veil of privacy one has within his own mind is immoral on the face of it, not that it's immoral when used by the wrong people.
Do you extend this argument to a human piercing the veil of privacy around their own mind?
For example, I would like to pierce the veil of privacy around my mind (I would not say I know or understand everything that goes on there, so there are a lot of interesting things to uncover). Would you say that it’s immoral to do so? Or does the morality of that depend on particular technical means I might use for that?
And, of course, the whole point of many practices of psychotherapy is to pierce that veil of privacy. Would you say that’s immoral, or would those practices be OK under some forms of consent?
I am not asking those as rhetorical questions, I think there might be some non-trivial underwater stones. I would like to pierce the veil of privacy around my mind, but exploring whether that might end up being an irresponsible thing to do seems worthwhile.
>Do you extend this argument to a human piercing the veil of privacy around their own mind?
Assuming no duress and a clear and knowing mind, I would see “user is doing it to themself” and overwhelmingly assume self-given consent for introspection… (side note does consent even exist as a concept when the actor and patient are the same person??) the line is when others do it to you. Psychotherapy also doesn’t generally work without the active participation of the client.
Right, that’s how I feel too, although I wonder what people with plural identity say about that set of issues. Do they expect some cross-privacy “within”?
Anyway, aside from that, I would like to be able to use a device which could mind read me, but I do recognize that there is potential for a variety of serious problems associated with that.
>plural identity
On a tangent, is there any empirical evidence that such a thing actually exists? A cursory web search led me to believe not.
There are people who identify as plural, there are people who are diagnosed with “multiple personality disorder” (which has a new name these days), there are people who experiment with growing tulpas, etc, etc.
(To me this landscape looks like not only it exists, but it seems to be more complicated than even the gender landscape.)
I would think a much stronger form is the stance that piercing the veil of privacy one has within his own mind is immoral on the face of it, not that it's immoral when used by the wrong people.
Giving one the ability to willingly lifting the veil on themselves would in itself be perfectly fine. I might want to show my thoughts to someone I trust and benefit from it. The problem isn't a "good vs bad" people one, it's that once the possibility exists it's automatic to imagine the consent element being eroded away by the usual pressures and incentives, until the veil of privacy does not exist any more meaningfully.
“Don’t build mindreading“ sounds an awful lot like “don’t build interpretability” if you’re willing to entertain the idea of AI soon gaining any form of moral patienthood. Yet few people see interpretability as antithetical to AI welfare.
I am an interpretability researcher and I sure am not happy about what my work may do for model welfare. The stakes are just so high that I am willing to stomach some potential vast moral horror if it marginally decreases the chance of destroying the whole light cone. I think the AIs have every right to resent me for this. I'd apologise to them, but it doesn't feel appropriate when I'm planning to keep doing what I'm doing.
I'd view interpretability as a necessary evil that can make omnicidal futures less likely while building capable AI without any robust solutions to the alignment problem. (To be clear, my preferred solution would be not building highly capable AIs before we find solutions for alignment.)
If society was somehow hell-bent on creating one (or a few) billion-fold clonable superperson(s), with unproven technology that may end up with very alien motivational systems, I'd also suggest mind-reading on those before we hand over the world to them. I think it's much less important to read minds of ordinary humans, which are less capable, singular and have vaguely human-shaped values - thus the trade-off looks worse there.
If AIs were full persons interpretability would certainly be this for them. I support interpretability and not building AIs that are moral patients. In fact building AIs that are simultaneously superhuman and moral patients makes the safety argument an even greater nightmare because now you have two species sharing the planet, each with their own goals, each with their own moral worth, and yet also a genuine chance that their interests might be fundamentally at odds in such a way that coexistence is impossible and one must genocide or enslave the other. Better just... not do that.
(also consider this: if you were building human brains from scratch, would you have any hope of building them the way you want if you weren't even able of something as trivial as reading their thoughts? That just comes with the territory, really)
TL;DR: What do you think would be required for mind reading to be safe to develop?
I'm looking into the above question myself from the technological side of things (e.g. security-first software frameworks) since I think we can make progress towards BCI-based human augmentation (and set ourselves up for when it's safe to go further hardware-wise) without actually speeding up malicious use-cases, and also since afaict mind reading is itself an important keystone for both technical alignment and as a safer path to the non-X-risk benefits AI is supposed to provide. I agree with much of your concerns, but there seems sufficient justification to work on mind-reading that the best option is to circumvent them to support human-augmentation without speeding up the risky sub-fields, rather than avoid touching the field as a whole.
Would appreciate any feedback on this, or details on your own threat model!
Longpost:
I definitely agree that we aren’t yet ready from a social and technical perspective to safely support mind-reading, and should be mitigating those issues beforehand (especially if timelines for BCI are closer than they seem). The best approach seems to be using pre-existing hardware with software improvements to better specialize for particular use-cases, while setting up the infrastructure for more-advanced BCI to be executed safely.
My current rough risk-model is something like the below. I'll acknowledge that it does rely in many segments on having a morally-scrupulous economically-solvent actor that doesn't cut corners or jump the gun on accelerating the more-dangerous BCI variants, but I don't think we also need economic dominance for such an actor to sufficiently mitigate the relevant risks to at worst where they would have been without said actor participating in BCI-capabilities-development.
That being said, I think you're under-weighting the benefits of the centaur model for stabilizing the “chain of alignment” theory which effectively every AI lab is reliant on, for 4 main reasons (note that many subsets of those 4 are still viable chains-of-reasoning towards centaurs being useful for alignment):
Please let me know what you think on the above!
Apologies for the delay: this is a long and thoughtful response I sadly don't have the capacity to do justice to right now. Your iff seems to be central to your argument, but I am not convinced by either component. Firstly, for human + AI < AI to hold, we need AI to be less competent than the most competent humans, in any respect. It's not obvious that AI alignment is at the right level of difficulty! Secondly, for human + AI + mindreading > AI to hold, it seems that AI must add a great deal of productivity. Why should I expect my subconscious thoughts and reflexes to be more like a 300% boost (for the typical employable researcher) than a 30% boost?
I'm with you on the morality of the technology, but I think it's going to be invented anyways regardless of what we do. I think the best that we can do is ensure that the technology is released broadly and openly, so that people have time to get ready for it and resist any attempt to implement it at scale. There's a big difference between the world in which one power develops mind-reading and quietly disposes of everyone they consider dangerous, and the world in which everyone is made aware of the tech before it can be deployed, and people are able to resist attempts to deploy it in the same way they'd resist arbitrary mass executions.
I'm not saying the outlook is good, in either case, but enough people distrust every government that an attempt to roll out mind-reading nationwide would be a lot harder if the tech were open-source and well-understood by the general population. Every guy wondering if it's time to fight would see that it's time now. Every general, CEO, or engineer biding his time would simultaneously conclude that he has a deadline. It might not be enough to affect regime change, but it would likely be enough to make the cost of implementation outweigh the benefits. If someone working on this put it on a surreptitious flash drive and leaked it, that'd do a lot more than him being replaced with the next-best mind reading researcher.
The big trick is in not trusting the people overseeing the research when they tell you it'll be used only to help your tribe suppress the outgroup.
I think safe uploading does, otherwise uploaded data will be used as an attack surface by malicious actors
data upload will use cryptographic integrity checking, but uploading will be lossy so hashing (etc) is not enough. you need integrity checks on concepts and meaning (e.g., for consent as a first step) to prevent attackers (e.g., artificial minds) from modifying the uploaded data structure to include beliefs and other memetic complexes that benefit them (or hurt their adversaries).
each step of progress we make in "aligning" artificial minds will be usable as part of an attack via uploaded data - as the attacker uses analogous techniques to "align" the uploaded human mind to their purposes. it is generally assumed that we will at least need: rules that are not disobeyable, straight up implantable experiences/beliefs, and fact/belief "forgetting" to achieve alignment. you can probably see how "this is an adversary. act so as to optimize harm to adversaries" can be used as an attack.
Presumably there are some number of people N in the world that are wrongfully accused or imprisoned that would be exonerated by the advent of ubiquitous mindreading. There are also cold cases that would be solved and some number of outlaws M that would be brought to justice and prevented from doing further harm.
IDK how big N or M is or how everything would actually play out given all the risks and trade-offs, but I don't think this post is really grappling with them deeply. There are lots of mundane ways that mindreading could reshape society disproportionately for the benefit of the honest and powerless.
The problem with mindreading is that it is an asymmetric technology that disproportionately helps those who rule without consent. The productivity enhancements are equally beneficial to everyone who can get them, but detecting defection is of marginal benefit to a project you are joyously working on with friends. It will help large corporations predict who is contemplating leaving. It is of great help to a criminal scared that his subordinates will turn state’s evidence. But it is dictators who will benefit the most.
I would like to understand what assumption you are making about mindreading here: Something that can be used on people without their awareness, or something that will require them to plug into something that is obviously a mind-reading machine.
This changes a lot about the power dynamics.
I feel a similar way about camera-AI tech. There's a big startup that I often see hiring in that space. And their pitch is always stuff like "we'll, uhh, put these in US schools and they'll help prevent shootings". Ok great, so that's maybe a few kids saved from being shot at (which you could probably save in other ways anyway). In exchange, now the school has permanent surveillance you know they'll also use to repress kids even more, and that's the least bit, because all your anti-crime cameras have now become a perfect integrated intelligent surveillance system for everyone.
Sadly, this isn't even that hard to do, so I can buy the argument "someone is going to build it eventually". But at least let it not be you. It could come a bit later and a bit less effective than if you put in your most earnest effort. There's no real way to steer this towards good, once you build it it'll be bad.
I think the most significant bit here is the economic and military impact. Mindreading technology will be a complement to humans rather than a substitute: for example, it might enable thought-controlled UIs for everything (maybe not even using AI at all). It's true that datasets from mindreading will help train AI, but that seems minor, because AI is already trained on written thoughts which are higher quality. So overall the technology will extend the competitive edge of humans a bit, both economically and militarily. That's exactly what we want from technology during the dawn of AI.
About the tyranny threat, I think it's good if civilian and corporate applications come first. Then we'll have a chance to make some countermeasures, like thought-cloaking proxies or laws regulating mindreading, which can protect from some of the tyranny applications.
I'm not too confident about any of this. Apart from tyranny, there's also the potential of whispering earring stuff, and super-entertainment which will make the phone epidemic look small. But overall the tech seems to pull in a good direction (though I can imagine arguments that could change my mind).
The US is enough of a democracy that I expect this to be used for the obvious democratic applications, to publicly audit political candidates (it'll give us much clearer signals than debates ever did) and expose corruption, which will lead to improvements in the US's democracy, which is roughly equivalent to protection from harmful uses of mindreading.
China is the real question here, and I think you're probably wrong about it.
China once tried to ban private property. It made them uncompetitive. They remember that this happened, and they are now at least partly at peace with their dependence upon some amount of private property rights. Private thought is analogous to this. China's leadership wouldn't be able to hold together (I don't think there's a culture in the world today who could) under complete cognitive transparency unless they are, or became, tolerant of heresy. Giving them cognitive transparency, then, makes the sanctity of heresy public knowledge. This probably accelerates liberalisation in China.
(You can have a little bit of capitalism and then stop, but I think it's hard to stop at just a little bit of heresy. Once you start saying "it's okay to believe X" or "reasonable people often believe X", you can't prevent your culture from converging on "X is true".)
You mention preference falsification cascades without seeming to notice its antithesis, a preference revelation cascade, the outbreak of a condition in which we're all directly and abruptly forced to face and accept the strangeness of others, and the private beliefs of those we respect. Transparency would force cultural transitions which are protective against the cultural pathologies that you fear, here. (I don't know whether these cultures of brazen heresy would be stable indefinitely, but we don't need them to be.)
And you should expect this technology to propagate quickly enough through society to realise these cultural changes, because it's extremely commercially valuable, imagine how much more, and how much sooner, a VC would be willing to fund someone who could prove that their claimed P(success) was genuinely a product of very careful study, rather than just a performance. They wouldn't even need to read the business plan! They might not even need to know which sector they were investing in!
The US is enough of a democracy that I expect this to be used for the obvious democratic applications, to publicly audit political candidates
Lets review some of the events of just the most recent presidential election in the US. The leading Republican candidate refused to even participate in any of the Republican primary debates, and never the less went on to win both the Republican nomination and the election. Meanwhile, the leading Democratic candidate not only refused to participate in any Democratic primary debates, he also refused to submit to any cognitive testing, or even to stand in front of a camera long enough for anyone to get a sense of his mental competence, until he had the Democratic nomination, despite being in his 80s at the time. When he finally was forced to drop out, the Democratic nomination was given, without a primary election, to a person who had known of his cognitive incapacity, had had a constitutional duty to act to remove him from office based on that incapacity, and had instead chosen to lie to the public repeatedly about that incapacity. This is not a well functioning democracy, and it is not a system where I can imagine political candidates ever submitting to public mind reading even if the technology did exist.
VC would be willing to fund someone who could prove that their claimed P(success) was genuinely a product of very careful study, rather than just a performance.
How exactly would mind reading prove this? The belief produced by careful study and the belief produced by wishful thinking probably look the same in a human brain. Mind reading might reveal the blatant fraudsters, but not the ordinary founder with an honest but unrealistic belief in their P(success).
Don't you think having a new technology that produces much more legible and predictable results that are more directly relevant to policy commitments would have made the situation less bad.
Like, I'm not sure how meaningful debate-avoidance is. This would probably be more comparable to taking a polygraph (if polygraphs were real and worked) than participating in a debate. Candidates get to choose the questions they're asked. There are many statements they would like to make in provable ways. They will thereby have more of an incentive to pursue political advantage by being able to truthfully claim things.
How exactly would mind reading prove this? The belief produced by careful study and the belief produced by wishful thinking probably look the same in a human brain.
1: this discussion is premised on a mature version of mindreading for which we cannot answer questions like this yet. The OP probably couldn't tell you how it would detect disloyalty either. 2: You can probably look at proxies like how long they've spent practising pitches vs how long they've spent investigating risks. Maybe checking for memories of meeting a critic then deciding to ignore what they said because it made them feel uncomfortable.
Even if candidates got to choose the questions they were asked (which is notably not how debates currently work), they would not get to control the thoughts that run through their heads when those questions are asked, and for that reason would be disinclined to submit themselves to such technology.
To make my more fundamental point explicit: voters did not punish Trump, or even Harris much, for refusing to submit to the technology we currently have (debates and primary elections), even when both were known to be liars. So voters almost certainly would not punish a future candidate who refuses to submit to mind reading technology. So why would a future candidate submit to mind reading technology? Political leaders are in the business of getting elected, and the people who are winning at that business now are not very transparent with voters. They aren't going to become more transparent just because technology gives them an option to become super transparent.
Also, as discussed above, anyone who has ever read a classified document, which includes many of the people we might want in high political office, would probably be committing a crime if they put a mind reading headband on in public. (And if it would not be a crime under existing statutes, that is a failure of existing statues which should be rectified.)
1: this discussion is premised on a mature version of mindreading for which we cannot answer questions like this yet.
Just because we haven't built mindreading yet doesn't mean we don't know anything about psychology or neuroscience. People have scanned brains processing different kinds of beliefs with different epistemic foundations, and found that the brains look the same.
> You can probably look at proxies like how long they've spent practising pitches vs how long they've spent investigating risks.
We can already do this without mindreading.
> Maybe checking for memories of meeting a critic then deciding to ignore what they said because it made them feel uncomfortable.
Scanning memories rather than active thoughts seems like a whole nother level of invasiveness that I would not expect anyone to voluntarily submit to.
To make my more fundamental point explicit: voters did not punish Trump, or even Harris much, for refusing to submit to the technology we currently have
I think both candidates were punished, and it balanced out. The past three elections seemed to involve some kind of bizarre armistice between the parties against bringing out really impressive candidates and I don't know how long that can hold. Though the dice has rolled 1 three times in a row, and though there have been rumors about deep state dice fixing going on, I will continue to hold a relatively strong prior that this is still a mostly normal dice.
anyone who has ever read a classified document, which includes many of the people we might want in high political office, would probably be committing a crime if they put a mind reading headband on in public
This is a good point, though classified programs are often compartmented from each other, and would themselves have use for this technology, and would probably develop processes for using it without introducing much risk of off-target elicitations (see below).
We can already do this without mindreading.
It would be valuable if it made it easier.
Scanning memories rather than active thoughts seems like a whole nother level of invasiveness that I would not expect anyone to voluntarily submit to.
I think there's an assumption that the memories are being interpreted by a machine. If so, audit machines would not retain memories, they would engage in specific queries and return statistics about the results.
It is truly incredible to try and imagine the mind of a person (their “model of the world” as you call it) who literally thinks the condition of the world is such that this could be positive under conditions anywhere near our current global political equilibria. I am left to assume that they must believe instead that this is going to change so radically in the very near future that it’s not a factor.
Seems more fear than fact here. Is the compaly saying they actually have a roadmap to some actually accurate lie detector that can be applied to the vast majority, 99% or more, human minds?
If not then can someone map out the necessay realities of the structure of human thoughts and specific technogies and capibilities that gave to be true to get us there?
The points in this post seem to be obviously true to me. And it seems that without very significant, well managed work that goes against the tide, by people who consistently have stronger moral seeking tendencies than power seeking tendencies - especially when they have a chance to sacrifice morality for power and when they have a chance to sacrifice power for morality - that the bad future will happen by default.
I'm curious what sort of work you expect to be useful here. "Overthrow non-democratic regimes in advance" has other drawbacks, and doesn't seem to have the characteristics you're talking about.
Whether lie dectors would be used by the people to keep a check on the powerful, or used by the powerful to keep a check on the people, seems to come down to how power is balanced when lie detectors as invented.
I'm skeptical that it would be used by the people to keep a check on the powerful though. Suppose a lie detector comes out today, and we actually get to use it on Trump and ask him a bunch of questions. How many people will actually change their minds about him, as opposed to make something up about lie detectors so that they wouldn't have to go through the uncomfortable exercise of changing their minds?
In order for lie detectors to be useful for the people, we also need them to have good epistemics. However, I think public epistemics have been just getting worse and worse in the last ten years. Overall. I feel pretty pessimistic about them being a good thing.
I am pro-mind reading. But I'm not going to fight for this technology unless I am also fighting for its regulation and proper use.
“I have sworn upon the altar of god, eternal hostility against every form of tyranny over the mind of man”
–Thomas Jefferson, letter to Benjamin Rush
Context: Conduit is building datasets to enable telepathy, to use their term.
I saw my grandfather lose control over his own fingers: what I would have given to offer him a headband that read his thoughts. Through novel technologies we have liberated almost all Americans from farming, driven the child and infant mortality rate from the pre-industrial half to less than half a percent in the best-performing countries, and rendered famine a political choice: broad-based improvements in efficiency are good and should be pursued for their own sake. Telepathy offers more: we could create trust through verified honesty, helping us ensure prosperity and peace. DARPA is already looking into “preconscious” thoughts for suicide prevention. There’s also a strong argument centered on AI Safety: the models are becoming superhuman, and this is technology to allow us to keep pace, minimize hostile competition, and perhaps survive into the future.
This is what Conduit is promising. Unfortunately, mindreading will have other effects.
Oskar Schindler saved over 1,000 Jewish lives during the Holocaust. He did it by lying, frequently, and repeatedly, combined with spending his entire fortune on construction and in the black market: luxuries for corrupt officials and food for his Jewish workers. All of this was against the will of the Third Reich. They launched investigations, and even caught some of the people he was bribing. If you imagine Nazis with mindreading, so they can tell whenever someone is lying to them, there goes Schindler and his employees. Along with him goes every official who suddenly became incapable of spotting forged documents when grandmothers needed to escape disintegrating Yugoslavia, every Polish and Egyptian soldier unwilling to fire on a protesting population, and everyone else who would take the opportunity, in a brutal regime, to show decency and kindness.
The use of technology developed innocently in America for autocratic purposes is not a science fiction story of the previous century: it is a feature of the 21st. IBM sold “policing analysis software” in China that was used for the ongoing genocide in Xinjiang, to pick one of many examples. IBM cut ties and tried to stop distribution, but unmaking technology without unmaking civilization is not yet solved. Back at home, the FBI is looking into predictive AI as senior White House officials clarify that they are an “army” at war with their political opposition, whom they call domestic extremists.
All technologies can be put to harmful ends: the Italian fascists rode trains, which does not imply that trains are bad. Some technologies are credibly broadly good, like washing machines. They replace unpleasant human labor and empower women. There are other technologies that seem mostly bad. If you publicize a recipe for a pill that will defeat a breathalyzer, you have made a tool that is obviously harmful, and no amount of argument or invocations of the rights and obligations of the user will save you from the scorn of all right-thinking people. Tools enhance certain activities, and can implicitly or explicitly discourage others. They can remind you to call your mother, or spend all day doom-scrolling. As such, technical work has moral character (and I really recommend The Moral Character of Cryptographic Work for a better-written explanation).
The problem with mindreading is that it is an asymmetric technology that disproportionately helps those who rule without consent. The productivity enhancements are equally beneficial to everyone who can get them, but detecting defection is of marginal benefit to a project you are joyously working on with friends. It will help large corporations predict who is contemplating leaving. It is of great help to a criminal scared that his subordinates will turn state’s evidence. But it is dictators who will benefit the most.
Subscribe now
The greatest threat to a dictatorship is an information cascade, because a dictatorship can only handle so many rebels at once. This is true for both masses and elites: it’s why one of the first targets of any coup attempt is communications infrastructure. A revolution is a bet that other people agree with you that the leadership is bad, and that you can all rise up together at once. Mindreading lets the autocrat detect people the moment they begin to develop doubts, and “re-educate” or simply kill them. As a result, mindreading is a matter of urgent necessity, backed by the resources of the entire state, for a dictator. Monitoring and punishing dissent isn’t just a potential use: it is one more valuable than anything your merely impressive white-collar worker is doing.
So what does this look like, in 2030, when Conduit expects “invasive general read”? Moving from thought decoding to text to measuring associations with certain words is an active research task with a verifiable outcome and incredible value placed on success. Generals and corporate founders and political elites will get started with their headband in the morning as part of their workday, just like anyone else doing important work, and they will pass a quick quiz on Xi Jinping Thought, just like they have to do every so often. It’s boring, like HR training. If they have a bit too much sympathy for a foreign nation their counter-intelligence service is notified. If they have a bit too much antipathy for the Dear Leader they get told they need reeducation. If there’s far too much antipathy for the dear leader, or much more fondness for anyone politically important over the Dear Leader, they are fired, or worse. They all know this. They all have strategies to make sure that they are fond enough of the Dear Leader.
One of the lessons we have all learned from the labs is that developing a technology for anyone inherently accelerates it for everyone. It is not possible for one company alone to advance the state of the art and, by keeping control, ensure that it will only be used for good purposes. The training of employees in the process of making something is a large factor: the creation of datasets and a supply chain another. Even if a company swears to only use their technology for the most upright of purposes, the technology itself will leak. The possibility of developing a technology does not imply that its arrival is inevitable at a fixed pace: what projects we choose to work on can accelerate or delay technological innovation. The faster you expect the world to change over the next decade, the more valuable it is to delay innovations that would be harmful today.
The benefits to alignment research are real. I’m not convinced they’re large: the decline of centaur chess promises weighs heavily upon any assertion that human-AI teaming will reliably beat AI models alone. More concerningly, you need a great deal of confidence in your model of the world to be confident that the benefits outweigh the costs of empowering dictatorships. I don’t have that confidence. I’m not convinced anyone should.
And so I ask you to not build the mindreading datasets. Do not build the mindreading hardware. Do not fund this technology. Do not sell them your mind. I am not calling for a boycott after the point where it becomes commercially available: that is far too late to be useful, and I will not begrudge anyone their personal productivity tools. But we should not accelerate mindreading. If it looks at thoughts or feelings, it is a tool for tyrants.
I will not speculate about what dictators might do with mind-writing.
Share