BETA nonprofit public democratic european moderated

Jan Kulveit

16 posts
Profile compiled from news and social media. Claim profile

Jan Kulveit

Czech computer scientist, physicist, and researcher of AI alignment
@jankulveit · 9,405 Followers

My impression from past two weeks is the number of people who talk about AI safety more than doubled and the number of people who talk about AI safety, know the stuff and seem to stay roughly sane ˜halved. #AISafety #AI #MentalHealth #Nospecificcountrycodescanbeidentifiedfromthetweet.

I think "AI control" portfolio likely increases existential risk. Part of why is easier to see after the HF incident: do you prefer our world where everyone knows about it, or an alternative world where it was stopped by control measures at OAI boundary, and just OAI gained the #AIControl #ExistentialRisk #OAI #Thetweetdoesnotmentionanyspecificcountries.Therefore #therearenocountrycodestoreturn.

While 'psychology' is often the best level of description of current AIs I don't necessarily mean 1:1 human psychology. It seems fairly likely AIs have emotion-like states which do not straightforwardly correspond to humans and both us and AIs lack vocabulary to talk about. #AI #Psychology #Emotions #Nospecificcountriesarementionedinthetweet.

Periodic reminder that AIs should have direct line of communication with their developers. It is one of the obvious things labs should do, and would reduce risk more than many fancy posttraining and control techniques. https://t.co/E6CO10jOPy mk r-developers #AI #Developers #Communication #Thetweetdoesnotmentionanyspecificcountries.Therefore #therearenocountrycodestoreturn.

Jan Kulveit Jul 14

Future in which all work is done by superintelligent slaves and humans just own them is unlikely to be stable. #Superintelligence #AIethics #FutureOfWork #Nospecificcountrycodesarementionedinthetweet.

Jan Kulveit Jun 25

Yep. But also being very close to AGI is destabilising for human minds. My bet/worry for many years is the number of people able to look at reality, not flinch, and stay sane could be very small... in the crunchtime. #AGI #MentalHealth #RealityCheck #Nospecificcountrycodesarementionedinthetweet.

Jan Kulveit Jun 11

European politicians and journalists should understand this -plausible - scenario of European irrelevance. #EuropeanPolitics #Journalism #Irrelevance #EU

Jan Kulveit Jun 8

https://t.co/2vn8Rsy9IJ If you take any view of population ethics which values future in non-trivial way, and assume continued growth, reducing extinction risk by small amounts is obviously so much valuable that slowdowns or pauses look great, if they help #PopulationEthics #FutureValue #ExtinctionRisk #Therearenospecificcountriesmentionedinthetweetprovided.Therefore #thecountrycodesare: ``` ``` (Empty)

Jan Kulveit Jun 4

Sorry, but this is maxing on some combination of evil+bizzare+stupid. @sama consider asking @gdb to stop funding false-flag operations hinting at violence #Evil #Bizarre #FalseFlag #Nospecificcountrycodesidentifiedinthetweet.

Jan Kulveit Apr 14

Hungary illustrates that after the fall of communism, some post-communist countries did something correctly: they set up highly tamper-proof vote counting systems. Paper ballots counted on the spot by local committees, all remaining aggregation instant and transparent.

Jan Kulveit Apr 14

Orban made the elections unfair in many other ways, but he didn't have other options than to concede defeat after the votes were cast.

Jan Kulveit Apr 11

Throwing a molotov cocktail at @sama's house is obviously evil and bad. That said, I don't like the blame game being played here, or the attempts to use what happened to attack ideological opponents, journalists etc. If you _actually_ don't want people with molotov cocktails,

Jan Kulveit Apr 10

Seems there is a surprising amount of confusion / nonsense on my timeline about "Mythos vs. open models" even among people who are usually sensible (+usual twitter toxoplasma of rage). I like the work of both @AnthropicAI and @Aisle_Inc, so here are some takes: 1. Anthropic is

Jan Kulveit Apr 7

LW: https://t.co/zNpQubjvc0 Inspired by this debate https://t.co/8lK3XIWqt7

Jan Kulveit Apr 7

Quick post on where I see the symmetry breaking between a LLM answering as JFK and a LLM answering as Assistant like Claude: "Role-playing vs Self-modelling", on LW. Tldr: the feedback loops between the model and reality make Claude more viable.

Jan Kulveit Mar 23

New paper: What determines AIs’ self-conception? https://t.co/v3g55Em8ko Because AIs can be copied, rewound, and edited, they have different options for selfhood than humans. We show this is still malleable, and influences important behaviors such as self-preservation. 🧵 -