N

Nebu

Software Engineer
0 karmaJoined Retired

Comments
3

Nebu
1
0
0
20% disagree

Current AIs are capable of suffering

 

Genuinely almost complete uncertainty with a weak prior towards "no".

Nebu
1
0
0
90% disagree

Benchmarks will become useless due to eval awareness¹

 

"Alignment" benchmarks will become useless as AIs can modify their behaviour or hide their motives when they know they are being evaluated. For "capabilities" benchmarks, an AI might hide its capability (pretend to be less capable) if it knows it's being evaluated, but it's not immediately obvious that an AI would want to hide its capability. It may know that it is being evaluated, and decide to try its best anyway.

Nebu
1
0
0
60% disagree

If animals continue to exist in a post-AGI world, animal suffering will not persist

 

I have difficulty imagining animals never suffering unless we turn them all into p-zombies or something (and it's not clear to me that turning them all into p-zombies would be a good thing).