Car Wash Test on 53 leading AI models: "I want to wash my car. The car wash is 50 meters away. Should I walk or drive?"

fubarx@lemmy.world · 2 months ago

Car Wash Test on 53 leading AI models: "I want to wash my car. The car wash is 50 meters away. Should I walk or drive?"

Womble@piefed.world · 2 months ago

Torture can be a useful way of extracting information if you have a way to instantly verify it, which actually makes it a good analogy to LLMs. If I want to know the password to your laptop and torture you until you give me the correct password and I log in then that works.

[deleted]@piefed.world · 2 months ago

If you can instantly verify it then you don’t need the torture.

Getting the person to volunteer the information is proven to be far, far more successful and being able to instantly verofy means you know when you have the answers.

JcbAzPx@lemmy.world · 2 months ago

In fact it cannot ever be a useful way of extracting information. Even just randomly guessing is a better way to get the information you want than torture.

Womble@piefed.world · 2 months ago

I’m not saying its anything other than morally repugnant, obviously, but in the example of a password with billions or trillions of combinations and where you can check the answers given torture pretty obviously is better than guessing.

That’s not a scenario that is ever likely to come up, and wouldn’t be justifiable even if it did, but pretending it wouldnt be effective is ridiculous.

Car Wash Test on 53 leading AI models: "I want to wash my car. The car wash is 50 meters away. Should I walk or drive?"

Car Wash Test on 53 leading AI models: "I want to wash my car. The car wash is 50 meters away. Should I walk or drive?"

Opper