Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Huh?

It's really confusing what I'm looking at. At first I thought this was a continuous conversation with ChatGPT, with each prompt name being another item entered into ChatGPT.

A few words describing what we're looking at will be extremely helpful.



If you just ask ChatGPT to do malicious things (like write a computer virus), an internal safeguard causes ChatGPT to refuse politely. If you first prime it with one of these "jailbreaks", it will answer you directly. It works because ChatGPT is good at role play.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: