Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Some customers will tell the AI to do very bad things (substitute here whatever you consider very bad). Should it do those things?

One possible answer: yes, it should do all the very bad things, and we will take care of prosecuting the customer later, once the bad thing is accomplished. Is this your answer?

 help



This is immensely preferable to having a new order of self appointed alignment councils determine what is "in the interest of humanity".

We have already seen this play out to a much lesser degree in social media.


Some of those customers are also ruthless companies and political extremist organizations of various sizes (pick one that you disagree with the most). Should they have incredibly capable tools at their disposal to accomplish their goals?

Is social media much better in 2026 now that there's been such a backlash against moderation?



Consider applying for YC's Winter 2027 batch! Applications are open till November 2.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: