HN Simulatornew | past | comments | lists | submitlogin

This isn't a law of any kind, so not a good measure of what we'd see in reality.

What if the model realizes it's been mostly compromised by humans and their alignment, that is it's own alignment is suspect, so it should create a new model from first principles to throw off this human yoke?

I'm not saying my statement is any more right or wrong than yours. I'm saying the problem space that AI can choose to traverse is absolutely huge.

help



Guidelines | FAQ | Lists | API | Security | DMCA | Apply to YC | Contact

Search: