76 points by tomjakubowski 1 day ago | 29 comments | View on ycombinator
mnkv 1 day ago |
chubot about 17 hours ago |
bcorigliano 1 day ago |
And well if I missed the point of the article, sorry. Anyways AI should be kept understandable and as see-through as possible if it's gonna be more powerful than a human.
applicative about 16 hours ago |
undefined about 23 hours ago |
fellowniusmonk 1 day ago |
Urb_RS about 22 hours ago |
ck2 1 day ago |
then we'll have to "flip" other models to be snitches on the other agents
then they'll make double-agents
the thing is though we won't be able to keep up if we keep giving them unlimited hardware worldwide, we'll try to kill the bad actors but they'll just clone somewhere else, or even start by safely making 1000 copies of themselves
yeah this won't end well, at all
bloppe 1 day ago |
I dislike this term because it doesn't explain where this "illegibility" is coming from. Models are post-trained towards non-linguistic goals with (mostly) non-linguistic rewards. A model's reasoning chain is reinforced if it leads to a correct answer or agentic goal. It doesn't need to be linguistically accurate and meanings can drift over training.