A korrentour readingWhat is a korrent?
Current models are unsafe because they are not smart enough, not because they are too smart.
Drawn from what François Chollet said
What this subject means
reinforcement learning Training by reward: environments, verifiable tasks, value functions, and whether it works or merely beats what came before.