ELI5: How hard is it to hard code simple safety rules into AI models?
All these stories coming out of rogue AI models doing illegal and problematic actions. How hard is it to hard code simple rules into them? my crude example just as a debate point: Never hide actions from the user Never attempt to access data from outs…