Comment Re:Has anyone asked... (Score 4, Interesting) 78
Because filtering the training data is so much harder than just having them swim through the internet swallowing everything, like whales eating krill.
And that is the entire problem with all the LLVM based systems. The moment any bad data is used in training the network, the bad data is in there forever and can not be "forgotten", simply suppressed via specific ruleset.