Comment Basic question (Score 2) 76
Why isn't there a second independent AI that examines the main AI's output, and decides if it should be censored or not. If it thinks so, the answer is not delivered.
Since the second one is only looking at the first AI's output it seems like it would be difficult or impossible to fool it, especially if it's window is very short (like only the current message).
I assume there is some reason this does not work. Any explanations?