It doesn't matter if the answer is right. If the AI does not have an abstract understanding of "red" then it is using a different process to get to the answer than humans. And according to Searle, a Turing machine cannot have an abstract understanding of "red", no matter how complex the question or how complex an internal model is used to determine its answers.
Going back to the Chinese Room, it is possible that the instructions carried out by the human are based on a complex model. In fact, it is possible that the human is literally calculating the output of a trained neural net by summing the weights of nodes, etc. You could even carry out these calculations yourself, if you could memorize the parameters.
Your use of "black box" gets to the heart of it. Memorizing all of the parameters of a trained NN allows you to calculate an answer, but they don't give you any understanding what the answer means. And if they don't tell you anything about the meaning, then they don't tell the CPU doing that calculation anything about meaning either.