Isn't this just science fiction?
Parts of the far future are genuinely speculative, and it's worth saying so. But evaluations, interpretability, robustness, reward hacking and AI control are empirical research problems today — you can write code about them, run experiments and measure things. Uncertainty is not evidence for catastrophe; it is also not evidence for safety.