So, it seems that your AI research bot took a while to complete your research for you. BTW, please read our rules on posting AI output, particularly without attribution.
It is good that you have now provided some empirical evidence for your claim. Now, one thing that is common in science (and less common elsewhere) is to look at the evidence skeptically. This is particularly important when you are using AI to do your research for you.
In particular, the claim that the empirical evidence is supposed to support is:
danieltanfh95 said:
AI cannot replace humans, and humans remain crucial for any leap in capability of AI for the foreseeable future
First,
danieltanfh95 said:
I am pretty surprised that you posted this as evidence supporting your position. In their summary they say
"Overall, we believe that internal agents at the time of our assessment plausibly had the
means,
motive, and
opportunity to start small rogue deployments, but they did not have the
means to make them highly robust." So, AI agents can
already go "rogue", they just don't do so very robustly. I would say that "forseeable future" includes a future where already existing capabilities become more robust.
Their figure 4, in particular, indicates that the capability of AI models increases between 8- and 16-fold per year and they already handle tasks that take humans >16 hours. So it is foreseeable that within two years, AI should be handling tasks that take a human half a year of full time work.
So a more accurate statement would be "AI cannot yet replace humans for tasks that take humans more than about 100 hours". This explicitly and empirically includes tasks involving using and modifying other AI models.
danieltanfh95 said:
This one is interesting. I agree that it supports the idea that, as of 2025, there was limit of about 8 hours on the complexity of tasks that can be handled by AI, specifically in the context of R&D of new AI. However, they also did show that AI could replace humans in the development of AI for tasks of 2 hours or less. 2 hours is not a lot, but it is also not a blanket "cannot replace". And certainly, the foreseeable future would include a steady increase in that time horizon.
danieltanfh95 said:
Here 70% of the planning decisions were attributed to humans, which means that AI did 30% of the planning decisions. That is not a majority of the decisions, but also does not support a "cannot replace" claim. It did replace the human in 30% of the empirically observed planning decisions.
danieltanfh95 said:
I am not clear why you thought this one was relevant.
danieltanfh95 said:
Same here
danieltanfh95 said:
Of all your empirical evidence, this one seemed the most credible and supportive of your claim. It is also relatively narrow, dealing specifically with model optimization tasks. I was surprised because this type of optimization problem is a type of problem that I would have assumed that AI would be well suited to.
So overall, your empirical evidence is not non-existent, but it is not particularly strong either.