If that’s true, crazy times lie ahead:
GPT-6 at the end of the month or early August, and Fable 5.1 also on the verge of release.
That would mean timelines are shortening even further, and if GPT-6 also features new pre-training, things will get really exciting!…
GPT-6 will presumably be announced or released in around four weeks.
New pre-training, much larger, significantly better. Mythos was essentially the origin story, the myth of AI, because it continued to prove one thing: much larger models create much more intelligence and…
> DeepSeek are preparing for V4 GA, which seems likely to be ≥ GLM-5.2, and have begun work on a new, larger model that will compete with the upcoming 2.7T MiniMax Pro
sounds like pure wishcasting based on Minimax news
but then again, why not. V4 is smol and undertrained…
1e25 is not a whole lot of compute. At most 1 month of 1 SuperPOD work (fp8). They have admitted that V4 is overcomplicated, they probably knew what how to make it better by the time it was completed. 100T tokens, AttnRes, repair the router etc
But: doubt