• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Reaction

    Gemini reinforcement-learning work reportedly brings bug fixes and improvements

    Arthur Douillard describes work since March to scale reinforcement learning for Gemini, with bug fixes, improvements and more still to build.

    NL
    JB
    AC
    18 Sources, ,

    TLDR

    Arthur Douillard wrote on Sept. 30 that he had worked on scaling reinforcement learning for Gemini since March, fixed a few bugs and made improvements. He expressed pride in the team’s progress and optimism about its direction, while emphasizing that there was still a lot to build.

    Combined views

    27.6K

    18 Sources, first seen 4h ago

    Combined views

    27.6K

    18 Sources, first seen 4h ago

    329 likes
    4h ago
    first seen 4h ago
    329 likes
    34 comments
    18 saves
    29 reposts

    Arthur Douillard described months of work on scaling reinforcement learning for Gemini in a Sept. 30 post.

    He wrote that he had been working on it since March, had fixed a few bugs and had made improvements. He characterized the bugs as “nasty” and the improvements as “fun.”

    Douillard expressed pride in the team’s progress and optimism about its direction, while emphasizing that the work continues: “Still a lot to build.”

    Gemini

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Featured Source
    34 comments
    18 saves
    29 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    #4

    Today's Rank

    #4

    18 Sources

    @finbarrtimbers@Ar_Douillard Very impressive work!
    @infoxiaoin @archit_sharma97 i trust <3
    @samsja19@Ar_Douillard diloco then scaling rl, sounds familiar congrats on the release and great work !
    @jon_barronUsing this model has been an absolute joy. I am feeling very good about Gemini these days, the progress made from 3.7 Flash onward has been tremendous.
    @allenainie@finbarrtimbers Hope it feels great to use as well. This is a new generation of models so it will feel different to Gen-3 models for sure.
    @zacharynadoRT @theBuoyantMan: @andersonbcdefg This is assuming Deepmind is working on Gemini. That's a big assumption.
    @Ar_Douillard@samsja19 At the end of the day, it’s all about putting more flops in :)
    @andrew_n_carr@jon_barron Has it been helpful in 3d for you? Opus and Astra are almost usable, so I'm excited for gemini to take a swing here
    @natolambert@zacharynado Congrats on the awesome looking model! Excited to try it :)
    @ziv_ravidLove the idea of people sitting in the room, tuning the parameters by hand; this is the real RSI!

    Related

    Google introduces Gemini 4 Argon with limited cyber rollout

    Google announces a 1 million-token output limit, up from 64,000, for its new model. Initial access goes to trusted cyber defenders through the Fairwind Program.

    Gemini evaluated with an MLCommons benchmark in an OpenMined secure enclave

    A post quoting AVERI’s standards director says Google DeepMind did not see the benchmark, and its owner did not see the Gemini model. The setup allowed the evaluation to produce results.

    Gemini's claimed vision lead despite a perceived lag behind Claude and GPT

    A user says Gemini is now behind or outdated compared with Claude and GPT, but still the best at vision.

    18 Sources

    @finbarrtimbers@Ar_Douillard Very impressive work!
    @infoxiaoin @archit_sharma97 i trust <3
    @samsja19@Ar_Douillard diloco then scaling rl, sounds familiar congrats on the release and great work !
    @jon_barronUsing this model has been an absolute joy. I am feeling very good about Gemini these days, the progress made from 3.7 Flash onward has been tremendous.
    @allenainie@finbarrtimbers Hope it feels great to use as well. This is a new generation of models so it will feel different to Gen-3 models for sure.
    @zacharynadoRT @theBuoyantMan: @andersonbcdefg This is assuming Deepmind is working on Gemini. That's a big assumption.
    @Ar_Douillard@samsja19 At the end of the day, it’s all about putting more flops in :)
    @andrew_n_carr@jon_barron Has it been helpful in 3d for you? Opus and Astra are almost usable, so I'm excited for gemini to take a swing here
    @natolambert@zacharynado Congrats on the awesome looking model! Excited to try it :)
    @ziv_ravidLove the idea of people sitting in the room, tuning the parameters by hand; this is the real RSI!