• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    OpenAI’s cost and performance on a bug-hunting evaluation

    The post also describes V4.1 as catching up, but says it still falls short of Grok’s level.

    AW
    T(
    2 Sources, 17d ago, first seen 17d ago

    TLDR

    One user praises OpenAI’s showing on a bug-hunting evaluation, calling it “cheaper and better than anyone.” The same post says V4.1 is catching up, though “not even Grok tier.”

    Combined views

    54.9K

    2 Sources, first seen 17d ago

    478 likes

    Combined views

    54.9K

    2 Sources, first seen 17d ago

    478 likes
    58 comments
    69 saves
    19 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    58 comments
    69 saves
    19 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    2 Sources

    @teortaxesTexAbsurd OpenAI dominance on this bug hunting eval. Cheaper and better than anyone V4.1 is catching up but not even Grok tier…
    @alexandr_wangmuse code + muse spark 1.3 max is really good at real-world software engineering

    2 Sources

    @teortaxesTexAbsurd OpenAI dominance on this bug hunting eval. Cheaper and better than anyone V4.1 is catching up but not even Grok tier…
    @alexandr_wangmuse code + muse spark 1.3 max is really good at real-world software engineering