• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Researcher Shares OpenAI Anthropic Updates and Benchmark Views

    Tweet from security researcher discusses recent AI lab announcements and benchmark priorities.

    BD
    XI
    2 Sources, 29d ago, first seen 29d ago

    TLDR

    A post by @0x10n, a CMU CSD PhD student focused on security research, highlights material from frontier AI labs. It references separate announcements from OpenAI and Anthropic. The same message connects those items to work on ExploitBench and addresses memorization alongside the viability of open source benchmarks. The author states that benchmarks should still strive to be open source. The visible text is labeled as the first entry in a thread and offers no additional details or external verification.

    Combined views

    6.3K

    2 Sources, first seen 29d ago

    Combined views

    6.3K

    2 Sources, first seen 29d ago

    86 likes
    86 likes
    1 comments
    41 saves
    9 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 comments
    41 saves
    9 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    2 Sources

    @0x10nInteresting updates from frontier AI labs! https://openai.com/index/path-to-astra/ and https://www.anthropic.com/claude-fable-and-mythos-5-1. With ExploitBench, we've been thinking a lot about memorization and the feasibility of open-source benchmarks. TL;DR: benchmarks should still strive to be open-source. (1/n)
    @moyixRT @0x10n: Interesting updates from frontier AI labs! https://openai.com/index/path-to-astra/ and https://www.anthropic.com/claude-fable-and-mythos-5-1. With ExploitBench, we've been th…

    2 Sources

    @0x10nInteresting updates from frontier AI labs! https://openai.com/index/path-to-astra/ and https://www.anthropic.com/claude-fable-and-mythos-5-1. With ExploitBench, we've been thinking a lot about memorization and the feasibility of open-source benchmarks. TL;DR: benchmarks should still strive to be open-source. (1/n)
    @moyixRT @0x10n: Interesting updates from frontier AI labs! https://openai.com/index/path-to-astra/ and https://www.anthropic.com/claude-fable-and-mythos-5-1. With ExploitBench, we've been th…