• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Report

    Language-model learning paper wins COLM Outstanding Paper; diagnostic-error abstract takes DEX26 People’s Choice

    The poster describes ordered skill emergence and says top models caught many physician diagnostic errors, though big gaps remain.

    Mark DredzeMD
    1 Source, 1h ago, first seen 1h ago

    TLDR

    A post announces that “What do Language Models Learn and When?” received COLM’s Outstanding Paper award and says language-model skills emerge in a consistent, compositional order. In a follow-up, the poster says Ahmed Hassoon’s benchmark of AI for diagnostic error correction won DEX26’s People’s Choice Best Abstract. The poster says top large language models caught many physician diagnostic errors, but big gaps remain.

    Combined views

    26

    1 Source, first seen 1h ago

    Combined views

    26

    1 Source, first seen 1h ago

    1 likes
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 likes
    1 comments
    Featured Source
    1 comments

    1 Source

    Mark Dredze@mdredze🏆 Best Abstract (People’s Choice) at DEX26: Ahmed Hassoon’s “Evaluating the Potential of AI for Diagnostic Error Correction: A Cross-Sectional Benchmark of Large Language Models.” Top LLMs caught about lots of physician diagnostic errors, but big gaps remain.1h

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 Source

    Mark Dredze@mdredze🏆 Best Abstract (People’s Choice) at DEX26: Ahmed Hassoon’s “Evaluating the Potential of AI for Diagnostic Error Correction: A Cross-Sectional Benchmark of Large Language Models.” Top LLMs caught about lots of physician diagnostic errors, but big gaps remain.1h