• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Andrew Gordon Wilson Aligns AI Capabilities With Safety

    The NYU professor cites reduced hallucinations as proof the goals reinforce each other.

    AG
    YD
    BP
    8 Sources, 27d ago, first seen 27d ago

    TLDR

    Machine learning professor Andrew Gordon Wilson at New York University has commented on the relationship between AI system capabilities and safety measures. He noted a common perception of a trade-off between the two but suggested they can align instead. As an illustration, Wilson highlighted how efforts to decrease hallucinations in models serve as a case where enhanced performance coincides with improved safety. The statement appears in a post discussing these topics among researchers in the field. Wilson focuses his work on areas including Gaussian processes and Bayesian deep learning.

    Combined views

    157.1K

    8 Sources, first seen 27d ago

    Combined views

    157.1K

    8 Sources, first seen 27d ago

    2.4K likes
    2.4K likes
    71 comments
    367 saves
    156 reposts
    71 comments
    367 saves
    156 reposts

    Sentiment

    Positive80.9%19.1%Negative

    Summary

    Sentiment

    Positive80.9%19.1%Negative

    Many accounts welcomed Astra's progress on reducing hallucinations as a major step toward trustworthiness, while some replies questioned the benchmarks or argued hallucinations are inherent to LLMs.

    Based on 52 sentiment-bearing replies from 47 accounts across 3 conversations.

    Summary

    Many accounts welcomed Astra's progress on reducing hallucinations as a major step toward trustworthiness, while some replies questioned the benchmarks or argued hallucinations are inherent to LLMs.

    Based on 52 sentiment-bearing replies from 47 accounts across 3 conversations.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    8 Sources

    @seanjtaylorEveryone's gotta have an Astra take, so I'll give you mine: we are making fast progress on eradicating hallucinations. This eval is not some academic benchmark; it very realistically captures user experience. https://deploymentsafety.openai.com/gpt-6-astra/performance-in-cases-flagged-by-users
    @andrewgwilsThere is often a perceived trade-off between capabilities and safety. Decreasing hallucinations is a great example of how they can be aligned.
    @BorisMPowerI’m glad that hallucinations aren’t top of mind any more as they’ve largely been eliminated!
    @AdamHolterererHuge progress on hallucinations with Astra.
    @sanderstedRT @AdamHoltererer: Huge progress on hallucinations with Astra.
    @haider1Astra's biggest improvement is barely being noticed Hallucinations for some reason, openai has buried this deep inside the blog post/system card, and nobody is talking about or reporting it hallucinations have always been my biggest issue, and i've wanted this for ages
    @steipeteRT @haider1: Astra's biggest improvement is barely being noticed Hallucinations for some reason, openai has buried this deep inside the b…
    @yanndubsWe made a big push on decreasing hallucination & making Astra much more trustworthy! (Nit: to interpret those benchmarks you really need to control for the amount of claims the model make, which neither of those plots show.)

    8 Sources

    @seanjtaylorEveryone's gotta have an Astra take, so I'll give you mine: we are making fast progress on eradicating hallucinations. This eval is not some academic benchmark; it very realistically captures user experience. https://deploymentsafety.openai.com/gpt-6-astra/performance-in-cases-flagged-by-users
    @andrewgwilsThere is often a perceived trade-off between capabilities and safety. Decreasing hallucinations is a great example of how they can be aligned.
    @BorisMPowerI’m glad that hallucinations aren’t top of mind any more as they’ve largely been eliminated!
    @AdamHolterererHuge progress on hallucinations with Astra.
    @sanderstedRT @AdamHoltererer: Huge progress on hallucinations with Astra.
    @haider1Astra's biggest improvement is barely being noticed Hallucinations for some reason, openai has buried this deep inside the blog post/system card, and nobody is talking about or reporting it hallucinations have always been my biggest issue, and i've wanted this for ages
    @steipeteRT @haider1: Astra's biggest improvement is barely being noticed Hallucinations for some reason, openai has buried this deep inside the b…
    @yanndubsWe made a big push on decreasing hallucination & making Astra much more trustworthy! (Nit: to interpret those benchmarks you really need to control for the amount of claims the model make, which neither of those plots show.)