• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Report

    VERA paper describes training AI agents and their harnesses together

    A user says VERA turns benchmark trajectories into more than 9,000 restartable sandboxes with rubric scoring.

    Machine Learning Street TalkML
    1 Source, 1h ago, first seen 1h ago

    TLDR

    A user says Nvidia’s VERA paper describes updating model weights and agent harnesses together using verifiable environments. Harness edits must pass self-tests and add at least five points on the development set; model checkpoints are rejected if scores drop by more than 20%, the user says. They report that the 9B agent beat the strongest single-axis baseline by 10.3 points on AutoCoWorkBench and 13.0 on AutoMedBench. The environment corpus is open-sourced, they say.

    Combined views

    77

    1 Source, first seen 1h ago

    Combined views

    77

    1 Source, first seen 1h ago

    16 reposts
    16 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Featured Source

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    Machine Learning Street Talk@MLStreetTalkRT @omarsar0: Banger paper from NVIDIA. One exciting trend I am seeing is building verifiable environments for your agents and training th…1h

    1 Source

    Machine Learning Street Talk@MLStreetTalkRT @omarsar0: Banger paper from NVIDIA. One exciting trend I am seeing is building verifiable environments for your agents and training th…1h