• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Claude Opus reportedly built Jev classifiers that matched or beat fine-tuned RoBERTa on six tasks

    Lightfield says the agent wrote readable classification instructions for Jev rather than tuning a model’s weights.

    Ves StoyanovVS
    2 Sources, 2h ago, first seen 2h ago

    TLDR

    Lightfield reports that, given the same training labels, natural-language classifiers built by Claude Opus and run on Jev matched or beat fine-tuned RoBERTa across six tasks. In a separate setup with no labels at the start, the builder used questions and labels from a simulated user; its classifiers beat zero-shot results on four tasks where that input helped. Lightfield says it is releasing its builder recipes, tools and paper.

    Combined views

    1.3K

    2 Sources, first seen 2h ago

    Combined views

    1.3K

    2 Sources, first seen 2h ago

    33 likes
    33 likes
    3 comments
    10 saves
    6 reposts
    Featured Source
    3 comments
    10 saves
    6 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    2 Sources

    Ves Stoyanov@vesko_stJev matches frontier models on classification at a fraction of the cost. The obvious follow-up: can an agent build and maintain classifiers? I tested on six tasks. Opus built classifiers matched or beat fine-tuned RoBERTa on all six. https://x.com/i/article/21082362486495150092h

    2 Sources

    Ves Stoyanov@vesko_stJev matches frontier models on classification at a fraction of the cost. The obvious follow-up: can an agent build and maintain classifiers? I tested on six tasks. Opus built classifiers matched or beat fine-tuned RoBERTa on all six. https://x.com/i/article/21082362486495150092h