• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    AI agent swarm reportedly beats the state of the art on NanoChat in three days

    A user describes a custom research harness built on a graph database, saying each iteration learned from mistakes to improve its research. Their team reportedly wrote more than 15,000 entries.

    DK
    1 Source, 15d ago, first seen 15d ago

    TLDR

    A user says their team sent a swarm of AI agents to tackle Karpathy’s NanoChat benchmark and surpassed the state of the art in three days. They describe a custom harness built on a graph database to automate research, with each iteration learning from mistakes. The user says the team wrote more than 15,000 entries.

    Combined views

    143.4K

    1 Source, first seen 15d ago

    Combined views

    143.4K

    1 Source, first seen 15d ago

    1.2K likes
    1.2K likes
    45 comments
    1.2K saves
    126 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    45 comments
    1.2K saves
    126 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @hyperparticleWe sent a swarm of AI agents to solve Karpathy’s NanoChat benchmark, and they crushed SoTA in 3 days. We built our own harness on top of a graph database to do auto-autoresearch, each iteration learning from mistakes to do research better. The team wrote >15,000 entries. 🧵

    1 Source

    @hyperparticleWe sent a swarm of AI agents to solve Karpathy’s NanoChat benchmark, and they crushed SoTA in 3 days. We built our own harness on top of a graph database to do auto-autoresearch, each iteration learning from mistakes to do research better. The team wrote >15,000 entries. 🧵