Alternating prompt updates and model retraining reportedly lifted a science agent’s accuracy from 42.2% to 73.3% · Digg