Some users express excitement about neural net architecture gains despite bad data as fun experimentation toward scalable methods, while others dismiss the claims as untrustworthy or manic marginal pursuits.
Based on 5 visible X reactions from 16 accounts; directional sample.
Ask a question below.
Published answers will appear here.
They argue simple data cleaning achieves the same performance gains.
@chris_j_paxton True when learning with current methods but it should be obvious that algorithms that require humans to clean the data don't scale well. There is no reason we can't design algorithms that don't rely on humans to provide them with clean datasets.
@yacineMTB You would make a solid speedrunner. Same level of rewarding mania with marginal improvements that make the dopamine flow.
@yacineMTB yep, that's part of the fun... hey maybe this will get worse if I try, lets see.
@yacineMTB Anyone who enjoys wall clock as authority cannot be trusted … 🔥
@yacineMTB complicated its over drop it
there is no Glory in cleaning data (look ma! i removed artifacts from broken machine translated text! that's cute, sweetie, but Anyone Can Do That) instead, all the Glory goes to the Breathrough-Chasers of architecture and algorithmic novelty for Future Lightcone Flagplant Purposes (fame! power! money! adoration from the parasocial masses!)
Coming up with a complicated new neural network architecture and measuring its improvement not on wallclock, but on overall trainer steps against an unturned baseline to get a 3% gain which could have been done by deleting one bad piece of data mixed into the training corpus
the things that will probably improve your policy more than any cool new science or technical innovation ever could: - cleaning the data yourself - using more parameters - improving sensor placement (in robotics)
Some users express excitement about neural net architecture gains despite bad data as fun experimentation toward scalable methods, while others dismiss the claims as untrustworthy or manic marginal pursuits.
Based on 5 visible X reactions from 16 accounts; directional sample.
Ask a question below.
Published answers will appear here.
@suchenzang someone doing data work is the biggest green flag