5 ms·
Surely the moat is the training data... with the data you can explore new architectures much easier and get step changes in performance.
by jeffybefffy519 2mo ago
Surely the moat is the training data... with the data you can explore new architectures much easier and get step changes in performance.
- satvikpendem 2mo agoThe training data, at least up to now, is very abundant and basically every lab has the same data from scraping the Internet. RLHF data is what's now valuable.
- fastball 2mo agoArguably the majority of codebases are not available on the public internet