It's cautionary tale for LLM use though. Data aren't democratic; some are better. Reminds me of something a biologist recently said to me of an AI colleague: "His models are fantastic. It's just a shame he doesn't realize the output's garbage because half the papers fed into it are wrong.." 5/n