The Financial Times announced a deal with OpenAI on Monday to license its world-class journalism for training and informing ChatGPT’s models. It joins Axel Springer and the Associated Press who struck similar deals, where OpenAI reportedly offers millions for the right to use content. However, ChatGPT was trained on lots of other web-scraped content that OpenAI did not pay for. So why is OpenAI paying for some datasets and not others?
