Using AI for text generation is the new normalcy, and we have to accept that it's really efficient and helps us to save a lot of time. Yet, I personally feel that it's kind of polluting the originality of the data on the World Wide Web. I mean, the LLMs are trained on the original data created by humans in the first place; now the same data is being generated by AI, and before long there will be a (I think we're already there) time when it won't be able to distinguish human-created data from AI-generated data. By data, I don't mean only text, it could very well be images, audio, and video of all formats. Can LLMs change our language, culture as we know it? Let's try to imagine the bigger picture here. The first generations of LLMs were trained on human-generated data, no doubt. Then the next few iterations might have some polluted data generated from the first few generations of AIs coz they're constantly getting trained on the public Internet. Which means...
Just another tech blog, also my personal blog. I'll be posting interesting articles about tech, programming and tricks, dev life, and any timely topics.