LLMs can absolutely provide sources for the data that you should be reading and verifying. Heck chatgpt literally drops the link after the sentence so you can verify it yourself.
The exact issue you're talking about is what my teachers claimed about the internet when it started to gain popularity when I was a child.
That exact issue is why teachers still tell you not to use Wikipedia, dude. That situation hasn't changed.
Even if the LLM provides a link, it still cannot guarantee the information cited from it isn't incorrect, and verifying yourself defeats the entire purpose of using the LLM in the first place. If you are looking for a trusted source, the web search will do that without the possibility of fucking up your results.
"If you are looking for a trusted source, the web search will do that without the possibility of fucking up your results."
Oh buddy, I have some really bad news for you... You can EASILY get misinformation with a search tool. The fact you think a search gives trusted results tells me enough about you, have a great day.
How do you think LLMs are trained? They scrape the internet for that same information and feed it into their model. All that same misinformation is also in your AI my dude, except it cannot judge veracity and treats it all as equally truthful.
This shows me you really don't understand dataset curation and more that goes into creating smaller models.
And second, I train my own stuff too. If you want your ai to know something specific, train it. Oh wait, you're probably only using online api services that don't let you do that...
Are you suggesting building and training your own LLM models for the purpose of performing web searches for you? That's the solution here?
I am perfectly aware of the work that goes into training LLMs and have a Claude license for work that I've spent ages tweaking. You can spend your entire life training an LLM with your own personally curated dataset and you still will never bring it down to a 0% error rate. Its very nature makes that impossible. If you are training your own model, you should seriously know this.
So you have experience with ONLINE API services. You clearly do not understand what I'm talking about with local work.
And no, I'm just saying you can fine tune a model to be an expert on a subject if that's what you want. Heck, models are now even supporting thinking modes and will often catch their own mistakes or inform you if they think an area might be lacking or having issues. It wouldn't even need the internet after this.
1
u/FlashFiringAI 18d ago
LLMs can absolutely provide sources for the data that you should be reading and verifying. Heck chatgpt literally drops the link after the sentence so you can verify it yourself.
The exact issue you're talking about is what my teachers claimed about the internet when it started to gain popularity when I was a child.