87
I really think that the term "Reverse Centaur" is underused
(www.youtube.com)
"We did it, Patrick! We made a technological breakthrough!"
A place for all those who loathe AI to discuss things, post articles, and ridicule the AI hype. Proud supporter of working people. And proud booer of SXSW 2024.
AI, in this case, refers to LLMs, GPT technology, and anything listed as "AI" meant to increase market valuations.
There are many things that are capital intensive, and we don't hold their product to be particularly special (cars, drugs)
That being said, it's true that we consider information to be a special case.
But what you say is a problem for the "deep research" use of chatbots, or the "AI summaries" of search engines.
But the same could well be said about search engines themselves. That still didn't get the level of hysteria that LLMs get. AI summaries on search should go away though, they destroy the viability of the websites they source information from.
I would like a truly open (incl. open training) model, that could be produced through public funding by a large collaboration of academics and private parties.
I also believe that once we give up on this "machine god" nonsense, there is a lot to be gained in highly optimized, sector specific, lightweight models, which will be less capital intensive to produce. If I need help writing python code, I don't a model that can translate Japanese into Hungarian.
I don't think it's impossible that the underlying tech of LLMs will be evolved into something undeniably useful. But I don't think "helping with Python" will be that.
I think on !fuck_ai I can suggest, you could also just learn the relevant bits of Python? Learning to express your thoughts in code and understanding the tools you use feels way more empowering than relying on blackbox-extruded code.
I do know the relevant bits of Python. Learning basic Python is just marginally harder than shouting at the screen, and adding numpy and some pyqt on top of that was also not impossible.
It was simply an example of a field where LLMs are having a lot of success, all the more reasons that this should be done as efficiently as possible rather than using the same 10 trillion parameters model that is used for everything else.
Secondly, I have heard positive things from people using LLM for instance to review their code, and they found it very very helpful, without the soul sucking effects of having it write all their code.
That I would happily try on my own code, but not in Dario's cloud. And the price to get a local rig right now is way too high.