The developer behind Teach My Little Sister How to Drive, a driving instruction sim where you instruct a generative-AI-powered, dynamic NPC, has been spending $1,000 dollars a day to keep the demo working.
The developer behind Teach My Little Sister How to Drive, a driving instruction sim where you instruct a generative-AI-powered, dynamic NPC, has been spending $1,000 dollars a day to keep the demo working.
Why can’t they just fine tune a small locally hosted model? Are they dumb?
Most gamers have a GPU anyways
Exactly my thought. Surely you could and achieve at least a similar result.
I would say yes, yes they are dumb. Taking out a loan for a game that isn’t even out yet according to the article to keep an ai going is pretty stupid. Even if the game does well, the overhead is massive for what could be done with a smaller specialized model. Hell, I bet payig someone to train an existing model would be significantly less than a $1000 a day when you aren’t even making money yet.
Not everyone has a GPU with enough RAM. I have a 3060 with 12GB, and I cannot run ollama and a 3D game at the same time. As soon as ollama starts processing, the GPU is at 100% and game and ollama start swapping data in and out of the GPU RAM, which means the whole computer becomes unresponsive for a long time. I had to set up a script that disables ollama if steam is running.
12GB is a lot of ram. By small local models I am talking about fine tuning something like < 2GB LLM
Sorry, all of your what?
Because this is vibe-coded. You think an LLM is going to give up how to make your own?
I dislike AI slop as much as the next guy but that’s a stupid take
yes. it’s not self aware.
I literally let deepseek + opencode modify the nanogpt code for my own custom model and suggest training datasets lol
Ran out of patience after a while because crappy GPU, but it was training and words were being worded when I prompted it after the first day of training
It will though