cross-posted from: https://piefed.ca/c/games@lemmy.world/p/1004331/ai-driving-game-players-are-burning-through-so-many-tokens-its-developer-had-to-take-out-a

The developer behind Teach My Little Sister How to Drive, a driving instruction sim where you instruct a generative-AI-powered, dynamic NPC, has been spending $1,000 dollars a day to keep the demo working.

  • Sergio@piefed.social
    link
    fedilink
    English
    arrow-up
    105
    ·
    2 days ago

    Teach My Little Sister How to Drive relies on Google Gemini and ChatGPT to power its dynamic NPC, which follows your instructions, reacts, and disobeys, but the developers are looking into ways of allowing local AI models to step in for players “with sufficiently capable PCs.”

    They’re using two cloud-based models for a task that is usually handled by if-then-else statements? I think I know who programmed this NPC…

    • QuantumEraser@fedia.io
      link
      fedilink
      arrow-up
      54
      ·
      2 days ago

      Why use an if statement that compiles down to a conditional branch instruction and takes a nanosecond or two to execute, when you can outsource the decision making to a paid cloud service with expensive hardware that multiplies matrices with billions of elements to get your answer?

      • neomachino@lemmy.dbzer0.com
        link
        fedilink
        arrow-up
        2
        ·
        edit-2
        7 hours ago

        This past week at work we had a server that kept getting overloaded because the customer started heavily using a feature that directly runs multiple files from the command line for each interaction. The feature was built long before I was there, and not thought out very well. Each files loads the same dependencies every time it’s run which takes ~5 seconds under heavily load x 50 scripts trying to run at the same time.

        I pointed this out for the 3rd time that day while my coworker who has seniority was trying to figure out thy the server was stalling and kept just deploying updates that chatgpt suggested. Suggested to turn those scripts into reusable modules so we load the dependencies once and fork them to run in the background, but the coworkers robot said forking would actually be more resource intensive than running the scripts, which isn’t true if you think about it for half a second.

        In the end, he said “Astra says we should upgrade the instance to <some aws one with crazy resources> and avoid updated the legacy script since they’re working as is”. My boss finally chimed in when he realized the coworker upgraded the instance from one that cost ~$200/m to one that would have cost ~$2,000 with like 60 cores and a gpu for some reason. His excuse was the robot suggested it and he stood by that decision.

        Thanks for listening to me rant.