• @[email protected]
    link
    fedilink
    1731 day ago

    I know people are gonna freak out about the AI part in this.

    But as a person with hearing difficulties this would be revolutionary. So much shit I usually just can’t watch because open subtitles doesn’t have any subtitles for it.

    • @[email protected]
      link
      fedilink
      106
      edit-2
      1 day ago

      The most important part is that it’s a local LLM model running on your machine. The problem with AI is less about LLMs themselves, and more about their control and application by unethical companies and governments in a world driven by profit and power. And it’s none of those things, it’s just some open source code running on your device. So that’s cool and good.

      • technomad
        link
        fedilink
        English
        381 day ago

        Also the incessant ammounts of power/energy that they consume.

            • @[email protected]
              link
              fedilink
              15 hours ago

              Any paper about any neural network.

              Using a model to get one output is just a series of multiplications (not even that, we use vector multiplication but yeah), it’s less than or equal to rendering ONE frame in 4k games.

            • @[email protected]
              link
              fedilink
              10
              edit-2
              16 hours ago

              I don’t have a source for that, but the most that any locally-run program can cost in terms of power is basically the sum of a few things: maxed-out gpu usage, maxed-out cpu usage, maxed-out disk access. GPU is by far the most power-consuming of these things, and modern video games make essentially the most possible use of the GPU that they can get away with.

              Running an LLM locally can at most max out usage of the GPU, putting it in the same ballpark as a video game. Typical usage of an LLM is to run it for a few seconds and then submit another query, so it’s not running 100% of the time during typical usage, unlike a video game (where it remains open and active the whole time, GPU usage dips only when you’re in a menu for instance.)

              Data centers drain lots of power by running a very large number of machines at the same time.

              • @[email protected]
                link
                fedilink
                210 hours ago

                From what I know, local LLMs take minutes to process a single prompt, not seconds, but I guess that depends on the use case.

                But also games, dunno about maxing GPU in most games. I maxed mine for crypto mining, and that was power hungry. So I would put LLMs closer to crypto than games.

                Not to mention games will entertain you way more for the same time.

        • Sixty
          link
          fedilink
          English
          116 hours ago

          Curious how resource intensive AI subtitle generation will be. Probably fine on some setups.

          Trying to use madVR (tweaker’s video postprocessing) in the summer in my small office with an RTX 3090 was turning my office into a sauna. Next time I buy a video card it’ll be a lower tier deliberately to avoid the higher power draw lol.

    • @[email protected]
      link
      fedilink
      401 day ago

      Yeah, transcription is one of the only good uses for LLMs imo. Of course they can still produce nonsense, but bad subtitles are better none at all.

    • @[email protected]
      link
      fedilink
      201 day ago

      Indeed, YouTube had auto generated subtitles for a while now and they are far from perfect, yet I still find it useful.