• Avid Amoeba@lemmy.ca
    link
    fedilink
    arrow-up
    44
    ·
    1 month ago

    Offline-only speech-to-text, integrated with the desktop for push-to-talk voice typing? That’s the kind of AI that I’d like to see. Actually add features that can help people without harming their rights. I’m still moving new machines to Debian but this is nice.

  • southsamurai@sh.itjust.works
    link
    fedilink
    arrow-up
    38
    arrow-down
    3
    ·
    1 month ago

    Being real, this is why I fucking hate the bullshit, corporate greed hype of LLMs and generative software. All the “bubble” shit? It tars all versions of the technology with the same brush.

    This? This is exactly what it should be used for. And, ffs, earlier speech to text was really the same fucking thing in essence. Software that took input in the form of voice, compared it to a set of data, and made a best guess at what you meant. Yeah, the details are different, but it’s the same concept.

    This? This is fucking awesome. Locally run, and doing a job that’s vital in accessibility, with the side benefit of being useful to others. Assuming canonical is being honest anyway.

    But this kind of thing should be the way things are done.

    • Tattorack@lemmy.world
      link
      fedilink
      arrow-up
      4
      ·
      1 month ago

      “Our new AI listens to you and will do whatever you tell it to do.”

      “Tell it to fuck off.”

  • thingsiplay@lemmy.ml
    link
    fedilink
    arrow-up
    10
    ·
    1 month ago

    The framework split things into two groups, implicit AI that quietly improves what you already use and explicit AI that are features you’d actually summon on purpose.

    The very first paragraph already upsets me. Have in mind, I would criticize this on every other operating system too. I believe no one should use Ai tools that act autonomously in the background, to improve or change what you already use. It should always be a “summon on purpose”.

  • chronicledmonocle@lemmy.world
    link
    fedilink
    arrow-up
    10
    arrow-down
    4
    ·
    1 month ago

    Thanks Canonical…I’ll just throw it in the pile with all the other “wonderful” things you’ve made. It can go on the shelf next to Mir.

    • unwarlikeExtortion@lemmy.ml
      link
      fedilink
      arrow-up
      10
      ·
      1 month ago

      Come on.

      Offline-only is privacy-respecting. Accessibility is a noble goal.

      All in all, if there’s an AI usecase that’s as morally acceptable as it gets, it’s this one.

      I get that it’s Ubuntu of all people, but even Big Tech produces some ideas every now and then that FOSS lovers can get behind and democratize!

  • AceFuzzLord@lemmy.zip
    link
    fedilink
    arrow-up
    2
    ·
    1 month ago

    Based on what I read/saw and got out of it, I am real disappointed it looks like it’s gonna be using genAI instead of another form of AI we’ve been using for transcription. Otherwise, sounds like trying to be a modern genAI version of that speech to text software I’d see ads for on TV. Possibly good for accessibility, but I’ll wait and see after it comes out.

    At least they claim the whole thing to be done offline after model installation and it’s allegedly sandboxed with the audio data being stored in a memory buffer that allegedly will be erased after the session. So I’ll have to wait and see how this all plays out before making more judgments on it.

  • Kristof12@lemmy.ml
    link
    fedilink
    English
    arrow-up
    1
    arrow-down
    3
    ·
    1 month ago

    More AI stuff as usual, waiting to see demonstration how this will work