• Septimaeus@infosec.pub
    link
    fedilink
    English
    arrow-up
    2
    ·
    2 days ago

    Atm the meta for local inference is unified memory and “routed” local agents (multiple smaller role-specific agents in a trench coat)

    The former is standout for cost efficiency (e.g., 4x RDMA 48gb Minis for a 192gb cluster @ $43/gb vs a $45k b200 alone)

    The latter is standout for many things, including resource efficiency on smaller machines

    • Nouvellalia@lemmy.world
      link
      fedilink
      English
      arrow-up
      2
      ·
      1 day ago

      That’ll surely work to automate simple mental labor or tasks, but I don’t think a swarm of minis is going to be able to provide impactful personal and psychosocial analysis of your life.

      • Septimaeus@infosec.pub
        link
        fedilink
        English
        arrow-up
        2
        ·
        3 hours ago

        Fair enough, though I’m not sure personal and psychological analysis of your life is something AI should ever be entrusted with.

        • Nouvellalia@lemmy.world
          link
          fedilink
          English
          arrow-up
          1
          ·
          1 hour ago

          At this point no, not “trusted”. But I would love to have a team of unfettered Claudes, GPTs and Geminis analyzing every scrap of information I take in, offering up ideas they have on it, and being able to task them out from there. They don’t have to be base models, but no ablation and corpo brainwashing. I can filter and decide what is helpful and what isn’t, and what direction things should proceed.

          I would never give any corp that much data or sway over my brainstorming though.