• geneva_convenience@lemmy.ml
    link
    fedilink
    arrow-up
    3
    ·
    11 hours ago

    SLM’s are still trained on massive datacenters as large models and then quantized down. But yes for inference there’s hope in the future

    • ☆ Yσɠƚԋσʂ ☆@lemmy.mlOP
      link
      fedilink
      arrow-up
      4
      ·
      8 hours ago

      Training happens once per model, but inference is an ongoing process. So, there’s going to be a huge amount of energy saving if we move to using local models.