[–] 0 points 1 week ago

While it may seem counterintuitive, cloud LLMs use less electricity than local LLMs. When serving a single user, an inference engine spends very little time and power doing calculations, and most of it reading in the model (weights) from memory

Cloud LLMs serve multiple users and therefore batch requests, so a model that has been copied once from memory can be used to generate hundreds of tokens

  • source
  • parent
  • context
  •  

    These scammers copy the text from new issues verbatim, and paste them in a new issue in a "support" repo. They tag the original author so they get notified.

    They then use GitHub Actions to reply with a phishing link and email.

    This particular repo has been up for a week and has done this to 113 people.

    The link leads to a page that impersonates GitHub support. Every link on that page leads to a crypto scam.

    If you stumble across such a repository, please report it. You can report this one here.

    view more: next ›