local models

Running models on your own hardware.

5 tagged items

Related terms

  • Inference

    Running a trained model to produce output - what happens every time you send a prompt.

  • Quantisation

    Storing a model with lower-precision numbers so it fits in less memory and runs on ordinary hardware.

Other tags