Gisting compresses context into a set of learned tokens, preserving its quality while making the model faster and cheaper.