Temperature influences how far the model deviates from the most probable next word choice. At temperature 0 it almost always picks the most likely token, producing consistent, factual answers. Higher values allow more surprising phrasing but increase the risk of errors and drift. Use low values for factual questions and higher values for creative writing; temperature is often tuned together with top-p.
