My guess is that RL training being done with particular generation parameters makes models much more brittle to changes in these parameters, and that's why we're seeing changes like this across model providers. But I don't really know.
by tolugenius
0 subcomment
> To improve determinism, define a system instruction with explicit rules for your specific use case.
Is this guaranteed to work any better than top_k or top_p? This just sounds like making a smaller version of a Agent.md doc.
by tough
0 subcomment
fwiw sonnet-5 also drops temperature (sonne-4 had it)
by impulser_
0 subcomment
Good. These have been basically useless for the past few generations of models, and most of the time made the model perform worst.
temperature, top_p, and top_k are deprecated and ignored. In future model generations, supplying these parameters returns an HTTP 400 error. Remove these parameters from all requests.