logoalt Hacker News

Der_Einzigeyesterday at 6:36 AM1 replyview on HN

There is for $/creativity/token. LLM sampling settings are poorly supported even in open source serverless providers but are the single best lever you have for getting better outputs in regards to creativity (and quality for long context or highly quantized models).


Replies

aurareturnyesterday at 6:46 AM

I'm pretty sure you can adjust the creativity for many Chinese model inference providers.

show 1 reply