logoalt Hacker News

Aurornisyesterday at 10:49 PM2 repliesview on HN

I tried some 1-bit, 2-bit, and bonsai quants against closed eval sets. They were essentially useless for my case. The little errors accumulate and send the whole output off track quickly.

If you had some use case with very small output sequences they could be interesting to try. I think dropping down to a 9B-class model would produce better results for most cases.


Replies

andaitoday at 4:55 AM

I wonder if this would help, or if it solves different kinds of errors.

Show HN: Forge – Guardrails take an 8B model from 53% to 99% on agentic tasks

https://news.ycombinator.com/item?id=48192383

Havocyesterday at 11:35 PM

What setup are you using to do said private evaluation? Software wise I mean