logoalt Hacker News

pplonski86today at 8:36 AM0 repliesview on HN

I was able to fit only small models Qwen3.5-4B in RTX3070 which is not very useful for Python and SQL generation thought. When I wan to test larger open LLM models I often just use cloud resources.