logoalt Hacker News

washadjeffmadyesterday at 12:04 PM0 repliesview on HN

Abilt models typically perform worse than their bases at the same tasks, so while I'd use one to evaluate content knowledge, I'd probably ultimately stick to one from a family I could fool with abstraction or coerce through system prompt.