logoalt Hacker News

minimaltomyesterday at 9:40 PM1 replyview on HN

Was really cool to see yous use Engrams to cut down compute!

Given its basically an O(1) lookup with disk space being the main constraint, I was curious if you've tried ablating engram layers and sizes across your setup?

Also, why mHC over attention residuals?


Replies

HenryNdubuakuyesterday at 9:44 PM

Yes, we ablated Engrams rigorously and found that it returned world knowledge like FFN without without compute expenditure.

show 1 reply