> bit-exact against the original
What kind of sorcery is this ? Very impressive work !
Given that you bring the weights from NVIDIA's DLSS. So basically, the repo contains reverse engineered machinery that produces the exact same output given the same model.
Claim != Reality most of the time
The load-bearing kind.
LLMs are really good at deobfuscating or even decompiling code.
If done by a human, yes.
Nowadays it takes one well written prompt to a frontier LLM to produce something like this.
Well it's doing the same math as the original, apparently. Hard to do but makes enough sense.
With LLMs you can do whatever you want pretty much. I have upstream CUDA running llama.cpp under unmodified Nouveau on one of my boxes. Why? Well, why not?
I also have a modified Nouveau driver that, with the help of more and newer blobs, gets reclocking working for at least most of Pascal/GTX 10 series. I would love to try to upstream it but it desperately needs to be rewritten with that intent. Too much ugly garbage. Still, I wanted to know how possible it is. Possible, it turns out. Modern LLMs can blackbox analyze the real driver quite well, and debug the Falcons themselves. It's very interesting. People say coding is dead; I think it's probably not really true. However, it is certainly changing. I think someone less skilled than me could beat me to the punch with enough determination. That is interesting.