logoalt Hacker News

cavoirom • yesterday at 1:29 PM • 5 replies • view on HN

counting 'r' in 'raspberry' to the LLM is similar to 4-dimension space to human. Their world's unit is token, not character, although they could use indirect method such as "run code" to find out. It will stay that way until they change the fundamental of the token that the LLM can perceive characters.


Replies

mdp2021 • yesterday at 9:25 PM

I hope you understand: it is a core point that systems that answer questions must have the ability to internally represent the objects they assess in a way that allows reliability. Whatever the object.

omneity • yesterday at 2:55 PM

I’m working on this problem using a vocab-free, byte-based approach. It’s definitely solvable.

https://huggingface.co/posts/omarkamali/593639295164067

https://huggingface.co/blog/omarkamali/tokenization

➕ show 2 replies
kevhito • yesterday at 4:19 PM

How many 'r's are there in the next 30 seconds of this [1] song?

[1]: https://youtu.be/l7vRSu_wsNc?si=SndkB6GBaRyhvNNA&t=61

estearum • yesterday at 1:42 PM

It's not even fair to call "run code" to be indirect compared to what a human would do. The word raspberry has no Rs in it in human language either. We have a written representation of it, which we can then write down either in our head or on paper, and then we can "run the algorithm" of counting each of the letters.

Nothing intrinsically more or less direct about the LLM's method than ours.

➕ show 1 reply