Hacker Timesnew | past | comments | ask | show | jobs | submitlogin

>DNNs/LLMs can only predict next tokens based on training data.

How do they decide between using 'a' or 'an'?



I don't get the argument; how do you decide between using 'a' or 'an'?


You use 'an' when the word that comes after it begins with a vowel.


They pick random top-k next token based on their amazing 4chan/reddit training data, duh.


So you you think a model if asked.

"There is an animal very similar to a crocodile but I cannot remember it's name"

and the model responds with

"I believe the animal you are thinking of might be " ("a" / "an")

Are you saying that it would pick the result fairly randomly and then based on it's choice pick an animal that starts with a consonant or a vowel?




Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: