Consider other variations of this rule.
In French, one would say "l'hôpital" rather than "le hôpital" because the 'h' is an aspirated breath (try saying "a aspirated" rather than "an aspirated") leading to an initial vowel sound.
In Irish, the word for father is "athair". To say "her father" it becomes "a hathair" https://en.wikipedia.org/wiki/Prothesis_(linguistics)
In Turkish, it can be seen with araba (car). "I saw the car" -> "Arabayı gördüm." (note the 'y' to avoid two vowels). Book is kitap and "I saw the book" is "Kitabı gördüm." Which doesn't need the additional consonant sound between the final sound and the "ı".
Whatever language you speak, it likely has some form of this rule.
And other languages have the quirk too (like the "a unicorn" or "a university"). For example, in French there's l’hôtel and l'hôpital. But if you've got the hedge... it's la haie... or even more egregiously, the marmoset is le ouistiti.
An uber-rich man riding in an Uber installed an Ubuntu package.
I'm actually a bit surprised by this rule. Is this line there entirely due to the word "ubiquitous"? I think it is.
Pre-LLM I was joking that real AI is whatever can make that decision correctly 100% of times.
Dialects and syntax are weirdly fascinating .
https://historyofenglishpodcast.com/
Happily sourced in Wikipedia so I don't have to figure out which episode I learned that from:
https://en.wikipedia.org/wiki/Adder <- see below the table in Taxonomy.
Firefox's local translation is pretty decent but not amazing. It did render this article's title as "Italiano: A vs. An", which I found pretty funny.
Why is the aa-oo combination at the beginning of "hour" considered a vowel but the ee-oo in "uniform" is a consonant?
> [LLM It doesn’t need to be clean or maintainable. It only needs to be correct
But correctness is precisely the issue with LLMs?
But also
> I would’ve spent more time on the trie simplification algorithm and less time on parsing cmudict and re-learning d3.js.]
You could've spent nothing on D3 because the result is bad/poorly readable - much worse than a simple table alphabetically-aligned interception table, which I could use to answer curious questions such as "how can I compare "A" and "O" in terms of exceptions? or "which letter has the most exceptions"?")
(and the blue color is out of place, vowel/consonant split should visually be the same "category", not black/colord. Red for exceptions is fine)
In Spanish, to say for example, "he is a doctor," you would say "él es médico." Which more literally translates to "he is doctor," which feels unintuitive coming from English. Spanish DOES have an equivalent for "a" - "un/una", so the natural thing for an English speaker to say would be "él es UN médico", but I understand that to be awkward/unnatural.
But then I realized, when you say the same sentence in plural - "they are doctors" - you don't have an article. You don't say "they are some doctors" (I had to Google, apparently "some" isn't an article anyway?)
So Spanish ends up more consistent - Singular: "Él es médico." Plural: "Ellos son medicos."
Unlike English - Singular: "He is a doctor." Plural: "They are doctors."