-
Notifications
You must be signed in to change notification settings - Fork 129
New issue
Have a question about this project? Sign up for a free GitHub account to open an issue and contact its maintainers and the community.
By clicking “Sign up for GitHub”, you agree to our terms of service and privacy statement. We’ll occasionally send you account related emails.
Already on GitHub? Sign in to your account
What is each phoneme in IPA terms? #29
Comments
@obones I'm using this conversion for my own work. (
|
@iamanigeeit Thanks so much, this was incredibly helpful! There are a few small typos in your dictionary though, where very a very similar looking character like ":" is used instead of "ː". For anyone else wanting to use this, here's a version without the typos:
|
If anyone comes across this issue, I noticed there is one more issue with the mapping for The ARPABET Wikipedia page has this unicode g. The full mapping with that change is below.
I have a more actively maintained fork of this project that supports generating International Phonetic Alphabet tokens. The mapping above is built-in. |
Describing the pronunciation of words is usually done using the International Phonetic Alphabet (IPA) which uses Unicode characters.
However, this package outputs ASCII character and there exists multiple mappings from Unicode IPA to ASCII.
Which one are you using?
The readme file says you are using The CMU Pronouncing Dictionary, which in turns says it's based upon ARPABET. Can I thus safely assume you are using this subset of ARPABET for all results?
The text was updated successfully, but these errors were encountered: