I like to spur my nieces' and nephew's interests in language. Their grades have gone up overall since they have more confidence and they are doing something they find enjoyable. The only problem is that my nieces are into Spanish and my nephew is interested in Japanese. I try to help them with the languages and am also tying to improve the Russian I learned in college.
In the spirit of Larry Ferlazzo and his list of tools I'll make lists of some of the books, websites, podcasts, etc. that are indispensable to me. Here are the Spanish books that I've bought for myself, the kids, or borrowed from the library.
I know plenty of people don't have the funds to buy all the books and some people just hate books versus on-line materials. I've always liked books as a primary or secondary source of instruction since they are portable, easily browsed, and can be a nice change of pace.
I think a comprehensive textbook is always a good place to start. The Ultimate Spanish Beginner-Intermediate gives a decent coverage of Spanish. It does cover vosotros and the intricacies of the Spanish of Spain, as the reviewer on Amazon points out the Spanish of Latin America is the most common. It's very comprehensive and worth the expense.
Spanish Now! is a flawed but serviceable textbook series. There are some typos and inconsistencies which can get in the way of a true understanding but it's good for the price. If you are really tight on funds you might be able to find Ultimate Spanish or another text in the library or use the Internet sites I'll list as primary references.
Once you have a decent textbook books on verbs and grammar are a good way to refine you understanding. The Big Red Book of Spanish Verbs is a great way to familiarize yourself with common verbs and their conjugations. The included CD is good for getting some practice with the verbs.
501 Spanish Verbs is an old standby, it is widely available in most bookstores. There isn't a CD just common verbs conjugated for study.
Only a brave few actually enjoy grammar but it is a necessary evil. Spanish Grammar for Independent Learners is a good reference in an easy to follow format. If you can deal with a more regimented format Spanish Grammar is an inexpensive option.
Once one has a decent foundation and has refined their understanding of verb conjugation and grammar the last part needed for the basics is a good dictionary. The cock of the walk would be the Oxford Spanish Dictionary which has CDs, Latin American and European Spanish words and phrases, etc. The only problem is unless you have a lot of money lying around you'd do better with a cheap dictionary and some web pages.
Short stories, poems, novels are all great ways to see the language in use. Spanish Short Stories is a good parallel text suitable for a intermediate learner. For the more advanced learner there is Short Stories in Spanish. Poetry is a great way to see the possibilities in meter and turn of phrase of a language.
Of course who someone would prefer to read is very subjective, personally I love reading Frederico Garcia Lorca. There are quite a lot of manuals, novels et al. that have been translated into Spanish so one isn't limited to works created in Spanish but can also pick up works that you've already read in English. Whether one is looking for Oscar Wilde, Jane Austen, or Stephen King.
I'll write up the Japanese and Russian books as well as the podcasts, web sites, computer programs etc. If you have any suggestions leave a comment.
Showing posts with label language. Show all posts
Showing posts with label language. Show all posts
Thursday, January 03, 2008
Wednesday, October 24, 2007
Google: Statistical machine translation
ArsTechnica has a mini review of Google's translation service. They have switched from the rules based machine translation that they used to have and most machine translation services use.
I remember studying linguistic rules and statistical machine translation methods in college. Like the article suggests neither one is great but they can work well enough for someone to feel their way to the actual translation. The linguistic rules approach parses texts into an intermediary state using former grammar rules of the source language. The intermediary text is transformed into the target language former grammar rules of the target language.
The main problem with the linguistic rules approach is that it is similar to taking a sentence and marking it a parse tree and rearranging it into the parse tree of the target language and then changing it word for word. Another major problem is the grammars do not do too well with slang since there may not be a direct translation. The other problem is one of syntax there may be structures missing from the source that are needed in the target. For example, to properly translate "I went to the store" from English to Russian one needs to know if I traveled on foot or in a vehicle, since that changes the verb.
The statistical approach basically uses an algorithm to weigh the probability of the part of speech and/or meaning of a word. The statistics can be modified with the help of volunteers marking up a sentence or providing a more accurate translation. Given enough corrections and a large enough corpora the system can improve. Google appears to be using their index of web pages as a potential corpora and users of the service as the volunteers instead of the usual college student looking for beer money.
Googles approach reminds me of a few journal articles on using web pages as an inexpensive means to develop a corpus. Most corpori are rather expensive proprietary collections of text of language in everyday use. The statistical approach seems like a no brainer for Google since they have a corpori lying around and harnessing users even a poor algorithm is bound to get better. The linguistic rules approach only gets better with the development of more elaborate syntactical and transformative rules. The only question is what took Google so long to figure this out?
I remember studying linguistic rules and statistical machine translation methods in college. Like the article suggests neither one is great but they can work well enough for someone to feel their way to the actual translation. The linguistic rules approach parses texts into an intermediary state using former grammar rules of the source language. The intermediary text is transformed into the target language former grammar rules of the target language.
The main problem with the linguistic rules approach is that it is similar to taking a sentence and marking it a parse tree and rearranging it into the parse tree of the target language and then changing it word for word. Another major problem is the grammars do not do too well with slang since there may not be a direct translation. The other problem is one of syntax there may be structures missing from the source that are needed in the target. For example, to properly translate "I went to the store" from English to Russian one needs to know if I traveled on foot or in a vehicle, since that changes the verb.
The statistical approach basically uses an algorithm to weigh the probability of the part of speech and/or meaning of a word. The statistics can be modified with the help of volunteers marking up a sentence or providing a more accurate translation. Given enough corrections and a large enough corpora the system can improve. Google appears to be using their index of web pages as a potential corpora and users of the service as the volunteers instead of the usual college student looking for beer money.
Googles approach reminds me of a few journal articles on using web pages as an inexpensive means to develop a corpus. Most corpori are rather expensive proprietary collections of text of language in everyday use. The statistical approach seems like a no brainer for Google since they have a corpori lying around and harnessing users even a poor algorithm is bound to get better. The linguistic rules approach only gets better with the development of more elaborate syntactical and transformative rules. The only question is what took Google so long to figure this out?
Wednesday, January 03, 2007
The Becky
The Language Log Blog just announced the winner of the Goropius Becanus Prize. More accurately, Jeff Nunberg announced it on Fresh Air. The award goes to a person or organization that has "outstanding contributions to linguistic misinformation.
I'd have to agree that this is a very good choice. I found the assertions that I heard in her radio interview on Radio Times to be just someone trying to cash in on stereotypes in disregard to actual scientific studies. For the sake of convenience here are the search results for articles talking about the "scientist" on the LLB.
The fact that she got so much attention and that actual linguistic research doesn't get too much of love is probably due to the preference for sound bite research that affirms stereotypes. The Northern City Vowel Shift didn't exactly burn up the phone lines on the radio. Still people have an interest in the sideshows of linguistics and science, researchers just have to do a better job of bringing people into the real show. A couple of letters to the editor or some such to say, "that was interesting load of crap but the true research is even more fascinating." I mean the fact that most studies show men talk just as much or more than women raises a number of interesting questions about gender roles and power. The stereotypes are not only wrong but pretty boring.
I'd have to agree that this is a very good choice. I found the assertions that I heard in her radio interview on Radio Times to be just someone trying to cash in on stereotypes in disregard to actual scientific studies. For the sake of convenience here are the search results for articles talking about the "scientist" on the LLB.
The fact that she got so much attention and that actual linguistic research doesn't get too much of love is probably due to the preference for sound bite research that affirms stereotypes. The Northern City Vowel Shift didn't exactly burn up the phone lines on the radio. Still people have an interest in the sideshows of linguistics and science, researchers just have to do a better job of bringing people into the real show. A couple of letters to the editor or some such to say, "that was interesting load of crap but the true research is even more fascinating." I mean the fact that most studies show men talk just as much or more than women raises a number of interesting questions about gender roles and power. The stereotypes are not only wrong but pretty boring.
Subscribe to:
Posts (Atom)