Speech and Language Processing : An Introduction to Natural Language Processing, Computational Linguistics and Speech Recognition (ISBN10: 0130950696; ISBN13: 9780130950697)
ISBN10: 0130950696
ISBN13: 9780130950697
Edition/Copyright: 00

Share This Book:

Publisher: Prentice Hall, Inc.
Cover: Hardback
Year Published: 2000
Weight: 3.1lbs.

Speech and Language Processing : An Introduction to Natural Language Processing, Computational Linguistics and Speech Recognition

by Daniel Jurafsky and James H. Martin

Marketplace
$36.00
Lowest Marketplace Price!
More?
Over 8,000 Sellers
Compare Pricing

Jurafsky, Daniel : University of Colorado, Boulder


Martin, James H. : University of Colorado, Boulder

Preface

This is an exciting time to be working in speech and language processing. Historically distinct fields (natural language processing, speech recognition, computational linguistics, computational psycholinguistics) have begun to merge. The commercial availability of speech recognition and the need for Web-based language techniques have provided an important impetus for development of real systems. The availability of very large on-line corpora has enabled statistical models of language at every level, from phonetics to discourse. We have tried to draw on this emerging state of the art in the design of this pedagogical and reference work:

  1. Coverage
    In attempting to describe a unified vision of speech and language processing, we cover areas that traditionally are taught in different courses in different departments: speech recognition in electrical engineering; parsing, semantic interpretation, and pragmatics in natural language processing courses in computer science departments; and computational morphology and phonology in computational linguistics courses in linguistics departments. The book introduces the fundamental algorithms of each of these fields, whether originally proposed for spoken or written language, whether logical or statistical in origin, and attempts to tie together the descriptions of algorithms from different domains. We have also included coverage of applications like spelling-checking and information retrieval and extraction as well as areas like cognitive modeling. A potential problem with this broad-coverage approach is that it required us to include introductory material for each field; thus linguists may want to skip our description of articulatory phonetics, computer scientists may want to skip such sections as regular expressions, and electrical engineers skip the sections on signal processing. Of course, even in a book this long, we didn't have room for everything. Thus this book should not be considered a substitute for important relevant courses in linguistics, automata and formal language theory, or, especially, statistics and information theory.
  2. Emphasis on Practical Applications
    It is important to show how language-related algorithms and techniques (from HMMs to unification, from the lambda calculus to transformation-based learning) can be applied to important real-world problems: spelling checking, text document search, speech recognition, Web-page processing, part-of-speech tagging, machine translation, and spoken-language dialogue agents. We have attempted to do this by integrating the description of language processing applications into each chapter. The advantage of this approach is that as the relevant linguistic knowledge is introduced, the student has the background to understand and model a particular domain.
  3. Emphasis on Scientific Evaluation
    The recent prevalence of statistical algorithms in language processing and the growth of organized evaluations of speech and language processing systems has led to a new emphasis on evaluation. We have, therefore, tried to accompany most of our problem domains with a Methodology Box describing how systems are evaluated (e.g., including such concepts as training and test sets, cross-validation, and information-theoretic evaluation metrics like perplexity).
  4. Description of widely available language processing resources
    Modern speech and language processing is heavily based on common resources: raw speech and text corpora, annotated corpora and treebanks, standard tagsets for labeling pronunciation, part-of-speech, parses, word-sense, and dialogue-level phenomena. We have tried to introduce many of these important resources throughout the book (e.g., the Brown, Switchboard, callhome, ATIS, TREC, MUC, and BNC corpora) and provide complete listings of many useful tagsets and coding schemes (such as the Penn Treebank, CLAWS C5 and C7, and the ARPAbet) but some inevitably got left out. Furthermore, rather than include references to URLs for many resources directly in the textbook, we have placed them on the book's Web site, where they can more readily updated.

The book is primarily intended for use in a graduate or advanced undergraduate course or sequence. Because of its comprehensive coverage and the large number of algorithms, the book is also useful as a reference for students and professionals in any of the areas of speech and language processing.

Overview of the Book

The book is divided into four parts in addition to an introduction and end matter. Part I, "Words", introduces concepts related to the processing of words: phonetics, phonology, morphology, and algorithms used to process them: finite automata, finite transducers, weighted transducers, N-grams, and Hidden Markov Models. Part II, "Syntax", introduces parts-of-speech and phrase structure grammars for English and gives essential algorithms for processing word classes and structured relationships among words: part-of-speech taggers based on HMMs and transformation-based learning, the CYK and Earley algorithms for parsing, unification and typed feature structures, lexicalized and probabilistic parsing, and analytical tools like the Chomsky hierarchy and the pumping lemma. Part III, "Semantics", introduces first order predicate calculus and other ways of representing meaning, several approaches to compositional semantic analysis, along with applications to information retrieval, information extraction, speech understanding, and machine translation. Part IV, "Pragmatics", covers reference resolution and discourse structure and coherence, spoken dialogue phenomena like dialogue and speech act modeling, dialogue structure and coherence, and dialogue managers, as well as a comprehensive treatment of natural language generation and of machine translation.

Using this Book

The book provides enough material to be used for a full-year sequence in speech and language processing. It is also designed so that it can be used for a number of different useful one-term courses:


 

NLP
1 quarter

NLP
1 semester

Speech + NLP
1 semester

Comp. Linguistics
1 quarter

1. Intro 1. Intro 1. Intro 1. Intro
2. Regex, FSA 2. Regex, FSA 2. Regex, FSA 2. Regex, FSA
8. POS tagging 3. Morph., FST 3. Morph., FST 3. Morph., FST
9. CFGs 6. N-grams 4. Comp. Phonol. 4. Comp. Phonol.
10. Parsing 8. POS tagging 5. Prob. Pronun. 10. Parsing
11. Unification 9. CFGs 6. N-grams 11. Unification
14. Semantics 10. Parsing 7. HMMs & ASR 13. Complexity
15. Sem. Analysis 11. Unification 8. POS tagging 16. Lex. Semantics
18. Discourse 12. Prob. Parsing 9. CFGs 18. Discourse
20. Generation 14. Semantics 10. Parsing 19. Dialogue
  15. Sem. Analysis 12. Prob. Parsing  
  16. Lex. Semantics 14. Semantics  
  17. WSD and IR 15. Sem. Analysis  
  18. Discourse 19. Dialogue  
  20. Generation 21. Mach. Transl.  
  21. Mach. Transl.    

Selected chapters from the book could also be used to augment courses in Artificial Intelligence, Cognitive Science, or Information Retrieval.

This book takes an empirical approach to language processing, based on applying statistical and other machine-learning algorithms to large corpora.

Methodology boxes are included in each chapter.

Each chapter is built around one or more worked examples to demonstrate the main idea of the chapter. Covers the fundamental algorithms of various fields, whether originally proposed for spoken or written language to demonstrate how the same algorithm can be used for speech recognition and word-sense disambiguation. Emphasis on web and other practical applications. Emphasis on scientific evaluation.

Useful as a reference for professionals in any of the areas of speech and language processing.

1. Introduction.

I. WORDS.

2. Regular Expressions and Automata.
3. Morphology and Finite-State Transducers.
4. Computational Phonology and Text-to-Speech.
5. Probabilistic Models of Pronunciation and Spelling.
6. N-grams.
7. HMMs and Speech Recognition.

II. SYNTAX.

8. Word Classes and Part-of-Speech Tagging.
9. Context-Free Grammars for English.
10. Parsing with Context-Free Grammars.
11. Features and Unification.
12. Lexicalized and Probabilistsic Parsing.
13. Language and Complexity.

III. SEMANTICS.

14. Representing Meaning.
15. Semantic Analysis.
16. Lexical Semantics.
17. Word Sense Disambiguation and Information Retrieval.

IV. PRAGMATICS.

18. Discourse.
19. Dialogue and Conversational Agents.
20. Natural Language Generation.
21. Machine Translation.

APPENDICES.

A. Regular Expression Operators.
B. The Porter Stemming Algorithm.
C. C5 and C7 tagsets.
D. Training HMMs: The Forward-Backward Algorithm.

Bibliography.
Index.


The Marketplace on Textbooks.com
The Marketplace on Textbooks.com
We've assembled a growing list of independent, authorized sellers, giving you even more choices when looking for your textbooks. Learn more about the Marketplace>
Don't forget, Marketplace orders do NOT qualify for FREE shipping. Why?

Filter by:   All (15)  |  New (2)  |  Very Good (8)  |  Good (5)

Price
 
Condition
Seller
Comments
$36.00
+$3.99 s/h
Add
VeryGood
Bingo Books WA
Vancouver, WA
Seller Rating: 4.67
1999 Hardcover Very Good Hardcover in very good + condition.
$39.95
+$3.99 s/h
Add
Good
savethetrees
Lookout Mountain, GA
Seller Rating: 3.17
No comments from the seller
$41.75
+$3.99 s/h
Add
Good
Crashing Rocks Bookstore
Punta Gorda, FL
Seller Rating: 0
LIBRARY BINDING Good 0130950696 Used, in good condition. Book only. May have interior marginalia or previous owner's name.
$41.75
+$3.99 s/h
Add
Good
Sandman Book Company
Punta Gorda, FL
Seller Rating: 4.48
0130950696 Used, in good condition. Book only. May have interior marginalia or previous owner's name.
$41.95
+$3.99 s/h
Add
Good
BOOKDEALZ
Lookout Mountain, GA
Seller Rating: 0
2000-02-05 Library Binding Good
$43.83
+$3.99 s/h
Add
VeryGood
Book Buggy
Marietta, GA
Seller Rating: 1
2000 Trade paperback Very Good. Trade paperback (US). Glued binding. 934 p. Practical Resources for the Mental Health Professional (Paperback).
$47.10
+$3.99 s/h
Add
VeryGood
Time For Books
Clermont, FL
Seller Rating: 4.24
0130950696 Very Nice Copy--SPEEDY SHIPPING/100% Money BACK Guarantee!
$74.10
+$3.99 s/h
Add
VeryGood
Extremely_Reliable
Richmond, TX
Seller Rating: 4.51
Buy with confidence. Excellent Customer Service & Return policy.
$74.75
+$3.99 s/h
Add
VeryGood
More Books
MIAMI, FL
Seller Rating: 4.27
2000 Library Binding Very good
$79.99
+$3.99 s/h
Add
VeryGood
booklab
Schenectady, NY
Seller Rating: 4.42
2000 Library Binding Very good Great customer service. You will be happy!
Page:   1     2     |   Next >
Would you like to edit your cart? (0 items)
view / edit
$0

Up to 90% off
millions of
textbooks daily


FREE SHIPPING
on orders
over $25*
(excludes rental and marketplace offerings)



$0.00
(you save $0.00!)


Close