nltk pos_tag usage

Tags:

nltk

pos-tagger

I am trying to use speech tagging in NLTK and have used this command:

>>> text = nltk.word_tokenize("And now for something completely different")

>>> nltk.pos_tag(text)

Traceback (most recent call last):
File "<pyshell#4>", line 1, in <module>
nltk.pos_tag(text)
File "C:\Python27\lib\site-packages\nltk\tag\__init__.py", line 99, in pos_tag
tagger = load(_POS_TAGGER)
File "C:\Python27\lib\site-packages\nltk\data.py", line 605, in load
resource_val = pickle.load(_open(resource_url))
File "C:\Python27\lib\site-packages\nltk\data.py", line 686, in _open
return find(path).open()
File "C:\Python27\lib\site-packages\nltk\data.py", line 467, in find
raise LookupError(resource_not_found)
LookupError: 
**********************************************************************
Resource 'taggers/maxent_treebank_pos_tagger/english.pickle' not
found.  Please use the NLTK Downloader to obtain the resource:

However, I get an error message which shows:

engish.pickle not found.

I have download the whole corpora and the english.pickle file is there in the maxtent_treebank_pos_tagger

What can I do to get this to work?

818

asked Dec 30 '12 10:12

Ashish Singh

2 Answers

Your Python installation is not able to reach maxent or treemap.

First, check if the tagger is indeed there: Start Python from the command line.

>>> import nltk

Then you can check using

>>> dir (nltk)

Look through the list to see if maxent and treebank are both there.

Easier would be to type

>>> "maxent" in dir(nltk)
>>> True
>>> "treebank" in dir(nltk)
>>> True

Use nltk.download() --> Models tab and check to see if the treemap tagger shows as installed. You should also try downloading the tagger again.

NLTK Downloader, Models Tab

answered Jan 01 '23 10:01

Ram Narasimhan

If you don't want to use the downloader gui, you can just use the following commands in a python or ipython shell:

import nltk
nltk.download('punkt')
nltk.download('maxent_treebank_pos_tagger')

answered Jan 01 '23 08:01

jjinking

Related questions
                            
                                Getting "bad escape" when using nltk in py3
                            
                                Splitting sentences with nltk while preserving quotes
                            
                                How to define special "untokenizable" words for nltk.word_tokenize
                            
                                Advantages of creating own corpus in NLTK
                            
                                Sentence Structure identification - spacy
                            
                                RegEx Tokenizer: split text into words, digits, punctuation, and spacing (do not delete anything)
                            
                                How to extract relationship from text in NLTK
                            
                                Python 3.5 UnicodeDecodeError for a file in utf-8 (language is 'ang', Old English)
                            
                                How to add punctuation to text using python?
                            
                                Train NER model in NLTK with custom corpus
                            
                                Can I use NLTK to determine if a comment is a positive one or a negative one?
                            
                                Using StanfordParser to get typed dependencies from a parsed sentence
                            
                                Extracting 'useful' information out of sentences?
                            
                                NLTK named entity recognition in dutch
                            
                                search similar meaning phrases with nltk
                            
                                Web Scraping Rap lyrics on Rap Genius w/ Python
                            
                                Lemmatizing Italian sentences for frequency counting
                            
                                General synonym and part of speech processing using nltk
                            
                                What is the difference between mteval-v13a.pl and NLTK BLEU?
                            
                                scikit-learn: don't separate hyphenated words while tokenization

Donate For Us

If you love us? You can donate to us via Paypal or buy me a coffee so we can maintain and grow! Thank you!

Donate Us With