Indexer for Word documents with Python
- Stato: Closed
- Premio: €35
- Proposte ricevute: 2
- Vincitore: ashki98
Descrizione del concorso
I have a MS Word 2010 document and want to automatise the creation of the index with a human component, so the index is tailor-made.
The python code with NLTK I need should do the following steps:
1. Extract only the words which start with a capital letter of a given Microsoft Word document.
2. Tokenize only the words and create a MS Excel datasheet with two rows (word, frequency).
3. [I want to edit the Excel datasheet, to make sure, only the desired words get into the index]
4. Afterwards the words in the Excel datasheet should be the source for creating the index. This should be done with inserting { XE “word” } after the particular word in the original MS Word file.
Perhaps you can create two different snippets of code for automation.
Competenze consigliate
Feedback del Datore di Lavoro
“It was a great pleasure to work with ashki98. He copes with every problem in a very successful way. Thanks”
FasaniVerlag, Germany.
Le migliori proposte per questo concorso
-
abdohusseinelab2 Egypt
Bacheca pubblica per chiarimenti
Come iniziare a usare i concorsi
-
Pubblica il tuo concorso Con facilità e in pochi istanti
-
Ottieni una Miriade di Proposte Da tutto il mondo
-
Seleziona la proposta migliore Scarica i file - Facile!