Update README.md

388a85d4 · Meiqi Guo · 211f63e1 · 388a85d4
Commit 388a85d4 authored Mar 16, 2017 by Meiqi Guo
--- a/README.md
+++ b/README.md
 # Big Data Process Assignment 2 
-I first try this
\ No newline at end of file
+## Pre-processing the input
+For the part of pre-procesing, the input consists of：
+* the document corpus of [pg100.txt](https://gitlab.my.ecp.fr/2014guom/BigDataProcessAssignment2/blob/master/input/pg100.txt)
+* the [Stopword file](https://gitlab.my.ecp.fr/2014guom/BigDataProcessAssignment2/blob/master/input/Stopwords) which I made in the assignment 1
+* the [Words with frequency file](https://gitlab.my.ecp.fr/2014guom/BigDataProcessAssignment2/blob/master/input/wordfreq) of pg100.txt that I obtained by runnning the assignment
+1 with a slight changement of [MyWordCount](https://gitlab.my.ecp.fr/2014guom/BigDataProcessAssignment1/blob/master/MyWordCount.java).
+
+
+
+
+All the details are written in my code Process.java.
+