This dataset has 6 lines comprising an empty line and two first lines with 2 sentences.
First line from test split. This is a line with multiple sentences. 
Second line from test split. This text is included to make sure Unicode is handled properly: 力加勝北区ᴵᴺᵀᵃছজটডণত

Third line from test split after an empty line
Fourth line with two spaces  from test split after an empty line