This dataset has 6 lines comprising an empty line and two first lines with 2 sentences.
Dev split: First lin. This is a line with multiple sentences. 
Dev split: Second lin. This text is included to make sure Unicode is handled properly: 力加勝北区ᴵᴺᵀᵃছজটডণত
Dev split: 
Dev split: Third lin after an empty line
Dev split: Fourth line with two spaces  .