Skip to content

fread heuristic predicts wrong column count on a valid file #2196

Description

@st-pasha

Minimally reproducible example:

fread('1,2,"3,a"\n4,5,"6,b"', verbose=T)

produces parsing log

Detecting sep ...
  sep==','(ascii 44)  with 2 lines of 3 fields using quote rule 0
  sep==','(ascii 44)  with 2 lines of 4 fields using quote rule 3
Detected 4 columns on line 1. This line is either column names or first data row (first 30 chars): <<1,2,"3,a">>

and the output dataset:

   V1 V2 V3 V4
1:  1  2 "3 a"
2:  4  5 "6 b"

Activity

  1. added this to the milestone on Oct 23, 2017
  2. added a commit that references this issue on Oct 23, 2017
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Type

    No type

    Projects

    No projects

      Milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions