hive-dev mailing list archives

Site index · List index
Message view « Date » · « Thread »
Top « Date » · « Thread »
From "Shuaishuai Nie (JIRA)" <j...@apache.org>
Subject [jira] [Commented] (HIVE-5795) Hive should be able to skip header and footer rows when reading data file for a table
Date Fri, 03 Jan 2014 22:39:51 GMT

    [ https://issues.apache.org/jira/browse/HIVE-5795?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=13861969#comment-13861969
] 

Shuaishuai Nie commented on HIVE-5795:
--------------------------------------

Hi [~leftylev], this property name cannot be changes, and also since this is a table level
property, it will apply to all partitions on the table. User cannot set this property on partition
level. It should also works for ALTER TABLE statement if is not set when creating the table.

> Hive should be able to skip header and footer rows when reading data file for a table
> -------------------------------------------------------------------------------------
>
>                 Key: HIVE-5795
>                 URL: https://issues.apache.org/jira/browse/HIVE-5795
>             Project: Hive
>          Issue Type: New Feature
>            Reporter: Shuaishuai Nie
>            Assignee: Shuaishuai Nie
>             Fix For: 0.13.0
>
>         Attachments: HIVE-5795.1.patch, HIVE-5795.2.patch, HIVE-5795.3.patch, HIVE-5795.4.patch,
HIVE-5795.5.patch
>
>
> Hive should be able to skip header and footer lines when reading data file from table.
In this way, user don't need to processing data which generated by other application with
a header or footer and directly use the file for table operations.
> To implement this, the idea is adding new properties in table descriptions to define
the number of lines in header and footer and skip them when reading the record from record
reader. An DDL example for creating a table with header and footer should be like this:
> {code}
> Create external table testtable (name string, message string) row format delimited fields
terminated by '\t' lines terminated by '\n' location '/testtable' tblproperties ("skip.header.line.count"="1",
"skip.footer.line.count"="2");
> {code}



--
This message was sent by Atlassian JIRA
(v6.1.5#6160)

Mime
View raw message