crunch-dev mailing list archives

Site index · List index
Message view « Date » · « Thread »
Top « Date » · « Thread »
From "Micah Whitacre (JIRA)" <j...@apache.org>
Subject [jira] [Resolved] (CRUNCH-331) Change default settings for CombineFileInputFormat
Date Wed, 05 Feb 2014 20:06:11 GMT

     [ https://issues.apache.org/jira/browse/CRUNCH-331?page=com.atlassian.jira.plugin.system.issuetabpanels:all-tabpanel
]

Micah Whitacre resolved CRUNCH-331.
-----------------------------------

       Resolution: Fixed
    Fix Version/s: 0.8.3
                   0.10.0
         Assignee: Micah Whitacre

Changes have been pushed to master and 0.8 branch.

> Change default settings for CombineFileInputFormat
> --------------------------------------------------
>
>                 Key: CRUNCH-331
>                 URL: https://issues.apache.org/jira/browse/CRUNCH-331
>             Project: Crunch
>          Issue Type: Improvement
>          Components: IO
>    Affects Versions: 0.9.0, 0.8.2
>            Reporter: Josh Wills
>            Assignee: Micah Whitacre
>             Fix For: 0.10.0, 0.8.3
>
>         Attachments: CRUNCH-331.patch, CRUNCH-331b.patch
>
>
> Currently, we default to enabling the CombineFileInputFormat settings for any extensions
of FileSourceImpl b/c it tends to improve performance for common file formats like text, sequence
files, and Avro files. However, this default has caused problems for formats like Parquet
and for custom file formats that have complex split logic.
> This JIRA is to track modifying the default combine file settings in at least some contexts,
such as with From.formattedFile for custom input formats.



--
This message was sent by Atlassian JIRA
(v6.1.5#6160)

Mime
View raw message