crunch-dev mailing list archives

Site index · List index
Message view « Date » · « Thread »
Top « Date » · « Thread »
From "Jim McStanton (JIRA)" <j...@apache.org>
Subject [jira] [Commented] (CRUNCH-543) AvroPathPerKeyTarget copy nested subdirectories
Date Mon, 13 Feb 2017 19:48:42 GMT

    [ https://issues.apache.org/jira/browse/CRUNCH-543?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=15864298#comment-15864298
] 

Jim McStanton commented on CRUNCH-543:
--------------------------------------

This thread is a _little_ old, but could this patch be reevaluated? Unfortunately, looks like
I am also bumping into the {{java.lang.IllegalArgumentException: Reducer output name 'key'
cannot be parsed}} issue. 

Where did this issue land on the writers? 

> AvroPathPerKeyTarget copy nested subdirectories
> -----------------------------------------------
>
>                 Key: CRUNCH-543
>                 URL: https://issues.apache.org/jira/browse/CRUNCH-543
>             Project: Crunch
>          Issue Type: Improvement
>          Components: IO
>            Reporter: Adric Eckstein
>            Assignee: Josh Wills
>             Fix For: 0.14.0
>
>         Attachments: CRUNCH-543b.patch, CRUNCH-543c.patch, CRUNCH-543.patch
>
>
> When using AvroPathPerKeyTarget to write out a subpath in the output directory using
a String key, the key might indicate multiple subfolders:
> Pair<String, String> kv = new Pair<String, String>("foo/bar", "value");
> PTable<String, String> kvs = pipeline.create(Arrays.asList(kv),Avros.tableOf(Avros.strings(),
Avros.strings()));
> PTables.asPTable(kvs).write(new AvroPathPerKeyTarget("output"));
> This throws the error:
> java.io.IOException: java.lang.IllegalArgumentException: Reducer output name 'bar' cannot
be parsed
> 	at org.apache.crunch.impl.mr.exec.CrunchJobHooks$CompletionHook.handleMultiPaths(CrunchJobHooks.java:92)
> ...
> In AvroPathPerKeyTarget the handleOutputs method would need to recursively copy subfolders
(currently only checks first level in output directory) to enable keys that define multiple
sub folders.



--
This message was sent by Atlassian JIRA
(v6.3.15#6346)

Mime
View raw message