Return-Path: Delivered-To: apmail-hadoop-pig-dev-archive@www.apache.org Received: (qmail 42442 invoked from network); 19 Jul 2010 22:24:36 -0000 Received: from unknown (HELO mail.apache.org) (140.211.11.3) by 140.211.11.9 with SMTP; 19 Jul 2010 22:24:36 -0000 Received: (qmail 6605 invoked by uid 500); 19 Jul 2010 22:24:36 -0000 Delivered-To: apmail-hadoop-pig-dev-archive@hadoop.apache.org Received: (qmail 6565 invoked by uid 500); 19 Jul 2010 22:24:35 -0000 Mailing-List: contact pig-dev-help@hadoop.apache.org; run by ezmlm Precedence: bulk List-Help: List-Unsubscribe: List-Post: List-Id: Reply-To: pig-dev@hadoop.apache.org Delivered-To: mailing list pig-dev@hadoop.apache.org Received: (qmail 6549 invoked by uid 99); 19 Jul 2010 22:24:35 -0000 Received: from nike.apache.org (HELO nike.apache.org) (192.87.106.230) by apache.org (qpsmtpd/0.29) with ESMTP; Mon, 19 Jul 2010 22:24:35 +0000 X-ASF-Spam-Status: No, hits=4.4 required=10.0 tests=FREEMAIL_ENVFROM_END_DIGIT,FREEMAIL_FROM,HTML_MESSAGE,RCVD_IN_DNSWL_NONE,SPF_PASS,T_TO_NO_BRKTS_FREEMAIL X-Spam-Check-By: apache.org Received-SPF: pass (nike.apache.org: domain of dlieu.7@gmail.com designates 209.85.216.176 as permitted sender) Received: from [209.85.216.176] (HELO mail-qy0-f176.google.com) (209.85.216.176) by apache.org (qpsmtpd/0.29) with ESMTP; Mon, 19 Jul 2010 22:24:27 +0000 Received: by qyk12 with SMTP id 12so2959729qyk.14 for ; Mon, 19 Jul 2010 15:23:06 -0700 (PDT) DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=gmail.com; s=gamma; h=domainkey-signature:mime-version:received:received:in-reply-to :references:date:message-id:subject:from:to:cc:content-type; bh=wveYWCPUQdUjts6GbCSunpoOH5CrZ4z8Gq6XpTmRciY=; b=UDvdfteqO8cgjLCzw1x1PfO2/UREGmTu8EpHM77qSYHL0YvqfIlrb1ZdJ+RVVSaYAx fyNhRT8c5cAn+cD/3aa3npoLjQ2pkDpzOaqr6sy9H+AdrvUeNwDpvtQPTGCGPiNYGk79 iaoFKUpVwUQj1mAZRIqN3tj3xZ7shfsuamqgA= DomainKey-Signature: a=rsa-sha1; c=nofws; d=gmail.com; s=gamma; h=mime-version:in-reply-to:references:date:message-id:subject:from:to :cc:content-type; b=ep8XoMv2m+/P7r1qIwimR+4mKt+cmWOk6i7lFFVcdhxj2JZsHFSS49W5NuSgjjiKP8 +G7jM3RV94R1m6dFmw+SwlUzfG3wMAnokoXUhmwoReDkiyeqFMGpd+fb2WoBhapwtEi4 i9TOrCO25g5T4BLbDwui0IzhkrNBZDCJJH+18= MIME-Version: 1.0 Received: by 10.224.96.160 with SMTP id h32mr4654372qan.269.1279578185967; Mon, 19 Jul 2010 15:23:05 -0700 (PDT) Received: by 10.229.251.15 with HTTP; Mon, 19 Jul 2010 15:23:05 -0700 (PDT) In-Reply-To: References: Date: Mon, 19 Jul 2010 15:23:05 -0700 Message-ID: Subject: Re: Pig filter by fails at backend , what am i doing wrong? From: Dmitriy Lyubimov To: pig-user@hadoop.apache.org Cc: pig-dev@hadoop.apache.org Content-Type: multipart/alternative; boundary=00c09f89957e3ca9de048bc5032a X-Virus-Checked: Checked by ClamAV on apache.org --00c09f89957e3ca9de048bc5032a Content-Type: text/plain; charset=ISO-8859-1 Seems like ressurrected *PIG-550 * On Mon, Jul 19, 2010 at 3:04 PM, Dmitriy Lyubimov wrote: > > PPS. the pig version is 0.7.0 > > Thanks. > > On Mon, Jul 19, 2010 at 3:02 PM, Dmitriy Lyubimov wrote: > >> i guess i need to add that contentRatings in IMP_F2 is a bag of tuples >> (mapped so by load function). >> >> On Mon, Jul 19, 2010 at 3:00 PM, Dmitriy Lyubimov wrote: >> >>> Hi, >>> >>> I would greatly appreciate somebody's help with the following pig error >>> during MR >>> >>> all mappers fail with the following stack trace >>> >>> java.lang.ClassCastException: java.lang.Integer cannot be cast to org.apache.pig.data.Tuple >>> >>> at org.apache.pig.backend.hadoop.executionengine.physicalLayer.expressionOperators.POProject.getNext(POProject.java:389) >>> at org.apache.pig.backend.hadoop.executionengine.physicalLayer.expressionOperators.POIsNull.getNext(POIsNull.java:152) >>> >>> >>> >>> at org.apache.pig.backend.hadoop.executionengine.physicalLayer.expressionOperators.PONot.getNext(PONot.java:71) >>> at org.apache.pig.backend.hadoop.executionengine.physicalLayer.expressionOperators.POAnd.getNext(POAnd.java:67) >>> >>> >>> >>> at org.apache.pig.backend.hadoop.executionengine.physicalLayer.relationalOperators.POFilter.getNext(POFilter.java:148) >>> at org.apache.pig.backend.hadoop.executionengine.physicalLayer.PhysicalOperator.processInput(PhysicalOperator.java:272) >>> >>> >>> >>> at org.apache.pig.backend.hadoop.executionengine.physicalLayer.relationalOperators.POLimit.getNext(POLimit.java:85) >>> at org.apache.pig.backend.hadoop.executionengine.physicalLayer.PhysicalOperator.processInput(PhysicalOperator.java:272) >>> >>> >>> >>> at org.apache.pig.backend.hadoop.executionengine.physicalLayer.relationalOperators.POLocalRearrange.getNext(POLocalRearrange.java:255) >>> at org.apache.pig.backend.hadoop.executionengine.mapReduceLayer.PigMapBase.runPipeline(PigMapBase.java:232) >>> >>> >>> >>> at org.apache.pig.backend.hadoop.executionengine.mapReduceLayer.PigMapBase.map(PigMapBase.java:227) >>> at org.apache.pig.backend.hadoop.executionengine.mapReduceLayer.PigMapBase.map(PigMapBase.java:52) >>> at org.apache.hadoop.mapreduce.Mapper.run(Mapper.java:144) >>> >>> >>> >>> at org.apache.hadoop.mapred.MapTask.runNewMapper(MapTask.java:621) >>> at org.apache.hadoop.mapred.MapTask.run(MapTask.java:305) >>> at org.apache.hadoop.mapred.Child.main(Child.java:170) >>> >>> >>> >>> >>> >>> the pig script fragment causing this is as follows : >>> >>> >>> >>> IMP_F2 = foreach IMP_F1 generate ... , FLATTEN(contentRatings) as contentRating; >>> IMP_F3 = filter IMP_F2 by contentRating is not null and contentRating.vendorId==1 >>> >>> if i remove IMP_F3 line then the job goes thru but adding IMP_F3 filtering causes this. >>> >>> >>> >>> describe IMP_F2 produces >>> >>> IMP_F2: {... ,contentRating: (vendorId: int, ... ), ... } >>> >>> >>> i also tried casts like 'filter by ... (int)(contentRating.vendorId)==1 which did not change anything. >>> >>> Any ideas for workaround are appreciated. >>> >>> >>> >>> Thanks in advance. >>> -Dmitriy >>> >>> >>> >>> >>> >>> >>> >> > --00c09f89957e3ca9de048bc5032a--