Travis Woodruff created PIG-5033:
------------------------------------
Summary: MultiQueryOptimizerTez creates bad plan with union, split
and FRJoin
Key: PIG-5033
URL: https://issues.apache.org/jira/browse/PIG-5033
Project: Pig
Issue Type: Bug
Components: tez
Affects Versions: 0.16.0
Reporter: Travis Woodruff
This script produces incorrect results:
{code}
a = load 'file:///tmp/input1' as (x:int, y:int);
b = load 'file:///tmp/input1' as (x:int, y:int);
u = union a,b;
c = load 'file:///tmp/input3' as (x:int, y:int);
e = filter c by y > 3;
f = filter c by y < 2;
g = join u by x left, e by x using 'replicated';
h = join g by u::x left, f by x using 'replicated';
store h into 'file:///tmp/pigoutput';
{code}
Without the union, or with opt.multiquery=false, or with non-replicated joins,
it works as expected.
--
This message was sent by Atlassian JIRA
(v6.3.4#6332)