Mailing-List: contact yarn-issues-help@hadoop.apache.org; run by ezmlm
Precedence: bulk
Reply-To: yarn-issues@hadoop.apache.org
Date: Tue, 3 Jun 2014 02:36:02 +0000 (UTC)
From: "Sandy Ryza (JIRA)" <jira@apache.org>
To: yarn-issues@hadoop.apache.org
Message-ID: <JIRA.12712927.1399492399557.62546.1401762962547@arcas>
In-Reply-To: <JIRA.12712927.1399492399557@arcas>
References: <JIRA.12712927.1399492399557@arcas>
Subject: [jira] [Commented] (YARN-2026) Fair scheduler : Fair share for
 inactive queues causes unfair allocation in some scenarios
MIME-Version: 1.0
Content-Type: text/plain; charset=utf-8
Content-Transfer-Encoding: quoted-printable


    [ https://issues.apache.org/jira/browse/YARN-2026?page=3Dcom.atlassian.=
jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=3D14016=
155#comment-14016155 ]=20

Sandy Ryza commented on YARN-2026:
----------------------------------

Hi Ashwin, have been busy with other stuff and probably will be for the nex=
t week or two.  I see your point.  I need to think about it a little more -=
 the main aim of preemption is to provide enforce guarantees for purposes l=
ike maintaining SLAs.  While converging towards fairness more quickly in us=
er queues could be a nice property, it satisfies a slightly different goal.

> Fair scheduler : Fair share for inactive queues causes unfair allocation =
in some scenarios
> -------------------------------------------------------------------------=
-----------------
>
>                 Key: YARN-2026
>                 URL: https://issues.apache.org/jira/browse/YARN-2026
>             Project: Hadoop YARN
>          Issue Type: Bug
>          Components: scheduler
>            Reporter: Ashwin Shankar
>            Assignee: Ashwin Shankar
>              Labels: scheduler
>         Attachments: YARN-2026-v1.txt
>
>
> Problem1- While using hierarchical queues in fair scheduler,there are few=
 scenarios where we have seen a leaf queue with least fair share can take m=
ajority of the cluster and starve a sibling parent queue which has greater =
weight/fair share and preemption doesn=E2=80=99t kick in to reclaim resourc=
es.
> The root cause seems to be that fair share of a parent queue is distribut=
ed to all its children irrespective of whether its an active or an inactive=
(no apps running) queue. Preemption based on fair share kicks in only if th=
e usage of a queue is less than 50% of its fair share and if it has demands=
 greater than that. When there are many queues under a parent queue(with hi=
gh fair share),the child queue=E2=80=99s fair share becomes really low. As =
a result when only few of these child queues have apps running,they reach t=
heir *tiny* fair share quickly and preemption doesn=E2=80=99t happen even i=
f other leaf queues(non-sibling) are hogging the cluster.
> This can be solved by dividing fair share of parent queue only to active =
child queues.
> Here is an example describing the problem and proposed solution:
> root.lowPriorityQueue is a leaf queue with weight 2
> root.HighPriorityQueue is parent queue with weight 8
> root.HighPriorityQueue has 10 child leaf queues : root.HighPriorityQueue.=
childQ(1..10)
> Above config,results in root.HighPriorityQueue having 80% fair share
> and each of its ten child queue would have 8% fair share. Preemption woul=
d happen only if the child queue is <4% (0.5*8=3D4).=20
> Lets say at the moment no apps are running in any of the root.HighPriorit=
yQueue.childQ(1..10) and few apps are running in root.lowPriorityQueue whic=
h is taking up 95% of the cluster.
> Up till this point,the behavior of FS is correct.
> Now,lets say root.HighPriorityQueue.childQ1 got a big job which requires =
30% of the cluster. It would get only the available 5% in the cluster and p=
reemption wouldn't kick in since its above 4%(half fair share).This is bad =
considering childQ1 is under a highPriority parent queue which has *80% fai=
r share*.
> Until root.lowPriorityQueue starts relinquishing containers,we would see =
the following allocation on the scheduler page:
> *root.lowPriorityQueue =3D 95%*
> *root.HighPriorityQueue.childQ1=3D5%*
> This can be solved by distributing a parent=E2=80=99s fair share only to =
active queues.
> So in the example above,since childQ1 is the only active queue
> under root.HighPriorityQueue, it would get all its parent=E2=80=99s fair =
share i.e. 80%.
> This would cause preemption to reclaim the 30% needed by childQ1 from roo=
t.lowPriorityQueue after fairSharePreemptionTimeout seconds.
> Problem2 - Also note that similar situation can happen between root.HighP=
riorityQueue.childQ1 and root.HighPriorityQueue.childQ2,if childQ2 hogs the=
 cluster. childQ2 can take up 95% cluster and childQ1 would be stuck at 5%,=
until childQ2 starts relinquishing containers. We would like each of childQ=
1 and childQ2 to get half of root.HighPriorityQueue  fair share ie 40%,whic=
h would ensure childQ1 gets upto 40% resource if needed through preemption.


--
This message was sent by Atlassian JIRA
(v6.2#6252)