hive-issues mailing list archives

Site index · List index
Message view « Date » · « Thread »
Top « Date » · « Thread »
From "Sahil Takiar (JIRA)" <>
Subject [jira] [Commented] (HIVE-16484) Investigate SparkLauncher for HoS as alternative to bin/spark-submit
Date Tue, 02 Jan 2018 22:50:00 GMT


Sahil Takiar commented on HIVE-16484:

Attaching updated patch. Basically, just rebased the old patch and resolved all conflicts.

I'll keep this JIRA focused on migrating to {{SparkLauncher}} and work on integrating with
{{InProcessLauncher}} in a sub-task.

> Investigate SparkLauncher for HoS as alternative to bin/spark-submit
> --------------------------------------------------------------------
>                 Key: HIVE-16484
>                 URL:
>             Project: Hive
>          Issue Type: Bug
>          Components: Spark
>            Reporter: Sahil Takiar
>            Assignee: Sahil Takiar
>         Attachments: HIVE-16484.1.patch, HIVE-16484.2.patch, HIVE-16484.3.patch, HIVE-16484.4.patch,
HIVE-16484.5.patch, HIVE-16484.6.patch, HIVE-16484.7.patch, HIVE-16484.8.patch
> The {{SparkClientImpl#startDriver}} currently looks for the {{SPARK_HOME}} directory
and invokes the {{bin/spark-submit}} script, which spawns a separate process to run the Spark
> {{SparkLauncher}} was added in SPARK-4924 and is a programatic way to launch Spark applications.
> I see a few advantages:
> * No need to spawn a separate process to launch a HoS --> lower startup time
> * Simplifies the code in {{SparkClientImpl}} --> easier to debug
> * {{SparkLauncher#startApplication}} returns a {{SparkAppHandle}} which contains some
useful utilities for querying the state of the Spark job
> ** It also allows the launcher to specify a list of job listeners

This message was sent by Atlassian JIRA

View raw message