hive-issues mailing list archives

Site index · List index
Message view « Date » · « Thread »
Top « Date » · « Thread »
From "Bharathkrishna Guruvayoor Murali (JIRA)" <j...@apache.org>
Subject [jira] [Updated] (HIVE-19733) RemoteSparkJobStatus#getSparkStageProgress inefficient implementation
Date Thu, 05 Jul 2018 07:02:00 GMT

     [ https://issues.apache.org/jira/browse/HIVE-19733?page=com.atlassian.jira.plugin.system.issuetabpanels:all-tabpanel
]

Bharathkrishna Guruvayoor Murali updated HIVE-19733:
----------------------------------------------------
    Attachment: HIVE-19733.2.patch

> RemoteSparkJobStatus#getSparkStageProgress inefficient implementation
> ---------------------------------------------------------------------
>
>                 Key: HIVE-19733
>                 URL: https://issues.apache.org/jira/browse/HIVE-19733
>             Project: Hive
>          Issue Type: Sub-task
>          Components: Spark
>            Reporter: Sahil Takiar
>            Assignee: Bharathkrishna Guruvayoor Murali
>            Priority: Major
>         Attachments: HIVE-19733.1.patch, HIVE-19733.2.patch
>
>
> The implementation of {{RemoteSparkJobStatus#getSparkStageProgress}} is a bit inefficient.
There is one RPC call to get the {{SparkJobInfo}} and then for every stage there is another
RPC call to get each {{SparkStageInfo}}. This could all be done in a single RPC call.



--
This message was sent by Atlassian JIRA
(v7.6.3#76005)

Mime
View raw message