alibaba / alibaba/pilotscope

how to deploy IMBD on a Spark cluster?

Open
#8 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
167
Forks
27
PR merge metrics
No merged PRs in 30d

Description

Thank you for the great work!

Are there instructions on deploying the JOB-benchmark into Spark? Specifically, how to load IMBD to Spark and adapt the JOB queries to the Spark SQL syntax. Also, have you tested whether pilotscope supports a multi-node Spark cluster?

Contributor guide

No contributing guide indexed for this repository

Research direction

Start with the JOB-benchmark and Spark-related entry points mentioned in the issue, checking how IMBD is loaded and how JOB queries would be adapted to Spark SQL. Done would mean deployment instructions plus a confirmed answer about whether PilotScope supports a multi-node Spark cluster.

Written by the indexing model from the issue text.

Assessment

Tech stack
spark, sql
Domain
databases, distributed-systems
Issue type
Documentation
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
30/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.