how to deploy IMBD on a Spark cluster?
Open
- Dominant language
- Python
- Stars
- 167
- Forks
- 27
- PR merge metrics
- No merged PRs in 30d
Description
Thank you for the great work!
Are there instructions on deploying the JOB-benchmark into Spark? Specifically, how to load IMBD to Spark and adapt the JOB queries to the Spark SQL syntax. Also, have you tested whether pilotscope supports a multi-node Spark cluster?
Contributor guide
No contributing guide indexed for this repository
Research direction
Start with the JOB-benchmark and Spark-related entry points mentioned in the issue, checking how IMBD is loaded and how JOB queries would be adapted to Spark SQL. Done would mean deployment instructions plus a confirmed answer about whether PilotScope supports a multi-node Spark cluster.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- spark, sql
- Domain
- databases, distributed-systems
- Issue type
- Documentation
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 30/100