pyspark-emr
importedsoftware/pyspark-emr
A toolset to streamline running spark python on EMR
Machine-generated from the listed sources and not yet reviewed by a human.
- Category
- Software & Systems
- Subcategory
- unknown
- License
- MIT(osi)
- Status
- dormant
- Maturity
- deployed
- Organization
- yodasco
- Country
- unknown
- Homepage
- unknown
- Repository
- github.com/yodasco/pyspark-emr
- Documentation
- unknown
- Tags
- emr · pyspark-emr · python · spark
- Regulatory
- unknown
Top contributors by commit count, from the project’s public repository. Avatars are served by their origin, not stored here. To be removed from this list, open an issue.
Computed from shared tags, weighted so a rare tag counts for more than a common one. These are suggestions, not curated relationships.
- AWS_EMR_Pysparklingemr · spark
Set Up Python environment on AWS EMR cluster with H2O Sparkling Water (Pysparling)
- emr-spark-jupyteremr · spark
:notebook: Repository/Tutorial for initiallizing Jupyter Notebook and Spark cluster on Amazon EMR
- sbt-lighteremr · spark
SBT plugin for Apache Spark on AWS EMR
- sparksnakeemr · spark
Improving the development of Spark applications deployed as jobs on AWS services like Glue and EMR
- glowspark
An open-source toolkit for large-scale genomic analysis
- TileDB-VCFspark
Efficient variant-call data storage and retrieval library using the TileDB storage library.
- api.github.com/repos/yodasco/pyspark-emrretrieved 2026-08-05 · via github-api
Machine-imported from GitHub search. Last push 2016-11-16, 20 stars, license reported as MIT. Category and schematic were assigned by keyword heuristics and are unreviewed.
Not yet verified by a human. Correct this record →
/v1/entries/42.json→ .entries["pyspark-emr"]
Entries are sharded 64 ways by a stable hash of the id, so a consumer can find any record without an index.