spark
importedsoftware/spark-dataflint
Drop-in replacement for Apache Spark UI
Machine-generated from the listed sources and not yet reviewed by a human.
- Category
- Software & Systems
- Subcategory
- unknown
- License
- Apache-2.0(osi)
- Status
- active
- Maturity
- deployed
- Organization
- dataflint
- Country
- unknown
- Homepage
- www.dataflint.io
- Repository
- github.com/dataflint/spark
- Documentation
- unknown
- Tags
- apache-spark · big-data · data-pipeline · data-pipelines · databricks · dataproc · emr · etl
- Regulatory
- unknown
Top contributors by commit count, from the project’s public repository. Avatars are served by their origin, not stored here. To be removed from this list, open an issue.
Computed from shared tags, weighted so a rare tag counts for more than a common one. These are suggestions, not curated relationships.
- fhir-data-pipesdata-pipeline · etl
A collection of tools for extracting FHIR resources and analytics services on top of that data.
- Custom-Swarms-Spec-Templatedata-pipelines
Build your dream AI agent swarm with enterprise-grade reliability and scalability. This repository contains our official specification template for custom swarm development using the powerful Swarms…
- physioviewdata-pipeline
A signal quality assessment pipeline and dashboard for wearable physiological data
- beginner_de_projectemr · etl
Beginner data engineering project - batch edition
- DiagnosisExtraction_MLbig-data · emr
Pipeline for building Machine Learning Classifiers for the diagnosis of EHR text-data. We used this pipeline for our study, published here: https://doi.org/10.2196/23930.
- adambig-data
ADAM is a genomics analysis platform with specialized file formats built using Apache Avro, Apache Spark, and Apache Parquet. Apache 2 licensed.
- api.github.com/repos/dataflint/sparkretrieved 2026-08-05 · via github-api
Machine-imported from GitHub search. Last push 2026-07-02, 482 stars, license reported as Apache-2.0. Category and schematic were assigned by keyword heuristics and are unreviewed.
Not yet verified by a human. Correct this record →
/v1/entries/0.json→ .entries["spark-dataflint"]
Entries are sharded 64 ways by a stable hash of the id, so a consumer can find any record without an index.