openmedical/registry
← registry

spark

imported

software/spark-dataflint

Drop-in replacement for Apache Spark UI

Machine-generated from the listed sources and not yet reviewed by a human.

record
Category
Software & Systems
Subcategory
unknown
License
Apache-2.0(osi)
Status
active
Maturity
deployed
Organization
dataflint
Country
unknown
Documentation
unknown
Tags
apache-spark · big-data · data-pipeline · data-pipelines · databricks · dataproc · emr · etl
Regulatory
unknown
built by · 6

Top contributors by commit count, from the project’s public repository. Avatars are served by their origin, not stored here. To be removed from this list, open an issue.

similar by tags

Computed from shared tags, weighted so a rare tag counts for more than a common one. These are suggestions, not curated relationships.

  • fhir-data-pipesdata-pipeline · etl

    A collection of tools for extracting FHIR resources and analytics services on top of that data.

  • Build your dream AI agent swarm with enterprise-grade reliability and scalability. This repository contains our official specification template for custom swarm development using the powerful Swarms…

  • physioviewdata-pipeline

    A signal quality assessment pipeline and dashboard for wearable physiological data

  • Beginner data engineering project - batch edition

  • DiagnosisExtraction_MLbig-data · emr

    Pipeline for building Machine Learning Classifiers for the diagnosis of EHR text-data. We used this pipeline for our study, published here: https://doi.org/10.2196/23930.

  • adambig-data

    ADAM is a genomics analysis platform with specialized file formats built using Apache Avro, Apache Spark, and Apache Parquet. Apache 2 licensed.

sources
  1. api.github.com/repos/dataflint/spark
    retrieved 2026-08-05 · via github-api

    Machine-imported from GitHub search. Last push 2026-07-02, 482 stars, license reported as Apache-2.0. Category and schematic were assigned by keyword heuristics and are unreviewed.

Not yet verified by a human. Correct this record →

machine-readable

/v1/entries/0.json→ .entries["spark-dataflint"]

Entries are sharded 64 ways by a stable hash of the id, so a consumer can find any record without an index.