Arc is an opinionated framework for defining data pipelines which are predictable, repeatable and manageable.
Arc-Jupyter is an interactive Jupyter Notebooks Extenstion for building Arc data pipelines via Jupyter Notebooks.
Provides the MongoDBExtract and MongoDBLoad stages
Provides KafkaExtract, KafkaLoad and KafkaCommitExecute stages
Provides the DeltaLakeExtract and DeltaLakeLoad stages
arc-dataquality-udf-plugin defines a set of data quality/validation user defined functions.
Provides the CypherTransform and GraphTransform stages
Provides GeoSpark UDFs functionality to Arc.
Provides the CassandraExtract, CassandraExecute, and CassandraLoad stages
Creates a list of formatted dates to easily calculate delta processing periods.
Plugin to support extract and load using the spark-bigquery-connector
Provides ElasticsearchExtract and ElasticsearchLoad stages
Provides the SASExtract stage
Provides the DebeziumTransform stage
Provides the XMLExtract and XMLLoad stages