Skip to content
#

incremental-processing

Here are 22 public repositories matching this topic...

A production-grade cryptocurrency data pipeline built on GCP that ingests real-time market data from the CoinGecko API, implements Medallion architecture (Raw → Staging → Curated), and supports idempotent backfill and metadata-driven incremental processing for reliable, scalable analytics.

  • Updated Mar 25, 2026
  • Python

❄️ 🔨End-to-end data engineering project built in Snowflake using a Medallion Architecture (🟫 Bronze → 🟦 Silver → 🟨 Gold). The project demonstrates ELT pipeline design, data ingestion from AWS S3, data cleaning and transformation, incremental processing, and dimensional modelling using a star schema.

  • Updated Aug 6, 2026

Apache Hudi — independent third-party profile of a public API surface, by API Evangelist. Apache Hudi is a data lake platform that provides incremental data processing primitives including upserts and incremental queries. It manages storage of large analytical datasets on distributed file systems with ACID transactions, timeline-based versioning, a

  • Updated Sep 12, 2026

Add this topic to your repo

To associate your repository with the incremental-processing topic, visit your repo's landing page and select "manage topics."

Learn more