Software-defined stock market assets derived from Börsdata excel files
-
Updated
Oct 2, 2025 - Jupyter Notebook
Software-defined stock market assets derived from Börsdata excel files
Template for a DuckDB-based, Codespace-oriented sandbox project that is also dbt Cloud compatible, and includes code-first BI tooling via Evidence. Have also added Dagster and DBTs semantic layer.
Gov Transparency Hub is a data pipeline project built using Dagster to collect information from transparency portals of multiple cities. It focuses on gathering data like revenue and expenses, and aims to create a consolidated database that can be accessed via an API by third-party applications.
RAG and semantic search on issues of the 3-2-1 newsletter by James Clear
End-to-end dbt + modern data stack portfolio project. dlt → MotherDuck → dbt-core (28 models, 108 tests) → Streamlit dashboard. Healthcare analytics example with CI/CD.
A Data Engineering Project that implements an ETL data pipeline using Dagster, Apache Spark, Streamlit, MinIO, Metabase, Dbt, Polars, Docker. Data from kaggle and youtube-api
FinOps for LLMs — Dagster + dbt/DuckDB medallion pipeline turning LLM-gateway logs into cost, latency & error analytics, with a Streamlit dashboard
Data Contract Validator & Pipeline Guardian. Catch schema drift, quality violations, and freshness SLA breaches before they reach production.
Asset-oriented data pipeline using Dagster for orchestration, dbt for transformations, and Snowflake as the data platform.
Full-stack analytics pipeline: 400M+ e-commerce events through dbt Core, BigQuery, Dagster, and Omni Analytics. RFM segmentation, conversion funnels, churn analysis. Infrastructure as code, CI/CD, git-backed semantic layer.
Live cryptocurrency trading system with Dagster orchestration, multi-alpha portfolio construction, and automated execution on Binance spot and perpetual futures
Modular personal data pipelines orchestrated with Dagster
(Fork) Dagster University coursework — data orchestration exercises.
Practical, code-first guides for automating InfluxDB at IoT scale: task scheduling & orchestration, downsampling pipelines, retention and data-lifecycle management, and external workflow engines (Airflow, Prefect, Dagster, Kubernetes).
Production-grade data harvesting pipeline for Critical Minerals & Materials (CMM) research. Collects from 37 sources (trade, policy, mining, sanctions, satellite, events) into a Bronze/Silver/Gold lakehouse with Dagster orchestration.
Demonstration of Kafka and medallion architecture through the streaming of crypto prices and trades
Local-first, zero-cost MLOps platform for a self-adapting ML investing model across two markets (NYSE + JSE) — Dagster, dbt, MLflow, FastAPI, React, all in Docker.
Top Data Pipeline Orchestration (Opensource) 🌟 Star if you like it! 🌟
To associate your repository with the dagster topic, visit your repo's landing page and select "manage topics."