Computer Science
Data Engineering & Pipelines Challenges
Data Engineering & Pipelines challenges put you inside the work of moving data reliably from source to insight. You'll develop skills in ETL Fundamentals, Data Pipeline Design, and Data Wrangling, and you'll write SQL for Analytics and dbt Models while building Airflow DAGs that orchestrate the flow.
From there you'll handle the harder edges — Kafka event streaming, Streaming-first design, Lakehouse architecture, and Data observability — working with Apache Spark and Snowflake or BigQuery query optimization the way data teams actually do. Each challenge you solve earns a verified credential you can share with recruiters.
- AnalysisBeginnerNew
Diagnose Churn Drivers for a B2B SaaS Workflow Tool
You receive three CSV exports: 18 months of weekly product-usage events for about 1,800 accounts, the full support-ticket history, and account firmographics (industry, size, pla…
- Exploratory Data Analysis
- Data Wrangling
- Feature Engineering
Applied Data Analysis and Practical Data Science - CodeIntermediateNew
Migrate a Legacy Warehouse to a Lakehouse for an Edtech AI Platform
You receive a Postgres dump of around 50 GB and the current dbt models that produce the student-attempts mart. Land the raw data in object storage (S3 or GCS) as Parquet partiti…
- Lakehouse Architecture
- Delta Lake
- Spark
Data Engineering and Big Data Systems - DesignIntermediateNew
Design Schema Evolution for a Multi-Tenant Event Platform
Design a schema-evolution model covering: schema-registry topology (per-tenant subjects vs shared), compatibility modes per topic class (strict vs forward-only vs none), tenant-…
- Schema Evolution
- Kafka
- Schema Registry
Event-Driven Architecture - CodeIntermediateNew
Bridge a Healthcare Messaging System from RabbitMQ to Kafka
Working from the provided RabbitMQ topology specification and the 30-day message-volume dataset, map each of the 24 queues to a Kafka topic and choose a sound partition key for …
- Kafka
- Rabbitmq
- Message Broker Migration
Open coursework Practice your coursework on real scenarios.
Every challenge is shaped from real-world context — not generic exercises. The work mirrors what your degree prepares you for.
Why Ewance
- StrategyIntermediateNew
Designing a BI Strategy for a Regional Retail Chain
Your challenge is to design a BI architecture including a data warehouse (conceptual and logical models), an ETL strategy to integrate data from multiple sources (POS, inventory…
- Data Warehousing
- Etl
- Olap
Business Intelligence - AnalysisBeginnerNew
Build a Reproducible Pricing Analysis for a DTC Skincare Brand
You receive 24 months of order-line data (around 480,000 lines), a Shopify-style customer export, and a discount-code log. Build a Python pipeline that produces: SKU-level price…
- Data Wrangling
- Exploratory Data Analysis
- Cohort Analysis
Applied Data Analysis and Practical Data Science - DesignIntermediateNew
Replace Nightly Batch Loads with Near-Real-Time Reporting at a Pump Manufacturer
Using the system inventory in the provided case-file, classify all nine source systems by the most appropriate way to capture their changes: log-based change data capture where …
- Change Data Capture
- Kafka
- Debezium
Open coursework - CodeIntermediateNew
Build a Saga Orchestrator for Cross-Border Payment Creation
Using the service contract specification and the failure scenario catalog provided, model the payment-creation flow as a saga with six forward steps and six compensations, and i…
- Saga Pattern
- Temporal Workflow
- Compensation Logic
Open coursework - Browse challenges
Explore role
Product Manager
Ship product that solves real user problems. Combine user research, prototyping, and stakeholder alignment to turn ambiguous briefs into measurable wins — the role at the centre of modern software teams.
- CodeIntermediateNew
Reproducible Patient-Cohort Analysis for a Pharma AI Vendor
You receive a written cohort definition (type-2 diabetes patients on metformin for at least 90 days, aged 40-70) and a target output: 12-month HbA1c change distribution plus a K…
- Reproducible Analysis
- Cohort Analysis
- Survival Analysis
Open coursework - DesignIntermediateNew
Event-Driven Architecture for an Order-to-Cash Pipeline
Design an event-driven architecture using a Kafka-like message bus (anonymized: assume a Kafka-compatible event log). Define: 9 event types with Avro schemas, partition strategy…
- Event Driven Architecture
- Kafka
- Saga Pattern
Software Architecture - CodeIntermediateNew
Instrument Network Telemetry for an ISP's Backbone
Receive the backbone topology (12 routers across 4 PoPs, mix of Cisco IOS XR + Juniper Junos), the current SNMP-based monitoring stack, and 4 weeks of customer-complaint tickets…
- Network Telemetry
- Gnmi
- Kafka
Advanced Computer Networks - DesignIntermediateNew
Stand Up a Feature Store for a Series-B Fintech
Pick one priority feature group (recommend the 25 transaction-history features used by the fraud model). Define the offline source-of-truth (likely Snowflake or BigQuery), the o…
- Feature Store
- Feature Engineering
- Airflow
Open coursework Build a verifiable portfolio.
Submissions become evidence. Reviewers with shipping experience score against a rubric; the result becomes a credential anyone can verify.
Why Ewance
- DesignIntermediateNew
Build a Feature Store for a Fintech Fraud Team
You will design a feature-store layer covering 12 representative fraud features (account-level, merchant-level, transaction-level), with both batch (Spark) and online (low-laten…
- Feature Stores
- Data Pipelines
- Spark
Machine Learning at Scale - CodeIntermediateNew
Automate Retraining with a Drift-Triggered MLflow Pipeline
Stand up the pipeline end to end with the team's existing stack (MLflow tracking + model registry, Airflow orchestration). Wire Evidently to compute weekly drift; when drift cro…
- Mlflow
- Airflow
- Data Drift Detection
ML Engineering and Production ML - DesignIntermediateNew
Migrate a 200TB Data Lake from Parquet to Iceberg
Receive an inventory of the 200TB hot tier (around 1,200 tables, around 38 PB of historical data referenced), the current Spark + Trino read patterns, and 6 months of schema-cha…
- Iceberg
- Parquet
- Data Lake
Big Data and Data-Intensive Systems - DesignIntermediateNew
Design a Real-Time Order Pipeline for a Fintech Payments Platform
You receive a synthetic Kafka stream of around 500 transactions per second, a static merchant dimension table (about 80,000 rows), and a daily FX rate snapshot. Design an end-to…
- Streaming Data
- Kafka
- Stream Processing
Data Engineering and Big Data Systems - DesignBeginnerNew
Scaling a Sydney D2C Cosmetics Startup's Data Pipeline
You are tasked with designing a cloud-based data pipeline for GlowUp. The pipeline must ingest real-time user events (page views, purchases, returns) from web and mobile apps, p…
- Cloud Computing
- Apache Spark
- Nosql
Open coursework - DesignBeginnerNew
Optimizing Inventory for a Toronto D2C Cosmetics Brand
Your task is to design a multidimensional data model (star schema) for inventory management, create an ETL pipeline to load sample data (provided as CSV files), and develop an O…
- Data Warehousing
- Etl
- Olap
Open coursework - AnalysisIntermediateNew
Halve a Daily Spark Bill Without Breaking the SLA
Work only from the materials in this file. Read the representative PySpark module (PULSE-JOB) and the cluster configuration (PULSE-CLUSTER) to understand how the nightly job is …
- Spark
- Cost Optimization
- Etl Pipelines
Open coursework - CodeIntermediateNew
Build a Feature Store Backbone for a Healthtech ML Team
You receive synthetic wearable telemetry (heart rate, accelerometer, sleep stages) for around 5,000 patients across 90 days, plus the existing scattered feature scripts from the…
- Feature Engineering
- Data Modeling
- Python
Data Engineering and Big Data Systems - AnalysisBeginnerNew
Build a Public Open-Data Dashboard for Urban Mobility
Pull the city's open-data cyclist-collision dataset (10 years of incidents, geocoded). Define a clear before/after window around the protected-lane rollout, control for traffic-…
- Exploratory Data Analysis
- Data Wrangling
- Geospatial Analysis
Applied Data Analysis and Practical Data Science - CodeIntermediateNew
Build a Real-Time Streaming Pipeline for Card-Fraud Scoring
Using the provided transaction-event sample, the cardholder feature snapshot, the reference scoring service, and the risk team's experiment specification, build a stream-process…
- Stream Processing
- Kafka
- Flink
Open coursework - DesignIntermediateNew
Build a Scalable-System-Design Spec for a Streaming-Ingest Pipeline
Receive the current architecture (Kafka 3.6, Flink 1.18, ClickHouse 24.x), 4 weeks of production metrics (per-topic throughput, partition skew, Flink operator backpressure, Clic…
- Scalable System Design
- Kafka
- Flink
Open coursework - DesignSeniorNew
Designing a Data Warehouse for a Renewable Energy Firm
You are given sample data from three sources: energy production logs, weather data, and equipment maintenance records. Your task is to: (1) design a star schema with fact and di…
- Data Warehousing
- Star Schema
- Etl
Database Systems
How it works
From brief to credential, in six steps.
Step 01
Browse challenges aligned to your studies.
Step 02
Accept the one that fits your goals.
Step 03
Work through it with AI Copilot guidance.
Step 04
Submit for structured evaluation.
Step 05
Earn a verified credential.
Step 06
Add it to LinkedIn with one click.
Related skill families
Browse all skillsIndustry teams behind a decade of practitioner briefs
Hiring from this pool?
Sponsor a challenge and meet candidates through actual work.
Industry teams can shape briefs around the skills they hire for, then evaluate students on rubric-scored deliverables — not resumes.



















































































