This project demonstrates a complete data engineering workflow: extracting data from the Harvard Art Museums API, transforming it into relational tables, loading into SQL databases, and building interactive analytics dashboards with Streamlit.
What This Project Does
API Integration
Fetches artifact data from Harvard Art Museums API with pagination and rate limiting
ETL Pipeline
Transforms nested JSON into structured relational tables (metadata, media, colors)
SQL Storage
Loads data into MySQL/TiDB Cloud with proper schema design
Analytics
Executes 20+ predefined analytical queries
Visualization
Interactive Plotly dashboards for data exploration
Installation
Show more
Installs
840
Repository
aradotso/data-skills
GitHub Stars
5
First Seen
Jun 15, 2026
Security Audits
Gen Agent Trust Hub
Pass
Socket
Warn
Snyk
Pass