Data Engineer · Seunghyeon Yang

I build and operate data pipelines

A data engineer who keeps looking for ways to automate repetitive pain points and build resilient data infrastructure.

  • 4+ yrsData Engineering
  • Batch · CDCPipelines
  • GCP · K8sInfrastructure
  • End-to-endIngest → Mart → Use
  • 01

    Pipeline optimization & automation

    Hardened batch pipelines on Airflow and automated cloud infrastructure to keep data platforms reliably running in production.

  • 02

    Building the data environment

    Owned the full path from raw data processing to a BigQuery-based warehouse, so the organization can make decisions grounded in data.

  • 03

    System design & operations

    Handled high-traffic workloads with gRPC/FastAPI and shipped MSA services reliably through GitOps (ArgoCD, Helm) — that combination is my strong suit.

  1. 2022.06 — Present

    SOCAR Data Engineer

    Batch platform, operational DB → warehouse pipelines, real-time pricing server, CDC lakehouse

  2. 2021.06 — 2022.06

    Haezoom Software Engineer

    Solar power data ingestion platform, weather data pipeline automation

Orchestration
  • Airflow
  • dbt
Warehouse & Lakehouse
  • BigQuery
  • Apache Iceberg
Streaming & CDC
  • Kafka
Infrastructure
  • Docker
  • Kubernetes
  • GKE
  • EKS
  • Helm
  • Datadog
Language
  • Python
  • SQL

Let's talk

Always up for a conversation about data engineering, data infrastructure,
or interesting technical problems and opportunities.

ysh410@gmail.com