Skip to content
View hey-Orion's full-sized avatar

Block or report hey-Orion

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
hey-Orion/README.md

πŸ‘‹ Hi, I'm Harsh

DataOps & Data Engineering | Building Production-Oriented Data Pipelines


πŸ“– About Me

I build production-oriented data pipelines that don't break silently. My focus is on quality, observability, and deterministic execution.

Validate early β†’ Fail loudly β†’ Monitor continuously

  • 🧊 Data quality enforcement using robust schema validation rules
  • πŸ”­ Deterministic and reproducible execution across all environments
  • πŸ§ͺ Clear observability and failure signals to eliminate silent pipeline drops
  • πŸ“‚ Testable and maintainable data workflows built for long-term scalability

πŸ“¦ Featured Project


A production-style DataOps pipeline designed to prevent silent data corruption and stale data issues.

πŸ› οΈ Tech Stack

Core & Validation: Python Pandas Pydantic SQLAlchemy Requests Python-DotEnv

Data & Config: YAML Logging

Database: PostgreSQL

Infrastructure & Testing: Docker Docker Compose pytest GitHub Actions Git GitHub Makefile Apache Airflow

DevOps Basics & Monitoring: Bash Linux Shell Sentry Networking & Security Basics

What it demonstrates:

  • πŸ—οΈ Medallion Architecture (Bronze β†’ Silver β†’ Gold)
  • πŸ“ Schema validation using Pydantic
  • πŸ”„ Idempotent pipeline execution
  • 🧹 Data cleaning & transformation with Pandas
  • πŸ—„οΈ SQL-based storage & querying
  • πŸ§ͺ Automated testing with Pytest
  • 🐳 Containerized execution with Docker
  • πŸ‘· CI pipeline via GitHub Actions
  • 🧱 Orchestration with apache-Airflow

πŸ› οΈ Working Stack

🟒 Core

Python Pandas Pydantic SQLAlchemy Requests Python-Dotenv

🟑 Data & Config

PostgreSQL YAML Logging

πŸ›  Infrastructure & CI/CD

Docker Docker Compose GitHub Actions GitLab CI/CD Makefile Apache Airflow

🧰 Tooling & DevOps Basics

Bash Linux Git GitHub AWS GCP Sentry n8n


πŸ›οΈ Engineering Approach

I focus on building systems that are predictable, debuggable, and production-ready.

  • βœ… Validate data before processing
  • ⚠️ Handle edge cases and invalid inputs
  • ✍️ Write tests for transformations
  • πŸ”„ Keep pipelines deterministic
  • πŸ“’ Surface failures clearly

πŸ’Ό Open to Work

Actively seeking a remote EU/UK DataOps or Data Engineering role β€” contract or full-time.

I build reliability into pipelines from the ground up: schema validation, idempotency, and observability baked in rather than bolted on afterward.

  • 🎯 Role Focus: Data Engineer / DataOps Engineer β€” Open to remote & on-site (relocation with sponsorship), full-time or contract
  • πŸ•’ Availability: Full-time, overlapping with CET/CEST business hours
  • 🧩 What I bring: Production-minded pipeline design, strong test coverage, fast onboarding
  • πŸ›‘οΈ Approach: Fail loudly, validate early, keep systems debuggable under pressure

πŸ“¬ Connect with me

Gmail LinkedIn GitHub

Pinned Loading

  1. Dataflow-Sentinel Dataflow-Sentinel Public

    This is Dataflow-Sentinel a production-inspired DataOps pipeline

    Python

  2. project-lab project-lab Public

    Experiments and hands-on DataOps & Data Engineering learning projects.

    Python

  3. hey-Orion hey-Orion Public

    GitHub profile and project showcase.

  4. workshop-2.0 workshop-2.0 Public

    Code-first learning program

    Python