I'm a B.Tech Artificial Intelligence graduate focused on Data Engineering and building reliable data pipelines.
My core stack includes Python, SQL, PySpark, AWS, Apache Airflow, Docker, and Snowflake.
I'm interested in data processing, cloud data engineering, data warehousing, pipeline orchestration, and scalable data systems.
ββββββββββββββββ
β Data Sources β
ββββββββ¬ββββββββ
β
ββββββββββββββββ
β Ingestion β
ββββββββ¬ββββββββ
β
ββββββββββββββββ
β Processing β
β PySpark β
ββββββββ¬ββββββββ
β
ββββββββββββββββ
βTransformationβ
ββββββββ¬ββββββββ
β
ββββββββββββββββ
β Storage β
β Snowflake β
ββββββββ¬ββββββββ
β
ββββββββββββββββ
β Analytics β
ββββββββββββββββ
- π Build ETL / ELT data pipelines
- β‘ Work with PySpark for data processing
- βοΈ Build pipelines using AWS data services
- ποΈ Work with Snowflake and SQL for analytical workloads
- π Orchestrate workflows using Apache Airflow
- π³ Use Docker for containerized environments
- π§Ή Work with structured and semi-structured data
- β‘ PySpark & distributed data processing
- βοΈ AWS Data Engineering
- ποΈ Snowflake & Data Warehousing
- π Airflow & pipeline orchestration
- ποΈ Data pipeline architecture
- π Scalable data processing
- π€ Data Engineering for AI systems
With a background in Artificial Intelligence, I'm also interested in how data engineering supports modern AI systems.
Raw Data
β
βΌ
Data Engineering
β
Clean & Reliable
β
βΌ
AI Systems
My interest is particularly around data pipelines, data quality, large-scale processing, and the data infrastructure behind AI applications.
Building my career as a Data Engineer and growing toward designing scalable, reliable, and production-ready data platforms.
