| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.
You must be logged in to block users.
Contact GitHub support about this user’s behavior. Learn more about reporting abuse.
Report abuseI'm a Data & AI Engineer passionate about designing and building end-to-end data solutions — from raw ingestion to AI-powered insights. I specialize in cloud-native data architectures, real-time streaming pipelines, and machine learning integrations using modern data stack technologies.
| Project | Description | Tech Stack |
|---|---|---|
| 🔥 Apache Spark Portfolio | End-to-end Spark data engineering solutions with local vs. global sort optimizations | PySpark, Scala |
| ☁️ Azure Data Engineer (DP-203) | Azure-based ETL/ELT pipelines — preparation for DP-203 certification | Azure Data Factory, Synapse, ADLS |
| 🤖 AI Chat RAG Workflow | Retrieval-Augmented Generation pipeline for intelligent document Q&A | Python, LangChain, OpenAI |
| 📰 News Trend Data Pipeline | Real-time news trend ingestion and analytics pipeline | Python, Airflow, Kafka |
| 🗄️ Dimensional Modeling - NBA | Star schema dimensional model for NBA analytics | SQL, PostgreSQL |
| ☸️ Kubernetes Data Engineer | Containerized data pipeline deployment with Kubernetes | Kubernetes, Docker, Python |
| 📊 SQL Deep Dive | Advanced SQL techniques: window functions, CTEs, optimization | SQL, Jupyter Notebook |
I'm always open to discussing data engineering, AI/ML projects, cloud architecture, or opportunities in consulting and technology.
⭐ "Turning raw data into actionable intelligence — one pipeline at a time."
Azure DP-203 Data Engineer certification prep: Azure Data Factory, Synapse Analytics, ADLS Gen2, Stream Analytics, Databricks & Delta Lake pipelines
Kubernetes-orchestrated data engineering platform: containerized ETL pipelines, Helm charts, pod autoscaling & cloud-native data workflow deployment
Python 1
Advanced SQL mastery: window functions, CTEs, recursive queries, query optimization, indexing strategies & analytical patterns for data engineering interviews
Jupyter Notebook 1
Production-grade PySpark data engineering solutions: ETL pipelines, sorting optimization, Spark SQL, Azure ADLS integration & dimensional modeling
RAG-powered AI chat workflow using LangChain & OpenAI for intelligent document Q&A — retrieval-augmented generation pipeline with vector embeddings
End-to-end containerized data pipeline for real-time news trend ingestion, transformation, data quality checks & alerting using Docker and Apache Airflow
| Back | FazBrowse Home | New Git URL |