Skip to content
View akshayvdoizode's full-sized avatar
🎯
Focusing
🎯
Focusing
  • https://github.com/For-Community
  • Banglore
  • 22:32 (UTC +05:30)

Organizations

@For-Community

Block or report akshayvdoizode

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
akshayvdoizode/README.md

Hi, I'm Akshay V Doyizode

Data Engineer | Cloud | Distributed Data | Streaming

I'm a Data Engineer focused on building scalable, reliable, and production-ready data platforms. I work across cloud data engineering, distributed processing, batch and streaming pipelines, data modeling, and modern data lakehouse architectures.

I enjoy solving problems where data volume, reliability, performance, and scalability actually matter.


About Me

  • Data Engineer with experience building and maintaining cloud-based data pipelines.
  • Strong focus on AWS, Azure, Apache Spark, PySpark, SQL, and distributed data processing.
  • Interested in streaming systems, data platforms, lakehouse architecture, and large-scale data engineering.
  • Currently deepening my expertise in Databricks, Spark, Kafka, Delta Lake, data modeling, and system design.
  • Preparing for opportunities where I can work on challenging, large-scale data systems.

Tech Stack

Languages

  • Python
  • SQL
  • Java
  • PySpark

Data Engineering

  • Apache Spark
  • Spark Structured Streaming
  • Apache Kafka
  • ETL / ELT
  • Batch Processing
  • Data Modeling
  • Data Warehousing
  • Data Lakehouse Architecture

Cloud

AWS

  • S3
  • Glue
  • EMR
  • Athena
  • Lambda
  • Step Functions
  • Redshift
  • RDS
  • ECR

Azure

  • ADLS Gen2
  • Azure Data Factory
  • Azure Databricks
  • Azure Synapse
  • Unity Catalog
  • Azure Purview

Databases & Storage

  • PostgreSQL
  • Oracle
  • MongoDB
  • Redshift
  • Delta Lake
  • Parquet

DevOps & Tools

  • Git
  • GitHub Actions
  • Docker
  • Kubernetes
  • Jenkins
  • ArgoCD
  • CloudWatch
  • Splunk

What I'm Currently Learning

  • Advanced Apache Spark internals and optimization
  • Spark Structured Streaming
  • Kafka and event-driven architectures
  • Delta Lake and Lakehouse architecture
  • Databricks
  • Advanced SQL and data modeling
  • Data-intensive system design
  • Distributed systems
  • AWS and Azure data platforms

Featured Areas

Data Engineering

Building scalable batch and streaming pipelines using cloud-native and distributed technologies.

Distributed Processing

Working with Spark and PySpark to process large datasets efficiently while understanding partitioning, shuffles, joins, caching, and performance optimization.

Streaming

Exploring real-time data pipelines using Kafka and Spark Structured Streaming.

Data Modeling

Practicing dimensional modeling, event modeling, state modeling, interval modeling, temporal modeling, and other analytical data patterns.

Cloud Data Platforms

Designing and working with modern data architectures across AWS and Azure.


GitHub Stats

GitHub Stats


Connect With Me


Building systems. Processing data. Learning every day.

I'm interested in data engineering, distributed systems, cloud platforms, and the engineering challenges that come with building reliable data products at scale.

Popular repositories Loading

  1. image-cropper-react image-cropper-react Public

    JavaScript 2

  2. interntest interntest Public

    JavaScript 1

  3. validation-form-react validation-form-react Public

    JavaScript 1

  4. react-larvel-project1 react-larvel-project1 Public

    PHP 1

  5. paypalCheckoutPage paypalCheckoutPage Public

    HTML 1

  6. authentication authentication Public

    JavaScript 1