[portfolio]
✉ vibhavari.bellutagi@gmail.com--:--
~/blog — ls -la *.md11 posts
--- date: January 16, 2026 · slug: inside-git ---
5 min readJanuary 16, 2026

In the previous post, we learnt [How Git manages version control](post.html?post=git-essentials) through its internal structures. In this continuation, we will delve deeper into h

cat inside-git.md →
--- date: January 13, 2026 · slug: git-essentials ---
5 min readJanuary 13, 2026

If you have ever played a video game, you know the anxiety of facing a difficult boss. What is the first thing you do before walking through that boss door? You save your game. W

cat git-essentials.md →
--- date: January 10, 2026 · slug: why-version-control-exists ---
4 min readJanuary 10, 2026

## Why Version Control Exists - The Pendrive Problem **Analogy: The Pendrive Problem** Let's imagine a world where version control systems like Git do not exist. Instead, develo

cat why-version-control-exists.md →
--- date: Feb 6, 2025 · slug: spark-execution-modes ---
6 min readFeb 6, 2025

In this post, we will discuss the different execution modes available in Apache Spark. Apache Spark provides three execution modes to run Spark applications: Cluster mode, Client mode, and Local mode.

cat spark-execution-modes.md →
--- date: Jan 21, 2025 · slug: spark-job-anatomy ---
7 min readJan 21, 2025

Understanding the internal execution flow of a Spark application is key to optimizing performance and debugging. This blog dives into the details of Spark jobs, stages, and tasks, providing a thorough exploration of how Spark handles distributed execution.

cat spark-job-anatomy.md →
--- date: Jan 13, 2025 · slug: handling-nulls-in-spark ---
8 min readJan 13, 2025

In SQL null is a special marker used to indicate that a data value does not exist in the database. A null should not be confused with a value of 0. Let's deep dive into handling nulls in Spark.

cat handling-nulls-in-spark.md →
--- date: Jan 10, 2025 · slug: columns-and-expressions ---
7 min readJan 10, 2025

Apache Spark's Column and Expression play a big role in making your pipeline more efficient. In this blog we will look into ALL the possible ways to select columns, use built-in functions and perform calculations with column objects and expressions in PySpark.

cat columns-and-expressions.md →
--- date: Jan 1, 2025 · slug: spark-basics ---
6 min readJan 1, 2025

Welcome to my Apache Spark series! I'll dive deep into Apache Spark, from basics to advanced concepts. This series is about learning, exploring, and sharing—documenting my journey to mastering Apache Spark.

cat spark-basics.md →
--- date: Nov 26, 2024 · slug: welcome-to-my-blog ---
2 min readNov 26, 2024

Welcome to my blog! This is where I'll be sharing my thoughts, experiences, and insights on technology, development, and more.

cat welcome-to-my-blog.md →
cd ~ → back to portfolio
TERMINAL · bash · ~/portfolio
vibhavari@portfolio:~$
:home · grep spark · theme paper · term