Niraj Zade

Sr. Data Engineer | Architect | Consultant

~ whoever owns storage, owns computing ~

About

I build data systems - from ingestion pipelines for lakehouses, to distributed processing platforms and the infra around them.
Without breaking the bank.

If you have data coming in at scale, and having trouble with it - contact me.

I've spent over a decade solving problems via software, 4.5+ years of it within tech companies, including Coditas, Konverge.ai and DecisivEdge. I'm a very high-agency person and operate like a manager of one.

Tech stack: Spark, Databricks, Azure, AWS, FastAPI, Python, SQL, Linux, Docker etc
I typically work across stacks. Your existing tech choices won't be a barrier at all.

Contact

Email (preferred)
niraj[at]theniraj[dot]com
LinkedIn
linkedin.com/in/nirajzade/
Coffee
If you know me via a mutual friend, whatsapp me. We can simply meet over a nice cup of coffee in Pune.

Blog

Articles

created category title pages read
api HTTP API design handbook
API design guidelines
#spark #bigdata
3 4m
resources PVLDB - links only
A convenient centralized list of all PVLDB papers till date
#resources
162 4hr25m
resources PVLDB - links with abstracts (large document)
A convenient centralized list of all PVLDB papers till date
#resources
3377 92hr14m
data engineering The Spark Field Manual
An engineer-focused field manual on Spark internals for new data engineers and seasoned experts who need a refresher.
#spark #bigdata
23 37m
resources Bookmarks
A centralized collection of papers, talks, lectures around computer science, engineering, and the overall industry
#resources
4 5m
work Lecture - You and your research by Dr. Richard Hamming
The lecture by Richard Hamming
#advice
49 1hr18m
python Unicode string normalization schemes in Python
Universally applicable string comparison in python. For string comparisons, and for de-duplication in datasets.
#python
5 7m
7 rows · last built 2026-10-01

Tools

Day info