Senior software engineer at EvolutionIQ
Nadeem Khan
I build data infrastructure: change data capture, Postgres pipelines, and the retrieval systems behind AI products.
NowReal-time document retrieval at EvolutionIQ, where documents become searchable as they land.
Try
Experience
Mar 2026 to nowNew York, US
EvolutionIQ Senior Software Engineer
Moved document ingestion from a 30-minute scheduled scan to Pub/Sub events feeding Postgres-backed job queues, so documents become searchable as they land. Own the embedding pipeline, and rebuilt tracing that took a slow endpoint from about 2 minutes to under one at p95.
Jan 2022 to Mar 2026Chicago, US
Crowe Senior Software Engineer, previously Cloud Senior Engineer
Architected a real-time CDC integration from Postgres to Salesforce on Debezium and Kafka, carrying 60M+ row changes a day, and led a SQL execution platform with planning, validation and sandboxed runs across four databases.
Jan 2021 to Feb 2022Boston, US
Boston University MSc Computer Science, research and teaching
Led product development for a public data-visualisation platform on US racial disparities at the Center for Antiracist Research (React, D3). Teaching assistant for MET CS 677, Data Science with Python.
Mar 2018 to Dec 2020India
Crowe Backend Team Lead
Led a team of five building a horizontally scalable microservice platform for a SaaS product on Spring, Docker and Kubernetes, with autoscaling tuned to 70% average CPU.
Projects
All projects
nl2sql playground
Ask your database questions in English. The model emits a typed query plan, never SQL text - validated against the real schema and the caller's role before any SQL is generated.

Post-training
A collection of small, runnable implementations of LLM post-training and alignment methods, from RL basics to DPO, RLHF, and RLAIF.

Decision Arena
Decision Arena: TypeSafe's Jev vs open-source Laya playing highway-env, Snake and Blackjack with zero training, plus benchmarks and a Claude Code watchdog

RAG playground
A local-first bench for learning and demonstrating RAG by experiment
- logscribe (opens in a new tab)
AI-powered log analysis for Python logging: batch, scrub PII, and route logs to an LLM for insights.
Python, Aug 2026
- medalflow (opens in a new tab)
dbt, but in Python classes. Declare medallion (Bronze/Silver/Gold) models as Python classes; MedalFlow extracts dependencies from your SQL and compiles them into a staged execution plan.
Python, Aug 2026
Writing
All 120 postsSelected
- How Kafka Really Works: Lessons from a 60M+ Events/Day Production Pipeline (opens on Medium)
Backend & Infra,
- Protecting PostgreSQL Primaries from Replication Slot Failures (opens on Medium)
Postgres Series,
- PostgreSQL Logical Replication at Scale: Database-Side Guardrails for 60M+ Change Events (opens on Medium)
Postgres Series,
- How to Stop Your NL2SQL Agents From Crashing in Production: The Worker-Pool Pattern (opens on Medium)
AI System Design,
Latest
- Parsing Documents: PDFs, Tables, Scans and Layout (opens on Medium)
AI System Design,
- Cleaning and Deduplication Before Anything Is Embedded (opens on Medium)
AI System Design,
- I Tested Jev and Laya, Two New AI Decision Models, on Games They Were Never Trained For (opens on Medium)
LLM Architectures,
- Decision Models: How Jev and Laya Decide Without Writing Text (opens on Medium)
LLM Architectures,
Tools I use
Languages
- Python
- Java
- SQL
- TypeScript
Data
- PostgreSQL
- Kafka
- Debezium
- Azure Synapse
- SQL Server
- MySQL
AI
- LangGraph
- NL2SQL
- RAG
- Vector Databases
- LLM Serving
Infra
- Azure
- Google Cloud
- Kubernetes
- Docker
- Spring Boot
- OpenTelemetry
- Azure DevOps