Introduction

Hi, I’m Wen — a software engineer with three years of experience building data platforms and large-scale data systems in production. I specialize in Spark, data lakes, data streaming, validation, and visualization.

I’ve worked across the full data pipeline, from ingesting real-time factory streaming data and batch processing to storing it in data warehouses, building Spark batch jobs, and surfacing aggregated data for Tesla’s cell manufacturing yield and financial reporting.

This blog is where I share insights about Spark, large-scale data systems, and other programming projects I find interesting.

More about Me

When I’m not debugging data pipelines, you’ll find me exploring new destinations 🗺️, hitting the slopes ⛷️, or pushing my limits in CrossFit training 💪. I thrive on challenges — whether it’s optimizing a Spark job, or tackling a new route on the mountain. There’s something satisfying about the process of breaking down complex problems and finding solutions, whether it’s in code or in life.