Stevens Qiu — software engineer @ Amazon

I build data infrastructure
that moves petabytes a day.

Data pipeline platform at Amazon — control plane to data plane, cloud-native compute, and the efficiency work that keeps petabytes flowing with less latency and less cost.

01 Experience

Software Engineer — Amazon

Data pipeline platform Control & data plane Cloud-native

I build and operate the platform that powers customers' data pipelines — from the control plane that orchestrates pipeline operations to the data plane that executes them at petabyte scale.

Petabyte-scale platform Design and operate systems that serve petabytes of data to customers every day.
Control → data plane Own pipelines end to end: orchestration and scheduling in the control plane, execution in the data plane.
Cloud-native compute Serverless functions, elastic fleets, and large-scale transforms on AWS.
Efficiency-first Improve compute and storage efficiency to cut latency and reduce cost to serve.
Petabytes+ / day
Scale

A pipeline platform serving petabytes of customer data, daily.

Control → data
Architecture

Orchestrate in the control plane, execute in the data plane.

Faster · cheaper
Efficiency

Compute & storage efficiency that reduces delays and cost to serve.

02 The platform

How a pipeline flows through the platform: ingest, transform, and export — orchestrated end to end and delivered to customers.

CONTROL & ORCHESTRATION orchestration · scheduling · monitoring Control plane pipeline operations DATA PLANE ingest → transform → export manages Sources customer data Ingest high-throughput Transform large-scale Export / Query on demand Customers 24/7 petabytes+ served daily ingest → export end to end control + data one platform lower latency lower cost to serve 24/7 pipeline operations

03 Stack

Languages

  • TypeScript
  • JavaScript
  • Go
  • Python

AWS & cloud

  • AWS Lambda
  • EC2
  • EMR
  • S3
  • Serverless
  • Cloud-native

Engineering

  • Distributed systems
  • Data pipelines
  • System design
  • Observability
  • Cost optimization

Practices

  • Control / data plane design
  • Performance tuning
  • Reliability
  • Efficiency

04 About

I'm a software engineer at Amazon, working on a data pipeline platform that serves petabytes of data every day. My work spans the control plane — the APIs, orchestration, and scheduling that run pipeline operations — and the data plane — the cloud-native compute that executes them: Lambda for serverless ingest, EC2 fleets for processing, and EMR for large-scale transforms.

I design systems that make pipelines easier to operate and cheaper to run. Most of my work comes back to efficiency: improving compute and storage so customers get their data faster, at lower cost, at every scale.

05 Contact

Have a role, a project, or an idea worth building? My inbox is open.

stevensqiu@gmail.com