About Me

I'm a software engineer with 4+ years of experience building AI infrastructure at AWS. My focus is the capacity layer — the systems that manage how GPU resources get reserved, scheduled, and delivered to ML workloads at cloud scale.

Before AWS, I interned at Apple working on strategic data infrastructure.

I studied Computer Science and Mathematics at Boston University (BA, 2020), then at Carnegie Mellon University (MS, 2022).

This Blog

I write about the infrastructure that makes large-scale AI possible:

  • GPU scheduling in Kubernetes (DRA, topology-aware placement, Karpenter)
  • Capacity planning and reservation systems for AI workloads
  • Distributed systems patterns (idempotency, workflow orchestration, state machines)
  • The scheduling–capacity interface: how cluster schedulers and capacity planners should talk to each other

Posts are bilingual (中文/English). Deep technical dives tend to be in English; industry commentary and career reflections often in Chinese.

Beyond Work

I host a Chinese-language podcast (杨思特的半熟电台) interviewing ordinary people with extraordinary stories. Topics range from AI and games to overseas life and more.

Contact

Email: mincany0708@gmail.com