Skip to content
View M4cr0Chen's full-sized avatar

Highlights

  • Pro

Block or report M4cr0Chen

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
M4cr0Chen/README.md

👋 Hi there! I'm Zhenghong (Marco) Chen

Twitter LinkedIn GitHub

I'm currently a third year Computer Science student at the University of Waterloo.

My interests: ML Infra, World Model, LLM Inference, GPU Programming, Distributed System.

Some of my experiences:

  • Lead Machine Learning Engineer at WAT.ai [Present]
  • Software Developer Intern at Geotab [Summer 2026]
  • Software Developer Intern at Geotab [Fall 2025]
  • Software Developer Intern at Octopodi Technologies [Winter 2025]

Pinned Loading

  1. mini-vllm mini-vllm Public

    A minimal vLLM-style LLM inference engine for Apple Silicon, built on MLX.

    Python

  2. llm-from-scratch llm-from-scratch Public

    Forked from stanford-cs336/assignment1-basics

    Implementation of a Large Language Model From Scratch - Stanford CS336

    Python

  3. llm-gateway llm-gateway Public

    Smart proxy layer between clients and LLM providers — routing, caching, rate limiting, cost control, observability

    Go

  4. CacheSystem CacheSystem Public

    High concurrency cache system implemented in C++. Utilized optimized LRU, LFU, ARC algorithms for data eviction.

    C++

  5. WLP4Compiler WLP4Compiler Public

    A compiler for WLP4 programming languages (a subset of C). The compiler implemented scanner, parser, semantic analyzer, and code generator from scratch using C++.

    C++

  6. KingDavidKo/Expense-App KingDavidKo/Expense-App Public

    Kotlin 1