👋 Hi, I'm Nathan Lemma · ECE @ UT Austin · Systems Engineer
🔬 Now: Agentic ML systems • Software Engineering Intern @ Caterpillar
⏪ Previously: SysML Research Group • Longhorn Racing (Battery Management) • FAST Research Group
| Project | What it is |
|---|---|
| EnerTune Baselines | Four published GPU-sharing serving systems rebuilt and ported to a 16-A100 cluster |
| MLLM-Energy | 1,080 profiled configs showing a one-size GPU split wastes energy across multimodal LLM phases |
| CHT-Radix | Seven concurrent hash tables/radix trees benchmarked against a real ShareGPT trace |
| Pintos | Teaching kernel extended with user processes, syscalls, demand paging, full filesystem |
"Beyond Utilization: Energy-Conscious GPU Sharing for Inference Serving" SOSP '26 · 3rd author Rebuilt and ported the four GPU-sharing serving systems EnerTune is benchmarked against onto a 16-A100 cluster — the evaluation behind its 1.4–2.3x energy reduction.


