Rutgers computer science and data science student targeting off-cycle SWE, AI infrastructure, ML systems, cloud backend, and MTS intern roles.
portfolio · merged open-source PRs · LinkedIn
// work
KVCacheForge-XKV-cache bottleneck lab · TTFT, latency, throughput, HBM stalls, GPU busy, baseline deltas |
RoboFleetOpsAWS-native robotics fleet control plane · Lambda, DynamoDB, SQS, IoT Core, API Gateway, CDK CI |
PosCacheBenchLong-context benchmark for positional-attention failure modes under fixed KV-cache budgets |
open-source systems PRsReviewed and merged work across NVIDIA, IBM, Microsoft, Hugging Face, FlashAttention, Kubernetes, Pulumi |
Current proof standard: every performance claim needs reproducible commands, baseline comparison, hardware/software environment, and an honest limitations section.
// research
satellite telemetry anomaly detection100K telemetry readings · 5 NASA/ESA fault modes · recurrence-plot CV · 0.91 F1 on Kepler-class wheel oscillation PDF · repo |
bell labs ml impact analysis71-paper corpus · semantic clustering · co-authorship networks · Gradient Boosting AUC 0.674 · SHAP attribution PDF · repo |
// open sourceNVIDIA/cuda-python#2087FIPS-safe hashes for program cache keys |
NVIDIA/cuda-quantum#4688nvqpp: discriminate measured-register bool iteration |
huggingface/accelerate#4054Aggregate profiler memory example |
Dao-AILab/flash-attention#2622weights_only=True across all torch.load sites |
ai-dynamo/dynamo#10281HTTP 415 for unsupported image formats |
linkedin/Liger-Kernel#1157Guard save_for_backward on grad_bias in fused linear CE |
// stack// metrics


