//02cv
↓ download

Adi Prabs

Computing @ Imperial · SRE @ Apple.

London, UKadiprabs19@gmail.comlinkedingithubapple sre · ml platforms
//summary

Computing student at Imperial with three years of production-level development experience across five startups and a current SRE role on the ML Platforms team at Apple.

I'm happiest at the seam between research and production — shipping things that demo well in a notebook and don't fall over under real traffic. Compilers, infra, agents, and the boring glue that turns prototypes into products.

//experience

Roles

2026 — present
London, UK
current

Apple

Site Reliability Engineer, ML Platforms
  • Built a Kubernetes capacity forensics platform end-to-end (collectors, scanners, delta-query UI) used by 30+ SREs to diagnose EC2 capacity exhaustion, saving up to $3M per AWS availability zone annually.
  • Built a cloud-agnostic capacity request and reservation management system — replaced ad-hoc Slack coordination with an auditable workflow handling hundreds of requests monthly.
  • Created a Kubernetes manifest validation framework detecting misconfigurations at deploy time; retrospective analysis shows it would have caught 72% of deployment-related incidents over the prior year.
KubernetesGoAWSSREDistributed systemsObservabilityLinux
Apr 2026 — May 2026

8x

Full-stack & AI/ML Developer
  • Optimized production analytics from 24s to sub-second via SQL-side aggregation and indexed Postgres RPC rewrites; cut /posts payloads 90%+ (22MB → ~1–2MB).
  • Simplified messaging architecture, deleting ~400 lines of legacy API code while enabling a new admin reply UX.
  • Resolved 3 critical production vulnerabilities: org takeover, exposed financial Server Actions, and DB search-path injection across 49 functions.
PostgreSQLNext.jsTypeScriptNode.jsSecurity
Mar 2026 — Apr 2026

Canopy Labs

General Engineer
  • Architected, built, and deployed the company's flagship full-stack web application, enabling real-time multi-user usage with Docker and Kubernetes orchestration.
  • Created agents to enable custom form filling from transcripts, cutting insurance resolution by 15 min per form.
  • Optimized backend concurrency and load balancing, reducing latency 500ms → 120ms, supporting 100+ active users.
  • Implemented AWS CI/CD pipelines with automated testing, shortening release cadence to 6 hours.
ReactTypeScriptNext.jsFastAPIRedisAWSKubernetesDocker
2025 — 2026

Vani

Full-stack & AI/ML Developer
  • Architected, shipped, and deployed the flagship multi-tenant web app on AWS with Docker + Kubernetes.
  • Cut p95 backend latency 500ms → 120ms; scaled to 100+ concurrent users.
  • Stood up CI/CD: release cadence 2 days → 6 hours, production bugs −76%.
  • Integrated LLM workflows via Model Context Protocol for clinical-admin automation.
  • Held 99.9% uptime through staged rollouts and load-balanced workers.
ReactTypeScriptNext.jsFastAPIRedisAWSKubernetesDocker
2024 — 2025

Trajex

Machine Learning Developer
  • Deployed LLama 3.2-7B-Instruct in production — 20% cost reduction vs OpenAI, 12% lower inference latency.
  • Led product design and built the inference backend.
  • Pitched investors and onboarded K3 Capital Group as a paying client.
LLama 3.2PythonInference optimizationProduct
2024

Altus Reach

ML Engineer (Contract)
  • Team of 3 — built a video saliency model improving prediction accuracy by 19%.
  • Shipped Azure-hosted inference pipeline for production traffic.
  • Full-stack work on company web app (TypeScript / Next.js / React).
Azure AIPythonTypeScriptNext.jsComputer Vision
//education
2023 — 2027 (expected)

Imperial College London

MEng in Computing (AI & ML)

  • On track for First Class · GPA 3.8
  • Coursework: compilers, OS, ML, distributed systems
  • Hackathons: 1st place — SwyftGesture (hands-free input)
//stack

What I reach for.

Languages
TypeScriptPythonCC++RustScalaHaskellKotlinJavaScript
Systems
LinuxDockerKubernetesCI/CDAWSAzureVercel
AI / ML
PyTorchLLama 3.2MediaPipeOpenCVMuJoCoMCPLocal LLMs
Web
Next.jsReactFastAPINode.jsTailwindMDX
Reliability
ObservabilityIncident responseLoad testingDistributed tracing
//languages
  • EnglishNative
  • FrenchNative
  • TamilNative
  • SpanishReading & speaking
  • HindiReading & speaking