Machine Learning Developer

Derryl Kevin

I build production ML systems — from research prototypes to low-latency inference. Currently shipping LLM tooling and MLOps platforms at a small, fast team.

12

Models shipped

4.2B

Tokens served

9ms

p99 latency

6

Publications

Selected work

(a) / 2023–2025

LLM · RAG

Retriever-X

Hybrid retrieval pipeline with reranking for enterprise search.

PyTorchFAISSFastAPI

+18% recall · 40ms p95

Vision · Edge

EdgeDetect

Quantized detection model running on-device for quality control.

ONNXTensorRTC++

0.91 mAP · 3× faster

MLOps · Infra

PipelineForge

Reproducible training orchestration with automatic model registries.

KubernetesRayMLflow

-60% deploy time

Stack

(b) / tooling

Languages & frameworks

PythonPyTorchJAXNumPyCUDA

LLMs & MLOps

TransformersvLLMDockerKubernetesMLflow

Experience

(c) / log
2023—now

Senior ML Engineer · Northwind AI

Leading LLM inference and retrieval systems.

2021—23

ML Engineer · Helios Labs

Built vision pipelines and MLOps tooling.

2019—21

Research Engineer · Cartograph

Geospatial models and data infrastructure.

Publications

(d) / papers

Contact

Let's build something precise.

© 2025 Derryl Kevin · Built with care