Ashraf Ahmed

Backend, Systems & ML-focused Engineer.

Contributor to Zalando Skipper & CNCF etcd — 6 merged PRs running in production, including a load-balancing algorithm.

Final-year CS · Bengaluru, India · open to SWE roles, 2026

Open Source

zalando/skipper3.1K+

weightedRoundRobin Load-Balancing Algorithm

Designed and shipped a smooth weighted round-robin algorithm with dynamic health-derived endpoint weights. Defended the locking design under maintainer review with a self-built C benchmark harness sustaining ~31k req/s at 1,000 endpoints — the harness was kept in the repo for testing all algorithms.

PR #4124+599/−4
View PR
zalando/skipper3.1K+

Prometheus Native Histogram Support

Implemented opt-in native histogram support with OTEL-recommended bucket tuning, deploying to production Kubernetes clusters while maintaining backward compatibility with existing dashboards.

PR #4108+245/−146
View PR
zalando/skipper3.1K+

Per-Data-Client Load Latency Metrics

Added instrumentation to measure route-loading latency per data source, giving SRE teams visibility into which integrations are slow.

PR #4087+76/−1
View PR
etcd-io/website48K+

TLS Certificate Matching Documentation

Clarified security documentation for certificate-matching semantics across v3.5–v3.7, resolving a long-standing confusion about wildcard support.

PR #1171+6/−3
View PR
etcd-io/gofailCNCF

Go Toolchain Security Updates

Patched three CVEs by bumping Go to 1.25.11 as part of etcd organization's tracked security effort.

PR #149+2/−2
View PR

Selected Projects

Aegis — AI-Native Distributed Systems Observability Platform

Automated microservice failure prediction and root-cause analysis, achieving a 228-second predictive warning lead time and 100% topological diagnosis agreement, by engineering an event-time stream correlator using Kafka and a shadow-mode MLOps pipeline with HistGradientBoosting and PSI drift monitoring.

PythonFastAPIKafkaScikit-LearnReactPrometheus

Go + Java Auth Platform — Polyglot Microservices Identity System

Built a two-service identity platform: a Go auth service with rotating refresh tokens, reuse detection, and a Redis JTI blacklist, consumed by a Java/Spring Boot resource API that validates every request over gRPC — with a token-hash Caffeine validation cache and a fail-closed Resilience4j circuit breaker. 61 tests (Testcontainers + in-process gRPC) in CI; k6-load-tested at ~92 req/s with 6ms p50 across the two-service auth path.

GoJavaSpring BootgRPCJWTRedisPostgreSQL

Custom Key-Value Storage Engine

Built a durable, crash-resilient storage engine from scratch in Python, achieving ~26,000 writes/sec in batched fsync mode and 100% recovery rates under simulated unclean shutdowns (kill -9), by designing an LSM-tree architecture with a Write-Ahead Log, Bloom filters, and size-tiered compaction.

PythonLSM-TreeWALTCPDockerPytest

ClearText API — Async AI Inference Platform

Optimized toxic comment ML inference throughput, scaling to 231 requests/sec at 106ms average latency under a 500 concurrent-user load, by building an asynchronous FastAPI serving backend utilizing Celery task queues, Redis model-versioned caching, and a custom worker micro-batcher.

FastAPIRedisPostgreSQLCeleryGroq LLM

Distributed Log Processing & Analytics System

High-throughput real-time log processing platform ingesting 343 logs/sec with 10ms P50 latency and 99.5% processing reliability, with real-time alerting and analytics.

FastAPIRedis StreamsPostgreSQL

Lightweight Radar–Camera Fusion System (RAMP-CNN + YOLO)

Multi-modal object detection system achieving >90% accuracy on 5,000+ COCO samples with CUDA acceleration. Published as peer-reviewed research paper (ASIANCONF 2026).

PythonPyTorchCUDA

About

Final-year Computer Science student and backend/distributed systems engineer with hands-on experience building production-grade distributed applications, high-performance storage engines, and real-time AI inference platforms. Proven expertise in designing scalable, reliable, and observable systems that handle high concurrency with low latency.

New Horizon College of Engineering, Bengaluru

B.E in Computer Science and Engineering | GPA: 7.71 | Nov 2022 – July 2026

Coursework: Data Structures & Algorithms, Design & Analysis of Algorithms, Operating Systems, Computer Networks, Database Management Systems, Linux System Programming, Computer Architecture (ARM), Machine Learning, Generative AI, Cloud Architecture & Security

Publication

ASIANCONF 2026 — "Lightweight Radar-Camera Fusion for Real-Time Drone Detection" (Peer-reviewed)

Skills & Tech Stack

CUDA
Performance Benchmarking
Java
Docker
gRPC
Stream Processing
TypeScript
Celery
Caching
Go
Fault Tolerance
RAMP-CNN
Spring Boot
Event-Driven Architecture
Pytest
Distributed Systems
Python
Redis
Task Queues
FastAPI
Async Processing
Gin
Load Testing
LSM-tree
Observability
PyTorch
AWS EC2
Kafka
YOLO
Git
SQLite
Concurrency
WAL
Locust
REST APIs
DSA
Kubernetes
C++
PostgreSQL
Rate Limiting
Groq LLM
BERT
System Design
GitHub Actions
CUDA
Performance Benchmarking
Java
Docker
gRPC
Stream Processing
TypeScript
Celery
Caching
Go
Fault Tolerance
RAMP-CNN
Spring Boot
Event-Driven Architecture
Pytest
Distributed Systems
Python
Redis
Task Queues
FastAPI
Async Processing
Gin
Load Testing
LSM-tree
Observability
PyTorch
AWS EC2
Kafka
YOLO
Git
SQLite
Concurrency
WAL
Locust
REST APIs
DSA
Kubernetes
C++
PostgreSQL
Rate Limiting
Groq LLM
BERT
System Design
GitHub Actions
CUDA
Performance Benchmarking
Java
Docker
gRPC
Stream Processing
TypeScript
Celery
Caching
Go
Fault Tolerance
RAMP-CNN
Spring Boot
Event-Driven Architecture
Pytest
Distributed Systems
Python
Redis
Task Queues
FastAPI
Async Processing
Gin
Load Testing
LSM-tree
Observability
PyTorch
AWS EC2
Kafka
YOLO
Git
SQLite
Concurrency
WAL
Locust
REST APIs
DSA
Kubernetes
C++
PostgreSQL
Rate Limiting
Groq LLM
BERT
System Design
GitHub Actions
Python
Load Testing
Observability
PyTorch
Concurrency
CUDA
Kubernetes
Rate Limiting
GitHub Actions
Locust
Pytest
TypeScript
Git
DSA
System Design
gRPC
Performance Benchmarking
BERT
YOLO
Async Processing
Go
WAL
Gin
SQLite
FastAPI
Task Queues
Celery
Caching
Kafka
Fault Tolerance
Redis
Java
Event-Driven Architecture
AWS EC2
Distributed Systems
Spring Boot
LSM-tree
Groq LLM
Stream Processing
REST APIs
C++
RAMP-CNN
Docker
PostgreSQL
Python
Load Testing
Observability
PyTorch
Concurrency
CUDA
Kubernetes
Rate Limiting
GitHub Actions
Locust
Pytest
TypeScript
Git
DSA
System Design
gRPC
Performance Benchmarking
BERT
YOLO
Async Processing
Go
WAL
Gin
SQLite
FastAPI
Task Queues
Celery
Caching
Kafka
Fault Tolerance
Redis
Java
Event-Driven Architecture
AWS EC2
Distributed Systems
Spring Boot
LSM-tree
Groq LLM
Stream Processing
REST APIs
C++
RAMP-CNN
Docker
PostgreSQL
Python
Load Testing
Observability
PyTorch
Concurrency
CUDA
Kubernetes
Rate Limiting
GitHub Actions
Locust
Pytest
TypeScript
Git
DSA
System Design
gRPC
Performance Benchmarking
BERT
YOLO
Async Processing
Go
WAL
Gin
SQLite
FastAPI
Task Queues
Celery
Caching
Kafka
Fault Tolerance
Redis
Java
Event-Driven Architecture
AWS EC2
Distributed Systems
Spring Boot
LSM-tree
Groq LLM
Stream Processing
REST APIs
C++
RAMP-CNN
Docker
PostgreSQL

Get in Touch

Interested in working together or have a question? I'd love to hear from you.

ashrafahmed1232@gmail.com