Available for Internships & Full-time · Summer 2026

Siddharth Sharma

Systems Engineer  ·  GPU / HPC  ·  Full-Stack

I build fast, correct systems from TCP stacks in C++ to GPU-accelerated Monte Carlo engines and production LLM pipelines.

Top 4%
LeetCode Global
871
DSA Problems
5
Systems Projects
3
Languages · C++ Python TS
Projects Each project solves a real problem. Every number is measurable.

Built from scratch. Shipped with intention.

Click on the project to open it. →

ColumnarDB
↑ Vectorized analytical query engine

Built a columnar OLAP engine from scratch with a Vectorized Volcano-model executor. Implements full query pipeline — scan, filter, project, aggregate.

C++17CMakeSIMDQuery Engine
GitHub ↗
Monte Carlo CUDA
American Options Pricer — GPU-Accelerated QMC

Reimplementation of Cvetanoska & Stojanovski — pricing American call options via Bermudan backward induction. 5 compute backends share a single math core: BS, LCG, Moro Inverse CND, Sobol/Halton + Brownian Bridge.

CUDAC++17OpenMPQMCSIMD
GitHub ↗
TCP from Scratch
Full RFC-compliant TCP state machine

Complete TCP/IP stack built in both C++23 and Rust, bypassing the OS. Full TCP state machine, connection teardown, retransmission — by-the-RFC.

C++23RustNetworkingSystems
GitHub ↗
Alfred — AI Butler
8 serverless functions · 3072-dim vectors

Personal knowledge base: ingests text/audio/images → semantic chunking → Qdrant vector DB → spaced repetition, RAG chat, live voice via Gemini WebSocket.

React NativeGemini 2.5QdrantAppwrite
GitHub ↗
KAIRA Voice AI
Real-time STT · TTS · facial sync

Intelligent voice assistant with real-time Speech-to-Text, TTS, synchronized facial animation rendering. Microservices-based architecture using ONNX/TFLite.

PythonONNXTFLiteMicroservices
GitHub ↗
Gemma 3 RAG
Full PDF ingestion pipeline · Edge LLM

Lightweight RAG app running Gemma 3 270M locally. Full PDF parsing, chunking, embedding, and retrieval pipeline with sub-300M param inference.

PythonLangChainGemma 3RAG
GitHub ↗
Technical Skills

Deep where it matters.

Languages
C++ (primary · 3 yrs)CUDA, SIMD, STL, templates
Python (2 yrs)ML, scripting, backends
TypeScript (2 yrs)React, Next.js, Node
Rust (learning)TCP stack, systems
HPC / GPU
CUDAOpenMPSIMDParallel Algos
AI / ML
LangChainONNXTransformersRAG
Infra
KubernetesDockerAWSGCP
Backend / DB
Node.jsFastAPIPostgreSQLQdrant
What I can do on Day 1
Write production C++ with clean abstractions and zero-overhead design
Profile GPU kernels and optimize for memory coalescing & occupancy
Ship a full-stack feature from API design to deployed UI
Set up a RAG pipeline with vector embeddings end-to-end
Debug low-level networking and systems issues independently
Engineering Principles
"Measure before optimizing. Optimize before scaling."
"Complexity is a liability. Every abstraction must earn its place."
"I build things I would actually ship and maintain."
Track Record

Consistency is the only edge that compounds.

254 consecutive days of problem solving. Not sprints. Sustained execution.

LeetCode Knight · Top 4%
LeetCode Heatmap
Codeforces Specialist
Rating1,432
Education
B.E. Computer Science
Thapar Institute of Engineering & Technology
GPA 8.08 · Batch 2027
Open to Internships · Full-time · Contract · Summer 2026

Let's build something worth shipping.

I'm looking for roles where I can work on hard problems at the systems level whether that's HPC, ML infra, backend systems, or full-stack product engineering.

HPC / GPU Compute Backend / Systems Eng ML Infrastructure Full-Stack Product

Thapar University · Punjab, IN · Batch 2027