• Backend Engineer with a strong interest in distributed systems and scalable infrastructure
• Passionate about ML infrastructure, inference optimization, and model serving
• Currently exploring CUDA, GPU programming, and high-performance computing
Pinned Loading
-
Disaster-Tweet-Classification-Using-Deep-Learning-LSTM-and-BERT
Disaster-Tweet-Classification-Using-Deep-Learning-LSTM-and-BERT PublicJupyter Notebook
-
-
Fleet-OS
Fleet-OS PublicFleetOS — an operating system for AI inference node fleets, handling orchestration, self-healing, and model deployment at scale
Python
-
LLM-inference-engine
LLM-inference-engine PublicA production-style LLM inference engine with dynamic batching, async queuing, and Redis semantic caching — inspired by systems like vLLM
JavaScript
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.