Skip to content
View bhavya-giri's full-sized avatar

Block or report bhavya-giri

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
bhavya-giri/README.md

Bhavya Giri

MLOps/AI Engineer at relevaince.ai. Exploring GPU programming, kernel optimization, and writing C++ on the side.

now

  • building ML infrastructure — training pipelines, model serving, deployment on k8s
  • learning CUDA — writing custom kernels for transformer inference
  • reading about computer architecture and parallel computing

projects

2026 cuda-kernels custom CUDA kernels for transformer inference — fused attention, quantized GEMM
2025 mlops-pipeline end-to-end ML pipeline with automated training, evaluation, and deployment
2025 gpu-bench benchmarking suite for GPU memory bandwidth and compute throughput
2024 vector-db lightweight vector database with HNSW indexing and CUDA-accelerated search
2024 model-serving high-throughput model serving with dynamic batching and TensorRT

connect

bhavyagiri.com · twitter · linkedin · email


// this readme runs on 1 thread. not optimal.

Pinned Loading

  1. model-app model-app Public

    Python

  2. retrieving-memes retrieving-memes Public

    Semantic Search for memes

    Jupyter Notebook 3

  3. spoiler-alert spoiler-alert Public

    Fine-tuned roberta base for predicting is a review a spoiler or not?

    Jupyter Notebook

  4. youtube-qa youtube-qa Public

    Jupyter Notebook