Hi, I'm Ujwal

AIML Engineer

Building from scratch is what I do for fun.

Career Path

Research Implementations

LLM Transformer Architecture from Scratch

LLM Transformer Architecture from Scratch

01
PyTorch·BPE·RoPE·RMSNorm
Training LLM to Reason with Reinforcement Learning(GRPO)

Training LLM to Reason with Reinforcement Learning(GRPO)

02
PyTorch·GRPO·SymPy·SmolLM2-135M
Distilling LLMs: A Reverse-KL Implementation

Distilling LLMs: A Reverse-KL Implementation

03
PyTorch·Reverse KL·Qwen2.5·On-Policy Distillation

Projects

Multi Agent RL Env for Disaster Relief (Meta Hackathon)

Multi Agent RL Env for Disaster Relief (Meta Hackathon)

04
OpenEnv·FastAPI·Docker·Python
Agentic AI Self-Healing CI/CD Pipeline

Agentic AI Self-Healing CI/CD Pipeline

05
Qwen-2.5-Coder·MCP·GraphDB·GitHub Actions
AdaptIQ — Agentic Employee Onboarding(IISC hack)

AdaptIQ — Agentic Employee Onboarding(IISC hack)

06
Qwen·GPT-OSS·FastAPI·Next.js
Multi Neural Network Mental Health Assessment

Multi Neural Network Mental Health Assessment

07
Python·TensorFlow·PyTorch·OpenCV

Now

Skills

Languages3 Technologies
ML / Deep Learning10 Technologies
LLM / GenAI17 Technologies
Big Data7 Technologies
Databases4 Technologies
Backend / Systems4 Technologies
Cloud / DevOps7 Technologies

Building

RL Environment to GRPO Fine-Tune LLMBUILDING
Base Inference System for LLM ServingBUILDING

Learning

Reinforcement Learning using openENVLEARNING
GPU Architecture and Data TransferLEARNING
GPU ProfilingLEARNING
Triton KernelsLEARNING
LLM InferenceLEARNING

Activity

GitHub Contributions

ujwal-s-r
2026 Activity
LessMore

Latest Repositories

Loading public repositories...

LinkedIn Feed

Scroll or drag verticallyTap button to view on LinkedIn