$Building production-grade AI systems

Hi, I'm Sameet Asadullah

AI/ML Engineer

AI/ML engineer building LLM, RAG and computer-vision systems that scale, from vector search and model serving to MLOps pipelines supporting up to 20M daily requests.

SA
Sameet Asadullah
0.0%
Platform uptime
0M+
Daily requests served
0+
ML models deployed
A quick introduction

About Me

SA
Sameet Asadullah

Sameet Asadullah

AI/ML Engineer

I'm an AI/ML engineer building and deploying production systems end to end, from LLM-powered features and RAG pipelines to vector search, model serving and observability. I work hands-on across the full lifecycle, using Python, FastAPI, Docker, Kubernetes and Triton on AWS and GCP.

I hold a Master's in Artificial Intelligence and Machine Learning from the University of Adelaide and a Bachelor's in Computer Science from FAST-NUCES. I bring a practical, product-minded approach to engineering with a strong focus on performance and reliability.

What I do

LLMs & RAG
Semantic Search
Generative AI
Computer Vision
MLOps & Serving
Backend APIs

Education

Master of Artificial Intelligence and Machine Learning

The University of Adelaide

Bachelor of Computer Science

National University of Computer and Emerging Sciences (FAST-NUCES)

Where I've worked

Experience

  • AI Engineer

    Jul 2024 - Jan 2026

    Stealth Startup · Sydney, Australia

    • Designed and rolled out a vector-based search engine with Elasticsearch and transformer embeddings, improving product search relevance by around 35–45%.
    • Integrated an LLM-driven recipe feature powered by RAG, matching required ingredients to supermarket products and adding them straight to the cart.
    • Built web scrapers for Coles, Aldi, IGA and Woolworths that keep thousands of product listings up to date in real time.
    • Created a product normalization pipeline that cleaned up messy product names and generated embeddings, improving both search ranking and product grouping.
    ElasticsearchTransformersRAGPython
  • Machine Learning Engineer

    Feb 2025 - Jul 2025

    Add Life Technologies · Adelaide, Australia

    • Optimized a real-time body tracking application by integrating MediaPipe with Unity and tuning performance for mobile, cutting processing latency by 40% on mid-range devices.
    • Led end-to-end development of a cross-platform AI-powered mobile app (backend APIs in FastAPI, front end in Flutter) for seamless real-time video stream handling.
    • Improved system reliability and scalability through stress testing, memory leak fixes and asynchronous data handling, stabilizing performance under 10+ concurrent video streams.
    MediaPipeUnityFastAPIFlutter
  • Machine Learning Engineer

    Aug 2022 - Jan 2024

    Vyro · Wyoming, United States

    • Built a bespoke serving architecture for ImagineArt, the second most popular AI art generator in the US, using FastAPI and Docker, supporting 30+ ML models at up to 20M requests/day with 99.5% uptime on Kubernetes.
    • Deployed ML models for Phototune (10M+ downloads) on Triton Inference Server, optimizing response time to 2–3 seconds.
    • Implemented a Docker-based serverless architecture for AvatarMe using Runpod and AWS (ECS, S3), cutting avatar-training turnaround from hours to 15 minutes.
    • Designed GitHub Actions CI pipelines achieving 97% test coverage, reducing production errors by 95%.
    • Engineered a Python SDK to host any Stable Diffusion workflow in production, cutting feature-hosting time by 80% through reuse.
    FastAPIDockerKubernetesTriton Inference ServerStable Diffusion
Some things I've built

Featured Projects

Stealth Startup

Vector-Based Product Search

Vector-based product search engine built with Elasticsearch and transformer embeddings, improving search relevance by 35–45% across the catalog.

ElasticsearchTransformersPython
Company project
Stealth Startup

RAG Recipe Copilot

LLM-driven recipe feature powered by RAG: enter a dish name and get a recipe with ingredients semantically matched to supermarket products and added straight to the cart.

RAGPythonElasticsearch
Company project
Add Life Technologies

Real-Time Body Tracking

Real-time body tracking app integrating MediaPipe with Unity, tuned for mobile deployment with a 40% latency reduction.

MediaPipeUnity
Company project
Vyro

ImagineArt Model Serving

Bespoke ML serving architecture for ImagineArt: 30+ models, up to 20M requests/day and 99.5% uptime on FastAPI, Docker and Kubernetes.

FastAPIDockerKubernetes
Company project
Vyro

Phototune on Triton

Triton Inference Server deployment for Phototune (10M+ downloads), optimized for a 2–3 second response time.

Triton Inference ServerDocker
Company project

ApplyGraph

Session-based agentic AI job copilot using FastAPI, LangGraph and PostgreSQL + pgvector to analyze job fit, tailor resumes, draft outreach and persist semantic memory across chat threads.

FastAPILangGraphPostgreSQL
View code

MergeWise

AI-powered pull request reviewer combining RAG with FAISS-based context retrieval, LLM reasoning and GitHub Checks, run inline or via a Celery/Redis queue.

RAGFAISSFastAPI
View code

AutomateIt

Home automation system using a CNN and React Native, recognizing Urdu voice commands with over 85% accuracy to control household appliances.

CNNReact NativeMongoDB
View code
Measurable results

Impact by the Numbers

0%Search relevance improvement from vector + hybrid search
0M+Daily requests served in production
0.0%Platform uptime at peak load
0%Latency reduction on real-time mobile computer vision
0%Faster time-to-ship for new GenAI features
0%Fewer production errors after a CI overhaul
My technical toolbox

Skills & Tools

Languages

PythonC++JavaSQL

Machine Learning

PyTorchTensorFlowScikit-learnOpenCVMediaPipeTransformers

Generative AI

LLMs (GPT, BERT)RAGLangGraphStable DiffusionGANs

Backend & Data

FastAPIREST APIsPostgreSQLMongoDBRedisElasticsearch

MLOps & Cloud

DockerKubernetesTriton Inference ServerAWSGoogle CloudGitHub ActionsLinux

AI Coding Tools

CursorClaude CodeGitHub CopilotChatGPTGemini
Contact

Let's Work Together

Have a project or role in mind? Drop me a line.

Prefer email?sameetassadullah744@gmail.com
Response timeUsually within 24 hours