Skip to content
View irajput215's full-sized avatar

Block or report irajput215

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
irajput215/README.md
Typing SVG Banner

Software Engineering • Data Engineering • AI Engineering

LinkedIn GitHub Email

Sydney UNSW Focus


Executive Summary & Engineering Focus

Software Engineer working across data and AI — designing the architecture first, shipping it to production, and staying calm when it breaks at 2am.

I work across three disciplines that reinforce each other: software engineering as the foundation, then data engineering and AI engineering built on top of it. That means I design typed, tested, well-bounded systems — and then apply them to pipelines and models that have to hold up in production, not just in a notebook.

I care about the parts most demos skip — idempotent pipelines, tested transformations, typed tools behind real boundaries, measured accuracy, and human-in-the-loop control before anything touches production data.

Software Engineering Data Engineering AI Engineering

  • Software Engineering: Python and TypeScript across the stack, typed contracts and clean module boundaries, FastAPI and REST API design, React front ends, pytest suites with enforced coverage, CI/CD quality gates, Docker packaging, and refactoring that treats readability as a feature.
  • Data Engineering: Medallion lakehouse design (S3 → Snowflake → dbt → Airflow), dimensional and SCD2 modelling, incremental/MERGE fact loading, keyless cloud storage integrations, and read/write boundary separation by role.
  • AI Engineering: LangGraph state machines, bounded tool-calling agents, hybrid RAG (pgvector + knowledge graph), QLoRA fine-tuning with execution-based eval harnesses, and LLM enrichment turned into queryable warehouse columns.
  • Production Reliability & Incident Response: MLflow experiment tracking, drift and data-quality monitoring, structured tracing, human-in-the-loop approval gates on anything that mutates production data, and root-cause analysis that reports UNRESOLVED rather than guessing.

Featured Projects

Zomato AI Data Platform

10M-Order Batch Lakehouse + LLM Analytics Layer

End-to-end batch platform: raw CSVs → Amazon S3 → Snowflake (medallion) → dbt → Airflow, over 10M orders and ~23M order items. An LLM lane enriches 300K free-text reviews into tested sentiment/topic columns, powering RAG chat, text-to-SQL and a Streamlit mart dashboard — every read-only consumer bound to a read-only role.

Repo dbt

Snowflake dbt Airflow S3

CourseLLM

Full-Stack Agentic RAG Platform with Citations & Evals

A complete application: FastAPI backend with a React 19 + TypeScript front end, built as a typed monorepo with Docker packaging and CI. Users upload course material, state a goal, and get answers grounded in those documents — with citations, a prerequisite-aware roadmap, quizzes and progress-adaptive planning. Hybrid retrieval over PostgreSQL + pgvector, a knowledge graph, a bounded LangGraph, and a measured evaluation pipeline, held up by 1,643 tests at 83% coverage under strict mypy and ruff.

Repo Tests mypy

LangGraph FastAPI React TypeScript Docker pgvector

Pipeline Incident Response Agent

Agentic Root-Cause Analysis for Failed Data Pipelines

Given a 2am alert that customer_claims_daily failed, a LangGraph state machine triages it across nine failure categories, then runs bounded rounds of investigation using nine typed diagnostic tools behind a real read-only SQL boundary. It commits to a root cause only when evidence supports one — otherwise it ends UNRESOLVED rather than guessing. Data-mutating fixes pause for human approval, and every node, tool call, token and millisecond is persisted.

Repo Evals

LangGraph LangSmith HITL SQL

GTSRB Traffic Sign Recognition

Production-Ready 43-Class Computer Vision System

Traffic sign classification taken from notebook to deployed service, with reproducible splits, calibrated confidence and structured error analysis. Checkpoint selection keys on macro F1 rather than accuracy — because accuracy hid a class sitting at 54.2% recall. 98.90% accuracy, 0.9837 macro F1 on the official 12,630-image test split, with MLflow tracking and a containerised FastAPI inference service.

Repo Tests

PyTorch MLflow FastAPI Docker


Technical Competencies

Domain Core Technologies & Tooling
Software Engineering Python TypeScript FastAPI React 19 pytest mypy Ruff Docker
Data Engineering & Warehousing Snowflake dbt Apache Airflow Apache Spark Databricks Amazon S3 Delta Lake
AI Engineering LangGraph LangChain Hugging Face vLLM OpenAI PyTorch scikit-learn
MLOps, Cloud & Reliability MLflow AWS GitHub Actions LangSmith PostgreSQL pgvector Streamlit
Deep-Dive: Comprehensive Technical Stack & Tooling
  • Languages: Python (3.12+), TypeScript, JavaScript, SQL (advanced — window functions, MERGE, recursive CTEs), Bash, Java (basics).
  • Software Engineering: Typed interfaces and clean module boundaries, FastAPI and REST API design, JWT auth, SQLAlchemy, React 19 / Vite / Tailwind front ends, pytest suites with enforced coverage, strict typing (mypy) and linting (ruff), code review and refactoring, Docker packaging, semantic versioning and Makefile-driven developer workflows.
  • Data Engineering: Medallion (Bronze/Silver/Gold) lakehouse design, star schemas, SCD Type 2 history, incremental fact loading, partitioning and clustering strategy, data contracts, idempotent DAG design, keyless S3↔Snowflake storage integrations, least-privilege role separation.
  • Transformation & Orchestration: dbt (models, tests, macros, incremental strategies, exposures), Apache Airflow (DAGs, sensors, retries, backfills), Apache Spark / PySpark, Databricks notebooks and workflows, Delta Lake.
  • AI Engineering: RAG (hybrid dense + lexical retrieval, knowledge graphs), LangGraph state machines, typed tool-calling agents, QLoRA / LoRA fine-tuning, vLLM serving, prompt and eval harnesses, execution-accuracy benchmarking, LLM enrichment into warehouse columns, text-to-SQL, guardrails and human-in-the-loop gates.
  • MLOps & Observability: MLflow tracking and registry, CI/CD quality gates on model accuracy, containerized serving, drift and data-quality monitoring, structured tracing (LangSmith).
  • Cloud & Infrastructure: AWS (S3, IAM, Lambda, EC2), Snowflake, Vercel, PostgreSQL + pgvector, SQLite, GitHub Actions, Linux.

Certifications & continuous learning: Databricks Fundamentals, Claude AI Fluency.


Core Engineering Tenets

"A pipeline you can't re-run isn't a pipeline. An agent you can't measure isn't a system. Good engineering is what makes both of them maintainable."

  • Software Craft First: Clear interfaces, low coupling and high cohesion, typed contracts end to end. The same discipline that makes an API maintainable is what makes a pipeline re-runnable and a model reproducible.
  • Measure, Don't Assert: Every AI claim belongs behind an automatic evaluation harness. If accuracy can't be reproduced on a held-out set, it isn't a result.
  • Idempotent & Observable by Default: Re-runnable pipelines, tested transformations, structured traces and honest failure modes — including saying "unresolved" rather than guessing.
  • Built for Failure, Not Just Success: Incident response, human-in-the-loop approval gates and graceful degradation are designed up front, not bolted on after the first outage.

Open to software, data and AI engineering roles, and to collaborating on production systems that need to be reliable, observable and well-architected.

Get in Touch LinkedIn



Profile Views

Pinned Loading

  1. coursellm2 coursellm2 Public

    AI Tutor

    Python

  2. incident-response-agent incident-response-agent Public

    An agentic AI system that investigates failed data pipelines — LangGraph orchestration, nine typed diagnostic tools behind a read-only SQL boundary, human-approved remediation, and a CI gate on roo…

    Python

  3. zomato-ai-data-platform zomato-ai-data-platform Public

    End-to-end batch data platform over 10M food-delivery orders: S3 → Snowflake (medallion) → dbt → Airflow, plus an LLM layer that turns 300K free-text reviews into tested sentiment/topic columns. 17…

    Python

  4. GTSRB-Traffic-Sign-Recognition-Deep-Learning-Project- GTSRB-Traffic-Sign-Recognition-Deep-Learning-Project- Public

    Traffic Sign Recognition for Autonomous Driving (COMP9444 Neural Networks & Deep Learning, UNSW)

    Python

  5. databricks-wanderbricks databricks-wanderbricks Public

    end to end data + AI/ML pipeline

    Jupyter Notebook

  6. fine-tuning-llms-text2sql fine-tuning-llms-text2sql Public

    fine tuning LLMs using vLLM and huggingface with evaluation harness

    Python