Hi There,
I'm Indir Lal

i am into

About Me
Indir Lal

About Me

Indir Lal

I'm Indir Lal

Data Engineer

I am a results-driven Data Engineer specializing in designing and implementing high-performance, scalable ETL/ELT pipelines and cloud data architectures (Azure & GCP). With a proven track record of managing large-scale data systems—such as architecting a 60M+ record enterprise data warehouse in Google BigQuery from over 500M+ raw records—I excel at transforming complex, multi-source data streams into optimized, business-critical assets. Expert in Python, advanced SQL, PySpark, and modern data modeling practices, I build fault-tolerant data infrastructures that empower organizations with reliable, real-time insights and analytics-ready datasets.

Skills & Abilities

PythonPython
SQL
PySparkPySpark
AzureAzure
GCPGCP
BigQueryBigQuery
PandasPandas
NumPyNumPy
PostgreSQLPostgreSQL
MySQLMySQL
MongoDBMongoDB
SQLiteSQLite
FastAPIFastAPI
DockerDocker
TerraformTerraform
GitGit
LinuxLinux
PyTorchPyTorch
scikit-learnscikit-learn
JavaScriptJavaScript
TypeScriptTypeScript
JavaJava

Projects Made

CryptoVolt AI — Trading Platform

Full-stack platform with PyTorch LSTM + XGBoost models and sentiment analysis. Improved Sharpe ratio 0.41 → 0.89 and cut max drawdown 34.2% → 19.4%.

FastAPIReactPyTorchXGBoost

Serverless AQI Forecasting

Serverless ML pipeline forecasting air quality 72h ahead — feature store, ensemble models, SHAP explainability, scheduled GitHub Actions, and a live dashboard.

Pythonscikit-learnHopsworksGitHub Actions

Azure Lakehouse Pipeline

End-to-end pipeline via Azure Data Factory through ADLS and Databricks into a Medallion (Bronze/Silver/Gold) Lakehouse, with Terraform IaC, CI/CD, and Key Vault secrets.

ADFDatabricksDelta LakeTerraform

VeriDocs — RAG Q&A Tool

LLM / retrieval (RAG) app answering questions over documents with citations, confidence scoring, and refusal behavior for low-confidence queries.

PythonHugging FaceTF-IDFStreamlit
View All

Experience

Noble Ventures Data Inc

Data Engineer Intern — Remote (USA)

Jun 2026 - Present

  • Built a 60M+ record B2B/EIN warehouse in BigQuery from 500M+ raw records.
  • Engineered fault-tolerant Python ELT pipelines (24GB+ CSVs, JSONL).
  • Built multi-phase SQL matching engines bridging 9.6M+ records.
  • Deployed a FastAPI microservice over BigQuery with auth & audit logging.

Freelance (Fiverr)

Data Engineer — Remote

Dec 2023 - Present

  • Built & maintained ETL/ELT pipelines in Python and SQL for international clients.
  • Standardized a 400M+ row dataset and resolved pipeline failures.
  • Maintained Level 1 Seller status across projects.

Get in Touch