whoami

Adil Imran Pintoo

AI Data Engineer

I build production data and AI systems — ingestion pipelines that don't fall over, orchestration platforms that run hundreds of jobs unattended, and LLM/RAG services that ship to Kubernetes instead of staying in a notebook. Most recently I built an AI-powered ESG scoring platform end to end at a mutual fund manager, from PDF ingestion to a production FastAPI service serving 2,000+ listed companies.

adil@portfolio — zsh
adil@portfolio:~$ welcome. this site is navigable by typing.
type "help" to see what you can do.

cat about.md

About

I build production data and AI systems — ingestion pipelines that don't fall over, orchestration platforms that run hundreds of jobs unattended, and LLM/RAG services that ship to Kubernetes instead of staying in a notebook. Most recently I built an AI-powered ESG scoring platform end to end at a mutual fund manager, from PDF ingestion to a production FastAPI service serving 2,000+ listed companies.

Based in Mumbai, India. I care about the unglamorous parts of data systems — idempotent pipelines, boring reliable orchestration, and services that fail loudly instead of silently.

git log --oneline --author=adil

Experience

  1. Data Engineer · DSP Investment Managers

    07/2025Present · Mumbai
    • Architected and shipped an AI-powered ESG scoring platform (LLMs, RAG, vector embeddings, Tavily web search, Gemini Search Grounding) on Kubernetes-deployed FastAPI microservices, automating ESG evaluation for 2,000+ listed companies and cutting manual analyst effort.
    • Built and now run the company's Apache Airflow orchestration platform from the ground up to production — designed the deployment, DAG conventions, and monitoring, and led a major-version upgrade in production. It orchestrates 100+ scheduled jobs across ingestion, ESG scoring, and analytics.
    • Built large-scale ingestion pipelines (Python, Airflow, AWS S3) that scraped and processed 2,300+ ESG PDFs, XBRL filings, and annual reports across multiple financial years, with idempotent processing and 99% pipeline reliability.
    • Designed and deployed a Typesense-based search engine — from schema design to containerized deployment as pods in the production Kubernetes cluster — powering fast, typo-tolerant search over companies (comps), portfolios, and financial metrics.
    • Designed PostgreSQL schemas, audit tables, and materialized views powering a real-time market dashboard (1-day to 10-year historical charting) and portfolio analytics across corporate actions, arbitrage funds, and index comparisons.
    • Engineered REST APIs and backend workflows for ESG score generation, analyst review, and version-controlled publishing, implementing a Generated → In Review → Finalized workflow with full audit history.
    • Debugged and hardened production Kubernetes services — root-caused a dependency-conflict outage traced to a stale Dockerfile pin and recovered via kubectl rollout undo — then rewrote the Capitaline ingestion pipeline to eliminate silent failures, corrupt files, and unsafe uploads.
    • Automated Snowflake permission auditing and redesigned ACL query generation for dynamic, enterprise-grade data access isolation.
  2. Co-Founder & CTO · Kevora Global Pvt. Ltd.

    02/202405/2025 · Delhi
    • Co-founded a D2C brand for 5 premium Kashmiri products, driving market-entry strategy with a target of 10% category share.
    • Built an NLP sentiment-analysis system over 10,000+ product reviews using NLTK and scikit-learn to guide product and messaging decisions.
    • Led technical strategy for Amazon and Flipkart marketplace integration to support expansion.
  3. Data Analyst Intern · Redemp Technologies Pvt. Ltd.

    05/202407/2024 · Pune, India
    • Designed an ETL pipeline (SQL, Python, Airflow) that improved ESG data integration and cut data latency by 26%.
    • Built automated data-validation and reconciliation checks across pipeline runs, catching schema drift and missing values before they reached client-facing dashboards.
    • Built a real-time KPI dashboard in Tableau that improved sustainability tracking and helped clients cut carbon footprints by 19%.
    • Ran data wrangling and lifecycle analysis (LCA) with Pandas and NumPy, driving a 15% reduction in client resource use.

cat education.md

Education

Indian Institute of Technology, Delhi

07/202108/2025

B.Tech in Biochemical Engineering & Biotechnology

Minor in Computer Science and Engineering

Data Structures & AlgorithmsOperating SystemsComputer NetworksDatabase Management Systems (DBMS)Programming LanguagesProbability Theory & Stochastic ProcessesProbability & Statistics

Certifications

Google Data AnalyticsGoogle IT SupportNYU CybersecurityIBM Full Stack Software Developer

pip list | grep -i relevant

Skills

AI / ML / LLM

LLMsRAGPrompt EngineeringVector DatabasesEmbeddingsGeminiTavilyAI EvaluationPyTorchscikit-learn

Backend & APIs

FastAPIREST APIsFlaskDjango

Data Engineering

ETL / ELT PipelinesData ModelingApache AirflowTypesensePandasNumPy

Databases

PostgreSQLMongoDBSnowflakeSQLite

Cloud & DevOps

AWS S3KubernetesDockerLinuxGit

Languages

PythonSQLJavaScriptCC++JavaGoR

Financial Systems

Market DataPortfolio AnalyticsESG AnalyticsCorporate Actions

Visualization & Other

Power BITableauPlotlyMatplotlibReactAdvanced Excel

mail -s 'hello' adil

Contact

Open to AI/data engineering roles and interesting problems. Fastest way to reach me is email — the form below works too.