Skip to content
View kandulanikhilvarma's full-sized avatar
🦁
🦁

Highlights

  • Pro

Block or report kandulanikhilvarma

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
kandulanikhilvarma/README.md

Nikhilvarma Kandula

Currently: Data Engineer in training → AI Engineering layer. Targeting Germany. Building in public, zero to market.

I build data products and research-grade analysis pipelines. Currently building Rankwell, a GEO/AI-visibility tracker that shows law firms how ChatGPT, Gemini and Perplexity talk about them versus competitors — daily scans, explainable 0–100 visibility scores, citation-gap recommendations.


Featured projects

Project What it shows Stack
Quantifying Data Quality Four-dimensional DQ scoring framework (completeness, consistency, accuracy, timeliness), validated against the UCI Air Quality dataset (9,357 observations); NER confidence scoring at 88.4% precision; K-Means clustering on quality metrics Python · spaCy · scikit-learn
German Tech Job Market Intelligence Production-scale NLP pipeline over 3,200 job postings: multi-source scraping, TF-IDF + spaCy EntityRuler skill extraction (30,324 skill-posting pairs), Levenshtein deduplication, 4 interpretable role archetypes (Silhouette = 0.61) Python · spaCy · TF-IDF
Rider Segmentation & Growth 5.5M bike-share trips analyzed in BigQuery SQL — casual vs. member behavior with 3 data-driven marketing recommendations SQL · BigQuery · Tableau
Flight Delay Analysis 3M U.S. domestic flights (2019–2023): delay drivers, seasonality, carrier comparison Python · Jupyter
Return Rate Analysis E-commerce return-risk segmentation by product category, price range and customer segment, framed against Germany's €92B e-commerce market Python · BI
ESG–GDP Regression Whether CO₂ emissions per capita and renewable-energy share predict GDP — hypothesis testing and regression diagnostics Python · Jupyter
Manufacturing Cost Variance Qlik Sense dashboard for manufacturing cost-variance monitoring Qlik · Python

By the numbers

  • 9,357 records analyzed for the data-quality framework · 96.7% overall quality score
  • 3,200 German tech job postings processed · 156 unique skills extracted · 30,324 skill-posting pairs
  • 5.5M bike-share rides examined · 3M flights analyzed
  • 88.4% NER precision · 0.61 clustering silhouette score

Skills

Data science — statistical modeling, NLP and text analysis, unsupervised learning, feature engineering Analytics — SQL/BigQuery, Tableau, Power BI, business intelligence Engineering — Python, data pipelines, web scraping; TypeScript/Next.js for product work Research — hypothesis testing, quantitative analysis, labour-market research, German market analysis

Links

Portfolio kandula.studio
LinkedIn linkedin.com/in/nikhilvarmakandula
Kaggle kaggle.com/nikhilvarmakandula
Email kandulanikhilvarma@gmail.com

Open to collaboration and research partnerships. Last updated: July 2026.

Pinned Loading

  1. rankwell rankwell Public

    Rankwell - GEO/AI visibility tracker for law firms. Tracks how ChatGPT, Gemini and Perplexity talk about a firm vs competitors.

    TypeScript 1

  2. esg-gdp-regression esg-gdp-regression Public

    Does environmental performance drive economic growth? OLS regression across 38 OECD countries using World Bank Sovereign ESG data (2010–2020).

    Jupyter Notebook 1

  3. manufacturing-cost-variance-qlik manufacturing-cost-variance-qlik Public

    Qlik Sense cost variance dashboard - manufacturing controlling case study

    Python 1

  4. skill-demand-deutschland-tech-market skill-demand-deutschland-tech-market Public

    NLP corpus of 3,200 German tech job postings - TF-IDF + spaCy skill extraction + K-Means clustering

    Python 2

  5. quantifying-data-quality quantifying-data-quality Public

    A Statistical Framework for Scoring and Monitoring Scientific Datasets

    Jupyter Notebook 1

  6. flight-delay-analysis flight-delay-analysis Public

    Analysis of 3 million U.S. domestic flights (2019–2023) to identify what drives departure delays and whether delay likelihood can be predicted before a flight departs.

    Jupyter Notebook 1