I will do data analysis, cleaning, and eda with python pandas and numpy


About this gig
Looking for clean data structures, reliable insights, and deployable code rather than just basic charts?
I am a Data Science Practitioner focused on transforming raw, messy data into functional machine learning pipelines and trustworthy business intelligence. My core methodology prioritizes clean data handling, honest analytical processing, and reproducible workflowsbecause the mathematical design behind your processing determines whether your results can actually be trusted.
What I Bring to Your Project:
Exploratory Data Analysis (EDA) Comprehensive parameter mapping, distribution analysis, and structural reporting.
Machine Learning Pipelines High-accuracy classifiers built with Scikit-learn and XGBoost, optimized using strict cross-validation and hyperparameter tuning to eliminate data leakage.
Complex Structural Cleaning Vectorized operations via Pandas and NumPy, handling missing data strategically through outlier-resistant median imputation rather than simple mean averages.
Advanced Text Parsing Custom string cleaning, Regex-based pattern filtering, and structural expansions (explode operations) for semi-structured text columns.
I document every processing loop and engineering
Get to know Mustajar Mehdi
Data Scientist, Core Python and ML Pipeline Developer
- FromPakistan
- Member sinceAug 2026
- Avg. response time1 hour
Languages
Urdu, English
My Portfolio
FAQ
What file formats can you work with?
CSV, Excel, JSON, and most common tabular data formats.
Can I see examples of your work before ordering?
Yes — check my portfolio notebooks and project walkthroughs linked in my profile description.
Do you handle large datasets?
Yes, within reasonable limits for the package tier. For very large datasets (100k+ rows) or complex pipelines, message me first to confirm scope.

