Job Description
Job Summary :
We are seeking a detail-oriented and analytically driven individual with a strong passion for data and machine learning. The ideal candidate will have hands-on experience across a range of predictive modeling techniques including classification, regression, and unsupervised learning. You will be responsible for the full ML pipeline from data preprocessing and exploratory analysis to model building and evaluation.
Key Responsibilities :
- Perform data cleaning , feature engineering , and preprocessing on structured datasets.
- Conduct exploratory data analysis (EDA) to uncover insights and inform model development.
- Build and evaluate supervised learning models for:
- Classification : Logistic Regression, Decision Trees, Random Forest
- Regression : Linear Regression, K-Nearest Neighbors, Gradient Boosting, Neural Networks
- Implement and evaluate unsupervised learning techniques , including K-means and other clustering algorithms.
- Apply performance metrics like accuracy, precision, recall, F1-score, RMSE, R² , and Silhouette Score to assess model effectiveness.
- Use techniques like cross-validation , grid/random search , and hyperparameter tuning to improve model performance.
- Visualize model outcomes and data trends using matplotlib , seaborn , or Plotly .
- Document experiments and communicate findings with stakeholders or team members.
Required Skills :
- Proficiency in Python and libraries such as pandas, numpy, scikit-learn, matplotlib, seaborn
- Experience with machine learning tools like XGBoost, LightGBM, and TensorFlow or PyTorch (for neural nets)
- Strong understanding of data preprocessing , EDA , and model evaluation metrics
- Familiarity with unsupervised learning concepts like clustering, dimensionality reduction (PCA, t-SNE)
- Ability to clearly present technical findings to non-technical audiences
Preferred Qualifications :
- Bachelor’s or Master’s degree in Computer Science, Statistics, Mathematics, Data Science, or a related field
- Experience with version control (Git) and Jupyter notebooks
- Exposure to cloud environments (AWS, Azure, GCP) is a plus
- Knowledge of SQL for data querying
Job Tags