Data Scientist – Time Series, Statistical Modelling
Capco Technologies Pvt LtdJob Description
Data Scientist – Time Series, Statistical Modelling
Job Title: Data Scientist
About Us
“Capco, a Wipro company, is a global technology and management consulting firm. Awarded with Consultancy of the year in the British Bank Award and has been ranked Top 100 Best Companies for Women in India 2022 by Avtar & Seramount. With our presence across 32 cities across globe, we support 100+ clients across banking, financial and Energy sectors. We are recognized for our deep transformation execution and delivery.
WHY JOIN CAPCO?
You will work on engaging projects with the largest international and local banks, insurance companies, payment service providers and other key players in the industry. The projects that will transform the financial services industry.
MAKE AN IMPACT
Innovative thinking, delivery excellence and thought leadership to help our clients transform their business. Together with our clients and industry partners, we deliver disruptive work that is changing energy and financial services.
#BEYOURSELFATWORK
Capco has a tolerant, open culture that values diversity, inclusivity, and creativity.
CAREER ADVANCEMENT
With no forced hierarchy at Capco, everyone has the opportunity to grow as we grow, taking their career into their own hands.
DIVERSITY & INCLUSION
We believe that diversity of people and perspective gives us a competitive advantage.
Key Responsibilities
- Design, develop, and deploy end-to-end machine learning solutions using Python and Databricks.
- Build and optimize scalable data processing pipelines using PySpark and Spark.
- Perform data extraction, cleansing, transformation, and feature engineering on structured and unstructured datasets.
- Develop, evaluate, and implement traditional machine learning models for classification, regression, clustering, and prediction problems.
- Work extensively with Databricks notebooks, Jupyter Notebooks, and collaborative development environments.
- Develop reusable, production-ready ML workflows and analytical frameworks.
- Collaborate with data engineers and business stakeholders to understand business requirements and translate them into scalable analytical solutions.
- Optimize model performance through experimentation, hyperparameter tuning, and validation.
- Ensure data quality, governance, and best practices throughout the data and model lifecycle.
- Document methodologies, model performance, and technical solutions for knowledge sharing and maintainability.
Required Skills
Programming
- Python
- PySpark
- SQL
Data Science & Machine Learning
- Strong understanding of traditional Machine Learning algorithms
- Supervised and Unsupervised Learning
- Feature Engineering
- Model Evaluation and Validation
- Statistical Analysis
- Predictive Analytics
- Data Preprocessing and Transformation
Big Data & Data Processing
- Apache Spark
- PySpark
- Large-scale data processing
- Data pipeline development and optimization
- ETL/ELT concepts
Databricks
- Strong hands-on experience with Databricks
- Experience working with Databricks Notebooks
- Building scalable ML workflows in Databricks
- Delta Lake (preferred)
- Databricks Jobs and Workflows (preferred)
Python Libraries
- Pandas
- NumPy
- Scikit-learn
- SciPy
- Matplotlib / Seaborn
- MLflow (preferred)
Cloud & Platform
- Azure Databricks
- Azure Data Services (preferred)
- Azure Machine Learning (good to have)
Development Tools
- Jupyter Notebook
- Git
- CI/CD concepts for ML pipelines (preferred)
Good to Have
- Experience implementing MLOps pipelines using MLflow or Azure ML.
- Time-Series Forecasting using Statsmodels or Prophet.
- Explainable AI frameworks such as SHAP or LIME.
- Experience with orchestration tools such as Azure Data Factory or Apache Airflow.
- Experience with Delta Lake and Unity Catalog.
- Knowledge of Docker and containerized deployments.
- Experience working in Agile environments.
Preferred Experience
- 3+ years of experience in Data Science and Machine Learning.
- Strong experience implementing ML solutions on Databricks.
- Experience building production-grade data and ML pipelines using PySpark.
- Experience with Azure cloud ecosystem is preferred.
- Ability to communicate technical concepts and analytical insights to business stakeholders.
- Experience in Energy, Utilities, Manufacturing, or other data-intensive industries is an advantage
If you are keen to join us, you will be part of an organization that values your contributions, recognizes your potential, and provides ample opportunities for growth. For more information, visit www.capco.com. Follow us on Twitter, Facebook, LinkedIn, and YouTube.
Job role
Job requirements
About company
Similar jobs you can apply for
Software / Web DeveloperPHP Developer
Shyam Web DevelopersEmbedded Linux Engineer
Carrington Associates Technologies Private Limited
Quality Control Engineer
Efficient Precision and Systems Private LimitedQuality Control Inspector
Shri Krishna Industries
IT Sales Executive
Techtrix Solutions Pvt Ltd
Software Tester
Altroz TechnologiesYou can expect a minimum salary of 0 INR. The salary offered will depend on your skills, experience and performance in the interview.
The candidate should have completed the required education and people who have 5 to 31 years are eligible to apply for this job. You can apply for more jobs in Pune to get hired quickly.
The candidate should have sound communication skills and sound communication skills for this job.
Both Male and Female candidates can apply for this job.
No, it's not a work from home job and can't be done online. You can explore and apply for other work from home jobs in Pune at apna.
No work-related deposit needs to be made during your employment with the company.
Go to the apna app and apply for this job. Click on the apply button and call HR directly to schedule your interview.
The last date to apply for this job is . For more details, download apna app and find Full Time jobs in Pune . Through apna, you can find jobs in 64 cities across India. Join NOW!