- Design scalable data pipelines using PySpark
- Perform data transformation and exploratory analysis using Pandas, Numpy and SQL
- Build, train and fine tune machine learning and deep learning models using TensorFlow and PyTorch
- Implementing big data, streaming AI/ML and prediction pipelines.
- Translate complex business problems into data driven solutions.
- Promote best practices in data science, and model governance.
- Use tools like Python, TensorFlow, PyTorch, SQL, and cloud platforms
- Stay ahead with evolving technologies and guide strategic data initiatives.
- Build, deploy and operationalize models on AWS SageMaker — including training jobs, hyperparameter tuning, model registry, real-time and batch inference endpoints, and monitoring for drift and model performance in production
- Develop and orchestrate ETL/ELT workflows across AWS services such as Glue, S3, Lambda, Step Functions, Redshift and Athena to deliver clean, reliable, analysis- and feature-ready datasets for downstream modeling.
JobID_Data Scientist826 in california at Unknown Company
This position is listed as full time and onsite.