ACSI India is looking for a senior data scientist who will leverage advanced analytical methods to solve critical business challenges, driving growth and efficiency. You will collaborate with stakeholders across CAS to develop innovative data-driven solutions, using consultative problem-solving to address complex challenges such as knowledge extraction, classification, search algorithms, and prediction models.
About ACS-I India
ACS International India Pvt Ltd. (ACSI India) is a wholly owned subsidiary of ACS International Ltd, USA and a part of the American Chemical Society. ACSI India represent products and services provided by ACS divisions, including (CAS) to the world's most important scientific companies, government organizations, global patent offices and academic institutions to promote research and discovery.
About CAS
CAS uses unparalleled scientific content, specialized technology and unmatched human expertise to help R&D organizations across Commercial, Government and Academic sectors create groundbreaking innovations that benefit the world. As the Scientific Information Solutions Division of the American Chemical Society, CAS manages the largest curated reservoir of scientific knowledge, and for 111 years, has helped innovators mine, assess and apply that information to keep businesses thriving. The CAS team is global, diverse, endlessly curious and strives to make actionable scientific insights accessible to innovators worldwide.
Job Responsibilities
- Efficiently communicate with other scientists on the project, actively and creatively develop solutions to support the overall project goals.
- Combine strong software development skills with a working knowledge of basic chemistry/physics/biology to develop sophisticated informatics solutions that drive efficiencies in data-based insights development.
- Build predictive models using machine learning algorithms and frameworks, such as TensorFlow, PyTorch, Scikit-learn etc.
- Apply NLP, machine learning, and deep learning in various domains.
- Present information and insights using data visualization techniques, such as matplotlib, plotly etc.
- Capable of self-directed research within broader goals set by group.
- Manage multiple projects at any given time along with tracking project milestones.
- Lead projects end-to-end, taking ownership of planning, execution and delivery while guiding and mentoring team members.
- Should be able to teach and train his/her team in all the above-mentioned aspects as and when required.
Ideal Candidate Will Have
- Experience with Agentic AI, Generative AI, LLMs, and prompt engineering.
- Experience with big data technology stack (Hadoop, Spark, HDFS, EMR, Glue).
- Experience with AWS, Azure or GCP
- Experience with Databricks/SageMaker/DataRobot, MLFlow or other ML and MLOps tools.
- Experience building applications using AWS Serverless technologies such as Lambda, SQS, Fargate, DynamoDB, S3.
- Experience with Neo4j and graph analytics.
- Demonstrated experience leading data science projects and cross-functional, cross-geographical teams spanning multiple regions and time zones.
Job Requirements
- Btech/Mtech/PhD in CS/ECE/AI/ML, Applied Statistics or a related field.
- 6+ years of post-degree experience working with large data sets/software development
- Experience building applications for public cloud environments (AWS preferred).
- Proficiency in programming languages such as Java/Scala/JavaScript/TypeScript/Python.
- Proficiency in Linux/Unix environments.
- Experience with databases technologies (relational, NoSQL, property graph, RDF/triple store).
- Self-motivated, proactive and excellent in communication skills