Search by job, company or skills

Senior ML Engineer (Remote, Full-Time)

Senior ML Engineer (Remote, Full-Time)

Smart Working
Fresher
Not Disclosed
  • Posted 5 hours ago
  • Be among the first 10 applicants

Job Description

About Smart Working


At Smart Working, we believe your job should not only look right on paper but also feel right every day. This isn't just another remote opportunity — it's about finding where you truly belong, no matter where you are. From day one, you're welcomed into a genuine community that values your growth and well-being.


Our mission is simple: to break down geographic barriers and connect skilled professionals with outstanding global teams and products for full-time, long-term roles. We help you discover meaningful work with teams that invest in your success, where you're empowered to grow personally and professionally.


Join one of the highest-rated workplaces on Glassdoor and experience what it means to thrive in a truly remote-first world.


About the Role


As a Senior ML Engineer, you will bridge the gap between applied machine learning and robust platform engineering. You will focus on translating AI capabilities into reliable and observable system components.


You will help ensure that machine learning models meet strict quality, cost and latency thresholds before reaching production. The role spans production ML architecture, workflow reliability, model evaluation, provenance, testing, human feedback integration and AI provider observability.


Responsibilities


  • Refactor and upgrade existing ML models, including NLP, generative AI and transcription models, into standardised, production-ready modular contracts.




  • Engineer resilient ML workflows using DAG-based orchestration tools such as Argo Workflows and Airflow.




  • Ensure robust retry logic, error handling and repeatable execution across ML workflows.




  • Define, automate and maintain strict ML evaluation pipelines using golden datasets.




  • Ensure new models meet baseline quality and performance thresholds before release.




  • Design and implement tracking mechanisms to capture precise model, prompt and input data provenance for full auditability and reproducibility.




  • Build infrastructure for shadow testing and A/B testing on live traffic.




  • Implement reliable fallback and kill-switch mechanisms to support safe deployment.




  • Engineer structured feedback pipelines that capture human reviews and corrections to continuously enrich and refine training datasets.




  • Integrate third-party AI APIs and manage adapter interfaces.




  • Implement granular telemetry to track compute costs, token usage and latency.





Requirements


  • Proven track record operating at the intersection of Machine Learning, MLOps and Platform/Backend Engineering.




  • Deep understanding of evaluation metrics for generative AI, LLMs and/or speech models, including Precision, Recall, F1, WER/CER, groundedness and hallucination rates.




  • Strong hands-on experience with containerisation using Docker and Kubernetes, alongside modern orchestration frameworks.




  • Experience building observable ML systems with logging, monitoring and strict cost-attribution tracking.




  • Familiarity with data provenance, compliance standards and building fail-safe mechanisms for sensitive data and AI outputs.






We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

More Info

Key Skills

Transcription models

Generative AI

DAG-based orchestration tools

Argo Workflows

Fail-safe mechanisms

AI provider observability

Data provenance

Platform Backend Engineering

ML evaluation pipelines

Shadow testing

Cost-attribution tracking

About Company

Similar Jobs

Hyderabad, India
Skills:
Sql, Tensorflow, Azure ML, Nlp, MLops, Azure Functions, Pytorch, Elasticsearch, Rest Apis, Azure, Ocr, Python, Computer Vision, AWS, LiteDB, Document Intelligence, ML deployment practices, annotation tooling, ADLS, ML engineering
Hyderabad, India
Skills:
snowflake , Python, Databricks, Java, Jira, Cucumber, Neural Networks, Azure, GraphRAG, Playwright, Agentic workflows, Text-to-SQL builders and executors, Multi-agent systems, Advanced RAG, LangChain, Knowledge Graphs, Generative AI, LangGraph, Chatbots, GitHub Copilot
Hyderabad, India
Skills:
snowflake , Java, Cucumber, Neural Networks, Azure Databricks, Python, Text-to-SQL, GraphRAG, Playwright, LangChain, Knowledge Graphs, Generative AI, GitHub Copilot, LangGraph
2-8 yrs
Hyderabad
Skills:
cloud platform , Machine Learning, AI ML, Python, Deep Learning