Worldwide
1 month ago

Job Overview

Job Type
Full Time
Pay
Not disclosed

Job description

Location:
Worldwide
Work arrangement:
On-site

Role Summary

We are looking for an MLOps Engineer with a strong interest in automation, cloud computing and efficient deployment of intelligent platforms.

If you are motivated to improve productive environments, guarantee scalable and reliable systems, and work with cutting-edge technologies in Machine Learning, AI Agents and Observability, this opportunity is for you.

🚀 What will your role be?

As an MLOps Engineer, you will be responsible for connecting the world of Data and ML with the infrastructure and operation in production, ensuring that models and agents work robustly and efficiently.

Responsibilities

Your main responsibilities will include

  • Design, implement and operate scalable cloud infrastructures in AWS, GCP or Azure, aimed at data and ML workloads.
  • Automate infrastructures and deployments using Terraform, Ansible, Helm and Kubernetes.
  • Create and maintain CI/CD pipelines for the continuous deployment of models, AI agents and observability platforms.
  • Manage containers and orchestration with Docker and Kubernetes (EKS), integrating data, AI and backend services.
  • Build and maintain reproducible training, validation and inference pipelines (Airflow, MLflow, DVC, Spark).
  • Implement and optimize monitoring, logging and observability solutions (Prometheus, Grafana, ELK/EFK, OpenTelemetry).
  • Ensure the security, availability and resilience of cloud platforms.
  • Integrate the complete Data → ML → Deployment → Monitoring cycle, working closely with Data, ML, DevOps and Backend teams.
  • Collaborate with Machine Learning teams to package and bring AI models and agents to production.
  • (Advanced Plus) Participate in the design of hybrid pipelines that combine traditional ML, generative AI and LLM-based agents (LangChain, LangGraph, CrewAI).

Requirements

We are looking for someone with

  • Training in Computer Engineering, Software, Telecommunications or related disciplines.
  • More than 3 years of experience in automation, deployment or management of cloud infrastructure.
  • Strong experience in infrastructure as code and automation (Terraform, Helm, Ansible).
  • Advanced knowledge of Linux, networks and distributed systems.
  • Hands-on experience in CI/CD pipelines (GitLab CI, Jenkins, ArgoCD, FluxCD).
  • Mastering Docker and Kubernetes.
  • Experience with model and data management tools such as MLflow, DVC or Vertex AI.
  • Knowledge of monitoring, logging and observability.
  • Experience with SQL and NoSQL databases (PostgreSQL, TimescaleDB, MongoDB).
  • Ability to diagnose and resolve incidents in high-traffic productive environments, optimizing costs and performance.
  • Good communication and teamwork skills in multidisciplinary environments.

🌟 We especially value if you also have

  • Experience with messaging and streaming systems such as Kafka, Redpanda or Benthos.
  • Knowledge of serverless and event-driven architectures (AWS Lambda, SNS/SQS).
  • Experience in advanced observability and model performance metrics.
  • Knowledge of SRE (Site Reliability Engineering) practices.
  • Experience with modern MLOps platforms (Kubeflow, MLflow).
  • Familiarity with automated deployments and good DevOps practices.
  • Knowledge of infrastructure for generative AI (GPU, optimized containers, RAG serving).
  • Advanced technical English, written and spoken, to collaborate with international teams and partners.

About the Company

About Our Client

We are a leading provider of nearshore staff augmentation services headquartered in New York. For over two decades, we’ve been delivering top-tier technology solutions to companies of all sizes, from innovative startups to industry leaders, helping them achieve their digital transformation goals. Our team of 600+ highly skilled tech professionals, based in Latin America, drives digital disruption by partnering with U.S. companies on their most impactful projects. Whether collaborating with Fortune 500 giants or scaling startups, we deliver results that make a difference.

By applying for this position, you’re taking the first step in joining a dynamic team that values your expertise and aspirations. We aim to align your skills with opportunities that foster exceptional career growth and success while contributing to transformative projects that shape the future.

Drive both external brand visibility and internal engagement, making a meaningful impact on the organization.

Work closely with leadership and passionate teams across the company.

Enjoy opportunities for professional growth and creative expression.

Role:
1225 - 405PIC | MLOps Engineer
Job Type:
Full Time

More jobs at TalentCross

Similar Machine Learning Engineer jobs at other companies