JobsBackendForward Deployed Engineer (Data, ML & AI)
C

Carbon60

·

Backend

Forward Deployed Engineer (Data, ML & AI)

RemotePrincipalPosted Aug 17, 2026
PythonSQLSparkLangChainMLflowKubeflowAirflowTerraform+64 more

About this role

A principal-level, customer-facing role embedding directly within client environments to build production-grade data pipelines, ML systems, and agentic AI solutions. You'll modernize legacy data estates using spec-driven development and lead technical delivery from discovery to deployment.

Must have

10+ years experience as Principal Data Engineer, Lead ML Engineer, or Enterprise Data Architect

Proven track record building and scaling distributed data and ML platforms

Strong executive presence and communication skills for direct client interaction

Mastery of distributed computing with AWS Glue, Apache Spark, and Databricks

Hands-on experience fine-tuning, evaluating, and deploying LLMs and embedding models

Advanced proficiency in Python and complex SQL

Fluency in at least two additional languages from Scala, Go, Rust, or TypeScript

Demonstrated skill using AI tools like Claude Code, Cursor, and Copilot for spec-driven development

Hands-on experience with cloud-native data services on AWS or Azure

Experience with Docker, Kubernetes, and Terraform for infrastructure as code

Willingness to travel occasionally to customer sites

Technologies

PythonSQLScalaAWSAzureSparkTypeScriptHadoopGitDockerRESTLinuxJiraJenkinsPostgreSQLTerraformKafkaMongoDBNoSQLMicroservicesMavenConfluenceDevOpsJUnitMachine learningCloudGitLabJSONSecurityVirtualizationGradleGrafanaAirflowPrometheusData engineeringGitHubTensorFlowBig dataSnowflakepandasCybersecurityData analyticsscikit-learnPySparkPyTorchOpen sourceAIKerasElasticsearchpytestNumPyData scienceServerless computingDeep learningPaaSIaCRustNLPDatabricksInfrastructure as a CodeKubernetesMLflowKubeflowGoCI/CDXGBoostLLMsLangChainMLOpsPrompt engineeringRedshiftContainerization

Responsibilities

Embed directly within customer environments to understand domain logic and data pipelines

Own end-to-end delivery from data discovery through pipeline deployment and model integration

Evaluate legacy technology estates and refactor into modern lakehouses and streaming systems

Apply spec-driven development using Generative AI tools to create production data models

Architect and deploy GenAI workflows, RAG pipelines, and autonomous AI agents

Establish evaluation frameworks for model accuracy, latency, and hallucination control

Build and maintain ML training and inference pipelines using MLflow, Kubeflow, or Airflow

Ensure CI/CD for models and data workflows

Design and audit production code across Python, SQL, Scala, Go, Rust, and TypeScript

Elevate customer teams through knowledge transfer of MLOps practices and SDD methodologies

Benefits

RRSP and 401k matchPaid parental leaveHealth insuranceDental insuranceVision insuranceWellness allowanceFlexible hoursCompetitive salaryRemote workProfessional development

Amenities

Remote first work environmentFlexible work hours and locationAccess to latest technologyGreenShield+ Counselling Mental Health$500 annual Health Care Spending Account

Recruitment process

1

High-ownership, hands-on role with immediate value demonstration at customer sites

2

Spec-Driven Development (SDD) and AI-assisted workflows are core methodologies

3

Partnership with Perkopolis Discounts available

4

Peer recognition rewards program

5

Equal-opportunity employer with accommodations available on request

6

Only qualified candidates will be contacted for interviews