Find your dream job faster with JobLogr
AI-powered job search, resume help, and more.
Try for Free
Motion Recruitment

Motion Recruitment

via LinkedIn

All our jobs are verified from trusted employers and sources. We connect to legitimate platforms only.

Data Engineer- Python, AI/ML

Anywhere
Contract
Posted 5/26/2026
Verified Source
Key Skills:
Python
SQL
Apache Spark
AWS

Compensation

Salary Range

$55K - 120K a year

Responsibilities

Build and maintain Python and SQL data pipelines with governance, quality checks, and AI/ML-assisted use cases.

Requirements

Requires 2+ years in Python and SQL data pipelines, experience with Spark, cloud platforms, data quality validation, and BI tools.

Full Description

Key Responsibilities • Build and maintain Python and SQL pipelines for governance-related ingestion, cleaning, transformation, and validation of structured and semi-structured data. • Implement and operate data quality checks, schema validation, and integrity rules across pipelines; investigate and resolve quality issues. • Contribute to master data workflows: standardization, deduplication, and consolidation of data from heterogeneous sources into consistent reference and golden-record datasets. • Instrument pipelines for data lineage, metadata, and catalog tooling. • Develop pipelines that feed governance dashboards and reporting in Tableau, Power BI, or Looker. • Build reproducible, well-documented pipelines for compliance and audit reporting. • Contribute to AI / ML-assisted governance use cases: embedding-based data classification, anomaly detection on quality metrics, LLM-assisted catalog search, and MCP-based exposure of governed datasets to AI assistants. • Partner with team leads, data stewards, and stakeholders to translate governance requirements into engineering work. • Follow team engineering practices: Git, code review, modular pipeline design, automated testing, CI/CD. Required Qualifications • Bachelor's or Master's degree in Computer Science, Data Science, Engineering, Statistics, or a related field. • 2+ years building data pipelines in Python (Pandas, NumPy, SciPy) and SQL. • Working experience with Apache Spark or PySpark and workflow orchestration (Apache Airflow). • Schema design across relational (PostgreSQL, MySQL, SQL Server) and analytical databases, including standardization across heterogeneous sources. • Experience implementing data quality validation, EDA, and integrity enforcement on production datasets. • Hands-on experience with at least one major cloud platform (AWS, Azure, or GCP). • Working familiarity with Python ML libraries (Scikit-Learn) for feature engineering and exploratory analysis. • Experience producing analytics-ready datasets for BI tools (Tableau, Power BI, or Looker). • Git, code review, and CI/CD practices. • Clear technical communication and collaborative working style.

This job posting was last updated on 6/1/2026

Ready to have AI work for you in your job search?

Sign-up for free and start using JobLogr today!

Get Started »
JobLogr badgeTinyLaunch BadgeJobLogr - AI Job Search Tools to Land Your Next Job Faster than Ever | Product Hunt