1,092,897open jobs
63,657companies
186,338added this week
Browse all
Salary
≈ $43k – $112k per year (Estimated)
Location
Remote (Argentina)
Seniority
Senior · 4+ years exp

Confirmed on the employer's own hiring board on Oct 2, 2026. First seen by Alion on Oct 1, 2026. Mutt Data scores B on the Alion truth index.

Overview
Company
Impact
Profile match
MUTT DATA is a consulting firm that specializes in leveraging AI and machine learning to create automated systems that drive revenue for their clients. They offer solutions in sectors such as Adtech, Martech, Fintech, and Telecommunication, focusing on optimizing advertising platforms, enhancing data-driven marketing strategies, and providing real-time insights in financial and telecommunications networks. MUTT DATA provides expert team augmentation, seamless integration of data architectures in the cloud, and tailor-made solutions to refine product strategies.

Join Our Remote Data Products & Machine Learning Startup!

At Muttdata, we build innovative Data Products and Machine Learning solutions that help companies solve complex business challenges. As a fast-growing, remote-first startup, we're passionate about technology, collaboration, and continuous learning.

We are looking for an experienced, ownership-driven Senior Data Engineer to join our team . You'll lead the architecture, design, and implementation of a next-generation, in-house clinical trial software platform built directly on Databricks, bridging the gap between software development and large-scale data engineering.

This role works closely with frontend developers, software architects, clinical research teams, and Clinical QA and Validation teams. It requires solid hands-on experience with Databricks and a strong understanding of clinical data standards and regulated environments. Ownership, clear communication, and the ability to build robust, compliant, and scalable solutions are essential to succeed in this fast-paced, collaborative environment.

What We Do

  • Leveraging our expertise, we build modern Machine Learning systems for demand planning and budget forecasting.
  • Developing scalable data infrastructures, we enhance high-level decision-making, tailored to each client.
  • Offering comprehensive Data Engineering and custom AI solutions, we optimize cloud-based systems.
  • Using Generative AI, we help e-commerce platforms and retailers create higher-quality ads, faster.
  • Building deep learning models, we enhance visual recognition and automation for various industries, improving product categorization, quality control, and information retrieval.
  • Developing recommendation models, we personalize user experiences in e-commerce, streaming, and digital platforms, driving engagement and conversions.

Our Partnerships

  • Amazon Web Services
  • Astronomer
  • Databricks

Our Values

  • We are Data Nerds
  • We are Open Team Players
  • We Take Ownership
  • We Have a Positive Mindset
  • Curious about what we’re up to? Check out our case studies and dive into our blog post to learn more about our culture and the exciting projects we’re working on!

Responsibilities

    • Design, build, and optimize enterprise data pipelines, lakehouse storage layers, and data models using Databricks (PySpark, Spark SQL, Delta Lake) to power custom clinical application backends.
    • Collaborate with frontend developers, software architects, and clinical research teams to build API-driven endpoints, data ingestion engines, and query layers for proprietary clinical trial software.
    • Build performant, standards-compliant data structures to store EDC outputs, audit trails, device telemetry, and patient-reported outcomes, enabling rapid querying and downstream analytics.
    • Partner with Clinical QA and Validation teams to ensure database structures, data pipelines, and clinical data repositories comply with GxP, 21 CFR Part 11, HIPAA, and GDPR.
    • Implement real-time and batch ingestion jobs connecting legacy clinical systems, central labs, EHRs, and wearable devices into a unified Databricks Lakehouse architecture.
    • Monitor, troubleshoot, and optimize Spark jobs, Delta Lake tables, and query execution times to support high-throughput, low-latency clinical platform workflows.

Required Skills

    • 4+ years of hands-on experience building production data pipelines and lakehouse architectures using Databricks, Delta Lake, and Apache Spark (PySpark or Scala).
    • Demonstrated experience building, extending, or maintaining custom software applications for clinical trials (e.g., custom EDC, CTMS, Clinical Data Repositories, or eCOA/ePRO platforms).
    • Deep understanding of clinical data standards and regulatory environments, including CDISC (SDTM, ADaM, CDASH), 21 CFR Part 11, GxP validation, and ICH-GCP guidelines.
    • Strong experience with relational schema design, dimensional modeling, and unstructured data handling within Delta Lake environments.
    • Proficiency in Python, SQL, RESTful API integrations, CI/CD pipelines, Git, and automated testing frameworks.
    • Experience working in cloud environments (AWS preferred, Azure or GCP).
    • Advanced English to discuss technical requirements and solutions with clients in the United States

Nice to have

    • Bachelor's or Master's degree in Computer Science, Data Engineering, Bioinformatics, or a related quantitative field.
    • Experience with Databricks Workflows, Delta Live Tables (DLT), and Unity Catalog governance.
    • Background working in a validated system environment (Computer System Validation / CSV).

Perks

  • Remote-first culture - work from anywhere!
  • AWS, DBT, Google Cloud, Azure & Databricks certifications fully covered
  • In-Company English Lessons.
  • Birthday off + an extra vacation week (Mutt Week! )
  • Referral bonuses - help us grow the team & get rewarded!
  • Maslow: Monthly credits to spend in our benefits marketplace.
  • Annual Mutters' Trip - an unforgettable getaway with the team!
  • Monthly Childcare Reimbursement - Because supporting families matters too
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,092,897 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
In your city
≈ $92k – $207k per year (Estimated) • Remote (likely Brazil)
Python
SQL
Scala
Python
pySpark
Databases
Apache Iceberg
Apache Kafka
Trino
AI/ML
Spark
dbt
DevOps
Terraform
CloudFormation
CI/CD
AWS
AWS Lambda
Amazon S3
Amazon Kinesis
Analytics
Metabase
AWS Glue
Apply
≈ $92k – $207k per year (Estimated) • Remote (likely Brazil)
Python
SQL
AI/ML
Scikit-learn
NLP
TensorFlow
PyTorch
LLM
Machine Learning
DevOps
AWS
Apply
≈ $90k – $202k per year (Estimated) • Remote (likely Brazil) • Full-Time • 5+ years exp
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
AI/ML
Spark
DevOps
Terraform
GCP
CI/CD
FinOps
Apply
ENGENHEIRO DADOS SR 1 hour ago
≈ $93k – $208k per year (Estimated) • Remote (likely Brazil) • Full-Time
Python
SQL
Databases
Databricks
Apache Kafka
AI/ML
Spark
Airflow
Dagster
dbt
LLMOps
Machine Learning
DevOps
GCP
Azure
CI/CD
AWS
Apply
$37k – $71k per year • Equity • Remote (Poland) • Full-Time • 3+ years exp • Bachelor's Degree
Python
SQL
Databases
Snowflake
Databricks
AI/ML
Anthropic
DevOps
GCP
Azure
AWS
Platform Engineering
Analytics
ETL/ELT
Management
Agile
Apply
Data Consultant 1 day ago
In office • Bachelor's Degree
Python
SQL
Databases
Databricks
AI/ML
Spark
Machine Learning
DevOps
AWS
Analytics
ETL/ELT
Apply
$116k – $209k per year • Equity • In office • Full-Time • 7+ years exp • Bachelor's Degree • New York • Herndon • Overland Park • Bellevue
Python
SQL
Analytics
Tableau
Power BI
Apply
$78k – $141k per year • Equity • Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Overland Park
Python
SQL
Apply
$135k – $140k per year • In office • Secret • 2+ years exp • Bachelor's Degree • Arlington
Python
SQL
AI/ML
NLP
Machine Learning
DevOps
CI/CD
Management
Agile
Apply
Data Scientist 1 day ago
$140k – $155k per year • In office • Secret • 5+ years exp • Bachelor's Degree • Washington
Python
SQL
AI/ML
NLP
Machine Learning
DevOps
CI/CD
Management
Agile
Apply
Remote (Mexico)
Python
Databases
Databricks
AI/ML
Spark
Model Context Protocol
AI Agents
Feature Store
Machine Learning
DevOps
GCP
Azure
AWS
Analytics
Tableau
Looker
Apply
≈ $43k – $114k per year (Estimated) • Remote (Mexico) • Full-Time
Python
Databases
Databricks
AI/ML
Model Context Protocol
MLFlow
XGBoost
Scikit-learn
AI Agents
Machine Learning
Apply
≈ $38k – $112k per year (Estimated) • Remote (Mexico)
Python
SQL
Databases
Databricks
AI/ML
Model Context Protocol
AI Agents
Feature Store
Machine Learning
DevOps
GCP
Azure
AWS
FinOps
Apply
≈ $56k – $125k per year (Estimated) • Remote (Argentina)
Python
Databases
Databricks
Amazon Redshift
AI/ML
Spark
Machine Learning
DevOps
GCP
Azure
AWS
Analytics
Tableau
Looker
Apply
≈ $37k – $109k per year (Estimated) • Remote (Mexico)
SQL
Databases
Databricks
AI/ML
Machine Learning
DevOps
GCP
Azure
AWS
Analytics
Power BI
Apply
See all jobs
This is one of many
1,092,897 more open roles from verified company boards, updated every day.