Overview
Company
Profile match
Impact
Conditions
Benefits
Hiring process
Similar jobs

Solvd

Over a decade implementing AI in production. Practitioners and NeurIPS/ICML/ECCV researchers partnering with enterprises across AI, data, cloud, and quality engineering.

Solvd Inc. is a rapidly growing AI-native consulting and technology services firm delivering enterprise transformation across cloud, data, software engineering, and artificial intelligence. We work with industry-leading organizations to design, build, and operationalize technology solutions that drive measurable business outcomes.

Following the acquisition of Tooploox, a premier AI and product development company, Solvd now offers true end-to-end delivery -from strategic advisory and solution design to custom AI development and enterprise-scale implementation. Our capability centers combine deep technical expertise, proven delivery methodologies, and sector-specific knowledge to address complex business challenges quickly and effectively.

Solvd is migrating a Tier-1 US payments platform from a single AWS region to a multi-region architecture in three phases (pilot light by EOY 2026, warm standby by EOY 2027, full active-active by EOY 2028). The data layer spans Aurora MySQL, Cassandra, ElastiCache (Redis / Valkey), DynamoDB, DocumentDB, OpenSearch, and Neptune.

We're hiring multiple Database Engineers to implement the per-engine replication, failover, and operational tooling for this program. You'll work under the Principal Database Architect and partner with the client's data and SRE teams.

What you'll do

  • Implement and operate cross-region replication for one or more of: Aurora MySQL Global Database, DynamoDB Global Tables, DocumentDB Global Clusters, Cassandra multi-DC topology, ElastiCache / MemoryDB / Valkey global datastore, Neptune Global Database, OpenSearch cross-cluster replication.

  • Codify everything in Terraform (or CDK).

  • Build the observability layer for replication health: lag metrics, consistency checks, alarm thresholds, and runbook automation. Datadog is the primary observability stack; CloudWatch is the baseline.

  • Author and execute failover and failback runbooks. Run game-days using AWS Fault Injection Service. Measure and report against RTO and RPO targets.

  • Implement KMS cross-account replication, IAM service principals for replication roles, and the secrets and parameter strategy for region-local dependencies.

  • Support data extraction from the monolith into microservices: build CDC pipelines (DMS, DynamoDB Streams, Kafka Connect, Debezium), validate data integrity, and own the cutover sequence for individual extractions.

  • Pair with application engineers on connection-string abstraction, retry and idempotency patterns, and graceful degradation during region failover.

  • Write design docs and runbooks the client team can take over after the engagement.

What you'll bring (required)

  • 5+ years in production database engineering on AWS, with hands-on responsibility for at least one of: payments, banking, brokerage, or comparable financial-services workloads.

  • Deep production experience with at least two of: Aurora MySQL, Cassandra, DynamoDB, ElastiCache (Redis or Valkey), DocumentDB, Neptune. Working knowledge of at least two more.

  • Practical experience with cross-region replication in production: you've configured Global Database, Global Tables, multi-DC Cassandra, or equivalent, and you've handled the failure modes (replication lag spikes, split-brain risk, version-skew issues).

  • Strong Terraform (or CDK). You can build a multi-region module from scratch and explain the trade-offs.

  • Comfort with CDC tooling (DMS, DDB Streams, Kafka Connect, or Debezium) and the operational gotchas (schema drift, ordering guarantees, replay).

  • Linux, networking fundamentals (VPC, Transit Gateway, peering, DNS), and IAM / KMS proficiency.

  • Solid written communication and ability to operate in a regulated client environment.

Nice to have

  • Cassandra multi-DC operations experience at scale (nodetool workflows, repair strategy, tombstone management, cross-region anti-entropy).

  • AWS MSK or Kafka MirrorMaker 2 experience for cross-region event replication.

  • Experience with Datadog Database Monitoring and custom CloudWatch metrics.

  • AWS Database Specialty or Solutions Architect Associate / Professional certifications.

When you join Solvd, you'll…

  • Shape real-world AI-driven projects across key industries, working with clients from startup innovation to enterprise transformation.

  • Be part of a global team with equal opportunities for collaboration across continents and cultures.

  • Thrive in an inclusive environment that prioritizes continuous learning, innovation, and ethical AI standards.

Ready to make an impact?

If you're excited to build things that matter, champion responsible AI, and grow with some of the industry’s sharpest minds. Apply today and let’s innovate together.

Solvd is an equal opportunity employer.

Recommended for you based on this role

Similar stack
Same company
In your city
Full-Time • 3+ year exp • Argentina
PowerShell
Python
SQL
DevOps
Rest API
Splunk
Cybersecurity
GDPR
ISO 27001
Least Privilege
Okta
SOC 2
Zero Trust
Management
Google Workspace
Slack
Marketing
Salesforce
Apply
Remote • Full-Time • 5+ year exp
Bash
JavaScript
Python
Databases
PostgreSQL
DevOps
AWS
Blue-Green Deployment
CI/CD
CircleCI
CloudFormation
Datadog
Docker
GitHub Actions
GitLab CI
Grafana
Jenkins
Kubernetes
New Relic
OpenTelemetry
Platform Engineering
Prometheus
Pulumi
Terraform
Apply
Remote • Full-Time • 7+ year exp
C#
JavaScript
Node JS
Python
TypeScript
Databases
PostgreSQL
Frontend
React.js
DevOps
AWS
CI/CD
Git
Rest API
Apply
Full-Stack Engineer 13 days ago
Remote • Full-Time • 3+ year exp
C#
JavaScript
Node JS
Python
SQL
TypeScript
Databases
PostgreSQL
Frontend
React.js
DevOps
AWS
Git
Rest API
Apply
Solution Architect 15 days ago
Remote • Full-Time
Java
TypeScript
Databases
Apache Kafka
AI/ML
RAG
DevOps
AWS
Azure
Apply
Remote • Full-Time • 8+ year exp
Databases
Databricks
Snowflake
AI/ML
AWS Bedrock
Gemini
LLM
Prompt Engineering
DevOps
AWS
Azure
CI/CD
GCP
Vector
Apply
Cloud Engineer 22 days ago
Remote • Full-Time • Bachelor's Degree
Ruby
TypeScript
Databases
Amazon Aurora
DevOps
Amazon EC2
AWS
AWS CDK
AWS Lambda
Chaos Engineering
CI/CD
CloudFormation
GitHub Actions
Terraform
Cybersecurity
PCI DSS
Apply
Remote • Full-Time • 3+ year exp
PowerShell
Python
SQL
DevOps
Rest API
Splunk
Cybersecurity
GDPR
ISO 27001
Least Privilege
Okta
SOC 2
Zero Trust
Management
Google Workspace
Slack
Marketing
Salesforce
Apply
Senior Data Scientist 3 months ago
Remote • Full-Time
Python
AI/ML
Computer Vision
PyTorch
Scikit-learn
TensorFlow
Time Series Forecasting
DevOps
AWS
Azure
GCP
Apply
In office • Full-Time • 1+ year exp
Node JS
Python
TypeScript
JavaScript
AI/ML
AI Agents
Fine-tuning
Knowledge Distillation
LangChain
LlamaIndex
LLM
Model Context Protocol
RAG
Frontend
React.js
DevOps
AWS
Azure
CI/CD
GCP
Rest API
Apply
Career impact
Discover how this job can transform your career
Get a personal career forecast for this job - salary uplift, next-level role, skill boost and a 3-year financial impact, all calculated from your profile.
Personal salary uplift vs. your current pay
Your 3-year career trajectory
Skills you will level up in this role
3-year financial impact in dollars
Create free account
Free forever • Less than a minute • No credit card

Work setup

Remote work
Remote
Employment
Full-Time

Compensation

Benefits
Continuous learning