415,615open jobs
14,704companies
74,833added this week
Browse all
Salary
$16k – $44k per year (Estimated)
Location
Remote/Hybrid (Hyderabad, India)
Seniority
Junior
Employment
Full-Time
Overview
Company
Impact
Profile match
Dun & Bradstreet is a global provider of business decisioning data and analytics headquartered in Jacksonville, Florida, and founded in 1841. The company offers a wide range of solutions including business credit reports, risk management tools, supply chain visibility, and sales and marketing data anchored by its proprietary D-U-N-S Numbering system. It operates globally, maintaining a commercial database of hundreds of millions of business records to serve clients in the public sector, finance, and enterprise markets.

The Data Engineer I work as part of an agile team supporting the organization's Research and Managed Services. This is a hands-on, hybrid, technical, and operational role. The team members will actively write and maintain code to build data tools, automations, and pipelines, while also performing day-to-day operational tasks that keep Research and Managed Services running reliably. The role includes identifying and evaluating internal and external data sources to feed AI-enabled and traditional research workflows. The team member will assess source relevance, quality, coverage, accessibility, reliability, compliance considerations, and applicability to defined business and technical use cases. The role drives best-in-class standards and continuous improvement and requires an adaptable professional who is willing and able to learn and adopt new technologies as they are introduced to the organization

Key Responsibilities:

    Coding & Development

  • Write, review, test, and maintain code, including SQL and Python, to build data tools, automations, and ingestion and transformation workflows that support Research and Managed Services
  • Automate manual processes and develop data tools to improve efficiency, accuracy, quality, and throughput
  • Develop and promote coding standards and contribute to code reviews within the agile team
  • Build and maintain web-scraping solutions, API integrations, and reusable data-processing components
  • Support scalable ETL/ELT pipelines for structured and unstructured data
  • Source Evaluation & AI Enablement

  • Identify prospective sources that can feed AI solutions and Research and Managed Services workflows
  • Define and apply source-evaluation criteria covering relevance, authority, freshness, completeness, coverage, consistency, accessibility, legal or licensing constraints, privacy, security, and technical compatibility
  • Perform source profiling, sample validation, proof-of-concept testing, and comparative assessments before recommending onboarding
  • Document source decisions, metadata, lineage, ownership, limitations, refresh expectations, and approved use cases
  • Implement and support AI-enabled workflows using LangChain or equivalent orchestration frameworks, large language models, embeddings, retrieval-augmented generation, vector databases, and prompt-engineering approaches where applicable
  • Monitor source and AI-workflow performance and recommend remediation, replacement, or additional sources when quality or coverage falls below requirements
  • Operational Tasks

  • Perform day-to-day operational activities supporting Research and Managed Services, including monitoring, exception handling, data maintenance, and issue resolution
  • Perform database administration activities, including performance tuning and implementation of best practices
  • Implement new data-maintenance processes and provide end-to-end process ownership
  • Ensure data integrity by validating, reconciling, and regularly cleaning data
  • Investigate and resolve production incidents, pipeline failures, data-quality issues, and operational exceptions
  • Follow applicable data governance, security, and operational standards
  • Collaboration & Continuous Learning

  • Evaluate and implement new technology solutions, and proactively learn and adopt new tools, platforms, and methodologies introduced by the organization
  • Communicate with stakeholders and conduct knowledge-exchange sessions for technical and non-technical audiences
  • Develop and maintain data documentation, including data dictionaries, source assessments, data-flow diagrams, data mappings, runbooks, and data lineage
  • Collaborate with cross-functional teams across Data & Analytics, Technology, Research Services, Managed Services, Product, and Data Governance
  • Additional duties as assigned.

Key Skills:

  • Strong SQL and Python skills, with demonstrated ability to write and maintain code as a core part of daily work
  • Experience with Playwright, Selenium, and other web-data collection techniques
  • Experience developing and supporting data-ingestion, transformation, and ETL/ELT workflows
  • Ability to collect and interpret data from multiple sources, including web scraping and GCS/S3, and formats including delimited files, XML, JSON, and PDF
  • Working knowledge of data systems and databases used to maintain data pipelines
  • Experience with Power BI, Tableau, or other dashboard tools
  • Experience managing stakeholders and project plans
  • Proficiency in Microsoft Office Suite
  • Willingness and demonstrated ability to learn new technologies as they are introduced
  • BigQuery experience and knowledge of AWS and/or GCP
  • Hands-on experience implementing AI solutions using LangChain or an equivalent orchestration framework
  • Exposure large language models, prompt engineering, retrieval-augmented generation, embeddings, vector databases, AI agents, or graph databases
  • Knowledge of Data Operations methodologies, data management approaches, ServiceNow, and/or Jira
  • Experience with NoSQL technologies, SQL Server administration, R programming, web technologies, and data mapping from multiple sources.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
415,615 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Hyderabad
$90k – $185k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Belfast
Java
Python
SQL
Databases
Apache Kafka
DynamoDB
Kafka
Snowflake
AI/ML
AI Agents
Claude
dbt
Function Calling
Multi-Agent Systems
OpenAI
Spark
Tool Use
DevOps
Amazon EKS
Amazon EventBridge
Amazon Kinesis
Amazon S3
AWS
AWS Lambda
AWS Step Functions
CI/CD
GitLab
Grafana
IAM
Terraform
Kubernetes
Analytics
ETL/ELT
QlikSense
Apply
$114k – $150k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree
Python
SQL
Databases
Amazon Redshift
MS SQL
Oracle
DevOps
Amazon S3
AWS
Azure
Azure DevOps
CI/CD
Git
Cybersecurity
HIPAA
Analytics
ETL/ELT
Power BI
Apply
$28k – $102k per year (Estimated) • In office • Full-Time • Madrid
AI/ML
AI Agents
Claude
Copilot
Copilot Studio
Management
Power Automate
SharePoint
Apply
$116k – $219k per year (Estimated) • Remote • Full-Time • 3+ years exp • Bachelor's Degree
Python
DevOps
AWS
Azure
GCP
Cybersecurity
SOC 2
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Ho Chi Minh City
Go
Go
Chi
Databases
MySQL
PostgreSQL
AI/ML
Claude
Claude Code
Copilot
Cursor
Mobile
Clean Architecture
DevOps
AWS
CI/CD
CloudFormation
GCP
GitHub
Jenkins
Terraform
Apply
$32k – $82k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Hyderabad
Python
AI/ML
AI Agents
Embeddings
Function Calling
LangChain
LLM Guardrails
LLMOps
Multimodal AI
Prompt Engineering
RAG
Tool Use
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
Vector
Apply
$59k – $150k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Stockholm
Analytics
Power BI
Marketing
Salesforce
Apply
$93k – $185k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Jacksonville
Apply
$73k – $187k per year (Estimated) • Remote/Hybrid • Full-Time • Jacksonville
Apply
$149k – $282k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Jacksonville
JavaScript
AI/ML
AI Agents
DevOps
CI/CD
Management
ServiceNow
Apply
$17k – $39k per year (Estimated) • Remote/Hybrid • 6+ years exp • Hyderabad
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
Angular
Angular Material
Mobile
Material Design
DevOps
CI/CD
Git
GitHub
GitHub Actions
GitLab
GitLab CI
Jenkins
QA
Playwright
Apply
$35k – $80k per year (Estimated) • Remote/Hybrid • Hyderabad
AI/ML
AI Agents
Apply
$14k – $204k per year (Estimated) • Remote/Hybrid • Hyderabad
AI/ML
AI Agents
Marketing
Salesforce
Apply
$28k – $68k per year (Estimated) • Remote • 8+ years exp • Bachelor's Degree • Hyderabad
JavaScript
Python
SQL
TypeScript
AI/ML
AI Agents
LLM
Frontend
GraphQL
DevOps
Amazon EKS
Amazon S3
API Gateway
AWS
AWS Lambda
CI/CD
Docker
GitHub
Rest API
WebSockets
Kubernetes
Cybersecurity
Checkmarx
FedRAMP
ISO 27001
NIST 800-53
OWASP Top 10
Snyk
SOC 2
Threat Modeling
Veracode
Apply
Technical Writer 1 day ago
$20k – $49k per year (Estimated) • Remote/Hybrid • 3+ years exp • Bachelor's Degree • Hyderabad
AI/ML
AI Agents
Apply
See all jobs
This is one of many
415,615 more open roles from verified company boards, updated every day.