430,878open jobs
14,717companies
59,947added this week
Browse all
Salary
$100k – $160k per year
Location
Remote (United States)
Seniority
Principal · 12+ years exp
Overview
Company
Impact
Profile match
Bright Vision Technologies is an IT consulting, enterprise technology services, and workforce solutions enterprise. Headquartered in Bridgewater, New Jersey, United States, the minority-owned firm specializes in technology staffing, cybersecurity, application management, and digital product engineering. Founded in 2020, the enterprise delivers specialized staffing and IT services alongside proprietary automation and AI software - including its flagship enterprise talent intelligence platform, Lumina - serving clients across information technology, defense, healthcare, government, and manufacturing sectors.

Observability Engineer - Remote

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.

This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title: Observability Engineer

Location: 100% Remote (U.S.)

Position Type: Full-time, Direct W2

Salary Range: $100,000-$160,000 Annually

Experience Required: 12+ years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary:

We are looking for an Observability Engineer to design and operate the metrics, logging, tracing, and alerting platforms that give engineering teams confidence in the systems they run. The role spans the full observability stack - from collection agents and pipelines to long-term storage, dashboards, and alerting workflows - with a strong focus on usability, signal quality, and operational ROI. The ideal candidate has built and operated observability platforms at scale, understands the trade-offs between open-source and SaaS approaches, and can translate noisy telemetry into actionable insight for both engineers and business stakeholders.

Key Responsibilities

  • Design and operate enterprise-grade observability platforms covering metrics, logs, traces, events, and synthetic monitoring.
  • Architect Prometheus / Thanos / Mimir, Grafana, Loki, Tempo, OpenTelemetry, and Datadog deployments for high availability and scale.
  • Develop standards for service instrumentation, including OpenTelemetry adoption, metric naming, label cardinality, and structured logging conventions.
  • Define and enforce SLOs, SLIs, and error budgets, and build the dashboards and alerts that operationalize them.
  • Build alerting strategies that minimize noise, surface actionable signals, and integrate cleanly with on-call workflows in PagerDuty, Opsgenie, or similar tools.
  • Operate large-scale time-series and log storage platforms, balancing retention, query performance, and cost.
  • Design distributed tracing pipelines and help teams use traces to diagnose latency and reliability issues.
  • Develop self-service tooling, paved-road libraries, and templates that make adoption of observability standards easy for product teams.
  • Drive cost management and label-cardinality discipline across the observability estate.
  • Lead incident response readiness improvements through better dashboards, alerting hygiene, and post-incident analysis tooling.
  • Partner with SRE and platform teams to integrate observability into deployment pipelines, canary analysis, and progressive delivery workflows.
  • Evaluate and recommend observability vendors and open-source tools based on cost, capability, and operational maturity.
  • Mentor engineering teams on observability fundamentals, debugging techniques, and SLO-driven operations.
  • Maintain documentation, onboarding guides, and runbooks for the observability platform.
Required Qualifications
  • Bachelor’s degree in Computer Science or a related field.
  • Ten or more years of experience in SRE, platform engineering, or observability roles.
  • Deep hands-on experience with Prometheus, Grafana, and at least one major commercial observability platform such as Datadog, New Relic, or Splunk.
  • Strong understanding of OpenTelemetry, distributed tracing, and structured logging.
  • Proficiency in at least one general-purpose language such as Go, Python, or Java.
  • Experience operating high-cardinality, high-throughput metrics and log pipelines.
  • Strong understanding of SLOs, error budgets, and SRE principles.
  • Experience integrating observability with CI/CD and incident management tooling.
  • Solid grasp of Linux internals, networking, and container platforms.
  • Excellent communication and collaboration skills.
Preferred Qualifications
  • Experience with Thanos, Mimir, Cortex, Loki, or Tempo at scale.
  • Contributions to OpenTelemetry or observability open-source projects.
  • Familiarity with eBPF-based observability tooling.
  • Experience driving observability cost optimization initiatives.
  • Exposure to regulated environments with audit-grade logging requirements.
How to Apply

Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at (908) 505-3899. Learn more about Bright Vision Technologies at www.bvteck.com.

Bright Vision Technologies is an Equal Opportunity Employer.

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
430,878 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$191k – $334k per year • Equity • In office • Full-Time • 10+ years exp • Bachelor's Degree • Santa Clara
Python
Java
Databases
Apache Kafka
AI/ML
Fine-tuning
Quantization
Prompt Engineering
AI Agents
TensorFlow
PyTorch
RAG
Anomaly Detection
Feature Store
LLM Guardrails
Cybersecurity
Zero Trust
Management
ServiceNow
Apply
$191k – $334k per year • Equity • In office • Full-Time • 12+ years exp • Bachelor's Degree • Santa Clara
Python
Java
Databases
ElasticSearch
OpenSearch
AI/ML
Embeddings
AI Agents
RAG
Semantic Search
Semantic Search
Recommender Systems
Mobile
Algolia
DevOps
Platform Engineering
Vector
Management
ServiceNow
Apply
$90k – $95k per year • Equity • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Chicago
Python
SQL
Databases
MySQL
DevOps
AWS
Analytics
Tableau
ETL/ELT
Alteryx
Apply
$25k – $73k per year (Estimated) • Remote • Full-Time • 4+ years exp • High School Diploma • Brazil
Python
PowerShell
DevOps
Terraform
Azure DevOps
GitHub Actions
Kibana
Datadog
FluxCD
Azure
CI/CD
GitOps
ArgoCD
AWS
Kubernetes
Grafana
Amazon EKS
Azure AKS
GitHub
Cybersecurity
OWASP ZAP
SonarQube
Mend
Apply
$156k – $234k per year • Remote/Hybrid • Full-Time • 10+ years exp • Irving • Jacksonville
Python
Python
Flask
FastAPI
Databases
Chroma
Milvus
Pinecone
AI/ML
LangChain
Vertex AI
Gemma
Fine-tuning
Prompt Engineering
AI Agents
NeMo Guardrails
Llama
Mistral
Pandas
NumPy
PyTorch
LLM
RAG
Google ADK
Hallucination
Hugging Face
NVIDIA NeMo
LLM Guardrails
Edge AI
Agentic Workflows
DevOps
OpenShift
CI/CD
Docker
Kubernetes
Vector
Apply
$100k – $150k per year • Remote • 5+ years exp • PhD
Python
PowerShell
DevOps
Terraform
Azure DevOps
GitHub Actions
Azure
CI/CD
Kubernetes
Service Mesh
Bicep
Azure AKS
FinOps
GitHub
Cybersecurity
PCI DSS
SOC 2
HIPAA
FedRAMP
Apply
$100k – $150k per year • Remote • 6+ years exp • Bachelor's Degree
Python
Bash
DevOps
Terraform
Ansible
GCP
Red Hat
OpenShift
Helm
Istio
Linkerd
Azure
CI/CD
GitOps
ArgoCD
Jenkins
AWS
Kubernetes
Service Mesh
Tekton
Cybersecurity
PCI DSS
SOC 2
HIPAA
Apply
$145k – $205k per year • Remote • 6+ years exp • PhD
Rust
SQL
C++
Rust
Actix Web
Axum
Databases
MySQL
PostgreSQL
Redis
NATS
DynamoDB
RabbitMQ
Apache Kafka
DevOps
Rest API
gRPC
Terraform
GCP
Istio
OpenTelemetry
Linkerd
Prometheus
Pulumi
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Grafana
Service Mesh
Apply
$140k – $200k per year • Remote • 6+ years exp • PhD
JavaScript
Java
TypeScript
SQL
Node JS
Solidity
Solidity
Truffle
Frontend
Vue.js
GraphQL
Angular
React.js
Remix
DevOps
Rest API
Terraform
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Web3
Foundry
Polygon
Hardhat
Remix
Avalanche
Optimism
Arbitrum
Layer 2
Smart Contracts
Ethereum
Hyperledger Fabric
Token Standards
Apply
$100k – $120k per year • Remote • 6+ years exp • Bachelor's Degree
Python
AI/ML
LLM
Cybersecurity
Threat Modeling
Apply
See all jobs
This is one of many
430,878 more open roles from verified company boards, updated every day.