Bright Vision Technologies
Bright Vision Technologies
·TodayBright Vision Technologies
·TodaySite observability engineer
Location
remote, United States
Salary
$100k – $150k/yr
Commitment
Full Time
Level
Senior (5+ years)
Required skills
Job Description
Site Observability Engineer
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Summary
We are looking for a Site Observability Engineer to design and operate the metrics, logging, tracing, and alerting platforms that give engineering teams confidence in the systems they run. The role spans the full observability stack — from collection agents and pipelines to long-term storage, dashboards, and alerting workflows — with a strong focus on usability, signal quality, and operational ROI.
Key Responsibilities
- Design and operate enterprise-grade observability platforms covering metrics, logs, traces, events, and synthetic monitoring.
- Architect Prometheus / Thanos / Mimir, Grafana, Loki, Tempo, OpenTelemetry, and Datadog deployments for high availability and scale.
- Develop standards for service instrumentation, including OpenTelemetry adoption, metric naming, label cardinality, and structured logging conventions.
- Define and enforce SLOs, SLIs, and error budgets, and build the dashboards and alerts that operationalize them.
- Build alerting strategies that minimize noise, surface actionable signals, and integrate cleanly with on-call workflows in PagerDuty, Opsgenie, or similar tools.
- Operate large-scale time-series and log storage platforms, balancing retention, query performance, and cost.
- Design distributed tracing pipelines and help teams use traces to diagnose latency and reliability issues.
- Develop self-service tooling, paved-road libraries, and templates that make adoption of observability standards easy for product teams.
- Drive cost management and label-cardinality discipline across the observability estate.
- Lead incident response readiness improvements through better dashboards, alerting hygiene, and post-incident analysis tooling.
- Partner with SRE and platform teams to integrate observability into deployment pipelines, canary analysis, and progressive delivery workflows.
- Evaluate and recommend observability vendors and open-source tools based on cost, capability, and operational maturity.
- Mentor engineering teams on observability fundamentals, debugging techniques, and SLO-driven operations.
- Maintain documentation, onboarding guides, and runbooks for the observability platform.
Required Qualifications
- Bachelor’s degree in Computer Science or a related field.
- Five or more years of experience in SRE, platform engineering, or observability roles.
- Deep hands-on experience with Prometheus, Grafana, and at least one major commercial observability platform such as Datadog, New Relic, or Splunk.
- Strong understanding of OpenTelemetry, distributed tracing, and structured logging.
- Proficiency in at least one general-purpose language such as Go, Python, or Java.
- Experience operating high-cardinality, high-throughput metrics and log pipelines.
- Strong understanding of SLOs, error budgets, and SRE principles.
- Experience integrating observability with CI/CD and incident management tooling.
- Solid grasp of Linux internals, networking, and container platforms.
- Excellent communication and collaboration skills.
Preferred Qualifications
- Experience with Thanos, Mimir, Cortex, Loki, or Tempo at scale.
- Contributions to OpenTelemetry or observability open-source projects.
- Familiarity with eBPF-based observability tooling.
- Experience driving observability cost optimization initiatives.
- Exposure to regulated environments with audit-grade logging requirements.
How to Apply
Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at (908) 505-3544. Learn more about Bright Vision Technologies at www.bvteck.com.
Equal Employment Opportunity (EEO) Statement
Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall. BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.
Ready to join the team?
Apply nowSimilar Jobs:
- Today
remote, United States
$75k – $85k/yr
Full Time
Senior (5+ years)
- Today
Bright Vision Technologies
Telemetry engineer
remote, United States
$100k – $150k/yr
Full Time
Senior (5+ years)
TodayCrum & Forster
Lead infrastructure & site reliability engineering
remote, Glastonbury, CT, United States
$105k – $198k/yr
Full Time
Lead / Manager
TodayMercury
Senior software engineer - sre
remote, United States
$201k – $251k/yr
Full Time
Senior (5+ years)
- 5 days ago
Bright Vision Technologies
Observability engineer
remote, United States
$89k – $110k/yr
Full Time
Senior (5+ years)
6 days agoTEKsystems
Grafana platform engineer
remote, VA, United States
Full Time
Junior (<2 years)
6 days agoBright Vision Technologies
Systems observability specialist
remote, United States
$135k – $155k/yr
Full Time
Senior (5+ years)