Clera home
·Dashboard

Jobs at Voltage Park (Now Hiring) — 2 open

Voltage Park logoVoltage Park

Infrastructure Engineer (Observability)

San Francisco, California, United States · Remote OK

Senior+$500M raised

Voltage Park is seeking an Infrastructure Engineer with a focus on Observability to join our Infrastructure Engineering team. Our engineers design and operate the systems that manage thousands of bare-metal servers, GPUs…

Skills: Observability, Telemetry, Prometheus, Grafana, ELK

Voltage Park logoVoltage Park

Infrastructure Operations Engineer

San Francisco, California, United States · Remote OK

Senior+$500M raised

Voltage Park is your enterprise AI factory. We offer scalable compute power, on-demand and reserved bare metal AI infrastructure using NVIDIA GPUs, with world-class service, performance and value. Founded with the missio…

Skills: Linux, AWS, Kubernetes, Terraform, Ansible

Voltage Park logo

Infrastructure Engineer (Observability)

Voltage Park

San Francisco, California, United States • Remote OK

Apply
Senior+

Tired of cold applications?

Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.

  • Full-time
  • Posted 8d ago
  • ~40 hrs/week
  • Remote in United States

Responsibilities

Design and maintain observability platforms for metrics, logs, and traces across bare-metal servers and GPUs. Create dashboards and noise-resistant alerting pipelines to provide actionable insights for internal and external stakeholders.

Requirements

Requires over 8 years of experience in infrastructure engineering or SRE with proficiency in monitoring tools and automation languages like Python or Go. Candidates must be based in the continental US and have experience with container observability and telemetry pipelines.

Full job description

Voltage Park is seeking an Infrastructure Engineer with a focus on Observability to join our Infrastructure Engineering team. Our engineers design and operate the systems that manage thousands of bare-metal servers, GPUs, and high-performance networks across multiple data centers.

This role combines the breadth of a core infrastructure engineer with a specialty in observability and telemetry. You’ll design and operate metrics, logs, traces, and alerting pipelines that provide actionable insights for both internal teams and external customers — helping to ensure reliability and transparency at scale.

This is a fully remote position, although candidates must be based in the continental United States. Unfortunately, we are unable to provide sponsorship for this role.

RESPONSIBILITIES

- Design, build, and maintain observability platforms spanning metrics, logs, traces, and events.

- Create dashboards and alerting for internal stakeholders (InfraOps, Engineering, Customer Success) and scoped visibility for external customers.

- Ingest and correlate telemetry from GPUs, CPUs, networking (Ethernet & InfiniBand), containers, APIs, and BMC/Redfish.

- Implement noise-resistant alerting pipelines that improve detection and reduce operational load.

- Collaborate with infrastructure, platform, and customer-facing teams to embed observability into workflows.

- Contribute to broader infrastructure engineering projects beyond observability.

QUALIFICATIONS

- 8+ years in infrastructure engineering, SRE, or observability roles.

- Strong experience with monitoring systems (Prometheus, Grafana, ELK, VictoriaMetrics, or similar).

- Proficiency in Python, Go, or bash for automation and data integration.

- Familiarity with container/Kubernetes observability.

- Understanding of streaming telemetry pipelines (Kafka, OTEL, Promtail, or equivalent).

- Strong written and verbal communication skills.

IDEAL EXPERIENCES

- Experience with GPU observability, particularly NVIDIA DCGM.

- Designing multi-tenant observability solutions with RBAC and scoped queries.

- Prior work with correlation engines for RCA, forecasting, or predictive alerting.

- Broader exposure to infrastructure domains (networking, storage, provisioning).

CULTURE

- You enjoy working with a small, highly motivated team.

- You’re comfortable balancing autonomy with company-wide priorities.

- You value clarity, documentation, and actionable insights in observability systems.

- You’re excited to specialize in observability while contributing as a core infrastructure engineer.

Voltage Park is an equal opportunity employer and makes employment decisions on the basis of merit. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, protected veteran status, or any other characteristic under federal, state, or local law. If you require an accommodation during the job application process, please notify your recruiter.. 

Related keywords

ObservabilityTelemetryPrometheusGrafanaELKVictoriaMetricsPythonGoBashKubernetesKafkaOTELPromtailNVIDIA DCGMGPUBare-metal

About Voltage Park

LinkedInVisit site
Industry
Technology, Information and Internet
Company size
51-200 employees
Headquarters
San Francisco, California
LinkedIn followers
6,614
Total funding
$500M

The Voltage Park AI Factory: a fully integrated hardware and software platform that eliminates the learning curve, vendor lock-in, and data privacy tradeoffs holding companies back. Preview it now.

Offices: 555 Montgomery St, San Francisco, California 94111, US

Cloud ComputingMachine LearningAI Infrastructure
View all jobs at Voltage Park

About Voltage Park

LinkedInVisit site
Industry
Technology, Information and Internet
Company size
51-200 employees
Headquarters
San Francisco, California
LinkedIn followers
6,614
Total funding
$500M

The Voltage Park AI Factory: a fully integrated hardware and software platform that eliminates the learning curve, vendor lock-in, and data privacy tradeoffs holding companies back. Preview it now.

Offices: 555 Montgomery St, San Francisco, California 94111, US

Cloud ComputingMachine LearningAI Infrastructure
View all jobs at Voltage Park

Similar companies hiring

Carvana (2409)Delivery Hero (997)Peraton (933)SFS (870)Celestica (817)Mindrift (751)BukuWarung (603)Cox Business (589)AUTO1 Group (511)Tieto (509)Lifted, an Upwork Company (317)Arrow Electronics (306)
Clera home

Your AI-talent agent. Connecting talents with dream jobs.

Earn $5,000

Tools

  • Salary Calculator
  • Resume Review
  • Startup Map

Explore

  • Jobs
  • Discover Jobs
  • Companies
  • Referral

Platform

  • Pricing
  • Integrations
  • Partners
  • Acquihire

Clera

  • Manifesto
  • Engineering
  • We are hiring!
  • FAQs
  • Blog
  • Press

Tools

  • Salary Calculator
  • Resume Review
  • Startup Map

Explore

  • Jobs
  • Discover Jobs
  • Companies
  • Referral

Platform

  • Pricing
  • Integrations
  • Partners
  • Acquihire

Clera

  • Manifesto
  • Engineering
  • We are hiring!
  • FAQs
  • Blog
  • Press

© 2026 Clera Labs, Inc.

PrivacyTermsBug Bounty