About Hark Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persisten…
About Hark Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persisten…
About Hark Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persisten…
Skills: Visual Design, Interaction Design, Systems Thinking, Design Systems, Web Design
About Hark Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persisten…
About Hark Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persisten…
About Hark Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persisten…
About Hark Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persisten…
Skills: Python, Test Automation, PCB Debugging, Fixture Design, Board Bring-Up
About Hark Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persisten…
Skills: Privacy Infrastructure, Software Engineering, System Architecture, DSAR Pipelines, Data De-identification
About Hark Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persisten…
About Hark Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persisten…
About Hark Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persisten…
Skills: Multimodal AI, Neural Networks, Distributed Machine Learning, Data Curation, Synthetic Data Generation
About Hark Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persisten…
Skills: Machine Learning, Reinforcement Learning, PyTorch, Python, Distributed Training
ADVANCE YOUR CAREER. ADVANCE THE WORLD. At AMD, we believe technology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technolo…
Credo is seeking a motivated and detail-oriented Application Engineering Intern to join our team. In this role, you will work closely with experienced validation, design, and firmware engineers to test and debug next-gen…
Senior Technical Program Manager, Enterprise Architect, Business Process Modeling Lead
San Jose, California, United States · On-site
$192k–$278k/yr
Senior$26M raised
Minimum qualifications: Bachelor's degree in a technical field, or equivalent practical experience. 8 years of experience in program management. 5 years of experience establishing and driving Business Process Modeling (B…
Skills: Program Management, Business Process Modeling, Process Mining, Enterprise Architecture, SAP Signavio
Senior UX Engineer, AI Systems Knowledge and Intelligence Engine
San Jose, California, United States · On-site
$159k–$230k/yr
Senior$26M raised
Minimum qualifications: Bachelor's degree or equivalent practical experience. 6 years of experience in front-end development, technical UX design, or prototyping. Experience in application development across multiple pla…
Minimum qualifications: Bachelor's degree or equivalent practical experience. 2 years of experience with security assessments or security design reviews or threat modeling. 2 years of experience with security engineering…
Senior Technical Program Manager II, Foundational Enterprise Architecture Capabilities
San Jose, California, United States · On-site
$240k–$333k/yr
Senior+$26M raised
Minimum qualifications: Bachelor's degree in a technical field, or equivalent practical experience. 10 years of experience in program management. 8 years of experience leading cross-functional architectural initiatives, …
Minimum qualifications: Bachelor's degree in a technical field, or equivalent practical experience. 8 years of experience in program management. Experience with technical program management and infrastructure systems. Pr…
Skills: Program management, Cybersecurity, Technical governance, AI/ML, Risk management
Minimum qualifications: Bachelor's degree or equivalent practical experience. 5 years of experience with software development and consumer facing products. 2 years of experience with mobile app development and Android ap…
Skills: Android, Java, Kotlin, Software Development, Mobile App Development
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
$180k–$450k/yr
Full-time
Posted 16d ago
~40 hrs/week
Responsibilities
Lead and manage large-scale GPU computing clusters to power AI training and deployment workloads. Design and maintain scalable infrastructure as code and optimize network fabrics for high-throughput training.
Requirements
Requires 5+ years of experience in infrastructure engineering with at least 2 years in ML or HPC environments. Must have demonstrated experience managing large-scale distributed compute infrastructure and GPU clusters.
Full job description
About Hark
Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persistent memory.
We're pairing that intelligence with next-generation hardware to create a universal interface between humans and machines. While today's AI largely operates through chat boxes and decade-old devices, Hark is focused on what comes next: agentic systems that interact naturally with people and the real world.
To get there, we're developing multimodal models and next-generation AI hardware together - designed from the ground up as a single, unified interface for a new era of intelligent systems.
About the Role
We are looking for a Member of Technical Staff, Infrastructure Compute to lead and manage large-scale GPU computing clusters powering our AI training and deployment workloads. You'll work at the intersection of systems engineering and machine learning infrastructure, owning the reliability, scalability, and efficiency of the compute platform that our research and engineering teams depend on. This is a high-impact, highly technical role suited for someone who thrives in complex distributed systems environments and cares deeply about infrastructure as a product.
Responsibilities
Design, implement, and maintain Infrastructure as Code (IaC) best practices to enable repeatable, auditable, and scalable cluster provisioning.
Enhance and harden CI/CD deployment pipelines to ensure robust, secure, and low-latency model service delivery across production environments.
Own and evolve stable training infrastructure operating at the scale of 10,000+ GPUs, including job scheduling, fault tolerance, and network fabric optimization.
Partner closely with ML researchers and engineers to understand compute bottlenecks and translate them into infrastructure improvements.
Monitor system health, define SLOs, and lead incident response for critical training and inference workloads.
Drive capacity planning, cost efficiency initiatives, and hardware lifecycle management across the GPU fleet.
Contribute to internal tooling and platform abstractions that improve developer experience for teams consuming compute resources.
Requirements
5+ years of experience in infrastructure, systems, or platform engineering, with at least 2 years working in ML or HPC environments.
Demonstrated experience managing GPU clusters or large-scale distributed compute infrastructure.
Strong proficiency in at least one systems or infrastructure programming language.
Deep understanding of networking fundamentals (RDMA, InfiniBand, or RoCE a plus) relevant to high-throughput training workloads.
Experience with container orchestration, job scheduling, and multi-tenant resource management.
Proven track record owning production systems with high reliability requirements.
Strong debugging and observability skills across the full infrastructure stack.
Rust and/or Go for systems-level tooling and performance-critical services.
Familiarity with PyTorch and Ray for understanding workload patterns and integration requirements.
Compensation
The US base salary range for this full-time position is between $180,000 - $450,000 annually.
The pay offered for this position may vary based on several individual factors, including job-related knowledge, skills, and experience. The total compensation package may also include additional components and benefits depending on the specific role. This information will be shared if an employment offer is extended.
Related keywords
GPUAI TrainingInfrastructure as CodeCI/CDDistributed SystemsRDMAInfiniBandRoCEKubernetesPulumiRustGoPyTorchRayHPCML Infrastructure
Building the Intelligence Behind the Next Generation of Products.
Industry
Business Consulting and Services
Company size
2-10 employees
Hark Labs is a GenAI and product innovation studio helping startups and enterprises harness the power of artificial intelligence. We design smarter products, automate complex workflows, and turn data into dynamic intelligence — bridging the gap between cutting-edge AI research and real-world business impact.
Based on 1749 listings with disclosed salaries, most software jobs in San Jose, CA pay between $120k–$286k per year. Individual offers vary with seniority, company size, and specialization.
How many Software jobs are open in San Jose, CA right now?
There are currently 1,934 open software positions in San Jose, CA listed on Clera. New openings are added daily as companies post roles.
Which companies are hiring for Software roles in San Jose, CA?
Companies currently hiring include Adobe, AMD, Google, Capital One, Cisco, among others. Browse the listings above to see every active employer.
Are there remote or hybrid Software jobs in San Jose, CA?
Yes — 657 of the 1934 open software positions offer remote or hybrid work (116 remote, 541 hybrid).
How do I apply for Software jobs in San Jose, CA?
Each listing links directly to the employer's application page. Apply early — fresh listings get the most recruiter attention in the first two weeks.