Cerebras
Senior SDET, Inference Platform
Part remote, part office, you need to live within commuting distance of a named location.
Hybrid · Sunnyvale, CA
Employer listed it 2 months ago · Added today
Been open since 2 months ago. Long-running listings are sometimes left up after the role is filled.
Salary
Not stated
Location
Hybrid · Sunnyvale, CA
Timezone
Not stated
Contract
Full-time
Experience
Senior
Category
Software
This employer didn't state pay. Jobs like this usually pay around $175k–$230k a year, a typical range taken from 595 senior-level software roles on Nomaders that do state pay. It's a guide, not an offer.
Remote flexibility
Hybrid
This role is only partly remote, the employer expects time in the office around Sunnyvale, CA, Toronto, CAN, Hybrid, so you need to live within commuting distance.
What the employer says
- Source listing states candidate location: "Sunnyvale, CA, Toronto, CAN, Hybrid"
- Listing mentions "Hybrid"
What Nomaders makes of it
- Not suitable if you plan to move between countries
The quotes above are the employer's own words; the reading is ours. Always check the original listing and employment terms before working from another country.
About the role
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.
About the Role
We are looking for an Inference Platform SDET to join the Inference Service Quality team at Cerebras and work on the inference platform. This team sits at the intersection of distributed systems, cloud and cluster infrastructure, and the software stack that serves the world's fastest AI inference.
In this role, you will own the quality and reliability of the infrastructure that deploys and runs the Cerebras Inference Platform — from CI/CD pipelines and Kubernetes-based deployments to ingress, load balancing, and service discovery. You will validate the platform both in cloud environments and on real Cerebras clusters, working side by side with the Inference Platform development team to catch issues before our customers do.
This is an excellent opportunity for engineers who enjoy infrastructure, automation, and debugging across the full deployment stack, and who want to ensure that a platform serving inference at massive scale stays fast, reliable, and production-ready.
Responsibilities
Design, build, and maintain test infrastructure and automation for deploying and validating the Cerebras Inference Platform.
Validate the platform across environments — from cloud-managed Kubernetes to deployments running on Cerebras hardware.
Test and verify deployment infrastructure including Kubernetes workloads, CI/CD pipelines, ingress and service discovery, NGINX, and load balancing.
Collaborate closely with the Inference Platform development team to ensure new features and platform capabilities ship reliably.
Investigate and debug complex issues spanning networking, orchestration, deployment, and distributed services.
Develop and maintain testbeds used to validate platform performance, scalability, and reliability.
Identify failure points, bottlenecks, and edge cases that impact platform stability and inference performance.
Contribute to test plans and validation strategies for new platform features and releases.
Improve observability, diagnostics, and debugging workflows across the inference platform stack.
Partner with engineering teams to ensure high-quality, production-ready releases of the Cerebras Inference Platform.
Minimum Skills & Qualifications
3+ years of experience in software engineering, QA/quality engineering, systems engineering, or infrastructure development.
Strong programming skills in Python and/or Go (experience with both is a plus).
Experience building automation tools, testing frameworks, or internal developer tooling.
Hands-on experience with CI/CD systems (e.g., Jenkins)
Experience debugging complex systems, distributed services, or networked infrastructure.
Familiarity with systems-level development, infrastructure tooling, or platform integration.
Strong problem-solving skills and the ability to investigate issues across multiple system and infrastructure layers.
Requirements
- ·3+ years of experience in software engineering, QA/quality engineering, systems engineering, or infrastructure development.
- ·Strong programming skills in Python and/or Go (experience with both is a plus).
- ·Experience building automation tools, testing frameworks, or internal developer tooling.
- ·Hands-on experience with CI/CD systems (e.g., Jenkins)
- ·Experience debugging complex systems, distributed services, or networked infrastructure.
Benefits
No benefits package published with this listing. Ask about it at first interview.
How to apply
- 1Check the flexibility label above, hybrid, matches where you plan to live and work.
- 2Tailor your CV to the role at Cerebras, mentioning your remote working experience.
- 3Apply directly on the employer's careers page using the button below. Nomaders never handles your application.
Found 22h ago. Last checked today. Always confirm the details on the original posting, salary and location can change after publication.
Listing sourced from Company boards.
Similar roles
Other open software roles with comparable remote rules.
Free to apply, no account needed.
Typically $175k–$230k · You'll be taken to the employer's careers page.