About the role
What will you do at Instrumental, Inc.?
Manufacturing advanced electronics requires understanding millions of signals generated across complex assembly processes. Instrumental builds systems that capture and analyze those signals — images, test results, and process data — enabling engineers to discover failures, identify root causes, and deploy production controls that improve yield and product maturity.
Leading companies such as NVIDIA, Cisco, and Meta rely on Instrumental to accelerate new product development and scale manufacturing across global factories. Instrumental has become mission-critical for manufacturers building and scaling the next generation of AI infrastructure hardware. The Instrumental platform collects, intelligently transforms, and contextually presents manufacturing data to technical end-users, enabling them to optimize their manufacturing process in real-time.
Our core technology is proprietary ML algorithms, packaged in an accessible, user-centric user interface – we believe we must have both the best technology and the best access to that technology to win.
Requirements
- * 5 or more years of DevOps or SRE experience deploying and operating commercial SaaS platforms on public cloud infrastructure, AWS preferred.
- * Expert knowledge with Linux, shell, containerization, Kubernetes, IaC (terraform preferred), monitoring, logging, and APM tools.
- * Proven ability to take initiative and drive impactful projects to completion efficiently and independently.
- * Comfort with ambiguity, pace, and frequent pivots inherent in a startup environment, with a track record of creating clarity for teams.
- * Experience introducing and integrating AI tools/processes into development and operation workflows.
- * Demonstrated skill in setting, iterating on, and measuring KPIs to ensure ongoing performance, reliability and efficiency.
- * Network/application security and compliance experience is a plus.
Who You Are
- * Dead serious about performance, scalability, and reliability (PSR): You care deeply about how systems behave in the real world and sweat the details around latency, uptime, and scale.
- * Systems engineering & infrastructure expertise: You’ve spent real time building and running distributed systems and know your way around cloud infrastructure, networks, and operating systems.
- * Automation, automation, automation: If something is repetitive or error-prone, your first instinct is to automate it and make it disappear.
- * Operating in ambiguity & high-growth environments: You’re comfortable making good calls without perfect information and adapting as the system and company grow fast.
- * Dependable, trustworthy: People trust you to own problems, show up when things are broken, and follow through.
- This position requires access to items and data that are developed under U.S. government contracts and subject to dissemination controls that limit access to U.S. citizens only.
- We’re a growing team that works collaboratively, is supportive of each other, and is highly energized by the opportunity for a large impact.
- We actively work to promote an inclusive environment, valuing passion and the ability to learn.
- You’re encouraged to apply even if your experience doesn’t precisely match the job description!
- The following is a representative annual base salary range for this position within the Bay Area: $158-225k.
- We consider candidates at multiple levels for this role.
- Job level and salary opportunities are evaluated through our interview process – we review the experience, knowledge, skills, and abilities of each applicant.
- Instrumental is proud to offer a highly-rated variety of benefits, including health, vision, dental, commuter plans, and parental leave.
- At Instrumental, protecting company and customer information is a shared responsibility.
- Employees are expected to comply with company engineering, security, access control, and privacy policies, and promptly report suspected security incidents or policy violations.
Which skills does this role require?
Make your next move
Build a shortlist and prepare
Identify the requirements you can demonstrate, then choose examples from your work to discuss with the hiring team.
- Build a focused shortlist before you applyCompare role requirements with your experience and give each application a clear reason.
- Platform engineering jobsCompare current openings and review what to look for in this role.
- Practice explaining your experience in an interviewRehearse your answers before meeting the hiring team.
Other roles to compare
Review the responsibilities and requirements before adding an opening to your shortlist.
Security Engineer - Infrastructure Security
Figure · San Jose, California, United States
AI Infrastructure AI/ML Engineer (100 % remote) (m/f/d)
EWOR · Capon Bridge, West Virginia, United States
Machine Learning Infrastructure Engineer
Bright Vision Technologies · Hillsboro, Oregon, United States
Senior Infrastructure Engineer - Data Protection
USAA · Tampa, Florida, United States
Senior AI Infrastructure Software Engineer - DGX Cloud
NVIDIA · Redmond, Washington, United States
Senior Machine Learning Engineer - ML Training Infrastructure
General Motors · Sunnyvale, California, United States
Role information can change. Confirm current details on the original application page.