Prepin
Log in
Advisor360°

engineering opportunity

DevOps Engineer - Agentic AI Platform

You will own the infrastructure powering AI systems, focusing on cluster operations, GitOps delivery pipelines, and platform reliability. Additionally, you will manage deployment strategies, security posture, and developer experience to ensure efficient, scalable AI workflows.

Needham, Massachusetts, United StatesonsiteFULL_TIME

Posted

About the role

What will you do at Advisor360°?

Our Agentic AI team builds the platform layer that makes AI systems truly

production-ready—and we're already live in production. This isn't a greenfield

initiative; it's a high-impact environment where real systems run at scale

today.

This is a DevOps role at the center of that platform, and it has two sides.

First, you'll provision and operate the infrastructure our AI and agentic

workloads run on—the Kubernetes, GitOps, identity, and gateway layers that let

LLM-powered services ship and scale safely. Second, you'll bring AI into how we

run operations itself—using agents and LLMs to cut toil in triage, deployment,

code review, and incident response (we already do this in production today and

want to push it much further).

You don't need to be an ML engineer. We're looking for a strong DevOps Engineer

who is genuinely excited to make AI a first-class part of the platform—and to

build deep AI-infrastructure skills here. You'll work alongside senior platform

engineers who partner with you on the harder architectural calls; this is a

hands-on build-and-operate role, not a solo "own everything" mandate.

Here’s

What You’ll Do

  • * Provision and operate AI infrastructure: the Kubernetes, identity, secrets,
  • and gateway layers that AI and agentic services depend on—built so teams can
  • ship LLM-powered features safely
  • * Apply AI to DevOps itself: build and operate agent-assisted automation that
  • reduces toil—triage, PR review, runbook generation, log and incident
  • analysis. We already run AI in our delivery pipeline and want a teammate
  • who'll take it further
  • * Cluster operations on AKS: node pool sizing, autoscaling policies, namespace
  • isolation, and day-two operational hygiene across environments
  • * GitOps delivery with ArgoCD: app-of-apps structure, environment promotion,
  • rollback strategy, and the guardrails that keep one team's bad deploy from
  • cascading
  • * Deployment strategies: rolling, blue-green, and canary patterns for agentic
  • services where a bad rollout has downstream effects on active workflows
  • * Platform reliability: SLIs, SLOs, alerting, and runbooks for the infra
  • layer—so when something breaks at 2am, there's a playbook to follow (and you
  • help write it)
  • * Cost and capacity management: AI workloads have spiky, non-linear cost
  • profiles. You'll instrument and enforce budgets, quotas, and rightsizing
  • across the cluster
  • What You Bring to the Table:
  • * 3+ years operating Kubernetes in production
  • * Hands-on GitOps with ArgoCD: multi-environment setups, sync waves, health
  • checks, and rollback under pressure
  • * Azure fluency: AKS, ACR, Azure Monitor, Key Vault, and managed/workload
  • identity
  • * Infrastructure-as-code as a default: Terraform for everything—no console
  • cowboys
  • * Scripting in Python, Go, or Bash for automation and tooling—maintained code,
  • not one-offs
  • * Solid incident-response instincts; you've been on-call, written postmortems,
  • and fixed the underlying conditions rather than just the symptom
  • * A real foothold in AI for infrastructure—either you've applied AI/LLMs to
  • operations work (automation, triage, code or PR review, log analysis), or
  • you've provisioned and operated infrastructure for AI workloads. You don't
  • need an ML background; you need to be the DevOps engineer who's already
  • reaching for AI and wants to go deeper

Bonus Points

  • & Where You'll Grow:
  • You won't have all of these on day one—several are exactly what you'll develop
  • in this role:
  • * * AI gateway / proxy patterns for AI workloads—centralized provider-key
  • management, rate limiting, quotas, cost attribution, and failover in front
  • of LLM providers
  • * Agentic AI frameworks (LangGraph, AutoGen, or similar) and the
  • infrastructure patterns they require
  • * LLM inference / serving infrastructure (vLLM, TGI, Triton, or managed
  • equivalents) and GPU capacity management
  • * Policy-as-code with OPA/Gatekeeper for cluster governance
  • * OpenTelemetry and distributed tracing across non-trivial services
  • * Service mesh (Istio or Linkerd) for service-to-service auth and traffic
  • management
  • * Multi-tenant platform expertise
  • Why You’ll Love Working Here:
  • It’s not just about work—it’s about building a career and enjoying the ride!
  • Here’s what you can expect:
  • We believe in recognizing and rewarding performance. Our compensation package
  • includes competitive base salaries, annual performance-based bonuses, and the
  • chance to share in the equity value you and your colleagues create during your
  • time with the company. We offer comprehensive health benefits, including dental,
  • life, and disability insurance. We also trust our employees to manage their time
  • effectively, which is why we offer an unlimited paid time off program to help
  • you perform at your best every day.
  • Join us on this journey. Advisor360° is an equal opportunity employer committed
  • to a diverse workforce. We believe diversity drives innovation and are therefore
  • building a company where people of all backgrounds are truly welcome and
  • included. Everyone is encouraged to bring their unique, authentic selves to work
  • each and every day. The way we see it, we are here to learn from each other.
  • --------------------------------------------------------------------------------
  • The estimated base salary range for this position is $160,000–$175,000 + bonus &
  • equity.
  • Advisor360° provides an estimate of the compensation for roles that may be hired
  • as required by state regulations.

Compensation

may vary based on factors

including, but not limited to, individual candidate experience, skills, and

qualifications.

Additionally, Advisor360° leverages current market data to determine

compensation, therefore posted compensation figures are subject to change as new

market data becomes available. The salary, other forms of compensation, and

benefits information is accurate as of the date of this posting. Advisor360°

reserves the right to modify this information at any time, subject to applicable

law.

Advisor360° is currently authorized to employ individuals remotely in the

following states: CA, CT, FL, GA, IL, MA, MD, ME, MI, NH, NJ, NY, NC, PA, RI,

SC, TN, TX, UT, WA, VA. Applicants must reside in one of these states at the

time of hire and throughout their employment unless otherwise approved by the

People Team. Candidates located within 50 miles of our Needham, MA headquarters

may be required to follow applicable hybrid work guidelines.

--------------------------------------------------------------------------------

While we are interested in qualified applicants who are permanently eligible to

work for any employer in the United States, we are unable to sponsor or take

over sponsorship for employment visas at this time.

To all recruitment agencies: We do not accept unsolicited agency resumes and are

not responsible for any fees related to unsolicited resumes.

It is unlawful in Massachusetts to require or administer a lie detector test as

a condition of employment or continued employment. An employer who violates this

law shall be subject to criminal penalties and civil liability.

Advisor360 is an Equal Opportunity Employer. We celebrate diversity and are

committed to creating an inclusive environment for all employees. All employment

decisions are based on business needs, job requirements, and individual

qualifications, without regard to race, color, religion, sex, sexual

orientation, gender identity, national origin, veteran, or disability status.

Advisor360 will not tolerate discrimination or harassment based on any of these

characteristics.

--------------------------------------------------------------------------------

As part of our recruiting process, Advisor360 uses Metaview during recruiter

screening conversations solely to transcribe interview notes and capture an

accurate record of the discussion. These transcripts help our recruiting team

focus on the conversation rather than manual note-taking.

Advisor360 does not use artificial intelligence to evaluate candidates, make

hiring decisions, rank applicants, or determine whether a candidate advances

through the interview process. All hiring decisions are made by our recruiting

team and hiring managers based on human review and assessment.

Which skills does this role require?

KubernetesArgoCDGitOpsAzureAKSTerraformPythonGoBashDevOpsCKACKSInfrastructure-as-codeAPI gatewayRBACAzure ADKey VaultCloud computingAutomationScalabilityNode.jsMachine LearningLLMs

Make your next move

Build a shortlist and prepare

Identify the requirements you can demonstrate, then choose examples from your work to discuss with the hiring team.

Review the responsibilities and requirements before adding an opening to your shortlist.

Role information can change. Confirm current details on the original application page.

Product

AI Candidate AgentCompaniesBrowse JobsDeep ProfileSkill AssessmentOpportunity Matching
Prepin.ai

© 2026 Prepin | All rights reserved.