About the role
What will you do at Advisor360°?
Our Agentic AI team builds the platform layer that makes AI systems truly
production-ready—and we're already live in production. This isn't a greenfield
initiative; it's a high-impact environment where real systems run at scale
today.
This is a DevOps role at the center of that platform, and it has two sides.
First, you'll provision and operate the infrastructure our AI and agentic
workloads run on—the Kubernetes, GitOps, identity, and gateway layers that let
LLM-powered services ship and scale safely. Second, you'll bring AI into how we
run operations itself—using agents and LLMs to cut toil in triage, deployment,
code review, and incident response (we already do this in production today and
want to push it much further).
You don't need to be an ML engineer. We're looking for a strong DevOps Engineer
who is genuinely excited to make AI a first-class part of the platform—and to
build deep AI-infrastructure skills here. You'll work alongside senior platform
engineers who partner with you on the harder architectural calls; this is a
hands-on build-and-operate role, not a solo "own everything" mandate.
Here’s
What You’ll Do
- * Provision and operate AI infrastructure: the Kubernetes, identity, secrets,
- and gateway layers that AI and agentic services depend on—built so teams can
- ship LLM-powered features safely
- * Apply AI to DevOps itself: build and operate agent-assisted automation that
- reduces toil—triage, PR review, runbook generation, log and incident
- analysis. We already run AI in our delivery pipeline and want a teammate
- who'll take it further
- * Cluster operations on AKS: node pool sizing, autoscaling policies, namespace
- isolation, and day-two operational hygiene across environments
- * GitOps delivery with ArgoCD: app-of-apps structure, environment promotion,
- rollback strategy, and the guardrails that keep one team's bad deploy from
- cascading
- * Deployment strategies: rolling, blue-green, and canary patterns for agentic
- services where a bad rollout has downstream effects on active workflows
- * Platform reliability: SLIs, SLOs, alerting, and runbooks for the infra
- layer—so when something breaks at 2am, there's a playbook to follow (and you
- help write it)
- * Cost and capacity management: AI workloads have spiky, non-linear cost
- profiles. You'll instrument and enforce budgets, quotas, and rightsizing
- across the cluster
- What You Bring to the Table:
- * 3+ years operating Kubernetes in production
- * Hands-on GitOps with ArgoCD: multi-environment setups, sync waves, health
- checks, and rollback under pressure
- * Azure fluency: AKS, ACR, Azure Monitor, Key Vault, and managed/workload
- identity
- * Infrastructure-as-code as a default: Terraform for everything—no console
- cowboys
- * Scripting in Python, Go, or Bash for automation and tooling—maintained code,
- not one-offs
- * Solid incident-response instincts; you've been on-call, written postmortems,
- and fixed the underlying conditions rather than just the symptom
- * A real foothold in AI for infrastructure—either you've applied AI/LLMs to
- operations work (automation, triage, code or PR review, log analysis), or
- you've provisioned and operated infrastructure for AI workloads. You don't
- need an ML background; you need to be the DevOps engineer who's already
- reaching for AI and wants to go deeper
Bonus Points
- & Where You'll Grow:
- You won't have all of these on day one—several are exactly what you'll develop
- in this role:
- * * AI gateway / proxy patterns for AI workloads—centralized provider-key
- management, rate limiting, quotas, cost attribution, and failover in front
- of LLM providers
- * Agentic AI frameworks (LangGraph, AutoGen, or similar) and the
- infrastructure patterns they require
- * LLM inference / serving infrastructure (vLLM, TGI, Triton, or managed
- equivalents) and GPU capacity management
- * Policy-as-code with OPA/Gatekeeper for cluster governance
- * OpenTelemetry and distributed tracing across non-trivial services
- * Service mesh (Istio or Linkerd) for service-to-service auth and traffic
- management
- * Multi-tenant platform expertise
- Why You’ll Love Working Here:
- It’s not just about work—it’s about building a career and enjoying the ride!
- Here’s what you can expect:
- We believe in recognizing and rewarding performance. Our compensation package
- includes competitive base salaries, annual performance-based bonuses, and the
- chance to share in the equity value you and your colleagues create during your
- time with the company. We offer comprehensive health benefits, including dental,
- life, and disability insurance. We also trust our employees to manage their time
- effectively, which is why we offer an unlimited paid time off program to help
- you perform at your best every day.
- Join us on this journey. Advisor360° is an equal opportunity employer committed
- to a diverse workforce. We believe diversity drives innovation and are therefore
- building a company where people of all backgrounds are truly welcome and
- included. Everyone is encouraged to bring their unique, authentic selves to work
- each and every day. The way we see it, we are here to learn from each other.
- --------------------------------------------------------------------------------
- The estimated base salary range for this position is $160,000–$175,000 + bonus &
- equity.
- Advisor360° provides an estimate of the compensation for roles that may be hired
- as required by state regulations.
Compensation
may vary based on factors
including, but not limited to, individual candidate experience, skills, and
qualifications.
Additionally, Advisor360° leverages current market data to determine
compensation, therefore posted compensation figures are subject to change as new
market data becomes available. The salary, other forms of compensation, and
benefits information is accurate as of the date of this posting. Advisor360°
reserves the right to modify this information at any time, subject to applicable
law.
Advisor360° is currently authorized to employ individuals remotely in the
following states: CA, CT, FL, GA, IL, MA, MD, ME, MI, NH, NJ, NY, NC, PA, RI,
SC, TN, TX, UT, WA, VA. Applicants must reside in one of these states at the
time of hire and throughout their employment unless otherwise approved by the
People Team. Candidates located within 50 miles of our Needham, MA headquarters
may be required to follow applicable hybrid work guidelines.
--------------------------------------------------------------------------------
While we are interested in qualified applicants who are permanently eligible to
work for any employer in the United States, we are unable to sponsor or take
over sponsorship for employment visas at this time.
To all recruitment agencies: We do not accept unsolicited agency resumes and are
not responsible for any fees related to unsolicited resumes.
It is unlawful in Massachusetts to require or administer a lie detector test as
a condition of employment or continued employment. An employer who violates this
law shall be subject to criminal penalties and civil liability.
Advisor360 is an Equal Opportunity Employer. We celebrate diversity and are
committed to creating an inclusive environment for all employees. All employment
decisions are based on business needs, job requirements, and individual
qualifications, without regard to race, color, religion, sex, sexual
orientation, gender identity, national origin, veteran, or disability status.
Advisor360 will not tolerate discrimination or harassment based on any of these
characteristics.
--------------------------------------------------------------------------------
As part of our recruiting process, Advisor360 uses Metaview during recruiter
screening conversations solely to transcribe interview notes and capture an
accurate record of the discussion. These transcripts help our recruiting team
focus on the conversation rather than manual note-taking.
Advisor360 does not use artificial intelligence to evaluate candidates, make
hiring decisions, rank applicants, or determine whether a candidate advances
through the interview process. All hiring decisions are made by our recruiting
team and hiring managers based on human review and assessment.
Which skills does this role require?
Make your next move
Build a shortlist and prepare
Identify the requirements you can demonstrate, then choose examples from your work to discuss with the hiring team.
- Build a focused shortlist before you applyCompare role requirements with your experience and give each application a clear reason.
- Platform engineering jobsCompare current openings and review what to look for in this role.
- Practice explaining your experience in an interviewRehearse your answers before meeting the hiring team.
Other roles to compare
Review the responsibilities and requirements before adding an opening to your shortlist.
Security Engineer - Infrastructure Security
Figure · San Jose, California, United States
Senior AI Infrastructure Software Engineer - DGX Cloud
NVIDIA · Redmond, Washington, United States
Machine Learning Infrastructure Engineer
Bright Vision Technologies · Hillsboro, Oregon, United States
Senior Infrastructure Security Engineer
A-Gas · Rhome, Texas, United States
Senior Infrastructure Engineer - Data Protection
USAA · Tampa, Florida, United States
Neural Data Infrastructure Engineer
Blackrock Neurotech · Salt Lake City, Utah, United States
Role information can change. Confirm current details on the original application page.
