Why this grade This listing scored 54/100, which is a D. It lost the most ground on pay transparency. See the breakdown
- Description depth 20 / 20 How much the posting actually says about the work, measured in characters of real text.
- Freshness 15 / 15 How recently it was posted. Older postings are likelier to be filled or abandoned.
- Remote clarity 8 / 15 Whether "remote" means anywhere, or is quietly restricted to one country.
- Role specificity 6 / 10 Whether the listing is tagged well enough to tell what the role actually is.
- Corroboration 5 / 10 Whether more than one source carries this listing.
- Pay transparency 0 / 25 A published salary range, worth more than any other single factor because it is what a candidate cannot find out without applying.
Every figure above is arithmetic over the posting itself — its salary field, its text, its age, its tags and how many sources carry it. How the grades work →
Developer Senior Full Time
About Dynamo AI
Dynamo AI helps enterprises deploy AI systems that are reliable, secure, and production-ready. Our market-leading technical controls span AI evaluations, guardrails, agentic risk management, and observability. Backed by world-class talent, we partner with innovative, highly-regulated global organizations — across financial services, government, and beyond — to deploy meaningful AI use cases at scale, securely and compliantly.
About the Role
We are looking for a Senior DevOps Engineer to help build, scale, and operate the infrastructure powering our AI platform. This is a hands-on, high-ownership role. We are looking for someone who can work independently, solve complex infrastructure problems, and thrive in a fast-paced startup environment. You should be comfortable designing systems, automating processes, troubleshooting production issues, and continuously improving reliability, scalability, and cost efficiency.
What You'll Own
- Design, build, and operate highly available production infrastructure on AWS, with strong expertise in EKS, EC2, VPC, S3, RDS/Aurora, IAM, ECR, ElastiCache, Load Balancers, and other core AWS services.
- Build and improve CI/CD and release automation using Jenkins, GitHub Actions, Helm, ArgoCD, and GitOps.
- Manage infrastructure using Terraform and Infrastructure as Code principles.
- Build and operate Kubernetes platforms at production scale, including cluster management, upgrades, autoscaling, networking, security, and troubleshooting.
- Run and operate AI/ML workloads in production, with a strong understanding of the infrastructure challenges associated with AI systems.
- Deploy, scale, monitor, and optimize AI inference and model-serving workloads across Kubernetes and cloud infrastructure.
- Work with GPU-based workloads, including GPU scheduling, utilization, autoscaling, capacity planning, and optimization.
- Drive infrastructure efficiency by balancing performance, reliability, scalability, and cost across AI workloads.
- Own monitoring, logging, and observability using tools such as Prometheus, Grafana, Thanos, and OpenTelemetry.
- Implement secure secrets management using technologies such as HashiCorp Vault and External Secrets Operator.
- Develop automation and internal tooling using Python and Bash.
- Drive improvements around reliability, security, scalability, performance, and infrastructure cost.
- Participate in production incidents, root-cause analysis, and drive long-term fixes rather than short-term workarounds.
- Work closely with Engineering, ML/AI, Security, and Product teams to solve infrastructure and platform challenges.
What We're Looking For
- 5+ years of strong hands-on experience in DevOps, SRE, Platform Engineering, or Cloud Infrastructure.
- Strong production experience with AWS and Kubernetes/EKS.
- Proven experience running AI/ML workloads or GPU-based workloads in production is highly valuable.
- Strong understanding of CI/CD, Infrastructure as Code, GitOps, observability, and cloud security.
- Excellent scripting and automation skills in Python and Bash.
- Experience operating production systems and troubleshooting complex infrastructure issues independently.
- Strong understanding of scaling, performance optimization, resource utilization, and cost management, particularly for compute-intensive workloads.
- Strong ownership mindset with the ability to take a problem from design to production.
- Experience working in a startup or fast-moving engineering environment is highly valued.
- Strong communication skills and the ability to work effectively across teams.
Nice to Have
- Experience with AI inference platforms, model serving, LLM infrastructure, or ML platforms.
- Experience with GPU infrastructure such as NVIDIA GPUs and Kubernetes GPU scheduling.
- Experience with multi-region or highly distributed systems.
- Experience with SOC 2, ISO 27001, or other security/compliance requirements.
- Experience with PostgreSQL, MongoDB, Redis, Kafka, or similar distributed systems.
The Kind of Engineer We Want
We're looking for someone who builds, automates, and takes ownership — not someone who simply operates existing infrastructure.
You should be comfortable with ambiguity, willing to dive deep into production problems, and constantly looking for ways to make our platform more reliable, secure, scalable, and cost-efficient.
Most importantly, you should understand that AI infrastructure has a different set of operational challenges. We want someone who can help us run AI systems efficiently at scale — making the right trade-offs between GPU utilization, performance, reliability, scalability, and cost.
This is a high-ownership role for someone who wants to make a meaningful impact on the infrastructure behind an AI platform as we scale.
Originally posted on Himalayas
Apply for this role Opens himalayas.app — the link as listed; we have not yet verified it is the employer's own page
Quick question · anonymous · one tap
Would you apply to this job?
Answer to see what other job seekers said.
Keep looking
Similar remote roles, still open
-
D
7h ago
Senior QA Engineer (Manual & Automation)
Salla Saudi Arabia
-
C
18h ago
Jobber Canada
-
D
17h ago
BetterMe Ukraine
-
A
19h ago
Nodeworthy Remote $4k - $6k/mo
See every "DevOps Engineer" role →
Get new “DevOps Engineer” roles by email
One email a day with what is new in "DevOps Engineer". Nothing new, no email.
We confirm the address first, and every mail carries an unsubscribe link. Alerts are ours, not a third party's.
Or follow the daily Telegram channel for the best new remote roles every morning.
Your turn · no account needed
Help the next applicant
You may know something about this listing that we cannot see from here. One tap. No account needed. Signed-in reports earn points once the evidence agrees with you.
I know what it pays
Sign in with Google to earn points for reports — 100 confirmed points buy a week of Early Access.
Where this listing came from
- 10 Oct 2026 Himalayas first sighting
Seen on 1 board over 0 days.