← 返回岗位列表美国IT/互联网fulltime

高级 DevOps 工程师

雇主

Akkadian Labs

地点

远程 · 美国

待遇

$面议

工作模式

远程

截止日期

12月12日

🤖 AI 简历匹配评估

检测你的简历与该岗位的匹配度,免费

免费评估

岗位摘要

IMPORTANT NOTE: This role is only available to residents of the United States.

岗位职责

IMPORTANT NOTE: This role is only available to residents of the United States. We are unable to employ anyone who is not physically located in the US, nor are we sponsoring Visas at this time. Please be aware that job offers are contingent upon a background check which includes identity & address verification and a criminal background check.
Who We Are
Akkadian Labsis a Collaboration Lifecycle Automation Platform that services some of the largest global enterprises and government agencies. Our platform currently reduces manual work by up to 90% and costs by as much as 50%, while improving accuracy and governance across leading Unified Communications platforms.
This ability to innovate at scale defines who we are: big enough to compete at the highest level, yet agile enough to stay ahead of a rapidly evolving market. Our culture is people-first, fully remote, and rooted in grit, clarity, trust and respect, because when our people thrive, so do our customers.
Who You Are
You are a hands-on Senior DevOps Engineer with deep experience designing, building, and strengthening the infrastructure that keeps platforms secure, compliant, and uncomplicated. You bring substantial experience maintaining production infrastructure and improving operational reliability through disciplined engineering practice. You have a track record of helping to scale systems with a focus on dependability, observability, deployment confidence, and repeatable processes that support long-term growth. You are energized by solving infrastructure challenges in complex enterprise environments and by building systems that reduce manual work, improve governance, and support operational accuracy at scale.
What You'll Do
This role supports Akkadian’s purpose to uncomplicate work by removing friction between people and their purpose — helping enterprise customers govern, automate, and continuously control complex collaboration environments across cloud, hybrid, and on-premises systems. You will design, implement, and maintain scalable and secure infrastructure and DevOps processes at Akkadian Labs. You will work closely with development, QA, and product teams to enable reliable deployments, automate workflows, and improve system observability across Rocky OS-based, AWS-hosted, and on-premises solutions.
This is a hands-on technical role focused on system-level execution, continuous improvement, and operational excellence within the DevOps function.
Key Responsibilities
Infrastructure and Environment Management
Deploy and maintain scalable infrastructure in AWS and hybrid cloud environments.
Manage infrastructure-as-code (IaC) using Terraform, CloudFormation, or similar tools.
Maintain Linux-based environments.
Design and implement containerization using Docker and orchestration via Kubernetes.
AI and Agent Infrastructure Implementation & Support
Design, deploy and manage AI agent workloads, including provisioning compute instances and managing resource scaling for inference-heavy tasks.
Build and maintain model deployment pipelines, including versioning, testing, and rollback of AI models in production environments.
Monitor AI API consumption and infrastructure costs, implementing alerting and controls to prevent runaway usage and support budget visibility.
Collaborate with the engineering team to implement infrastructure-level security guardrails for AI systems, including access controls and data isolation for model inputs and outputs.
Observability and Reliability
Manage monitoring and observability efforts using tools such as Prometheus, Grafana, and the ELK stack.
Troubleshoot system issues and contribute to incident response and root cause analysis.
Develop and execute strategies for improving system reliability, performance, and uptime.
CI/CDand Automation
Build, maintain, and optimize CI/CD pipelines using tools such as Jenkins, Bitbucket CI/CD, or similar.
Automate routine operational tasks including builds, testing, deployments, and system updates.
Collaborate with engineering teams to integrate pipelines with Akkadian tools.
Security and Compliance
Follow secure DevOps practices and implement and maintain security controls.
Support compliance initiatives and vulnerability remediation efforts.
Collaboration and Documentation
Work closely with DevOps, engineering, QA, and product teams to support deployments and releases.
Maintain documentation for infrastructure, processes, and operational procedures.
Participate in collaborative team processes and continuous improvement initiatives.
Requirements
Experience: 10+ years of experience in DevOps or Site Reliability Engineering (SRE).
Cloud Expertise: Expertise with AWS (e.g., EC2, ECS, S3, IAM, Lambda, CloudWatch).
Infrastructure as Code: Expertise with infrastructure-as-code tools, including Terraform and CloudFormation.
Linux Knowledge: Strong knowledge of Linux environments.
Containerization: Experienced with Docker and Kubernetes.
Scripting: Scripting ability in Python, Bash, or similar languages.
CI/CD: Experience building or maintaining CI/CD pipelines and related tools.
Observability: Experience in monitoring and observability tools such as Prometheus, Grafana, and ELK.
Security: Experience in implementing secure DevOps practices and compliance frameworks (SOC2, ISO, etc).
Experience supporting AI or machine learning workloads and compute environments.
Exposure to AI model deployment pipelines and model versioning practices.
Familiarity with hybrid cloud or on-premises environments.
Exposure to security best practices in DevOps contexts, including AI-specific concerns such as data isolation and access controls.
Experience supporting production systems and participating in on-call rotations.
IMPORTANT NOTE: This role is only available to residents of the United States. We are unable to employ anyone who is not physically located in the US, nor are we sponsoring Visas at this time. Please be aware that job offers are contingent upon a background check which includes identity & address verification and a criminal background check.
Benefits
We offer a fully remote environment, plus a competitive benefits package including medical, dental, vision, company-paid life insurance and disability policies, 401(k) with a generous matching program, and paid time off.
Originally posted on Himalayas

申请条件

- 必须居住在美国境内
- 不提供签证赞助
- 需通过背景调查,包括身份与地址验证及犯罪背景调查
- 具备动手实践的资深DevOps工程师经验
- 拥有设计、构建和强化基础设施的深厚经验
- 具备维护生产基础设施的大量经验
- 具备通过规范化工程实践提升运维可靠性的经验
- 有协助扩展系统的成功经验
- 注重可靠性、可观测性、部署信心和可重复流程
- 善于在复杂企业环境中解决基础设施挑战
- 能够构建减少人工工作、改善治理并支持大规模运营准确性的系统

雇主简介

Akkadian Labs provides a Collaboration Lifecycle Automation Platform that reduces manual work and costs for large enterprises and government agencies, improving accuracy and governance across Unified Communications platforms.

对这个岗位感兴趣?

该岗位暂未开放在线申请,顾问可为您推荐同类岗位或申请指导

咨询不收取任何费用,顾问将为您推荐合适的岗位与申请方式

申请海外岗位,英文简历符合当地格式规范吗?

AI 自动评估你与该岗位的匹配度,3 分钟出结果

免费评估简历匹配度

数据来源:Himalayas

岗位信息来源于公开渠道,版权归原作者所有