DevOps & AI Infrastructure Engineer
# DevOps & AI Infrastructure Engineer **Crost AI** · Remote · `Remote` 🕒 **Статус:** *Опубликовано: сегодня* · *Источник: Get on Board (LATAM)* --- ### About the Role - Advanced English proficiency (CEFR C1 or higher), with the ability to discuss technical issues clearly on calls and write clear documentation and work updates. - Hands-on experience deploying and maintaining production applications. - Strong working knowledge of Linux, Docker, Git, and CI/CD. - Practical understanding of networking, DNS, SSL, reverse proxies, and application hosting. - Experience administering PostgreSQL, Redis, or comparable services. - Ability to write scripts and automate tasks using Bash, Python, JavaScript, or similar languages. - Familiarity with AI coding agents and an interest in building repeatable workflows around them. - Ability to troubleshoot problems systematically and verify changes before considering work complete. - Sound judgment around credentials, access permissions, backups, and production changes. - Experience administering cloud-hosted and self-hosted infrastructure, including server setup, updates, patching, and ongoing maintenance. - Understanding of firewalls, VPNs, SSH, and secure remote access. - Practical knowledge of storage, persistent volumes, backup retention, and restoring services and data. - Ability to monitor system health and diagnose CPU, memory, disk, network, and service failures. - Ability to assess resource requirements and manage infrastructure capacity and costs. - Clear communication and the ability to turn agreed goals into working, documented implementations. - Deploy and maintain client applications across development, staging, and production environments. - Build and improve CI/CD pipelines, deployment automation, and rollback procedures. - Manage containers, databases, networking, DNS, SSL certificates, secrets, and access controls. - Set up monitoring, logging, alerts, backups, and tested recovery procedures. - Maintain internal Linux servers, development environments, and self-hosted services. - Help design and implement reusable AI-assisted development workflows for coding, testing, pull request review, and addressing feedback. - Connect AI agents and development tools with repositories, CI/CD pipelines, and internal systems. - Improve visibility into automated workflows, including progress, failures, and work requiring human input. - Troubleshoot infrastructure and deployment issues alongside developers. - Document configurations and procedures so successful setups can be reused across projects. Success means reliable applications, repeatable deployments, well-maintained internal systems, and development workflows that require less manual coordination. - Experience with Northflank, Railway, Buildkite, GitHub Actions, or Infisical. - Experience coordinating AI agents or automating development and code review workflows. - Familiarity with MCP, APIs, webhooks, and integrations between development tools. - Experience managing physical servers, virtualization, and self-hosted infrastructure. - Familiarity with infrastructure as code and reusable environment configuration. - Working knowledge of Node.js and TypeScript applications. - Experience deploying and running AI/ML models on self-hosted GPUs, including model serving, GPU configuration, and performance and resource optimization.
Наблюдалась 2026-10-09, впервые 2026-10-09, источник — Get on Board (LATAM).