Role Overview
Toradex is seeking a DevOps Engineer to strengthen our cloud operations and engineering practices. The focus will be on reliable website delivery, secure AWS foundations, and fast but controlled delivery of new services. This position combines AWS operations, infrastructure as code, CI/CD, automation, and pragmatic software engineering. The role also supports on-premises to cloud migration, global service optimization, and practical responses to increasing AI-driven traffic. We value candidates who can use modern AI-assisted development effectively while applying disciplined Git, review, security, and deployment practices.
Responsibilities
- Operate and continuously improve the AWS environment with clear ownership of reliability, availability, security, and cost awareness.
- Protect website availability through sound architecture for caching, traffic filtering, DNS, origin protection, failover, capacity, and observability.
- Maintain infrastructure as code with Terraform or CloudFormation, including reusable modules, peer review, state management, environment separation, and controlled change processes.
- Design CI/CD pipelines in GitLab or similar tooling that support automated testing, approvals, rollback options, segregation of duties, and resilient delivery.
- Develop scripts, automation, and small services in languages such as Python, JavaScript, TypeScript, or Ruby to reduce manual work and improve operational consistency.
- Use AI-assisted development methods responsibly to create proof-of-concept projects quickly, while keeping source control, security checks, documentation, and deployment discipline in place.
- Monitor systems end to end, investigate incidents, perform root-cause analysis, and turn recurring issues into engineering improvements.
- Apply practical security practices across cloud accounts, pipelines, repositories, secrets, network exposure, access rights, and auditability.
- Support migration of services from on-premises environments to the cloud, including dependency analysis, cutover planning, operational readiness, and post-migration optimization.
- Improve global delivery of services through latency reduction, regional placement, edge strategy, routing decisions, and performance testing.
- Develop approaches to handle growing automated and AI-originated traffic, including detection, rate limiting, bot mitigation, cost protection, and service-quality safeguards.
- Document cloud architecture, deployment workflows, recovery procedures, ownership, and operational standards so services can be maintained sustainably.
Requirements
- Strong hands-on AWS experience across web delivery, compute, storage, databases, networking, security, and operational tooling.
- Ability to design for high availability, resilience, disaster recovery, performance, and cost efficiency in production environments.
- Proficiency with Terraform or CloudFormation, including modular design, version control, environment promotion, and safe change management.
- Very good knowledge of Git, preferably GitLab, including merge request workflows, branching strategy, protected environments, and release governance.
- Experience building CI/CD pipelines with automated checks, approvals, rollback strategies, artifact handling, secrets protection, and segregation of duties.
- Programming and scripting ability in Python, JavaScript, TypeScript, Ruby, or similar languages.