About this Devops Engineer - Torizon Cloud role at Toradex
Toradex is a global company strongly focused on engineering & technology. We’re powered by a diverse & uniquely gifted workforce. We pursue the best people to propel our innovative vision of embedded computing and IoT. If you’re interested in being a driving force at an agile technology company, engineering clever computing solutions & helping other companies bring their products to life, we should talk.
About the team
Torizon Cloud is a platform that delivers over-the-air software updates to embedded Linux devices in the field. Our users are companies running fleets of industrial and robotics hardware. They use our platform to publish updates on physical machines they can't easily reach so reliability is a key selling point of our product.
The team is small, distributed, and motivated! You'd be joining a group of engineers who own the platform end to end - the backend services and the infrastructure they run on.
We have big aspirations as a team and our product is scaling!
What you'll do
- Own and care for the Kubernetes deployments of our Torizon OTA services. We have various services (mostly scala services), Kafka/Red-Panda messaging system.
- Manage infrastructure as code with Terraform/OpenTofu.
- Build and maintain CI/CD pipelines, container builds, automated testing, and promotion across staging and production environments.
- Push our devops practice forward: declarative, reviewed, reproducible environments, and fewer things that depend on someone remembering a manual step.
- Improve observability and how we handle incidents: metrics, logs, alerting, and the follow-through after something breaks.
- Help us expand and evolve our testing frameworks and pipelines.
- Work with fellow team-mates to evolve deployability, database migrations, safe rollouts, and outage investigations.
What we're looking for
- Advanced English communication skills, written and spoken
- Available to work Hybrid from Toradex Office in Campinas/SP.
- Kubernetes in production. Not just familiarity with
kubectlbut someone who has run real workloads, managed resources, and debugged something misbehaving under load. - Helm Management: Utilize Helm to automate and streamline the deployment of applications and services to Kubernetes clusters. Create, maintain, and manage Helm charts for production-ready deployments.
- Docker and container fundamentals, including building lean, reproducible images.
- Terraform or OpenTofu at a level where reviewing someone else's module feels routine.
- Kafka, Red-panda, or similar, this is the glue that holds our system together and is a key component to ensuring reliability.
- AWS Infrastructure Management: Build, manage, and optimize AWS cloud infrastructure, including EKS, S3, VPCs, RDS, IAM, R53 and more. Implement best practices for cost management, scaling, and security within AWS.
- CI/CD pipeline design and maintenance. We use GitLab CI; equivalent experience elsewhere transfers fine. Proficiency with Git and Git-based workflows.
- Working scripting ability — Python, Ruby, Bash
- Infrastructure as code and GitOps as your default, not as something you'd have to be talked into.
- Incident Management & Troubleshooting: Respond to incidents, troubleshoot, and resolve system issues related to performance, availability, and security in a timely and effective manner.
- Self-direction. The team is remote. You'll need to pick up a loosely-defined problem, decide on an approach, and keep people in the loop in writing - without being managed hour to hour.
Nice to have
- MYSQL and Postgresql experience, especially debuging and tuning for optimizations
- Nix or NixOS.
- Exposure to embedded Linux, Yocto, or OTA update systems.
- Identity and access tooling — Keycloak, OIDC.
- Secrets management (Hashicorp vault), supply-chain security, and artifact signing.
- Bachelor’s degree in Computer Science, Computer Engineering, or a related field.
Why this role
The work is genuinely interesting! And the team is genuinely fun!
Our team is small enough that you will quickly be working on very important pieces of our stack and you will have a big say in how things are built. There are many opportunities to learn new things, and bring in new ideas to the team. We try to embody the ethos of a technology-first team where we focus on building the best software and no single person dictates how everything is built.
There's real room to improve how we deploy and manage infrastructure, our product is scaling which makes for interesting work rather than some project stuck in “maintenance mode”.
What do we offer
- An agile and intercultural work environment
- Working with latest IoT computer technology in an international environment
- Participation in defining processes and your working environment
- Contemporary employment conditions, modern office space and a flexible working environment
- Opportunities for your personal development
- Meal allowance (flash card)
- Health and dental care
- Flexible working hours
- Parking space
- Anniversary day off
Disclaimer: This is not a temporary role. It is a long-term, full-time opportunity.
Toradex is a leader in embedded computing, serving innovative products to industrial, medical, automotive & IoT companies creating feature-rich & intelligent systems for demanding applications, e.g. supercars, self-driving tractors, patient monitoring systems – to name a few.
We provide effective & robust embedded computing solutions and strive for the best development experience in the industry with a focus on intelligent hardware design, innovative software solutions & free comprehensive support. Our relationship with our customers ensures that we all succeed and it allows us to participate in the realization of incredible new products. Our products are directly sold to more than 3'000 industrial customers in over 70 different countries worldwide.
For more information
- Visit our Homepage: www.toradex.com
- Connect with us and join our community!: https://community.toradex.com
- Our YouTube Channel