← WorkMundi · 1M+ jobs from around the world, liveSign inCreate free account

AI Infrastructure Engineer

Palona AI · Toronto, Ontario, Canada

🌐 Remote📅 12/08/2026
🔔 Alert me about jobs like this
No password, no sign-up. Just the email — and you can leave the list anytime.
🔓 Apply — free →
Opens this job on WorkMundi. The account is free and takes under a minute.

See the other 43,945 jobs in Canada →

🎁 Before you apply, rehearse this interview. Create your free WorkMundi account and get an Interview Training on HelpsYouSpeak — no cost, no card. I want my training →
Palona's AI agents operate continuously in production, handle real-time guest interactions, integrate with restaurant systems, and face sharp traffic peaks. Infrastructure is therefore part of the product: latency, reliability, deployment safety, observability, security, and cost directly shape the guest and operator experience. We are looking for an Infrastructure Engineer who combines cloud and reliability depth with strong software engineering judgment. You will build and operate the platform beneath Palona's AI products, improve how engineers ship, and turn production signals into durable system improvements. This is not a ticket-driven IT or operations role. You will write production code, design systems, automate repetitive work, and own outcomes across the full service lifecycle. Our current environment includes Python services, Docker, AWS and selected Azure services, ECS and Lambda workloads, API Gateway, load balancers, relational data systems, OpenTofu/Terraform, Datadog, and CI/CD automation. We value the ability to learn and make sound tradeoffs more than exact tool-for-tool matching. What you will own: Design, build, and evolve secure, scalable cloud infrastructure for real-time AI services and customer-facing applications Improve service reliability through clear SLOs, actionable observability, capacity planning, failure testing, and pragmatic incident prevention Build deployment and release systems that make production changes fast, repeatable, auditable, and safe Own infrastructure as code, environment consistency, and reusable platform patterns across development, staging, and production Partner with product and AI engineers on architecture, performance, data flows, and operational readiness for new capabilities Diagnose complex distributed-system failures across application, network, database, model-provider, and third-party integration boundaries Reduce infrastructure and model-serving cost without compromising customer experience or engineering velocity Strengthen secrets management, access controls, backup and recovery, vulnerability management, and other practical security foundations Build internal tooling and paved paths that let engineers ship and operate services with less manual work Participate in incident response and turn incidents into better systems, automation, documentation, and engineering judgment Requirements 3+ years industrial experience in relevant technical domain Strong software engineering fundamentals and experience building or operating production distributed systems Hands-on experience with a major cloud platform; AWS experience is especially relevant Experience with containers, infrastructure as code, CI/CD, monitoring, alerting, and production debugging Ability to write reliable automation and services in Python or another modern programming language Sound judgment around availability, latency, scalability, security, and cost tradeoffs A track record of taking ambiguous operational problems from diagnosis through durable resolution Clear communication during architecture reviews, launches, and incidents AI-native working habits and curiosity about the operational behavior of LLM- and agent-powered systems Benefits Competitive Salary and Stock Option Plan Medical, dental, vision, retirement, leave, and disability benefits as applicable Family Leave Short Term & Long Term Disability Paid time off and company holidays Learning and development support
Read the rest of the job →
For people searching Engineer

144,883 engineer jobs are open right now

Here's how to pick the right one and stand out in your application.

144.883Jobs
31.687IN
81%EN

That number is real. WorkMundi's database shows 144,883 open engineer roles across the world. India has the most with 31,687 jobs, followed by the United States with 30,084. If you just finished reading one job ad and felt paralyzed by choice, you're not alone—but this scale is actually an advantage. It means you can afford to be selective.

Start by geography and language. The majority of engineer ads—117,837 of them—have the job posting text written in English. Use that as one filter, but remember: the ad text language tells you nothing about whether the role actually requires you to speak English day-to-day. Read the job description carefully. Then check which countries have the volume you're targeting. Singapore, Poland, and Australia round out the top five after India and the US.

Next, learn who's hiring. Accenture has posted 2,801 engineer roles. andurilindustries, speechify, and jobgether are also actively recruiting. If you're applying to one of these names, research their hiring patterns and interview style before you apply. That homework pays off.

When you interview, expect the question every engineer hears: 'Tell me about a time you had to debug a problem that wasn't in your job description.' Have a specific story ready—not a general one. Name the tools, the deadline pressure, and what you learned. Hiring managers listen for whether you see problem-solving as part of the role itself, not a favour.

👁 21 have read this
0 comments
Want to comment?

Leave your e-mail to comment, react and follow the posts for your role. It is free.

Similar jobs

Job on WorkMundi — the world's largest job board. See more jobs from every continent, updated live.

📢
🎁

Before you apply, rehearse this interview.

Create your free WorkMundi account and get an Interview Training on HelpsYouSpeak — no cost, no card.

I want my training →