Staff Machine Learning Engineer

Zendesk

Melbourne, AUhybridPosted Jul 22, 2026
Posting intelligenceActively listedReposted 24×, possible evergreen/ghost posting

Skills

kubernetessalesforcepostgrespythonkafkaredisreactcicdawsecsllmgoml

About the role

Job Description

What we have built

The Custom Agent Service handles the full agent lifecycle: an admin configures an agent with instructions, knowledge articles, and actions through a UI; a customer ticket triggers execution; the agent plans, acts through real APIs via the Integration Action Platform, and resolves the issue. The backend is Python, the infrastructure is Kubernetes on AWS, and the agent architectures range from single-pass ReAct loops to our iterative multi-plan executor.

We are onboarding our first internal and external EAP customers, so the work ships to real accounts with real tickets.

What we need help with

Agent execution core. The planning loop, tool dispatch, memory integration, and error recovery that make up the main execution path. You will work directly on the code that decides what the agent does next and make it faster, more reliable, and more capable. This includes integrating new architectures into the production path and hardening them for real traffic.

Knowledge retrieval. Agents retrieve and reason over customer knowledge bases at runtime. The retrieval pipeline (embedding, reranking, context assembly) needs to balance answer quality against latency and token cost, across thousands of heterogeneous knowledge bases per deployment.

Actions and connectors. Agents call Zendesk APIs, third-party connectors (Shopify, Salesforce, etc.), custom actions configured by admins, and increasingly other agents via A2A. The execution layer needs reliable retries, timeouts, schema validation, and graceful degradation when connectors fail mid-execution. You would also self-service new connector integrations through the Connector SDK.

Production instrumentation for model training. Every agent execution generates a trajectory (reasoning steps, tool calls, outcomes, user feedback). We are building toward training domain-specialized models, and that requires clean production data. You would instrument the execution pipeline to capture implicit reward signals (resolution success, escalation patterns, user satisfaction) that feed into the ML team's training pipeline.

Security and compliance. PII filtering, audit logging, action versioning, and governance patterns that keep agents within admin-configured bounds. You would work directly with Product Security on security review items as the platform scales.

What we are looking for

5+ years of backend engineering with strong Python skills. You have shipped production systems, not just models. You understand the difference between getting an agent to work locally and running it across 100,000 accounts.

Comfortable across the full agent stack: LLM APIs, prompt engineering, tool calling, memory management, evaluation. You can build an agent loop from scratch, and you know when a framework helps vs. when it gets in the way.

You think about what happens when the model returns garbage, the connector times out, and the customer is waiting. You build for the failure case, not just the happy path.

You ship working code, review PRs carefully, and communicate clearly about what is done, what is blocked, and what is at risk.

Tech Stack

Languages: Python (primary), some Go for platform services

Agent Frameworks: Custom iterative architectures, ReAct, with integration points to open-source tooling

Infrastructure: Kubernetes, Spinnaker, AWS (ECS, S3, ElastiCache)

Data: Postgres, ElastiCache/Redis (vector + KV), Kafka

Evaluation: Braintrust (experiment tracking, scoring, CI/CD integration)

Protocols: MCP, REST, gRPC

Why Zendesk for this work

Zendesk has 100,000+ customers, billions of support interactions, and a live product surface where agents are already resolving tickets. The feedback loop from an agent action to a measurable customer outcome is minutes, not months. We are hiring 2+ engineers at each of these levels.

The intelligent heart of customer experience

Zendesk software was built to bring a sense of calm to the chaotic world of customer service. Today we power billions of conversations with brands you know and love.

Zendesk believes in offering our people a fulfilling and inclusive experience. Our hybrid way of working, enables us to purposefully come together in person, at one of our many Zendesk offices around the world, to connect, collaborate and learn whilst also giving our people the flexibility to work remotely for part of the week.

As part of our commitment to fairness and transparency, we inform all applicants that artificial intelligence (AI) or automated decision systems may be used to screen or evaluate applications for this position, in accordance with Company guidelines and applicable law.

Zendesk endeavors to make reasonable accommodations for applicants with disabilities and disabled veterans pursuant to applicable federal and state law. If you are an individual with a disability and require a reasonable accommodation to submit this application, complete any pre-employment testing, or otherwise participate in the employee selection process, please send an e-mail to peopleandplaces@zendesk.com with your specific accommodation request.

Questions about this role

Click "Apply with AI Applyd" above. We auto-fill the application from your resume and answer screening questions in seconds. No copy and paste, no juggling tabs.

Compensation for Machine Learning Engineer roles in Australia varies widely by seniority, employer size, and remote vs onsite arrangement. Check the salary range on this listing when published, or browse our Machine Learning Engineer hub for Australia medians across recent openings.

Most applications complete in under 90 seconds. You can track the status in your dashboard and watch the screenshot proof land the moment the application submits.

AI Applyd supports Greenhouse, Lever, Ashby, Workday, iCIMS, SmartRecruiters, Personio, Teamtailor and other major ATS platforms. If we can submit through the platform, we do.

Want AI Applyd to auto-apply to roles like this?

We tailor your resume per posting, fill the forms, and track replies for you.