Responsibilities
- Embed with enterprise and high-growth technology customers to understand technical requirements, AI use cases, and integration challenges.
- Architect and build custom production solutions using Telnyx Voice, Messaging, Wireless, WebRTC, AI, and Edge Compute platforms.
- Build and deploy AI Voice Assistants, conversational AI, LLM-powered applications, and production LLM infrastructure.
- Design model routing, fallbacks, observability, cost strategies, RAG pipelines, prompt architectures, evaluations, and model integrations.
- Deploy containerized services on Kubernetes and collaborate with customer infrastructure teams on production environments.
- Troubleshoot complex issues across APIs, networking, SIP, WebRTC, AI inference, and customer infrastructure.
- Lead technical discovery, architecture sessions, POCs, pilots, and production launches from whiteboard through go-live.
- Own customer outcomes through deployment, stabilization, and handoff.
- Partner with Product and Engineering to turn customer feedback into product improvements.
- Create technical documentation, runbooks, and reference architectures.
- Help customers address data residency, security, privacy, and compliance requirements across APAC.
Requirements
- A CS degree or equivalent practical experience.
- At least 3 years of experience building software or infrastructure, or working in technical consulting.
- Strong proficiency in one or more of Python, Node.js, or Go.
- Hands-on experience deploying LLM applications or AI infrastructure in production.
- Experience with LiteLLM or a comparable LLM gateway, including routing, fallbacks, rate limits, and observability.
- Comfort with Kubernetes, Docker, APIs, cloud infrastructure, and production operations.
- Exposure to SIP, WebRTC, VoIP, real-time communications, or messaging.
- Understanding of production AI concerns including latency, concurrency, provider limits, retries, streaming, and cost.
- Ability to diagnose complex technical problems and translate customer requirements into practical solutions.
- Ability to work directly with customers and lead technical architecture discussions.
- Strong communication skills for explaining complex technology to technical and non-technical audiences.
- Based in Sydney or willing to relocate.
- Legally authorized to work in Australia or eligible for sponsorship.
- Bonus experience with AI Voice Assistants, conversational AI, STT/TTS, real-time AI, open-weight models, vLLM, SGLang, TGI, Ollama, LLM fine-tuning, RAG, evaluation, inference optimization, GPU infrastructure, telecom, CPaaS, cloud or AI infrastructure, high-growth SaaS, enterprise or on-premises infrastructure, regulated industries, APAC data residency, DevOps, security, or multiple APAC markets.
Benefits
- Sydney-based hybrid work arrangement.
- Approximately 10–30% travel, primarily across Australia and New Zealand, with occasional travel elsewhere in APAC.
- Sydney relocation may be supported through willingness to relocate and potential Australian work sponsorship, subject to eligibility.
Tech Stack
Categories
About Telnyx
Infrastructure for real-time AI. Edge compute, voice AI agents, and global communications, all on one platform. We combine global telephony, dedicated AI infrastructure, and full customizability under one roof so you can design and deploy AI-powered agents that feel like part of your team. From low-latency voice streaming to scalable, multi-language support, Telnyx gives you everything you need to build real-time, intelligent voice experiences. Whether you’re enhancing customer support, automating outbound calls, or embedding voice AI into your product, Telnyx makes it easy to launch and scale with confidence. And you never have to worry about piecing together multiple providers or sacrificing performance. Build agile. Launch faster. Scale globally. All on one platform built for real-time engagement.