Lead AI Engineer
Posted 20 days ago · 0 applicants
Saving, applying or scoring takes a few seconds to set up your free account.
The role in plain words
This role involves designing and building a large-scale, low-latency streaming platform for Voice AI. You will deploy proprietary speech models into production, develop APIs and SDKs for partner integrations, and handle full-stack development across frontend applications and backend microservices.
- Triton Inference Server
- React, Python, Go
- Cloud infrastructure (AWS/GCP) and containerization (Docker)
- 3-5+ years in full-stack software development with a focus on scalable systems
- Hands-on experience integrating ML models or multi-model complex AI pipelines into production web applications
Extracted from the job description · kept up to date automatically
Who this suits
This role suits experienced full-stack engineers with 3 to 5+ years of experience who have hands-on expertise integrating machine learning models or complex AI pipelines into production. It is less ideal for those without experience in real-time streaming infrastructure or cloud technologies like AWS and GCP.
Full job description
Original listing · kept for referenceClaroAI is hiring a Lead AI Engineer with full stack responsibility to build our Voice AI streaming-platform.
You will design and architect a large-scale, very low-latency system integrating our proprietary AI models and develop API interfaces that power seamless communication across our partners.
About the Role
You will work closely with our speech-AI researchers to bring proprietary company models and AI pipeline into production, ensuring optimal performance. Additionally, you will build and maintain a client-side application.
About ClaroAI
ClaroAI clarifies speech and corrects accent‑related mispronunciations, in real-time, helping everyone succeed in today’s globalized professional world.
No matter the platform, global calls are still hard because of pronunciation differences between people. ClaroAI Speech-to-Speech technology enabling everyone to communicate effectively and confidently in English.
Key Responsibilities
• Platform engineering: architect and build large-scale, low-latency infrastructure to support real-time streaming Voice AI
• AI Integration: work closely with the AI research team to deploy proprietary models into production environments
• Partner integrations: develop, scale, and maintain robust API interfaces, webhooks, and SDKs to support ClaroAI's partners
• Full-stack development: optimize both frontend user-facing application and backend microservices, ensuring system reliability and scalability
• Monitoring & optimization: Implement tracing and monitoring for end-to-end latency to deliver a seamless, high-quality user experience
Qualifications & Requirements
• Tech Stack: Triton Inference Server, highly familiar with web frameworks (e.g., React, Python, Go), cloud infrastructure (AWS/GCP), and containerization (Docker)
• Experience: proven track record (3-5+ years) in full-stack software development with a focus on scalable systems
• AI/Backend expertise: hands-on experience integrating ML models, or multi-model complex AI pipelines into production web applications
• Communication: excellent collaboration skills to work seamlessly across both technical team members and business-oriented distribution partners
Questions about this role
- This listing did not state a salary. We only show pay when the employer publishes it.
- Triton Inference Server, React, Python, Go, Cloud infrastructure (AWS/GCP) and containerization (Docker), 3-5+ years in full-stack software development with a focus on scalable systems, Hands-on experience integrating ML models or multi-model complex AI pipelines into production web applications