Role: AI Software Engineering Intern – Multimodal Voice & Video
About the Role
Join the team building next-generation real-time conversational AI. You will work on cutting-edge multimodal infrastructure, merging live camera feeds and native voice-to-voice processing into modern web applications.
We are seeking an AI Software Engineering Intern to help build the future of real-time conversational intelligence. In this hands-on role, you will work on cutting-edge multimodal infrastructure, developing ultra-low-latency voice and video processing pipelines in Python and crafting fast, responsive web interfaces in Next.js. You will directly contribute to shipping native voice-to-voice and live camera capabilities, connecting custom APIs, WebSockets, and state-of-the-art AI models into production-ready software.
To excel in this position, you should have strong coding skills in Python and Next.js (React/TypeScript), alongside a solid grasp of APIs, WebSockets, or real-time data streaming.
We value high ownership, rapid execution, and a genuine curiosity for pushing the boundaries of what software can do with sight and speech. If you are eager to build at the bleeding edge of real-time multimodal AI and take full ownership of impactful features, this is the position for you.
What You’ll Do
- Develop ultra-low-latency voice and video processing pipelines in Python.
- Build fast, responsive streaming interfaces and user experiences in Next.js.
- Connect custom APIs, WebSockets/WebRTC, and multimodal AI models into core product features.
What We’re Looking For
- Strong hands-on coding skills in Python and Next.js (React / TypeScript).
- Familiarity with APIs, WebSockets, or real-time data streaming.
- High ownership, quick execution, and genuine curiosity for real-time AI tech.
📌 AI Software Engineering Intern (Montreal)
🏢 Eclatira
📍 Montreal
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.