Back to all jobs
H
Multimodal AI Systems Architect (AI Engineering)
Hyphen Connect Limited
San Francisco Bay Area1mo ago
About the role
<p>We are seeking a talented Multimodal AI Systems Architect to develop and optimize AI systems that seamlessly integrate vision and audio models. This role focuses on enhancing our voice-to-voice interactions and multimodal retrieval capabilities, ensuring our systems are efficient and innovative.</p>
<p> </p>
<p><strong>Responsibilities:</strong></p>
<ul>
<li>Integrate vision encoders and audio-native models into core agent reasoning loops.</li>
<li>Optimize streaming latency for voice-to-voice AI interactions.</li>
<li>Architect multimodal RAG systems capable of retrieving insights from videos and PDFs.</li>
</ul>
<p><strong>Qualifications:</strong></p>
<ul>
<li>Experience with Whisper, CLIP, and multimodal LLM integration.</li>
<li>Knowledge of streaming architectures and WebRTC.</li>
<li>Expertise in cross-modal alignment.</li>
</ul>
753,000+ hidden jobs like this
Hyphen Connect Limited and thousands of companies post here first — often days before LinkedIn or Indeed. Your first 5 applications are free; go Pro to apply without limits.
Everything Pro unlocks:
- Unlimited applications — free stops at 5
- Track every application in one place
- Apply straight to the source, one click
- Save & organize roles you love
- Roles pulled from company boards before the big sites