Multimodal Researcher
You will research how models learn from and respond across text, images, audio, and other input types.
About the role
The role covers research methods, datasets, training and evaluation, error analysis, and product-focused experiments for multimodal behavior.
What you'll do
- Develop and evaluate methods for multimodal learning and behavior.
- Design datasets, rubrics, and evaluations for realistic multimodal tasks.
- Run experiments, diagnose failures, and test targeted improvements.
- Collaborate with research, engineering, design, product, and safety partners.
What we're looking for
- A strong machine learning research background with experience in multimodal systems.
- Experience in one or more areas such as vision-language models, audio, video, image generation, or multimodal reasoning.
- The ability to design clean experiments, reliable evaluations, and useful metrics.
- Strong research and engineering judgment on open-ended problems.
Preferred experience
- Experience moving multimodal research into a product or production evaluation.
Equal opportunity
Entlegnant, the company behind Leiolai, is an equal opportunity employer. We consider qualified applicants without regard to race, color, religion, sex, gender identity or expression, sexual orientation, national origin, ancestry, age, disability, veteran status, genetic information, or any other status protected by law.
Applicants must be at least 18. If you need an accommodation during the application process, email talent@leiolai.com.
Apply
Submit your resume and a few details below. A cover letter is optional.