About the company
Pinterest is a platform where millions of people find creative ideas and plan for memories. The Pinterest Labs organization focuses on applied ML research and development, working on LLMs/VLM, agent design, core computer vision, multimodal representation learning, visual generative modeling, recommender systems, and graph learning.
Responsibilities
- Build state-of-the-art visual encoders, VLMs, and diffusion models that power Pinterest's visual AI capabilities.
- Experiment with billion-scale image datasets, backed by large-scale GPU computing.
- Build flexible visual reasoning tools such as composed image retrieval, promptable image feature computation, instruction-tuned embedding and generative editing models.
- Read research papers, participate in group discussions, and brainstorm the company's overall AI strategy.
- Help construct data agents to build training data shared across multimodal representation, composed image retrieval, image-editing generation, and visual language modeling.
- Collaborate with product and infrastructure engineers to ship new capabilities.
- Publish and share work through conferences like CVPR and KDD, paper submissions, and blog posts.
- Mentor junior researchers and research interns.
Requirements
- Experience building and training large scale vision models of all categories.
- Experience with multimodal representations and visual language modeling is strongly preferred.
- Track record of research contributions (e.g., publications, open-source work) and/or shipping ML models to production.
- Hands-on experience with large-scale model training and modern deep learning frameworks (e.g., PyTorch).
- Strong collaboration skills and ability to work effectively in a small, fast-moving team.
- M.S. or PhD in Machine Learning or related academic areas, or equivalent work experience.
- Experience using AI-accelerated research tooling akin to auto-research, data agents, etc.