About the role#
This internship position is within the algorithm team, which focuses on developing advanced computer vision, NLP, and multimodality models. The primary objective is to protect the platform and users by identifying content and behaviors that violate community guidelines. As a PhD intern, you will contribute to research and product development while collaborating with industry experts.
What you'll do#
- Develop computer vision or multimodal models to detect violation content.
- Explore cutting-edge large models, including CLIP, COCA, ALBEF, BLIP, Flamingo, ViT-G, ViT-22B, and EVA-enormous.
- Investigate the application of Large Language Models (LLMs) in business scenarios such as pre-training, zero-shot/few-shot learning, and hard case mining.
- Optimize training frameworks to improve the efficiency and adaptability of large model training.
What you'll need#
- Currently a PhD candidate in Computer Science or a related technical discipline.
- Research experience in computer vision, multimodality, or LLMs.
- Proficiency with at least one deep learning framework, such as PyTorch or TensorFlow.
- Strong analytical, problem-solving, logical thinking, and communication skills.
- Published papers in top AI conferences or journals (e.g., CVPR, ICCV, ECCV, NIPS, ICML, ICLR, TPAMI, IJCV) are considered a plus.
Location & details#
- Location: San Jose, California.
- Term: Fall 2026.
- Modality: On-site.
- Employment Type: Full-time, paid internship.
- Compensation: $60 per hour.


