San Francisco, California, 94102
Job description
As a key contributor, you will design and train innovative Vision-Language Models aimed at analyzing intricate construction site data while steering the research agenda for spatial intelligence. Your responsibilities will also involve developing scalable training pipelines and enhancing inference efficiency to manage extensive video data. Applicants should possess over 6 years of practical experience in deep learning and transformer architectures, demonstrating advanced skills in Python. A degree in Computer Science, Machine Learning, or a relevant discipline is essential, paired with familiarity with prominent deep learning frameworks.
STEMHUNTER is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, sexual orientation, gender identity, and gender expression), national origin, age, disability, genetic information, veteran status, or any other status protected by applicable federal, state, or local law. We comply with all applicable equal employment opportunity and affirmative action regulations.

