Job description
**Job Title: Vision-Language Model Researcher**
In this role, you will be responsible for crafting, training, and refining vision-language models to analyze intricate data from construction sites. You will also oversee the complete research process, covering everything from data preparation and performance evaluation to the deployment of models in production settings. Applicants should possess 2-4 years of practical research experience in deep learning, focusing on transformer-based architectures. A strong command of Python and familiarity with contemporary deep learning frameworks such as PyTorch or JAX is essential, along with a degree in Computer Science, AI, or a related discipline.
Ironsite AI provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.


