AI Data Engineer
Apply NowJubil ID: 9
Company: VirtualVocations
Location: Lowell, MA
Description:
Join our missionNebuly AI is a well-funded
fast-growing VC-backed startup in the emerging AI sector. Nebuly is not only pioneering but defining a new market category: LLMs user analytics. Recognizing the immense potential of Generative AI
we are ambitiously charting unexplored territories
crafting a unique niche that no company has ventured into before.Thanks to Nebuly
companies can automa
Qualifications:
Minimum of 3 years of professional experience in training and evaluating Neural Networks
with a particular emphasis on NLP tasks
Experience in training large deep learning models on multiple GPUs
Professional experience in using Python and common deep learning frameworks as PyTorch
HuggingFace Transformers and vLLM
Strong knowledge and experience with microservices architecture and building scalable systems
Strong knowledge of PostgresSQL and experience in working with ORMs like SQLAlchemy
Proficiency in designing and implementing RESTful APIs
Strong knowledge and experience with testing practices and frameworks
such as unit testing
integration testing
and end-to-end testing
using libraries like Pytest
Understanding of DevOps practices and CI/CD pipelines
For technical roles
this will include a practical or coding interview
This helps us understand how you approach unfamiliar problems
Behavioral Test
Benefits:
Regardless of your seniority
your contributions will have a direct impact—from the earliest ideas to product launch
This means your role won’t stay static—you’ll keep growing with the product
facing new technical and strategic problems as we scale
Competitive compensation & stock options
We offer a competitive salary
tailored to your experience and location
In addition
you’ll have access to stock options so you can share in the value we’re creating together
Remote-friendly & flexible
Responsibilities:
As an AI Engineer
you will oversee data labeling processes and ensure the quality and consistency of training datasets critical to model performance
You will fine-tune proprietary large language models
