Machine Learning & NLP Expert
United States
- Pay
$80–110/hour — pay source
This role is for one of our clients Compensation: $80-$110 per hour Join a cutting-edge AI research initiative and help shape the next generation of frontier AI models. We are seeking experienced Machine Learning & NLP Experts to contribute their technical expertise toward training and evaluating advanced AI systems. In this role, you will design challenging, real-world machine learning and natural language processing tasks, create high-quality reference solutions, and assess AI model performance to identify reasoning gaps and improve overall capabilities.
Read the full posting- Work setup
Remote stated — work setup source
You will be engaged as an independent contractor. This is a fully remote role that can be completed on your own schedule. Projects may be extended, shortened, or concluded early depending on business needs and performance.
Read the full posting- Employment
Part-time — employment source
Join a cutting-edge AI research initiative and help shape the next generation of frontier AI models. We are seeking experienced Machine Learning & NLP Experts to contribute their technical expertise toward training and evaluating advanced AI systems. In this role, you will design challenging, real-world machine learning and natural language processing tasks, create high-quality reference solutions, and assess AI model performance to identify reasoning gaps and improve overall capabilities. This is a part-time, fully remote opportunity requiring approximately 20 hours per week. Requirements
Read the full posting
What you’ll work on
Full postingDesign challenging, real-world machine learning and natural language processing tasks covering areas such as:
Develop accurate reference solutions and integrate tasks into agentic development environments using Python.
Build executable evaluation frameworks and testing components where appropriate.
From the employer’s posting
Key Responsibilities Design challenging, real-world machine learning and natural language processing tasks covering areas such as: Machine Learning Model Development and Evaluation
Transformer Models and Large Language Models (LLMs) Develop accurate reference solutions and integrate tasks into agentic development environments using Python. Build executable evaluation frameworks and testing components where appropriate.
Develop accurate reference solutions and integrate tasks into agentic development environments using Python. Build executable evaluation frameworks and testing components where appropriate. Evaluate AI model outputs for technical correctness, reasoning quality, and overall performance.
What you’ll bring
All qualificationsCore experience
- Strong proficiency in Python with practical experience developing ML or NLP applications.
- Experience with industry-standard frameworks such as PyTorch, TensorFlow, Hugging Face Transformers, or equivalent.
- Strong understanding of modern machine learning techniques, including:
- Ability to commit approximately 20 hours per week.
Preferred experience
- Experience in AI model evaluation, AI training data creation, or human-in-the-loop model assessment.
- Familiarity with Retrieval-Augmented Generation (RAG), vector databases, embedding models, or multimodal AI systems.
- Experience building benchmarking frameworks, automated evaluation pipelines, or testing infrastructure.
- Experience working with production-scale machine learning systems.
Qualification wording
Strong proficiency in Python with practical experience developing ML or NLP applications.
Experience with industry-standard frameworks such as PyTorch, TensorFlow, Hugging Face Transformers, or equivalent.
Strong understanding of modern machine learning techniques, including:
Ability to commit approximately 20 hours per week.
Experience in AI model evaluation, AI training data creation, or human-in-the-loop model assessment.
Familiarity with Retrieval-Augmented Generation (RAG), vector databases, embedding models, or multimodal AI systems.
Experience building benchmarking frameworks, automated evaluation pipelines, or testing infrastructure.
Experience working with production-scale machine learning systems.
Tools in this posting
- Python
- PyTorch
- TensorFlow
- Transformers
- Huggingface
Source — Tool mentions in context
- Transformer Models and Large Language Models (LLMs) - Develop accurate reference solutions and integrate tasks into agentic development environments using Python. - Build executable evaluation frameworks and testing components where appropriate.
- Deep hands-on experience in Machine Learning and/or Natural Language Processing through industry, research, or graduate/PhD-level work. - Strong proficiency in Python with practical experience developing ML or NLP applications. - Strong understanding of modern machine learning techniques, including:
- Feature Engineering and Model Optimization - Experience with industry-standard frameworks such as PyTorch, TensorFlow, Hugging Face Transformers, or equivalent. - Ability to commit approximately 20 hours per week.
Job description
This role is for one of our clients
Compensation: $80-$110 per hour
Join a cutting-edge AI research initiative and help shape the next generation of frontier AI models. We are seeking experienced Machine Learning & NLP Experts to contribute their technical expertise toward training and evaluating advanced AI systems. In this role, you will design challenging, real-world machine learning and natural language processing tasks, create high-quality reference solutions, and assess AI model performance to identify reasoning gaps and improve overall capabilities.
This is a part-time, fully remote opportunity requiring approximately 20 hours per week.
Requirements
Key Responsibilities
- Design challenging, real-world machine learning and natural language processing tasks covering areas such as:
- Machine Learning Model Development and Evaluation
- Natural Language Understanding (NLU)
- Natural Language Generation (NLG)
- Information Retrieval and Search
- Applied Machine Learning Pipelines
- Transformer Models and Large Language Models (LLMs)
- Develop accurate reference solutions and integrate tasks into agentic development environments using Python.
- Build executable evaluation frameworks and testing components where appropriate.
- Evaluate AI model outputs for technical correctness, reasoning quality, and overall performance.
- Identify capability gaps, classify model failure modes, and provide detailed written analyses.
- Create and refine evaluation guidelines, scoring rubrics, and quality standards for ML and NLP tasks.
- Collaborate with fellow subject matter experts to ensure consistency, accuracy, and high-quality training data.
Required Qualifications
- Deep hands-on experience in Machine Learning and/or Natural Language Processing through industry, research, or graduate/PhD-level work.
- Strong proficiency in Python with practical experience developing ML or NLP applications.
- Strong understanding of modern machine learning techniques, including:
- Model Training and Evaluation
- Transformer Architectures
- Large Language Models (LLMs)
- NLP Pipelines
- Feature Engineering and Model Optimization
- Experience with industry-standard frameworks such as PyTorch, TensorFlow, Hugging Face Transformers, or equivalent.
- Ability to commit approximately 20 hours per week.
- Excellent written communication skills and the ability to work independently in a remote environment.
Preferred Qualifications
- Experience in AI model evaluation, AI training data creation, or human-in-the-loop model assessment.
- Familiarity with Retrieval-Augmented Generation (RAG), vector databases, embedding models, or multimodal AI systems.
- Experience building benchmarking frameworks, automated evaluation pipelines, or testing infrastructure.
- Contributions to open-source ML/NLP projects or published research are a plus.
- Experience working with production-scale machine learning systems.
Role Details
- Employment Type: Independent Contractor
- Work Arrangement: Fully Remote
- Schedule: Approximately 20 hours per week
- Project Duration: Based on project requirements and performance, with opportunities for extension
Equal Opportunity
We consider all qualified applicants without regard to legally protected characteristics and provide reasonable accommodations upon request.
Contract & Payment Terms
- You will be engaged as an independent contractor.
- This is a fully remote role that can be completed on your own schedule.
- Projects may be extended, shortened, or concluded early depending on business needs and performance.
- Your work will not involve access to confidential or proprietary information from any employer, client, or institution.
- Payments are made weekly via Stripe or Wise based on services rendered.
- Please note: We are unable to support H1-B or STEM OPT candidates at this time.
Your next step
- Have your CV and examples of relevant work ready.
- Check the listed location, eligibility and core experience before starting.
Complete your application on apply.workable.com. The employer’s form will show what is required.
Already applied? Track this application
Source & posting history
Source notes
Source excerptsSelected passages from the saved posting. Check the full description for conditions and exceptions.
- Pay
This role is for one of our clients Compensation: $80-$110 per hour Join a cutting-edge AI research initiative and help shape the next generation of frontier AI models. We are seeking experienced Machine Learning & NLP Experts to contribute their technical expertise toward training and evaluating advanced AI systems. In this role, you will design challenging, real-world machine learning and natural language processing tasks, create high-quality reference solutions, and assess AI model performance to identify reasoning gaps and improve overall capabilities.
- Location & working pattern
United States
Join a cutting-edge AI research initiative and help shape the next generation of frontier AI models. We are seeking experienced Machine Learning & NLP Experts to contribute their technical expertise toward training and evaluating advanced AI systems. In this role, you will design challenging, real-world machine learning and natural language processing tasks, create high-quality reference solutions, and assess AI model performance to identify reasoning gaps and improve overall capabilities. This is a part-time, fully remote opportunity requiring approximately 20 hours per week. Requirements
More source context
- Ability to commit approximately 20 hours per week. - Excellent written communication skills and the ability to work independently in a remote environment. Preferred Qualifications
More relevant text appears in the full description.
- Work authorization
No clear work-authorization passage found. Eligibility is unconfirmed.
- Status in our records
- Active
- First seen by us
- Aug 15, 2026
- Recorded sightings
- 171
- Last seen by us
- Oct 9, 2026
- Employer says posted
- Jul 29, 2026
These dates show when we found the listing. Check the employer’s website to confirm it is still accepting applications.
Report an errorSee how this role fits your experience
Add your resume to compare the role’s scope, tools and requirements with your experience.