GenAI Research Engineer Intern
Appen
Job Overview
Who's the hiring manager?
Sign up to PitchMeAI to discover the hiring manager's details for this job. We will also write them an intro email for you.

Job Description
About Appen
Appen has been a leader in AI training data for over 30 years. We specialise in human generated data to train, fine tune, and evaluate models across generative AI, large language models, computer vision, and speech recognition. Our AI assisted data annotation platform and global crowd of more than 1 million contributors in over 200 countries support model pre training, supervised fine tuning, evaluation and benchmarking, safety and red teaming, and multilingual global expansion.
Why Join This Team
Appen’s GenAI research team advances how frontier models are evaluated, improved, and deployed in production environments.
The purpose of this role is to design and implement research and engineering workflows that strengthen model performance, create new benchmarks, and improve production models without regressing on core characteristics.
This role provides hands on ownership of training and evaluation pipelines, benchmark development, and model improvement initiatives that directly influence deployed systems.
Your Impact as a GenAI Research Engineer Intern
- Design and implement a lightweight supervised fine tuning training pipeline using open source LLMs.
- Create new benchmarks to evaluate frontier models across defined scientific and performance criteria.
- Analyze production models to identify measurable areas for improvement.
- Improve model performance through targeted retraining and hyperparameter search.
- Deploy improved models while maintaining core model characteristics and avoiding regression.
- Build Python tooling to automate training, evaluation, benchmarking, and experimentation workflows.
- Implement structured evaluation methods, including rubric based scoring and LLM as a judge workflows.
- Document experimental design, benchmark methodology, and performance results with clarity and precision.
- Iterate rapidly in a research driven environment to increase model quality and reliability.
What You Bring
- Current enrollment in or recent completion of a Master’s or PhD in Computer Science, AI, Machine Learning, Computer Engineering, or a closely related technical field.
- Strong experience working with large language models, including supervised fine tuning, prompt engineering, or model evaluation.
- Hands on experience building machine learning pipelines or research infrastructure.
- Experience improving model performance through retraining or hyperparameter tuning.
- Proficiency in Python and comfort working with machine learning frameworks and open source model ecosystems.
- Familiarity with cloud environments such as AWS or Azure.
- Strong technical problem solving ability, including use of LLMs as development aids for building and iteration.
- Ability to work independently with minimal hand holding.
- Strong written communication skills for summarising research and drafting technical documentation.
- Ability to collaborate effectively in a remote research environment.
Additional Details
- Duration: June-August
- Schedule: Full-time
- Work Type: Remote
Why You'll Love Working Here
At Appen, we foster a culture of innovation, collaboration, and excellence. We value curiosity, accountability, and a commitment to delivering the highest quality AI solutions for frontier models.
You’ll work on complex challenges that shape the future of AI across industries and geographies, alongside talented people in a culture that values humility over ego. You’ll have the flexibility to deliver in a way that works for you and your team, supported by tools, resources and development opportunities to continue to build your capability over time.
Key skills/competency
- Large Language Models
- Supervised Fine-tuning
- Prompt Engineering
- Model Evaluation
- Machine Learning Pipelines
- Python Programming
- Cloud Environments (AWS/Azure)
- Hyperparameter Tuning
- Research & Experimentation
- Technical Documentation
How to Get Hired at Appen
- Research Appen's culture: Study their mission, values, recent news, and employee testimonials on LinkedIn and Glassdoor, focusing on their AI leadership.
- Tailor your resume: Highlight ML pipeline experience, LLM expertise, cloud skills (AWS/Azure), and any research publications relevant to GenAI for Appen.
- Showcase GenAI projects: Provide portfolio examples of fine-tuning, benchmarking, or prompt engineering, demonstrating direct impact on model performance.
- Prepare for technical deep-dives: Expect rigorous questions on Python, ML frameworks, model optimization, and experimental design specific to Appen's GenAI focus.
- Demonstrate collaboration: Emphasize your ability to work independently, communicate complex ideas, and collaborate effectively in a remote research setting.
Frequently Asked Questions
Find answers to common questions about this job opportunity
Explore similar opportunities that match your background