Mirage is an AI-native video platform seeking a Research Engineer to develop agentic systems and large language models for multimodal creative tasks, particularly video. The role focuses on advancing agent capabilities, improving reasoning and control, and enabling new interactions between language models and time-based media. You'll design end-to-end agentic pipelines, fine-tune LLMs, and build evaluation frameworks for video analysis and editing.
Responsibilities
Design and build end-to-end agentic systems for creative tasks
Develop novel approaches for training and adapting the large language models that power these agents
Design new objectives, datasets, and fine-tuning strategies to improve agent behavior and reliability
Explore multimodal reasoning and structured generation for creative control
Run systematic experiments to evaluate and improve agent performance in real-world tasks
Design evaluation frameworks for agentic workflows in video analysis and editing
Analyze failure modes across the full agent loop (planning, tool use, execution) and iterate on improvements
Requirements
BS/MS/PhD in CS, ML, or related field
Strong track record building production ML systems or agentic pipelines
Deep understanding of transformers and modern LLM techniques
Experience with fine-tuning, alignment, or post-training methods, especially for adapting models to generate structured outputs or drive tool use
Mirage (formerly known as Captions) is an AI-focused video generation and editing platform that leverages frontier research to create photorealistic AI video. The company provides a suite of tools that allow users to generate videos from text prompts and edit content using AI-powered features.