AI Research Intern
full-time
junior
Posted 1 day ago
Apply Now
Stand out: build a proof-of-work pitch →
Free GitHub-based preview. Direct apply stays one click away.
Get weekly job alerts like this →Hiring for this role?
AI Market Demand Pack ยท $29 one-time
Compare this role's skills with the full AI hiring market. Get ranked demand, salary bands, leading companies, public source URLs, and a decision brief.
About this role
๐จ OpusClip https://www.opus.pro is the world's No.1 AI video agent, built for authenticity on social media.
We envision a world where everyone can authentically share their story through video, with no expertise needed. Within just 18 months of our launch, over 10 million creators and businesses have used OpusClip to enhance their social presence.
We have raised $50 million in total funding and are fortunate to have some of the most supportive investors, including SoftBank Vision Fund, DCM Ventures, Millennium New Horizons, Fellows Fund, AI Grant, Jason Lemkin (SaaStr), Samsung Next, GTMfund, Alumni Ventures, and many more.
Check out our latest coverage by Business Insider https://www.businessinsider.com/opusclip-softbank-vision-fund-2-funding-valuation-2025-3 featuring our product and funding milestones, and our recognition as one of The Information's 50 Most Promising Startups in 2024. https://www.linkedin.com/posts/yzopus_opusclip-has-been-recognized-by-the-information-activity-7255358200395243520-zhOz/?utm_medium=member_desktop&utm_source=share
Headquartered in Mountain View, we are a team of 100 passionate and experienced AI enthusiasts and video experts, driven by our core values:
- Be a Champion Team
- Prioritize Ruthlessly
- Ship fast, Quality Follows
- Obsess over customers
Be a part of this exciting journey with us!
ABOUT THE ROLE
We're looking for an AI Research Intern to join our AI team and explore cutting-edge research across multimodal AI, LLMs, computer vision, speech, and agent systems.
You'll work across OpusClip, AgentOpus, and our next-generation AI products, collaborating closely with AI researchers and engineers to investigate emerging technologies, build research prototypes, and ship features used by millions of creators worldwide.
WHAT YOU'LL DO
AI Research & Model Development
- Research and develop deep learning models in one or more of the following areas, depending on product priorities:
- Computer vision (e.g. video enhancement, super-resolution, restoration)
- Speech & audio (e.g. speech enhancement, voice cloning, voice generation)
- Multimodal understanding and generation
- LLM post-training (e.g. SFT, RLHF, DPO)
Applied AI Engineering
- Build AI-powered product features by integrating frontier foundation models into production systems through prompt and context engineering strategies and Agent workflows (e.g., using LangChain, RAG frameworks).
- Collaborate with product and engineering teams to rapidly prototype and ship new AI capabilities across OpusClip and AgentOpus.
Model Evaluation & Benchmarking
- Design scalable evaluation pipelines for multimodal AI systems.
- Develop domain-specific benchmarks using automated evaluation methods (e.g. LLM-as-a-Judge) together with task-specific visual, audio, and language quality metrics.
Stay at the Frontier
- Keep up with the latest AI research and open-source developments.
- Reproduce state-of-the-art research and translate new advances into production-ready systems.
WHAT WE'RE LOOKING FOR
BASIC QUALIFICATIONS
- Education: Currently pursuing or recently completing a Master's degree in Computer Science, Artificial Intelligence, Mathematics, or a related field.
- Deep Learning Foundation: Solid understanding of Transformer architecture and Attention mechanisms; familiarity with mainstream generative model families (GANs, diffusion models, autoregressive models).
- Media Processing: Familiarity with media processing fundamentals (video and/or audio โ e.g., ffmpeg, codecs, signal processing basics).
- Coding Skills: Strong programming skills in Python. Familiarity with Linux development environments, Git, and data structures.
- Fluent in English with strong technical reading and writing skills, including the ability to read research papers and write technical documentation.
HANDS-ON EXPERIENCE IN ONE OR MORE OF THE FOLLOWING
- Computer Vision (especially low-level vision): e.g., Real-ESRGAN, SwinIR, BasicVSR++, or diffusion-based SR; NTIRE / AIM challenge participation.
- Voice / speech: voice cleaning (speech enhancement / denoising / separation), voice cloning (TTS / voice conversion), or voice generation.
- LLM fine-tuning: SFT, RLHF / DPO, LoRA / PEFT, or post-training of open-source models.
PREFERRED QUALIFICATIONS
- Experience building Agent Systems or LLM-powered product features with frontier-model APIs (e.g., ChatGPT, Claude, Gemini) . This role contributes to both OpusClip and AgentOpus products.
- Familiarity with TypeScript is a bonus, helpful for shipping product features.
- Ownership & execution: Involvement in projects from inception to completion, with strong coding fundamentals; open-source contributions are a plus.
- Research breadth: Academic background or interest in adjacent areas โ video understanding and generation, multimodal systems, agents, and model evaluation / benchmarking.
- Publicatio
Similar Jobs
Related searches:
On-site Jobs
Junior Jobs
On-site Junior Jobs
Junior Generative AIJunior AI Agents & RAGJunior Machine LearningJunior Computer VisionJunior NLP & Language AI
AI Jobs in Mountain View
Generative AI in Mountain ViewAI Agents & RAG in Mountain ViewMachine Learning in Mountain ViewComputer Vision in Mountain ViewNLP & Language AI in Mountain View
agentsfine-tuningdeep-learningdiffusion-modelsgenerative-aicomputer-visionnlprag
Get jobs like this delivered weekly
Free AI jobs newsletter. No spam.