Why This Role Stands Out
Gain invaluable experience in cutting-edge AI research at NewsBreak, a recognized innovator in the content intelligence space, by contributing to the development of advanced large language models. This hands-on role is perfect for intellectually curious individuals eager to drive experiments and propose novel ideas, offering a unique opportunity to explore foundational AI applications. Apply now to be part of a high-growth company shaping the future of content.
Quick Overview
Job Description
About NewsBreak
Founded in 2015, NewsBreak is the Content Intelligence platform shaping the future content economy. With over 40 million monthly active users, our flagship platform delivers highly personalized local news and information powered by advanced AI, recommendation systems, and adtech.
Recognized by Fast Company as #32 on the Top Workplaces for Innovators, we're proud to be Great Place to Work® certified and home to a dynamic team of technologists, product innovators, and business leaders who are passionate about solving meaningful challenges at scale.
Together, we reached unicorn status in 2021, and we remain committed to continuing this high-growth trajectory with the right team to fulfill our mission: building the infrastructure layer for content intelligence.
If you’re inspired to dream big, innovate fast, and make a difference, we’d love to hear from you! For more information, visit www.newsbreak.com/about
About the Role
We are looking for a Research Intern to join our Agent RL Training team. You will be paired with a full-time employee as your mentor, working together to explore, from zero to one, how to apply large language models to NewsBreak’s core business, including content understanding, recommendation, agentic web browsing, and autonomous multi-step task completion.
This is a hands-on research role. You are expected to independently drive experiments, propose novel ideas, and iterate quickly. We value self-starters with deep intellectual curiosity and the drive to push boundaries in LLM post-training and agent capabilities.
Location: Onsite in Mountain View, CA office
What You’ll Work On
- Collaborate with your full-time mentor to identify high-impact research directions for applying LLMs to NewsBreak’s products
- Independently run end-to-end SFT experiments on LLM-based agents, and assist with RL-related exploration such as reward design and training iteration
- Curate and build high-quality training datasets: instruction-following, preference pairs, agent trajectories, and synthetic data
- Contribute to public publications; we encourage and support top-venue submissions during your internship
What We’re Looking For
Requirements
- Highly motivated and committed: willing to put in extra hours when needed to push projects across the finish line
- Genuine passion for research: you read papers for fun, tinker with models on weekends, and care deeply about advancing the field
- Independently capable of end-to-end model SFT: with basic understanding of RL-based post-training methods (RLHF, DPO, PPO, GRPO, etc.)
- Excellent taste in model behavior: able to reason about what “good” looks like across user-facing domains and articulate why
- Strong Python and PyTorch skills
Preferred Qualifications
- Publication at a top-tier venue (NeurIPS, ICML, ICLR, ACL, EMNLP, or equivalent)
- Experience with multi-node distributed training (FSDP, DeepSpeed, Megatron-LM)
- Proficiency in writing custom GPU kernels with Triton or CUDA
- Experience building synthetic data pipelines for agent training
- Familiarity with open-source RL frameworks: TRL, OpenRLHF, veRL/vLLM
Hourly Pay: $35- $50
Similar jobs
- PP
Summer 2027 Quantitative Research Intern - PhD/Postdoctoral
NewAuto ApplyPDT Partners
New York🇺🇸$200k/yrHybrid7 hours agoMATLABC++PythonHealthcare - NA
Senior Research Associate
NewAuto ApplyNatera
Austin🇺🇸$69.7k - $87.1k/yrHybrid22 hours agoMicrosoft ExcelMicrosoft Office - NA
Research Associate
NewAuto ApplyNatera
Austin🇺🇸Hybrid22 hours agoComplianceForecastingHIPAA+1 - HA
Research Specialist
NewAuto ApplyHarbor
Remote🇺🇸RemoteYesterdayIntellectual PropertyLegal ResearchLexisNexis+3 - EZ
Health Practice Researcher, Executive Search
NewAuto ApplyEgon Zehnder
Chicago, Illinois🇺🇸HybridYesterdayBusiness DevelopmentMarket ResearchLegal Research - NA
Sr Research Associate
NewAuto ApplyNatera
Austin🇺🇸$69.7k - $87.1k/yrHybridYesterdayComplianceHIPAA - NB
Researcher, TODAY (Contract)
NewAuto ApplyNBCUniversal
New York🇺🇸$26 - $31/hrHybridYesterdayManufacturing - SA
Research Fellow
NewAuto ApplySnorkel AI
New York City🇺🇸$85 - $165/hrRemoteYesterdayLLM - GE
Research Associate
NewAuto ApplyGenScript
Redmond🇺🇸$75k - $90k/yrHybridYesterday - BR
R&D Technician
Auto ApplyBrinc
Seattle🇺🇸On-site1 week agoDroneRoboticsAssembly+1 - EG
Quantitative Research Intern
NewAuto ApplyEngineers Gate
New York🇺🇸$100k - $130k/yrHybrid6 hours agoMachine LearningPythonRisk Management - BI
Research Associate II, High-Throughput Cell Biology
NewAuto ApplyBiohub
Chicago🇺🇸$70k - $87k/yrHybrid2 days ago