Skip to main content

Staff Machine Learning Engineer, AI Security

Reddit
Remote - United StatesUpdated 5d ago
Base salary
$230k–$322k
Published base salary range
Location
Remote - United States
Remote eligibility
Employment
Full-time
Staff / Principal
Role family
Engineering
AI / ML
Apply on job-boards.greenhouse.io
Job actionsApply now
Job actionsApply now

About the job

About Reddit

Reddit is a community of communities built on shared interests, passion, and trust. It hosts 100,000+ active communities and approximately 130 million daily active unique visitors, making it one of the internet's largest sources of information.

About the Role

The AI Security team within Reddit's Security Platform Engineering organization builds security into Reddit's products, engineering systems, AI platforms, and operational infrastructure. A core part of this work is developing practical, high-quality machine learning systems that detect and prevent risks such as prompt injection, jailbreaks, sensitive data exposure, and unsafe or unauthorized AI behavior. Building on Reddit's centralized LLM Guardrails Platform, the team is expanding ML-powered protections as Reddit's AI systems and threat landscape evolve.

We are looking for a founding Staff Machine Learning Engineer to lead the development, training, and optimization of models for AI security at Reddit. This is a strategic and hands-on individual contributor role, combining deep ownership of model architecture, training data, and experimentation with technical leadership across teams.

Responsibilities

  • Select, adapt, fine-tune, evaluate, and deploy pretrained models and lightweight classifiers for Reddit-specific security problems.
  • Build reproducible training and evaluation pipelines on Reddit's ML platform, partnering with platform engineers to improve inference performance, resource efficiency, and operational reliability.
  • Set the technical vision and multi-quarter modeling roadmap, partnering with cross-functional teams to gather requirements, define model architectures, and iterate on model development.
  • Conduct model evaluations and performance analysis to improve accuracy and adversarial robustness, and define launch criteria that balance false positives, latency, throughput, reliability, and cost.
  • Own training-data quality and the production model lifecycle, using monitoring, incident findings, and red-team feedback to guide dataset improvements, retraining, and safe rollout or rollback.
  • Establish best practices for responsible ML development and deployment, including reproducible experiments, testing, model and data lineage, and privacy-aware data use.
  • Stay current with research in NLP, large language models, and relevant multimodal techniques, translating promising advances into measurable model improvements.
  • Mentor engineers and lead technical discussions and reviews, shaping the team's long-term ML capabilities and AI security direction.

Qualifications

  • 8+ years of experience developing machine learning models, with substantial hands-on model training experience, demonstrated production impact, and a record of leading complex initiatives across teams.
  • Strong background in Python programming, software engineering, and deep learning frameworks and libraries such as TensorFlow, PyTorch, or Hugging Face Transformers.
  • Deep understanding of neural network architectures and optimization, with proficiency in data preprocessing, tokenization, embeddings, language modeling, and model calibration.
  • Expertise in scalable data pipelines and distributed training frameworks such as Ray Train or PyTorch Distributed, with a strong understanding of hardware and system tradeoffs.
  • Demonstrated rigor in experimental design and model evaluation, including representative holdouts, ablation studies, adversarial tests, and error analysis to diagnose training issues, bias, and generalization gaps.
  • Excellent written and verbal communication, with the ability to explain model behavior, security risk, uncertainty, and tradeoffs to technical and non-technical partners.

Preferred Qualifications

  • Experience applying ML to security, trust and safety, fraud, privacy, or related adversarial domains.
  • Experience with adversarial training, model distillation, active learning, or synthetic-data generation to improve model quality, robustness, and training efficiency.

Compensation & Benefits

Base salary range: $230,000—$322,000 USD. This role is eligible for equity in the form of restricted stock units, and may also be eligible for a commission. Reddit offers a wide range of benefits to U.S.-based employees, including medical, dental, and vision insurance, 401(k) program with employer match, generous time off for vacation, and parental leave.

Application Instructions

Apply through the provided link.

Skills & tags

What you can verify before applying

Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.