NextRaiseNextRaiseFind jobs
Sign inSign up free
Jobs / AI / ML Researcher in United States of America
3 days ago
Apply with autofill
Apply with autofill
Clera·3 days ago
3 days ago

Research Engineer, Privacy and Anonymization

San Francisco, United States of AmericaFull-timeOn-siteMid · 2+ yearsAI / ML Researcher

Sign up free to see how well your resume matches this role.

Boost your chances at clera

How you compare FREE

?
Your scoreYour score: not yet known
→
49
Top 10%Top 10%: 49 out of 100

Top 10% of NextRaise users matched against AI / ML Researcher roles in United States.

Must-have skills for this role

  • python
  • anonymization
  • information extraction
  • named-entity recognition

PDF or DOCX · no account needed

Apply faster with autofill FREEclera uses Ashby - autofill it instead of retyping.careers.example.com/applyAutofillingFull namePriya SharmaEmailpriya.sharma@example.comPhone+49 30 1234567LocationBerlGet the extension

What you'll do

  • Build systems to detect PII, quasi-identifiers, credentials, and other sensitive information, designing transformations based on data type and downstream use case.
  • Develop and benchmark detection approaches that combine rules, statistical models, classifiers, and LLM-based methods.
  • Build production pipelines that anonymize raw data before it enters downstream processing, training, evaluation, or synthetic data generation workflows.
  • Create evaluation frameworks that measure privacy risk and retained data utility, including recall-weighted metrics, leakage tests, and adversarial re-identification attempts.
  • Design systems that remain robust to new data sources, schema drift, unusual formats, and sensitive information embedded in unexpected fields.
  • Collaborate with engineering, research, operations, and customers to translate privacy requirements into practical technical policies and safeguards.

What they're looking for

  • 2+ years building production data or ML systems in Python, with strong proficiency in the language.
  • Hands-on experience with information extraction, named-entity recognition, classification, or related methods for detecting sensitive or rare content.
  • Strong experimental instincts and the ability to compare approaches across recall, precision, latency, cost, and downstream data utility.
  • Solid understanding of redaction, masking, pseudonymization, anonymization, and synthetic data generation, and when each technique is appropriate.
  • Experience designing systems that are robust to schema drift, unusual data formats, and edge cases.
  • End-to-end experience building data processing pipelines without a fully prescribed roadmap.

Nice to have

  • Familiarity with privacy-enhancing technologies such as differential privacy, k-anonymity, secure aggregation, or format-preserving encryption is a plus.
  • Experience with low-latency or high-throughput ML inference and data-processing systems is a plus.
  • Prior work with sensitive data in healthcare, finance, security, or related domains is a plus.

Summarised by NextRaise from the employer’s description, which follows in full below.

Full description from employer

About the Role

This is a Research Engineer role focused on building privacy and anonymization systems that make sensitive, real-world data safe and useful for AI training. You will own the full pipeline for protecting privacy without destroying the structure and signal that make data valuable, sitting at the intersection of applied research and production engineering in a fast-moving AI infrastructure company.

What You'll Do

  • Build systems to detect PII, quasi-identifiers, credentials, and other sensitive information, designing transformations based on data type and downstream use case.

  • Develop and benchmark detection approaches that combine rules, statistical models, classifiers, and LLM-based methods.

  • Build production pipelines that anonymize raw data before it enters downstream processing, training, evaluation, or synthetic data generation workflows.

  • Create evaluation frameworks that measure privacy risk and retained data utility, including recall-weighted metrics, leakage tests, and adversarial re-identification attempts.

  • Design systems that remain robust to new data sources, schema drift, unusual formats, and sensitive information embedded in unexpected fields.

  • Collaborate with engineering, research, operations, and customers to translate privacy requirements into practical technical policies and safeguards.

What We're Looking For

  • 2+ years building production data or ML systems in Python, with strong proficiency in the language.

  • Hands-on experience with information extraction, named-entity recognition, classification, or related methods for detecting sensitive or rare content.

  • Strong experimental instincts and the ability to compare approaches across recall, precision, latency, cost, and downstream data utility.

  • Solid understanding of redaction, masking, pseudonymization, anonymization, and synthetic data generation, and when each technique is appropriate.

  • Experience designing systems that are robust to schema drift, unusual data formats, and edge cases.

  • End-to-end experience building data processing pipelines without a fully prescribed roadmap.

  • Familiarity with privacy-enhancing technologies such as differential privacy, k-anonymity, secure aggregation, or format-preserving encryption is a plus.

  • Experience with low-latency or high-throughput ML inference and data-processing systems is a plus.

  • Prior work with sensitive data in healthcare, finance, security, or related domains is a plus.

Location

On-site in San Francisco, California. Visa sponsorship is available.

Company

Clera
San Francisco, United States of America

Company facts come from this company's own listings. We only show what the postings themselves carry.

Sourced from Clera's careers site·first seen 17 Sept 2026·last verified 17 Sept 2026·How we source jobs

Similar jobs

  • Applied Researcher – Network Expert at designworkstalentBellevue, United States of America–match not yet calculated
  • Applied Researcher – Data Centers at designworkstalentBellevue, United States of America–match not yet calculated
  • Applied Researcher – AI Expert at designworkstalentBellevue, United States of America–match not yet calculated
  • Research Scientist, Networking Research - PhD New College Grad 2026 at NVIDIASanta Clara, United States of America–match not yet calculated
  • Researcher Senior (Health Services) at elevancehealthWILMINGTON, United States of America–match not yet calculated

Browse more jobs

  • AI / ML Researcher jobs in United States
  • Machine Learning Engineer jobs in United States
  • AI Engineer jobs in United States
  • Computer Vision Engineer jobs in United States
  • AI / ML Researcher jobs in United Kingdom
  • AI / ML Researcher jobs in Canada