Member of Technical Staff, Data Engineering jobs in United States
cer-icon
Apply on Employer Site
company-logo

Cohere · 4 days ago

Member of Technical Staff, Data Engineering

Cohere is on a mission to scale intelligence to serve humanity by training and deploying frontier models for AI systems. As a Data Engineer specializing in pretraining data, you will develop and manage the data pipeline crucial for the performance of Cohere's advanced language models, ensuring high quality and diverse datasets.

Artificial Intelligence (AI)Foundational AIGenerative AIMachine LearningNatural Language Processing
check
Comp. & Benefits
check
H1B Sponsor Likelynote

Responsibilities

Design and build scalable data pipelines to ingest, parse, filter, and optimize diverse web datasets
Conduct data ablations to assess data quality and experiment with data mixtures to enhance model performance
Develop robust data modeling techniques to ensure datasets are structured and formatted for optimal training efficiency
Research and implement innovative data curation methods, leveraging Cohere’s infrastructure to drive advancements in natural language processing
Collaborate with cross-functional teams, including researchers and engineers, to ensure data pipelines meet the demands of cutting-edge language models

Qualification

PythonData pipelinesApache SparkData modelingCollaborationProblem solving

Required

Strong software engineering skills, with proficiency in Python and experience building data pipelines
Familiarity with data processing frameworks such as Apache Spark, Apache Beam, Pandas, or similar tools
Experience working with large-scale web datasets like CommonCrawl
A passion for bridging research and engineering to solve complex data-related challenges in AI model training

Preferred

Bonus: paper at top-tier venues (such as NeurIPS, ICML, ICLR, AIStats, MLSys, JMLR, AAAI, Nature, COLING, ACL, EMNLP)

Benefits

An open and inclusive culture and work environment
Work closely with a team on the cutting edge of AI research
Weekly lunch stipend, in-office lunches & snacks
Full health and dental benefits, including a separate budget to take care of your mental health
100% Parental Leave top-up for up to 6 months
Personal enrichment benefits towards arts and culture, fitness and well-being, quality time, and workspace improvement
Remote-flexible, offices in Toronto, New York, San Francisco, London and Paris, as well as a co-working stipend
6 weeks of vacation (30 working days!)

Company

Cohere

twittertwittertwitter
company-logo
Cohere is an enterprise AI firm developing secure and private AI technology to address real-world business challenges.

H1B Sponsorship

Cohere has a track record of offering H1B sponsorships. Please note that this does not guarantee sponsorship for this specific role. Below presents additional info for your reference. (Data Powered by US Department of Labor)
Distribution of Different Job Fields Receiving Sponsorship
Represents job field similar to this job
Trends of Total Sponsorships
2025 (11)
2024 (14)
2023 (13)
2022 (5)
2021 (2)

Funding

Current Stage
Late Stage
Total Funding
$1.71B
Key Investors
Government of CanadaTiger Global ManagementIndex Ventures
2025-09-24Series D· $100M
2025-08-14Series D· $500M
2025-06-17Secondary Market

Leadership Team

leader-logo
Aidan Gomez
cofounder + ceo
linkedin
leader-logo
Ivan Zhang
Co-Founder
linkedin
Company data provided by crunchbase