About this role
<div class="description">
<h3>About the Role</h3>
<p>We're hiring a Data & AI Program Manager to support a Fortune 100 technology company. This role sits at the intersection of data strategy and applied machine learning, owning the end-to-end data lifecycle that fuels natural language processing (NLP) applications and large language model (LLM) development. You'll define what "good data" looks like for a range of NLP use cases, build the guidelines and pipelines to get there, and partner closely with data engineering, data science, and ML teams to turn high-quality annotated data into measurable model performance gains.</p>
<p>This is a full-time W2 position. You would be employed by <a href="https://himalayas.app/companies/cyborg-mobile">Cyborg Mobile</a> but placed with our client working full-time there and reporting to a manager there. We're looking for someone who is a self-starter, self-sufficient, and able to execute with little direction. This is a remote position (Washington candidates preferred). </p>
<h3>What You'll Do</h3>
<ul>
<li>Conduct market and user research to understand data needs and expectations across various NLP applications and domains</li>
<li>Define data specifications, scope, and sourcing strategy for each use case and domain</li>
<li>Design judgment guidelines and instructions for data annotation, validation, and quality control</li>
<li>Deliver datasets, schemas, guidelines, and annotations that meet approved specifications, scope, and sourcing requirements</li>
<li>Manage the end-to-end data annotation process</li>
<li>Analyze data quality metrics (accuracy, consistency, coverage, and diversity) and turn findings into actionable recommendations</li>
<li>Implement approved data quality improvement actions</li>
<li>Collaborate with client data engineers, data scientists, and machine learning engineers to integrate annotated data into production pipelines, support LLM fine-tuning, and evaluate model performance and impact</li>
<li>Identify data gaps and opportunities; propose new data sources, methods, and features to improve data quality and model outcomes</li>
<li>Communicate data vision, strategy, and results to internal and external stakeholders, including product owners, engineers, researchers, customers, and partners</li>
<li>Prepare and deliver data reports, presentations, and demos using client-approved formats and channels</li>
</ul>
<h3>What We're Looking For</h3>
<ul>
<li>Experience in NLP data strategy, annotation program management, or a related data quality/linguistics role</li>
<li>Strong understanding of data annotation workflows, guideline design, and inter-annotator quality control</li>
<li>Familiarity with LLM training/fine-tuning pipelines and how annotated data feeds model evaluation</li>
<li>Analytical mindset with the ability to translate quality metrics into concrete improvement plans</li>
<li>Excellent stakeholder communication skills — comfortable presenting to both technical and non-technical audiences</li>
<li>Experience working in large, matrixed corporate environments </li>
</ul>
<h3>Must Haves</h3>
<ul>
<li>Customer centric</li>
<li>Good with follow through </li>
<li>Has skills to build Agents, PowerBI dashboards, Sharepoint list management, coding agents like Codex, Claude Code, Github copilot</li>
</ul>
</div><p>Originally posted on <a href="https://himalayas.app">Himalayas</a></p>