Senior/Staff FDE - CUA
Snorkel AI · New York City, NY (Hybrid); San Francisco, CA (Hybrid)
About this role
**Senior/Staff FDE - CUA** **About Snorkel** Snorkel AI helps enterprises transform expert knowledge into specialized AI at scale. We believe meaningful AI starts with the data—working with leading organizations to build custom AI faster than ever before. **About the Role** Forward Deployed Engineer focused on Computer Use Agents, partnering with leading AI labs and enterprises on critical agentic-AI initiatives. You'll lead technical execution of complex customer engagements involving agents that operate computers, browsers, and software environments. **Main Responsibilities** **Computer Use Agents, Data & Evaluation** • Design task environments, datasets, and evaluation workflows for computer-using agents • Translate customer goals and agent failure modes into representative multi-step tasks • Develop data-generation and quality-assurance pipelines for multimodal training/evaluation data • Build automated evaluators and measurement frameworks for task completion and agent performance • Diagnose agent failures and turn findings into improved tasks and evaluations • Design and run experiments measuring how data and task design affect agent performance • Deliver production-grade task suites and evaluation assets **Forward Deployed Engineering & Customer Partnership** • Lead technical workstreams from solution design through production delivery • Build and iterate on solutions addressing customer needs • Rapidly prototype and productionize solutions across models, frameworks, and environments • Communicate technical tradeoffs and recommendations to stakeholders • Serve as trusted technical partner resolving complex blockers **Technical Leadership & Scale** • Identify patterns across engagements and turn solutions into reusable frameworks • Define technical standards for agent task design and evaluation • Partner with engineering, research, and product teams on platform capabilities • Lead technical design reviews and provide guidance to other engineers • Stay current with emerging agentic-AI and evaluation techniques **Requirements** • 5+ years in ML engineering, software engineering, applied AI, or similar technical role • Strong Python skills and experience building reliable production systems • Hands-on experience building/evaluating LLM-based or agentic systems, including computer-use agents • Strong understanding of experimentation, evaluation, and LLM-as-a-judge approaches • Experience designing task environments, datasets, and verifiers for agents • Experience with APIs, automation, web applications, and developer tools • Ability to take ambiguous problems from definition through delivery • Strong technical communication and customer-facing experience • Demonstrated ability to set technical direction and influence product decisions **Preferred Qualifications** • Agent benchmarks, task suites, or simulator development • Multimodal models and GUI interaction evaluation • Agentic coding tasks and browser/computer-use agent experience • Data pipelines for fine-tuning, RL, or model evaluation • Fast-paced, customer-facing environment experience **Compensation** $180,000–$320,000 USD base salary + variable compensation, equity, and benefits. Actual compensation based on skills, experience, and location.
Listing freshness
CronJobs last confirmed this listing 11h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.