Prime Intellect, Inc. logo

Applied Research - Forward-Deployed

Job Overview

Location

San Francisco

Job Type

Full-time

Category

Engineering

Date Posted

July 10, 2026

Full Job Description

đź“‹ Description

  • • Embed directly with strategic customers to understand their agent architectures, failure modes, and product goals.
  • • Design and build custom RL environments, evaluation harnesses, and verifiers that capture what "good" looks like for each customer's domain.
  • • Architect agent scaffolding — tool use, multi-step reasoning, memory, sandbox execution — tailored to customer workflows.
  • • Configure and launch training runs on Lab, iterating on reward functions, rollout strategies, and evaluation criteria.
  • • Serve as the technical lead for engagements end-to-end: from discovery through deployed, improved models.
  • • Identify repeatable patterns from customer engagements and codify them into reference implementations, templates, and documentation.
  • • Serve as the voice of the customer internally, shaping the roadmap for Lab, verifiers, the Environments Hub, and training infrastructure.
  • • Build high-quality examples and "recipes" that make it easy for new customers and open-source contributors to extend the stack.
  • • Contribute to technical content (blog posts, tutorials, case studies) that demonstrates real-world platform usage.
  • • Develop novel evaluation methodologies for agentic behavior — multi-step reasoning, tool use correctness, recovery from failure, long-horizon task completion.
  • • Prototype and iterate on agent harnesses for real-world tasks: code generation, workflow automation, document processing, and more.
  • • Experiment with reward design, rubric construction, and environment shaping to improve training signal quality.
  • • Stay current on the frontier of agentic AI, evals, and post-training methods, and bring that knowledge directly into customer work.

🎯 Requirements

  • • Deep hands-on experience building, evaluating, or deploying LLM-based agents in the past 1–2 years — you've seen what breaks in production and know what good evals look like.
  • • Strong intuition for evaluation design: you can look at a customer's agent and quickly identify what to measure, how to construct a rubric, and where the reward signal is weak.
  • • Working understanding of RL and post-training concepts (GRPO, RLHF, reward modeling, SFT) — you don't need to have written a trainer from scratch, but you should understand what the knobs do and why they matter.
  • • Strong Python skills and comfort with the modern AI stack (Hugging Face, inference engines, agent frameworks).
  • • Experience in a customer-facing or consulting-adjacent technical role, or as a technical founder — you're comfortable in a room with a customer's engineering team figuring out what to build.
  • • Excellent written and verbal communication — you can write a clear environment spec, a compelling case study, and a useful Slack message to a frustrated customer.
  • • High agency and comfort with ambiguity. You don't wait for specs; you scope the problem, ship a solution, and iterate.

🏖️ Benefits

  • • Cash Compensation Range of $150-300k + equity incentives.
  • • Flexible Work (San Francisco or hybrid-remote).
  • • Visa Sponsorship & relocation support.
  • • Professional Development budget.
  • • Team Off-sites & conference attendance.

Skills & Technologies

Python
JavaScript
TypeScript
React
Next.js
Remote

Ready to Apply?

You will be redirected to an external site to apply.

AI Job Fit Analysis
Pro

See exactly how your profile matches this role — strengths, skill gaps, and what to do about them.

Prime Intellect, Inc. logo
Prime Intellect, Inc.
Visit Website

About Prime Intellect, Inc.

San Francisco–based startup building decentralized AI infrastructure that lets researchers pool compute and data to collaboratively train large models. Founded in 2023, the company offers open-source protocols and cloud orchestration tools that aggregate GPUs across providers, coordinate distributed training, and cryptographically verify contributions so participants share ownership and future rewards of the resulting models.

Get more remote jobs like this

Subscribe to the weekly newsletter for similar remote roles and curated hiring updates.

Newsletter

Weekly remote jobs and featured talent.

No spam. Only curated remote roles and product updates. You can unsubscribe anytime.

Similar Opportunities

Expires soon
Dubai
Full-time
Expires Sep 14, 2026 (Soon)
Python
REST
Senior
+1 more

2 months ago

Expired
Remote - Munro, Argentina
Full-time
Expired Sep 2, 2026
Remote

2 months ago

Expires soon
Argentina
Full-time
Expires Sep 12, 2026 (Soon)
Python
TypeScript
AWS
+4 more

2 months ago

Expired
Argentina
Full-time
Expired Jul 27, 2026
Python
JavaScript
TypeScript
+4 more

3 months ago