LLM Benchmarking AI Jobs
Active AI and machine learning roles that require LLM Benchmarking. Each listing includes salary, location, and personalized match scoring when you describe your background.
8 active listings · updated as gigs are ingested
$180k–$230k
Senior Manager, Infrastructure Tax
Anthropic · Remote-Friendly (Travel Required) | San Francisco, CA · Remote · Posted 6 days ago
Gather facts on site specs, counterparty structure, commercial terms, and timing; prepare the structure review and term-sheet input for the. Lead's sign-off early enough to shape the deal.
$190k–$230k
Strategic Sourcing Business Partner, R&D Operations
Anthropic · San Francisco, CA | Seattle · Onsite · Posted 12 days ago
Contract management and templates. Own the commercial redline for HDO and R&D services agreements (MSAs, SOWs, order forms, amendments, and renewals), partnering with Legal on legal terms, so deals move quickly while maintaining negotiation posture.
$405k–$485k
Staff+ Researcher, Cybersecurity Products
Anthropic · San Francisco, CA · Onsite · Posted 12 days ago
Prototype rapidly to find define the AI frontier for cybersecurity work Design evaluations that measure model performance on the work security teams actually do.
$500k–$850k
Pre-training Distributed Systems Tech Lead / Manager
Anthropic · San Francisco, CA · Onsite · Posted 12 days ago
Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.
$300k–$385k
Infrastructure Tax Lead
Anthropic · Remote-Friendly (Travel-Required) | San Francisco, CA | New York City · Remote · Posted 12 days ago
Gather facts on site specs, counterparty structure, review and own commercial contracting for tax clauses; provide structure review and term-sheet input early enough to shape the deal.
$146k–$146k
Engineering Manager, AI
Lattice · Anywhere in the World · Remote · Posted 25 days ago
Lead, coach, and grow a high-performing team of AI and software engineers, developing them into strong technical leaders while fostering a culture of ownership, technical excellence, experimentation, and continuous learning.
$196k–$230k
User Researcher, AI Evaluations
Notion · Remote · Posted 76 days ago
Establish clear, reusable evaluation criteria that reflect real user expectations—helpfulness, trust, tone, control, and transparency. You’ll translate qualitative insight into scoring guidance that can be applied consistently across teams and over time.
Salary not listed
Product Manager | Vendor Intelligence & Marketplace
Ramp · Remote · Posted 78 days ago
Own the strategy and roadmap for price benchmarks and contract intelligence, turning Ramp's aggregated spend and contract dataset into the most accurate price benchmarking product on the market.
Related searches
Frequently asked questions
- What LLM Benchmarking AI jobs are hiring now?
- Gigmash tracks live LLM Benchmarking job listings from employer career sites and job boards. This page shows active roles that list LLM Benchmarking as a required or preferred skill, updated as new gigs are ingested.
- How do I know if I qualify for a LLM Benchmarking role?
- Describe your skills on Gigmash to get a personalized match score for each listing. You'll also see reach opportunities and recommended courses to close any skill gaps.
- Are these LLM Benchmarking jobs remote?
- Listings include remote, hybrid, and on-site roles. Use the full job browser to filter by work model, salary, and location.