[All remote jobs](https://jobicy.com/jobs.md)Open role[![Sporty Group logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/8e6c246a-221.png)](https://jobicy.com/company/sporty-group.md)Remote opportunity at[Sporty Group](https://jobicy.com/company/sporty-group.md)

# Senior AI Scientist

Review the role, location requirements, compensation details, and application process before deciding whether this opportunity fits your next career move.

[Apply for this job](#job-application)[View company](https://jobicy.com/company/sporty-group.md)Share18 Aug 2026Published29Listing views3Application actions18 Sep 2026Apply before  Opportunity details

## About this role.

AI SummarySporty is seeking a Senior AI Scientist to improve the quality, accuracy, safety, and effectiveness of AI-powered conversational products. The role centers on analyzing AI interactions, designing prompt strategies, optimizing RAG retrieval, evaluating LLMs, and measuring outcomes through structured experiments and quality metrics. The successful candidate will work closely with AI Engineers to move validated improvements into production while maintaining benchmarks, evaluation frameworks, and documentation. This remote-first position requires at least five years of relevant AI, ML, NLP, or conversational AI experience, with strong Python, SQL, stakeholder communication, and ownership skills.

## Role DNA

A quick view of the complexity, pace, ownership and collaboration implied by the job description.

### Job Complexity

5/5EasyHard

### Pace & Pressure

4/5RelaxedFast-paced

### Autonomy Level

5/5GuidedFull ownership

### Communication Load

4/5IndependentCollaborative

AI insightThis is a senior, technically demanding role requiring deep practical expertise across LLM evaluation, prompt optimization, RAG systems, experimentation, Python, SQL, and production AI operations. The role also requires independent judgment to translate interaction data and user feedback into measurable product improvements.

## Salary analysis

Estimated compensation compared with the broader US market for similar roles.

Estimated job medianHighly competitive$210,000US market range$170k–$250k0$275k

AI insightNo salary range was provided. For the U.S. market, a Senior AI Scientist focused on LLMs, RAG, evaluation, and production AI commonly commands an estimated base salary range of $170,000 to $250,000 annually; a $210,000 annual median is a reasonable market estimate before bonus, equity, or location adjustments.

## Core skills

Skills and capabilities most closely associated with this opportunity.

[Artificial Intelligence](https://jobicy.com/jobs?search_keywords=Artificial%20Intelligence.md)[Machine Learning](https://jobicy.com/jobs?search_keywords=Machine%20Learning.md)[Large Language Models](https://jobicy.com/jobs?search_keywords=Large%20Language%20Models.md)[Conversational AI](https://jobicy.com/jobs?search_keywords=Conversational%20AI.md)[Prompt Engineering](https://jobicy.com/jobs?search_keywords=Prompt%20Engineering.md)[Retrieval-Augmented Generation](https://jobicy.com/jobs?search_keywords=Retrieval-Augmented%20Generation.md)[RAG Evaluation](https://jobicy.com/jobs?search_keywords=RAG%20Evaluation.md)[Python](https://jobicy.com/jobs?search_keywords=Python.md)[SQL](https://jobicy.com/jobs?search_keywords=SQL.md)[Responsible AI](https://jobicy.com/jobs?search_keywords=Responsible%20AI.md)

Cover letter sampleDear Hiring Team,

I am excited to apply for the Senior AI Scientist role at Sporty. My background in machine learning, conversational AI, and production evaluation systems aligns well with your focus on improving LLM response quality, retrieval performance, and user satisfaction.

I bring strong experience designing prompt experiments, analyzing interaction data, defining measurable quality metrics, and partnering with engineering teams to deploy validated improvements. I am particularly motivated by the opportunity to apply responsible AI practices and rigorous benchmarking to scalable AI-powered products.

I would welcome the opportunity to contribute an analytical, ownership-driven approach to Sporty’s remote-first team.

Copy   Sample interview questionsHow would you build an evaluation framework for a production conversational AI system?I would begin by defining the target user outcomes and a baseline across quality dimensions such as factual accuracy, relevance, task completion, safety, latency, and user satisfaction. I would create a representative evaluation dataset from real interactions, combine automated metrics with expert review, run controlled experiments, and monitor production results after deployment.

How would you diagnose and improve poor retrieval quality in a RAG pipeline?

I would inspect failed queries by intent, domain, language, and retrieval score to determine whether the issue is document coverage, chunking, embedding quality, metadata filtering, ranking, or answer-generation behavior. I would then test targeted changes such as revised chunking, hybrid retrieval, reranking, query rewriting, and better grounding prompts against a fixed benchmark before releasing them incrementally.

Describe how you would evaluate a new prompt-engineering strategy.

I would state a clear hypothesis, define primary and guardrail metrics, select a representative sample, and compare the candidate prompt with the baseline through offline evaluation and, where appropriate, an online A/B test. I would assess quality gains alongside latency, token cost, safety, and consistency, then document results and deploy only if the improvement is statistically and operationally meaningful.

Which metrics would you use to assess LLM quality and safety?

I would use a combination of automated checks, rubric-based human evaluation, adversarial test cases, and production feedback. Important measures would include factual grounding, instruction adherence, harmful-content rates, hallucination frequency, escalation behavior, and performance across different user segments and edge cases.

How would you recommend an LLM for a specific product use case?

I would first establish the business objective and constraints, including quality requirements, latency, cost, data privacy, deployment environment, and governance needs. I would evaluate shortlisted models on a domain-specific benchmark, review error patterns rather than relying solely on aggregate scores, and recommend the model that offers the strongest overall trade-off for the use case.

About the role

We are looking for a Senior AI Scientist to continuously improve the quality, accuracy and effectiveness of our AI-powered products. You will analyse AI interactions, identify improvement opportunities and optimise conversational behaviour, prompting strategies and retrieval mechanisms. Working closely with AI Engineers, you will drive measurable improvements in AI performance and user satisfaction.

What you’ll be doing

* Analyse AI conversations and identify opportunities to improve response quality and user experience.
* Design, implement and evaluate prompt engineering strategies.
* Optimise RAG pipelines and retrieval quality.
* Evaluate different LLMs and recommend the most effective models for specific use cases. Potentially train customer personal models.
* Define and monitor AI quality metrics, including accuracy, relevance, latency and user satisfaction.
* Design and execute experiments to validate AI improvements.
* Analyse user feedback and interaction data to continuously improve AI behaviour.
* Collaborate with AI Engineers to deploy validated improvements into production.
* Maintain evaluation frameworks, benchmarks and documentation.
* Stay current with advances in generative AI and recommend practical improvements to existing products.

What you’ll bring

* 5+ years of experience in AI, ML, NLP or a related field.
* Strong understanding of LLMs and conversational AI.
* Experience with prompt engineering, prompt evaluation and LLM optimisation.
* Experience evaluating conversational AI systems using qualitative and quantitative methods.
* Strong Python and SQL skills.
* Strong analytical and problem-solving skills.
* Excellent communication and stakeholder management skills.
* Proactive mindset with strong ownership and accountability.
* Experience with speech and voice AI technologies.
* Experience with AI evaluation frameworks and benchmarking.
* Experience with vector databases and embedding models.
* Knowledge of AI governance, safety and responsible AI practices.
* Experience working with production AI systems at scale.

What’s in it for you

* Sporty is a remote first company in pursuit of sustainability
* A competitive salary + individual performance based bonuses every quarter
* 28 days paid annual leave
* Our core working hours are 10am-3pm in your local time zone with flexibility outside of this
* Referral bonuses & flash bonuses
* Top of the line equipment
* Annual company retreats to provide great internal networking opportunities

Interview Process

* Remote video screening with our Talent Acquisition Team
* Online home assignment
* Remote video interview with Team Members (3×45 Mins)

If you’re interested, we encourage you to apply! Every application is reviewed by a member of our team (AI is not used in our recruitment process), and we aim to respond within 48 hours.

Show more

[Apply now >](https://jobicy.com/jobs/150999-senior-ai-scientist.md)

>  Annual salary information is not provided for this position. Explore salary ranges for similar roles in our [Salary Directory ›](https://jobicy.com/salaries.md)

*

![Upload CV](data:image/svg+xml;base64,PHN2ZyB4bWxucz0iaHR0cDovL3d3dy53My5vcmcvMjAwMC9zdmciIHdpZHRoPSI2NSIgaGVpZ2h0PSI2NSIgZmlsbD0ibm9uZSIgeG1sbnM6dj0iaHR0cHM6Ly92ZWN0YS5pby9uYW5vIj48ZyBjbGlwLXBhdGg9InVybCgjQSkiPjxwYXRoIGQ9Ik0wIDBINjVWNjVIMFYwWiIgZmlsbD0iIzAyOWFlYiIvPjxnIGZpbGw9IiNmZmYiIHN0cm9rZT0iI2ZmZiIgc3Ryb2tlLXdpZHRoPSIyIj48cGF0aCBkPSJNMzMuMDQ5IDE1LjQ1NGExLjQzIDEuNDMgMCAwIDAtMi4wOTcgMGwtNy41NzkgOC4xNDdhMS4zOCAxLjM4IDAgMCAwIC4wOSAxLjk3MyAxLjQ0IDEuNDQgMCAwIDAgMi4wMDgtLjA4OGw1LjEwOS01LjQ5MnYyMC42MWExLjQxIDEuNDEgMCAwIDAgMS40MjEgMS4zOTdjLjc4NSAwIDEuNDIxLS42MjUgMS40MjEtMS4zOTd2LTIwLjYxbDUuMTA5IDUuNDkyYTEuNDQgMS40NCAwIDAgMCAyLjAwOC4wODggMS4zOCAxLjM4IDAgMCAwIC4wOS0xLjk3M2wtNy41NzktOC4xNDZ6TTE2Ljc2OSAzOC40YzAtLjc3My0uNjItMS40LTEuMzg1LTEuNFMxNCAzNy42MjcgMTQgMzguNHYuMTAybC4yMTUgNi4yMjljLjIyMyAxLjY4LjcwMSAzLjA5NSAxLjgxMyA0LjIxOHMyLjUxIDEuNjA3IDQuMTcyIDEuODMzYzEuNi4yMTggMy42MzYuMjE4IDYuMTYuMjE4aDExLjI4bDYuMTYtLjIxOGMxLjY2Mi0uMjI2IDMuMDYxLS43MDkgNC4xNzItMS44MzNzMS41ODktMi41MzggMS44MTMtNC4yMThDNTAgNDMuMTEzIDUwIDQxLjA1NSA1MCAzOC41MDNWMzguNGMwLS43NzMtLjYyLTEuNC0xLjM4NS0xLjRzLTEuMzg1LjYyNy0xLjM4NSAxLjRsLS4xOSA1Ljk1OGMtLjE4MiAxLjM3LS41MTUgMi4wOTUtMS4wMjYgMi42MTJzLTEuMjI4Ljg1My0yLjU4MyAxLjAzOGMtMS4zOTUuMTktMy4yNDMuMTkzLTUuODkzLjE5M0gyNi40NjJjLTIuNjUgMC00LjQ5OC0uMDAzLTUuODkzLS4xOTMtMS4zNTUtLjE4NC0yLjA3Mi0uNTIxLTIuNTgzLTEuMDM4cy0uODQ0LTEuMjQyLTEuMDI2LTIuNjEyYy0uMTg3LTEuNDEtLjE5MS0zLjI3OS0uMTkxLTUuOTU4eiIvPjwvZz48L2c+PGRlZnM+PGNsaXBQYXRoIGlkPSJBIj48cGF0aCBmaWxsPSIjZmZmIiBkPSJNMCAwaDY1djY1SDB6Ii8+PC9jbGlwUGF0aD48L2RlZnM+PC9zdmc+)

### Upload your resume now

To unlock remote work opportunities and be discovered by global employers.

This job listing has been manually reviewed by the Jobicy Trust & Safety Team for compliance with our posting guidelines, including verification of the company's legitimacy, accuracy of job details, clarity of remote work policy, and absence of misleading or fraudulent content.

Next step

## Apply now.

Follow the employer’s application method and review Jobicy’s safety guidance before sharing personal information.

Keep exploring

## Related remote jobs.

Matched by job category10 related opportunities[Data Science & Analytics](https://jobicy.com/categories/data-science.md) [Browse all jobs](https://jobicy.com/jobs.md)
*
![Sporty Group logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/8e6c246a-221.png)
Sporty Group  Aug 18

### [Senior BI Engineer](https://jobicy.com/jobs/151002-senior-bi-engineer.md)

About the roleWe are looking for a Senior BI Engineer to design, build and maintain our business intelligence platform. You will be responsible for transforming data into trusted, actionable insights…

*
![Experian logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2021/09/dcc5b29a570bb19b9f5c3e150db2fdfe.jpg)
Experian  Aug 18

### [Senior Data Scientist, Innovation Lab](https://jobicy.com/jobs/149399-senior-data-scientist-innovation-lab.md)

Company DescriptionExperian is a global data and technology company, powering opportunities for people and businesses around the world. We operate across a range of markets, from financial services to healthcare,…

*
![Sona logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/740d9e4e-221.png)
Sona  Aug 17

### [Senior Machine Learning Engineer](https://jobicy.com/jobs/150930-senior-machine-learning-engineer.md)

Running a frontline business is an operational puzzle most software has never touched. Shift-by-shift labour costs, compliance that changes by region and by role, and margins thin enough that a…

*
![Clover Health logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/8d4e74b3-221.jpg)
Clover Health  Aug 17

### [Data Analyst, Value-Based Care Analytics](https://jobicy.com/jobs/150925-data-analyst-value-based-care-analytics.md)

Counterpart Health is an AI‑powered physician enablement platform that delivers clinical insights to providers at the point of care. Our flagship product, Counterpart Assistant, is embedded into clinicians’ workflows and…

*
![Clover Health logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2025/06/8d4e74b3-221.jpg)
Clover Health  Aug 17

### [Data Analyst, Clinical Data Effectiveness](https://jobicy.com/jobs/150923-data-analyst-clinical-data-effectiveness.md)

Counterpart Health is an AI‑powered physician enablement platform that delivers clinical insights to providers at the point of care. Our flagship product, Counterpart Assistant, is embedded into clinicians’ workflows and…

*
![Instructure logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2022/04/430057884baf037dd2aa1e1e017e72b6.jpg)
Instructure  Aug 17

### [Decision Scientist](https://jobicy.com/jobs/150911-decision-scientist.md)

At Instructure, we believe in the power of people to grow and succeed throughout their lives. Our goal is to amplify that power by creating intuitive products that simplify learning…

*
![Instructure logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2022/04/430057884baf037dd2aa1e1e017e72b6.jpg)
Instructure  Aug 17

### [Senior Decision Scientist](https://jobicy.com/jobs/150909-senior-decision-scientist.md)

At Instructure, we believe in the power of people to grow and succeed throughout their lives. Our goal is to amplify that power by creating intuitive products that simplify learning…

*
![Instructure logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2022/04/430057884baf037dd2aa1e1e017e72b6.jpg)
Instructure  Aug 17

### [Sr. Data Engineer](https://jobicy.com/jobs/150903-sr-data-engineer.md)

At Instructure, we believe in the power of people to grow and succeed throughout their lives. Our goal is to amplify that power by creating intuitive products that simplify learning…

*
![Cresta logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2022/08/4d686e40b22cf2510b627c59afe5fd95.jpg)
Cresta  Aug 16

### [Data Science Intern (Customer Success)](https://jobicy.com/jobs/150863-data-science-intern-customer-success.md)

Cresta unlocks the true potential of the customer experience, turning every conversation into a competitive advantage. Cresta’s unified AI platform combines conversational AI agents, real-time human agent augmentation, and comprehensive…

*
![Gopuff logo](https://jobicy.com/data/server-nyc0409/galaxy/mercury/2021/05/Jobicy-210510095548-598926.jpg)
Gopuff  Aug 16

### [Data Engineer](https://jobicy.com/jobs/150836-data-engineer.md)

At Gopuff, data sits at the heart of our strategy. We are reimagining how people purchase everyday essentials, from snacks to household goods to alcohol, all delivered in minutes. To…