The Learning House Logo

The Learning House

Senior Data Scientist (NLP + Applied AI)

Posted 23 Days Ago
Be an Early Applicant
Remote
Hiring Remotely in GBR
44K-63K Annually
Senior level
Remote
Hiring Remotely in GBR
44K-63K Annually
Senior level
Build and operate NLP enrichment pipelines over scientific literature, extracting entities, classifications, claims, and summaries. Compare classical NLP, retrieval, embedding, LLM, and fine-tuned model approaches using evaluations, cost, quality, and operational metrics. Create golden datasets, define evaluation strategies, contribute to agentic AI applications, and collaborate with editors, product managers, engineers, SMEs, and vendors to deliver production modeling systems.
The summary above was generated by AI

Job Description:


We believe in bold ideas, diverse perspectives, and the drive to transform knowledge into impact. Here, your curiosity fuels progress, your voice shapes innovation, and your ambition helps redefine what’s possible within science and learning. We are a culture that obsesses over impact, challenges, and drives what’s next to power infinite possibilities for our customers, colleagues and society at large.

About the Role:

About the role 

We're building the systems that turn one of the world's largest scientific corpora into research intelligence. That means production NLP pipelines running over millions of journal articles, extracting entities, classifications, claim tuples, and summaries optimized for use by downstream agentic applications. We're looking for a senior data scientist to own domain-specific content modeling work end to end, from the eval set through the pipeline stage that ships it. 

You'll join a small, senior team where data scientists own their models in production. You'll write the code, own the evaluations, ship the changes, and stay accountable for the outcomes. This is a hands-on role for someone who wants to see their models through to real users in a rapidly evolving market. 

What you'll do 

  • Design and build NLP enrichment pipelines that extract entities, classifications, claims, and summaries from scientific full-text at scale. 

  • Compare NLP approaches to extraction and enrichment against LLM-based approaches, and pick the right tool for each task. That means putting traditional NLP (NER, sequence labeling, classification), embedding-based retrieval, LLM prompting, and fine-tuned smaller models on the same table, and defending each choice with evaluation, cost, and operational tradeoffs. This is a core part of the job, not an occasional exercise. 

  • Own evaluation. Build the golden sets in consultation with SMEs and vendors, choose the metrics, and make productive tradeoffs between speed, quality, and cost. 

  • Contribute to agentic AI application work: tool-using systems that reason over the enriched corpus, where your NLP and evaluation background will shape how the agent grounds and defends its answers. 

  • Work directly with editors, product managers, and engineers. Bring the modeling perspective into product decisions, and translate stakeholder pushback into concrete modeling work. 

What you'll bring 

  • Strong NLP background across modern (LLMs, transformers, embeddings, retrieval) and classical (NER, classification, sequence labeling) approaches. You've built evaluations and learned from the results. 

  • Clean python.  You are comfortable in exploratory notebooks and production repositories, and an engineer taking over a modeling output from you has a good head start. 

  • A habit of comparing approaches and choosing the right one for the task. You can defend "prompt a large LLM" and "train a small classifier on 2,000 labels" with equal seriousness, back the choice with an eval and a cost estimate, and know what to do when performance drifts. 

Nice to have 

  • Experience working with scientific or scholarly text. 

  • Familiarity with AWS (S3, Batch, Lambda, SageMaker) and Parquet or Iceberg data lake patterns. 

  • Experience running LLMs under real cost and latency budgets in production. 

  • Some exposure to agentic AI applications: tool use, multi-step reasoning, guardrails, and evaluation of trajectories rather than single-turn outputs. 

Why us 

We publish some of the world's most-read research, and we're now in a rare position: applying modern AI to a corpus of trusted scientific knowledge that spans two centuries. Researchers will use the systems you build here to move faster and get closer to the answers they came for. That's the work: from knowledge to impact. 


We power infinite possibilities.


For more than 200 years, we've transformed knowledge into discoveries that shape the world. Today, our global team of innovators, creators, and experts is driving what's next in science, education, and publishing—creating impact that reaches everywhere. 


We're not just observers of progress. We're the ones accelerating scientific breakthroughs, advancing learning, and sparking innovation that redefines entire fields and improves lives. 


Here, your talent matters. Your ideas have room to grow. And your work creates breakthroughs that can change everything. 
Wiley is an equal opportunity/affirmative action employer. We evaluate all qualified applicants and treat all qualified applicants and employees without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, disability, protected veteran status, genetic information, or based on any individual's status in any group or class protected by applicable federal, state or local laws. Wiley is also committed to providing reasonable accommodation to applicants and employees with disabilities. Applicants who require accommodation to participate in the job application process may contact [email protected] for assistance.


We are proud that our workplace promotes continual learning and internal mobility. We offer meeting-free Friday afternoons allowing more time for heads down work and professional development, and through a robust body of employee programing we facilitate a wide range of opportunities to foster community, learn, and grow.
We are committed to fair, transparent pay, and we strive to provide competitive compensation in addition to a comprehensive benefits package. The range below represents Wiley's good faith and reasonable estimate of the base pay for this role at the time of posting roles in the United Kingdom, Canada, USA, Austria, Czechia, Denmark, France, Greece, Italy, Netherlands, Romania, or Spain. It is anticipated that most qualified candidates will fall within the range, however the ultimate salary offered for this role may be higher or lower and will be set based on a variety of non-discriminatory factors, including but not limited to, geographic location, skills, and competencies.
When applying, please attach your resume/CV to be considered.

Salary Range:

44,200.00 GBP to 63,400.00 GBP #LI-CW1

Similar Jobs

39 Minutes Ago
Easy Apply
Remote
United States
Easy Apply
153K-259K Annually
Expert/Leader
153K-259K Annually
Expert/Leader
Cloud • Security • Software • Cybersecurity • Automation
Leads the architecture and delivery of GitLab’s static analysis engine and evaluation tooling. Owns program modeling, vulnerability detection quality, SAST rules, AI-assisted development guardrails, and benchmark validation. Sets technical direction, mentors engineers, coordinates complex initiatives, contributes to research and open source, and participates in on-call rotations. The role requires deep program analysis, application security, systems programming, performance optimization, containerized CI/CD, and experience building trustworthy LLM tooling.
Top Skills: Ai AgentsAstsCall GraphsCi/CdControl-Flow GraphsCweData-Flow AnalysisDockerFuzzingGoIntermediate RepresentationsLlm ToolingMutation TestingOwasp Top 10RubyRustSastSsaStatic AnalysisTaint AnalysisType Inference
56 Minutes Ago
Remote
Senior level
Senior level
Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Lead and develop a team of 6–8 Mid-Market Account Executives across EMEA. Responsibilities include setting sales strategy, driving revenue and market share, coaching and mentoring sellers, recruiting and onboarding talent, establishing performance goals, analyzing sales trends, managing stakeholder relationships, and optimizing sales processes. The role focuses on building a high-performing sales organization and developing future sales leaders.
2 Hours Ago
In-Office or Remote
Senior level
Senior level
Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
Outbound Account Executive responsible for selling Clearpay to mid-market merchants across the UK. Duties include prospecting, managing the full sales funnel, building pipeline, engaging senior executives, demonstrating value propositions, negotiating contracts, exceeding targets, and maintaining accurate CRM and pipeline management. The role requires strong full-cycle sales, business development, negotiation, communication, collaboration, and experience in e-commerce, payments, fintech, or retail.
Top Skills: CRMSalesforce

What you need to know about the Seattle Tech Scene

Home to tech titans like Microsoft and Amazon, Seattle punches far above its weight in innovation. But its surrounding mountains, sprinkled with world-famous hiking trails and climbing routes, make the city a destination for outdoorsy types as well. Established as a logging town before shifting to shipbuilding and logistics, the Emerald City is now known for its contributions to aerospace, software, biotech and cloud computing. And its status as a thriving tech ecosystem is attracting out-of-town companies looking to establish new tech and engineering hubs.

Key Facts About Seattle Tech

  • Number of Tech Workers: 287,000; 13% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Amazon, Microsoft, Meta, Google
  • Key Industries: Artificial intelligence, cloud computing, software, biotechnology, game development
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Madrona, Fuse, Tola, Maveron
  • Research Centers and Universities: University of Washington, Seattle University, Seattle Pacific University, Allen Institute for Brain Science, Bill & Melinda Gates Foundation, Seattle Children’s Research Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account