The AI Engineer will design and deploy classification models for media, optimize data pipelines, lead R&D in voice and audio generation, and research image intelligence technologies.
This is a remote position.
Our client is looking for an innovative and driven AI Engineer to join their team. A leader in media intelligence and AI-driven content creation, they have recently expanded their work in AI voice and image technologies, driving the development of the next generation of cutting-edge products. This role will focus on the creation, classification, and organization of massive volumes of AI-generated media, along with spearheading R&D into AI voice and audio generation and advanced image intelligence capabilities.
Job Description
Responsibilities:
- Design, train, and deploy classification models for content pipeline, including style detection, quality scoring, content moderation, filtering, and semantic categorization of generated media.
- Develop and maintain automated tagging and organization systems for the media library: extracting attributes, detecting visual features, clustering similar content, and enabling intelligent search.
- Build and optimize training data pipelines: create annotation tooling, curate datasets, establish active learning loops, and ensure high-quality labeled data.
- Lead R&D into AI voice and audio generation, including voice cloning, text-to-speech, and audio synthesis; prototype integrations and create a production-ready pathway from research to features.
- Research and prototype image intelligence technologies such as face/body analysis, pose estimation, style transfer, and image-to-image consistency.
- Develop evaluation frameworks to measure the accuracy of classifiers, the quality of generation models, and model drift over time.
- Optimize inference pipelines for performance, cost, and latency—incorporating batching, quantization, caching, and model serving strategies.
- Integrate with GPU compute infrastructure and deliver models via production APIs.
Requirements
- 3+ years of experience building and deploying machine learning models in production, particularly in classification, tagging, or content understanding.
- Hands-on experience with model training, including dataset curation, experimenting with architectures, tuning hyperparameters, and debugging.
- Strong background in image classification and computer vision techniques (e.g., CNNs, vision transformers, CLIP).
- Experience or demonstrated interest in voice/audio AI (e.g., text-to-speech, voice cloning, audio classification).
- Proficiency in Python, with experience in PyTorch or TensorFlow.
- Experience with building data labeling pipelines, annotation workflows, or active learning systems.
- Understanding of model serving in production environments, including REST APIs and latency optimization.
Qualifications:
- Bachelor’s degree or higher in Computer Science, Engineering, or related field.
- Experience in AI/ML, particularly in content classification, tagging, and media organization systems.
- Proven experience with Python and ML frameworks like PyTorch or TensorFlow.
- Strong communication skills to collaborate with R&D teams and integrate new technologies into production.
Benefits
Similar Jobs
Fintech • Information Technology • Insurance • Financial Services • Big Data Analytics
Screen and process incoming Group Life claims and related mail, enter and update claims in the FEGLI system, handle claimant calls and correspondence, produce letters, perform peer reviews, and meet quality and production metrics.
Top Skills:
Fegli Claim SystemPc
Fintech • Information Technology • Insurance • Financial Services • Big Data Analytics
Reviews and approves escalated complaint responses, advises on complex life insurance and annuity transactions, investigates high-value disbursements and suspected fraud, and researches contract histories. The role develops resolutions for customers and partners, supports legal, compliance, product, and regulatory teams, identifies operational risks and trends, and recommends process improvements while maintaining strong controls.
Top Skills:
Microsoft CopilotMicrosoft Office Suite
Artificial Intelligence • Machine Learning • Natural Language Processing • Software • Conversational AI
Research Staff will develop foundational voice AI technologies, including low-bitrate neural audio codecs, steerable speech generation, disentangled audio representations, latent recombination, synthetic audio data generation, and multimodal speech-to-speech models. The role also involves designing scalable architectures, training methods, and inference algorithms optimized for hardware, billion-hour datasets, and real-time deployment. Candidates need strong mathematical foundations, foundation-model expertise, large-scale data pipeline experience, rigorous experimentation skills, deployment optimization knowledge, and publications or open-source contributions in speech or language AI.
Top Skills:
Data PipelinesFoundation ModelsGenerative ModelsGpu Hardware OptimizationLatent Space ModelsMultimodal LearningNeural Audio CodecsSelf-Supervised LearningSpeech-To-Speech SystemsStatistical Learning TheoryTriton Kernels
What you need to know about the Seattle Tech Scene
Home to tech titans like Microsoft and Amazon, Seattle punches far above its weight in innovation. But its surrounding mountains, sprinkled with world-famous hiking trails and climbing routes, make the city a destination for outdoorsy types as well. Established as a logging town before shifting to shipbuilding and logistics, the Emerald City is now known for its contributions to aerospace, software, biotech and cloud computing. And its status as a thriving tech ecosystem is attracting out-of-town companies looking to establish new tech and engineering hubs.
Key Facts About Seattle Tech
- Number of Tech Workers: 287,000; 13% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Amazon, Microsoft, Meta, Google
- Key Industries: Artificial intelligence, cloud computing, software, biotechnology, game development
- Funding Landscape: $3.1 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Madrona, Fuse, Tola, Maveron
- Research Centers and Universities: University of Washington, Seattle University, Seattle Pacific University, Allen Institute for Brain Science, Bill & Melinda Gates Foundation, Seattle Children’s Research Institute


