Pika Logo

Pika

Data Engineer

Posted 2 Days Ago
Remote
Hiring Remotely in US
Mid level
Remote
Hiring Remotely in US
Mid level
Design, build, and scale data pipelines, ETL workflows, analytics infrastructure, and machine-learning data systems. Ensure data quality, security, monitoring, and reliability while optimizing storage and processing performance. Collaborate with engineering, product, and analytics teams on data requirements, modeling, schemas, and scalable infrastructure. The role also establishes data engineering best practices and supports a data-driven culture.
The summary above was generated by AI

About Pika

 

At Pika, we’re building the next generation of AI creative tools to empower human creativity. Our mission is to make video creation seamless, intuitive, and accessible to everyone, leveraging the power of advanced AI. We believe that AI should amplify creative expression—enabling everyone to create, collaborate, and communicate across media. Our team includes engineers, artists, and product thinkers, all passionate about building tools that unlock new creative possibilities.

 

Pika has raised significant funding and is backed by leading investors, with a collaborative culture based in Palo Alto, CA. We prefer hybrid in-office, sharing ideas and launching products together.

 

About the Role

 

We are seeking a Data Engineer to design, build, and scale the data infrastructure powering Pika’s creative AI platform. As a Data Engineer, you will play a key role in architecting, implementing, and maintaining our data pipelines and analytics systems, enabling our team to make data-driven decisions and deliver world-class AI experiences. You will work closely with product, engineering, and data teams to ensure data is accurate, reliable, and accessible for users and internal business needs.

 

You will combine software engineering know-how with data architecture expertise, helping us build robust, scalable, and high-performance systems. Your contributions will directly support the success of millions of creators and help shape the future of AI-powered media tools.

 

What You’ll Do

 
  • Design, develop, and maintain scalable data pipelines and ETL workflows

  • Build, automate, and optimize our data infrastructure for analytics, reporting, and machine learning applications

  • Ensure data quality, consistency, and security across all sources and sinks

  • Collaborate with engineering, analytics, and product teams to define data requirements and deliver reliable datasets

  • Implement monitoring solutions and proactively resolve data pipeline issues

  • Optimize storage and data processing performance for growth and efficiency

  • Contribute to data modeling efforts and schema design for analytics and product needs

  • Help establish best practices and empower a data-driven culture across the organization

 

What We’re Looking For

 
  • 4+ years of experience as a data engineer or in a similar role designing, building, and maintaining data infrastructure

  • Strong software engineering background with proficiency in Python, SQL, and/or similar languages

  • Hands-on experience with data pipeline orchestration tools (Airflow, Prefect, Dagster, etc.)

  • Experience with cloud data platforms (AWS/GCP, Redshift, BigQuery, Snowflake, etc.)

  • Knowledge of database systems, data modeling, and data warehousing best practices

  • Familiarity with monitoring, logging, and data quality practices for data workflows

  • Excellent analytical and problem-solving skills with attention to detail

  • Great communication skills and ability to work cross-functionally in a collaborative environment

  • Self-motivated, curious, and comfortable in a fast-paced, high-growth startup

 

Nice to Have

 
  • Experience supporting data for machine learning or AI-powered applications

  • Familiarity with real-time or streaming data architectures (Kafka, Kinesis, etc.)

  • Prior work at high-growth startups or experience with rapid scaling

  • Open source, hackathon, or data engineering community experience

 

Our Stack

 

Python, Go, Node.js, Postgres, Redis, Docker, Kubernetes, AWS/GCP

 

What We Offer

 
  • Competitive salary in the AI industry

  • Substantial equity in a fast-growing startup defining the future of AI and creativity

  • Comprehensive health benefits, monthly stipends, and company retreats

  • Collaborative, high-growth culture—everyone contributes to growth and success

Similar Jobs

2 Days Ago
Remote
United States
121K-164K Annually
Junior
121K-164K Annually
Junior
Artificial Intelligence • Cloud • Consumer Web • Productivity • Software • App development • Data Privacy
Build and operate production data pipelines and dimensional models using Spark, SparkSQL, and cloud lakehouse technologies. Own pipelines from requirements through deployment, monitoring, and iteration; improve data quality, lineage, reliability, and cost efficiency. Partner with data scientists, analysts, product managers, and engineers to support datamarts, KPIs, reporting, and analysis. Participate in business-hours on-call rotations and improve runbooks and alerting.
Top Skills: AirflowC++DatabricksJavaKafkaKinesisMonte CarloPythonScalaSparkSparksqlSQLStructured Streaming
2 Days Ago
Remote or Hybrid
OH, USA
Junior
Junior
Financial Services
Develop and maintain secure, scalable data pipelines and production code using AWS, PySpark, Python, and ETL technologies. Extract and transform data, implement quality checks, optimize data processing workflows, support cloud data modernization, and ensure reliable data availability. Collaborate with cross-functional teams, troubleshoot technical issues, automate recurring remediation, and communicate with technical and non-technical stakeholders. The role also involves software testing, operational stability, continuous delivery, and potentially mentoring other engineers.
Top Skills: Ab InitioAirflowAWSCi/CdData LakesDatabricksETLInformaticaJavaPysparkPythonSnowflakeSQL
14 Days Ago
In-Office or Remote
124K-207K Annually
Senior level
124K-207K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Build and operate production data pipelines supporting analytics, AI, and agentic workflows. Responsibilities include implementing canonical data models, maintaining Databricks or Snowflake platforms, monitoring reliability, responding to incidents, validating healthcare data mappings, improving performance and cost, and documenting architecture. Requires strong SQL and Python skills, cloud data platform experience, ETL/ELT orchestration expertise, and healthcare or pharmaceutical data experience.
Top Skills: Ai/Ml WorkflowsDatabricksEltETLHedisOmopPythonSnowflakeSQL

What you need to know about the Seattle Tech Scene

Home to tech titans like Microsoft and Amazon, Seattle punches far above its weight in innovation. But its surrounding mountains, sprinkled with world-famous hiking trails and climbing routes, make the city a destination for outdoorsy types as well. Established as a logging town before shifting to shipbuilding and logistics, the Emerald City is now known for its contributions to aerospace, software, biotech and cloud computing. And its status as a thriving tech ecosystem is attracting out-of-town companies looking to establish new tech and engineering hubs.

Key Facts About Seattle Tech

  • Number of Tech Workers: 287,000; 13% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Amazon, Microsoft, Meta, Google
  • Key Industries: Artificial intelligence, cloud computing, software, biotechnology, game development
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Madrona, Fuse, Tola, Maveron
  • Research Centers and Universities: University of Washington, Seattle University, Seattle Pacific University, Allen Institute for Brain Science, Bill & Melinda Gates Foundation, Seattle Children’s Research Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account