Syllo Logo

Syllo

Staff Software Engineer, Search & Retrieval Infrastructure

Posted Yesterday
Remote
Hiring Remotely in USA
190K-230K Annually
Senior level
Remote
Hiring Remotely in USA
190K-230K Annually
Senior level
Own and scale the search, indexing, and data-scanning infrastructure for petabyte-scale data. Optimize hybrid lexical and vector retrieval for sub-second latency, design cost-effective hot/warm/cold data tiering and massive asynchronous scans, drive down latency and cost, and provide technical leadership for resilient, highly available indexing and retrieval systems.
The summary above was generated by AI

About Syllo 

Syllo is on a mission to transform litigation. Our product is a unified litigation platform that enables lawyers and paralegals to safely harness the power of language models and agentic AI throughout the litigation life cycle. Since going to market, we have gained a diverse group of enterprise customers, including some of the biggest law firms and corporations in the country, and we are quickly expanding. By reducing the expense of litigation industry-wide, we aim to improve access to high-quality representation and promote the alignment of legal outcomes with merit. 


About the Role 

We are seeking a Staff Software Engineer to take ownership of our advanced search, indexing, and data scanning infrastructure as we scale to the next echelon of data volume.

Our retrieval stack is robust and proven, but as our ingest sizes push into the multi-petabyte range, the complexity of balancing speed, availability, and cost increases exponentially. You will own this constellation of scaling challenges. Your mission is to continuously optimize and evolve our systems—ensuring our hot indexes maintain sub-second latency for interactive workflows, while simultaneously designing highly concurrent, cost-effective architectures for deep scanning and vectorizing massive volumes of cold-storage data. You will lead the design and implementation of sophisticated data tiering and retrieval strategies that keep our platform operating at peak performance without inflating cloud compute costs.


Responsibilities 

  • Scale the Retrieval Stack: Lead the optimization and architectural evolution of our existing hybrid search infrastructure, maximizing the throughput and efficiency of both lexical search (e.g., Elasticsearch, Lucene) and dense vector databases.
  • Advanced Data Tiering & Scanning: Design and implement intelligent, cost-effective tiering strategies across hot, warm, and cold data states. Evolve our distributed pipelines to efficiently execute asynchronous, massive-scale scans of petabytes of data in varying states of availability.
  • Relentless Optimization: Drive down latency and cost-to-serve. Deeply analyze system bottlenecks, tune indexing and querying algorithms, and optimize cloud infrastructure (compute, storage, and networking) for maximum efficiency at extreme scale.
  • Technical Leadership: Act as the domain expert and owner of the indexing and search ecosystem. Set the long-term technical vision for data storage and retrieval, guiding engineering teams on best practices for high-volume data modeling and performance tuning.
  • Resiliency at Scale: Ensure fault-tolerant, highly available operations during massive parallel ingest events and complex, concurrent querying across millions of documents.

Qualifications 

  • Extreme Scale Experience: 8+ years of software engineering experience, with a proven track record operating at the Staff/Principal level optimizing and scaling highly distributed, high-throughput systems to handle petabyte-level data.
  • Search & Vector Mastery: Deep, production-level expertise tuning and scaling Lucene-based search engines (Elasticsearch, Solr) and modern vector indexing infrastructure. You deeply understand index internals, chunking strategies, and embedding retrieval optimization.
  • Cost-Aware Architecture: A strong history of managing the compute vs. storage trade-off. You know how to design sophisticated cold-storage scanning solutions and hot-index architectures that are highly performant but fundamentally cost-effective.
  • Distributed Systems: Extensive experience managing complex data pipelines, high-throughput event streaming (Kafka, Kinesis), and distributed compute architectures handling billions of records.
  • Cloud Infrastructure: Expert command of cloud primitives (GCP preferred), Kubernetes, and infrastructure-as-code.
  • Languages: Expert-level proficiency in systems-level and backend languages (Go, Rust, Python, or Java/C++).

Salary Range ($190- $230K) plus health insurance and equity. 

United States - Remote Pay Range
$190,000$230,000 USD

Similar Jobs

5 Days Ago
Remote
US
190K-270K Annually
Senior level
190K-270K Annually
Senior level
Artificial Intelligence
Design and build scalable backend components and indexing pipelines for semantic and hybrid retrieval, build retrieval orchestration and knowledge-graph services, improve retrieval quality via evaluation and observability, design APIs, and optimize latency, throughput, cost, reliability, and security for large-scale AI inference and retrieval workloads.
Top Skills: C++ElasticEmbeddingsGoHybrid RetrievalJavaKnowledge GraphKubernetesLlmsObservability FrameworksOpensearchPineconePulumiPythonRagRustSemantic SearchTerraformVector Databases
9 Minutes Ago
Easy Apply
Remote or Hybrid
USA
Easy Apply
140K-170K Annually
Senior level
140K-170K Annually
Senior level
Artificial Intelligence • Big Data • Logistics • Machine Learning • Software • Transportation
Sell FourKites SaaS supply chain and logistics solutions to new and existing Fortune 1000 accounts. Manage 15-25 accounts, exceed quota, develop strategic account plans, map solutions to customer SOPs, coordinate cross-functional GTM efforts, update Salesforce, and leverage internal AI tools to drive growth and expanded ARR.
Top Skills: Fourkites Ai ToolsLinkedin Sales NavigatorSaaSSalesforceZoominfo
24 Minutes Ago
Remote or Hybrid
147K-278K Annually
Senior level
147K-278K Annually
Senior level
Cloud • Software
Design, deploy, and operate large-scale, multi-region cloud-native services to improve reliability, performance, and security. Partner with application teams to build automation, run SLO-driven incident response and on-call rotations, leverage Kubernetes and CNCF tooling, and implement scalable operations, chaos and scale testing, and infrastructure-as-code for a resilient SaaS platform.
Top Skills: ArgocdAWSGoKubernetesLinux/UnixOpentelemetryPrometheusPythonService Mesh

What you need to know about the Seattle Tech Scene

Home to tech titans like Microsoft and Amazon, Seattle punches far above its weight in innovation. But its surrounding mountains, sprinkled with world-famous hiking trails and climbing routes, make the city a destination for outdoorsy types as well. Established as a logging town before shifting to shipbuilding and logistics, the Emerald City is now known for its contributions to aerospace, software, biotech and cloud computing. And its status as a thriving tech ecosystem is attracting out-of-town companies looking to establish new tech and engineering hubs.

Key Facts About Seattle Tech

  • Number of Tech Workers: 287,000; 13% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Amazon, Microsoft, Meta, Google
  • Key Industries: Artificial intelligence, cloud computing, software, biotechnology, game development
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Madrona, Fuse, Tola, Maveron
  • Research Centers and Universities: University of Washington, Seattle University, Seattle Pacific University, Allen Institute for Brain Science, Bill & Melinda Gates Foundation, Seattle Children’s Research Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account