Syllo Logo

Syllo

Staff Software Engineer, Search & Retrieval Infrastructure

Posted One Month Ago
Remote
Hiring Remotely in USA
190K-230K Annually
Senior level
Remote
Hiring Remotely in USA
190K-230K Annually
Senior level
Own and scale the search, indexing, and data-scanning infrastructure for petabyte-scale data. Optimize hybrid lexical and vector retrieval for sub-second latency, design cost-effective hot/warm/cold data tiering and massive asynchronous scans, drive down latency and cost, and provide technical leadership for resilient, highly available indexing and retrieval systems.
The summary above was generated by AI

About Syllo 

Syllo is on a mission to transform litigation. Our product is a unified litigation platform that enables lawyers and paralegals to safely harness the power of language models and agentic AI throughout the litigation life cycle. Since going to market, we have gained a diverse group of enterprise customers, including some of the biggest law firms and corporations in the country, and we are quickly expanding. By reducing the expense of litigation industry-wide, we aim to improve access to high-quality representation and promote the alignment of legal outcomes with merit. 


About the Role 

We are seeking a Staff Software Engineer to take ownership of our advanced search, indexing, and data scanning infrastructure as we scale to the next echelon of data volume.

Our retrieval stack is robust and proven, but as our ingest sizes push into the multi-petabyte range, the complexity of balancing speed, availability, and cost increases exponentially. You will own this constellation of scaling challenges. Your mission is to continuously optimize and evolve our systems—ensuring our hot indexes maintain sub-second latency for interactive workflows, while simultaneously designing highly concurrent, cost-effective architectures for deep scanning and vectorizing massive volumes of cold-storage data. You will lead the design and implementation of sophisticated data tiering and retrieval strategies that keep our platform operating at peak performance without inflating cloud compute costs.


Responsibilities 

  • Scale the Retrieval Stack: Lead the optimization and architectural evolution of our existing hybrid search infrastructure, maximizing the throughput and efficiency of both lexical search (e.g., Elasticsearch, Lucene) and dense vector databases.
  • Advanced Data Tiering & Scanning: Design and implement intelligent, cost-effective tiering strategies across hot, warm, and cold data states. Evolve our distributed pipelines to efficiently execute asynchronous, massive-scale scans of petabytes of data in varying states of availability.
  • Relentless Optimization: Drive down latency and cost-to-serve. Deeply analyze system bottlenecks, tune indexing and querying algorithms, and optimize cloud infrastructure (compute, storage, and networking) for maximum efficiency at extreme scale.
  • Technical Leadership: Act as the domain expert and owner of the indexing and search ecosystem. Set the long-term technical vision for data storage and retrieval, guiding engineering teams on best practices for high-volume data modeling and performance tuning.
  • Resiliency at Scale: Ensure fault-tolerant, highly available operations during massive parallel ingest events and complex, concurrent querying across millions of documents.

Qualifications 

  • Extreme Scale Experience: 8+ years of software engineering experience, with a proven track record operating at the Staff/Principal level optimizing and scaling highly distributed, high-throughput systems to handle petabyte-level data.
  • Search & Vector Mastery: Deep, production-level expertise tuning and scaling Lucene-based search engines (Elasticsearch, Solr) and modern vector indexing infrastructure. You deeply understand index internals, chunking strategies, and embedding retrieval optimization.
  • Cost-Aware Architecture: A strong history of managing the compute vs. storage trade-off. You know how to design sophisticated cold-storage scanning solutions and hot-index architectures that are highly performant but fundamentally cost-effective.
  • Distributed Systems: Extensive experience managing complex data pipelines, high-throughput event streaming (Kafka, Kinesis), and distributed compute architectures handling billions of records.
  • Cloud Infrastructure: Expert command of cloud primitives (GCP preferred), Kubernetes, and infrastructure-as-code.
  • Languages: Expert-level proficiency in systems-level and backend languages (Go, Rust, Python, or Java/C++).

Salary Range ($190- $230K) plus health insurance and equity. 

United States - Remote Pay Range
$190,000—$230,000 USD

Similar Jobs

One Month Ago
Remote
US
190K-270K Annually
Senior level
190K-270K Annually
Senior level
Artificial Intelligence
Design and build scalable backend components and indexing pipelines for semantic and hybrid retrieval, build retrieval orchestration and knowledge-graph services, improve retrieval quality via evaluation and observability, design APIs, and optimize latency, throughput, cost, reliability, and security for large-scale AI inference and retrieval workloads.
Top Skills: C++ElasticEmbeddingsGoHybrid RetrievalJavaKnowledge GraphKubernetesLlmsObservability FrameworksOpensearchPineconePulumiPythonRagRustSemantic SearchTerraformVector Databases
19 Minutes Ago
Remote or Hybrid
65K-75K Annually
Mid level
65K-75K Annually
Mid level
Cloud • Real Estate • Software • PropTech
Supports clients and vendors within a specialized trade area, primarily HVAC. Reviews vendor quotes, resolves complex technical and operational issues, coordinates initiatives, maintains stakeholder relationships, and serves as a subject matter expert. The role is remote within the United States and requires weekend and/or evening availability.
Top Skills: HvacExcelMicrosoft OutlookMicrosoft PowerpointMicrosoft Word
23 Minutes Ago
Remote
United States
Senior level
Senior level
Edtech • Fintech • Payments • Social Impact • Financial Services • Big Data Analytics
Perform manual exploratory testing and build/maintain automated test suites (Playwright/Selenium). Integrate tests into CI/CD, collaborate with product and engineering on acceptance criteria, investigate and track defects, and advocate shift-left testing and continuous quality improvements.
Top Skills: Api TestingAWSCi/CdCucumberGherkinJIRAPlaywrightSeleniumTestrail

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account