Syllo Logo

Syllo

Staff Software Engineer, Search & Retrieval Infrastructure

Posted Yesterday
Remote
Hiring Remotely in USA
190K-230K Annually
Senior level
Remote
Hiring Remotely in USA
190K-230K Annually
Senior level
Own and scale the search, indexing, and data-scanning infrastructure for petabyte-scale data. Optimize hybrid lexical and vector retrieval for sub-second latency, design cost-effective hot/warm/cold data tiering and massive asynchronous scans, drive down latency and cost, and provide technical leadership for resilient, highly available indexing and retrieval systems.
The summary above was generated by AI

About Syllo 

Syllo is on a mission to transform litigation. Our product is a unified litigation platform that enables lawyers and paralegals to safely harness the power of language models and agentic AI throughout the litigation life cycle. Since going to market, we have gained a diverse group of enterprise customers, including some of the biggest law firms and corporations in the country, and we are quickly expanding. By reducing the expense of litigation industry-wide, we aim to improve access to high-quality representation and promote the alignment of legal outcomes with merit. 


About the Role 

We are seeking a Staff Software Engineer to take ownership of our advanced search, indexing, and data scanning infrastructure as we scale to the next echelon of data volume.

Our retrieval stack is robust and proven, but as our ingest sizes push into the multi-petabyte range, the complexity of balancing speed, availability, and cost increases exponentially. You will own this constellation of scaling challenges. Your mission is to continuously optimize and evolve our systems—ensuring our hot indexes maintain sub-second latency for interactive workflows, while simultaneously designing highly concurrent, cost-effective architectures for deep scanning and vectorizing massive volumes of cold-storage data. You will lead the design and implementation of sophisticated data tiering and retrieval strategies that keep our platform operating at peak performance without inflating cloud compute costs.


Responsibilities 

  • Scale the Retrieval Stack: Lead the optimization and architectural evolution of our existing hybrid search infrastructure, maximizing the throughput and efficiency of both lexical search (e.g., Elasticsearch, Lucene) and dense vector databases.
  • Advanced Data Tiering & Scanning: Design and implement intelligent, cost-effective tiering strategies across hot, warm, and cold data states. Evolve our distributed pipelines to efficiently execute asynchronous, massive-scale scans of petabytes of data in varying states of availability.
  • Relentless Optimization: Drive down latency and cost-to-serve. Deeply analyze system bottlenecks, tune indexing and querying algorithms, and optimize cloud infrastructure (compute, storage, and networking) for maximum efficiency at extreme scale.
  • Technical Leadership: Act as the domain expert and owner of the indexing and search ecosystem. Set the long-term technical vision for data storage and retrieval, guiding engineering teams on best practices for high-volume data modeling and performance tuning.
  • Resiliency at Scale: Ensure fault-tolerant, highly available operations during massive parallel ingest events and complex, concurrent querying across millions of documents.

Qualifications 

  • Extreme Scale Experience: 8+ years of software engineering experience, with a proven track record operating at the Staff/Principal level optimizing and scaling highly distributed, high-throughput systems to handle petabyte-level data.
  • Search & Vector Mastery: Deep, production-level expertise tuning and scaling Lucene-based search engines (Elasticsearch, Solr) and modern vector indexing infrastructure. You deeply understand index internals, chunking strategies, and embedding retrieval optimization.
  • Cost-Aware Architecture: A strong history of managing the compute vs. storage trade-off. You know how to design sophisticated cold-storage scanning solutions and hot-index architectures that are highly performant but fundamentally cost-effective.
  • Distributed Systems: Extensive experience managing complex data pipelines, high-throughput event streaming (Kafka, Kinesis), and distributed compute architectures handling billions of records.
  • Cloud Infrastructure: Expert command of cloud primitives (GCP preferred), Kubernetes, and infrastructure-as-code.
  • Languages: Expert-level proficiency in systems-level and backend languages (Go, Rust, Python, or Java/C++).

Salary Range ($190- $230K) plus health insurance and equity. 

United States - Remote Pay Range
$190,000$230,000 USD

Similar Jobs

5 Days Ago
Remote
US
190K-270K Annually
Senior level
190K-270K Annually
Senior level
Artificial Intelligence
Design and build scalable backend components and indexing pipelines for semantic and hybrid retrieval, build retrieval orchestration and knowledge-graph services, improve retrieval quality via evaluation and observability, design APIs, and optimize latency, throughput, cost, reliability, and security for large-scale AI inference and retrieval workloads.
Top Skills: C++ElasticEmbeddingsGoHybrid RetrievalJavaKnowledge GraphKubernetesLlmsObservability FrameworksOpensearchPineconePulumiPythonRagRustSemantic SearchTerraformVector Databases
7 Minutes Ago
Remote or Hybrid
55K-75K Annually
Junior
55K-75K Annually
Junior
Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Handle inbound and warm insurance leads remotely, consult with customers to recommend Property & Casualty coverage, close sales, complete paid training and licensing, follow defined shift schedule, and meet remote workspace and connectivity requirements.
2 Hours Ago
Easy Apply
Remote
USA
Easy Apply
75K-90K Annually
Junior
75K-90K Annually
Junior
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
Own performance marketing across Meta and direct-buy networks. Set up, monitor, and optimize Meta campaigns; test and scale direct-buy partnerships; develop creative; maintain tracking (pixel/S2S, macros); analyze ROAS/CAC and report insights; negotiate with network partners; collaborate with Creative, Analytics, and Partnerships teams to grow acquisition channels.
Top Skills: Ad Network MacrosCpa EmailEmail RemarketingExcelFraud MonitoringGoogle SheetsMeta Ads ManagerNewsbreakNextdoorPixel TrackingQuoraS2S Tracking

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account