SumerSports Logo

SumerSports

Data Engineer

Reposted 24 Days Ago
Remote
Hiring Remotely in United States
Mid level
Remote
Hiring Remotely in United States
Mid level
As a Data Engineer, you'll build and maintain data pipelines, work with ML teams, and ensure the data integrity for analysis and AI-driven products.
The summary above was generated by AI

SumerSports is a leading football intelligence technology company that specializes in providing an innovative suite of products for football fans and NFL clubs. We are a collection of executives, engineers, data scientists, and visionaries from NFL clubs, technology startups, finance, and academia. 


Our data-driven platform empowers teams with insights and tools to make informed decisions within salary cap constraints. The platform also serves the NCAA, offering insights around the transfer portal and more.


What sets us apart is our unique blend of big tech talent, data scientists, and former NFL personnel, who have a combined 600+ years of NFL experience. Our domain knowledge is augmented by AI and machine learning technologies to create a unique view into many aspects of Football.

As a Data Engineer, you’ll design, build, and maintain the data pipelines that power our deep learning and LLM systems. You’ll work across ingestion, transformation, and orchestration layers — from real-time feeds to analytics-ready datasets. 


Your mission is to make data reliable, discoverable, and scalable for use by model training, analytics, and AI-driven products across multiple sports. You’ll collaborate closely with our MLOps, LLMOps, and Sports Data teams to ensure seamless integration between data and AI. 


Responsibilities:

  • Build and operate robust data pipelines for ingestion, cleaning, and transformation using Databricks, Airflow, or Dagster. 
  • Develop efficient ETL/ELT workflows in Python and SQL to support both batch and streaming workloads.
  • Collaborate with ML and AI teams to deliver high-quality datasets for training, evaluation, and production features.
  • Model and maintain structured data assets (Delta, Parquet, Iceberg) for reliability, versioning, and lineage tracking. 
  • Implement orchestration and monitoring — schedule jobs, track dependencies, and automate recovery from failures. 
  • Ensure data quality and compliance through validation frameworks, schema enforcement, and audit logging. 
  • Contribute to data platform evolution — evaluate tools, standardize best practices, and improve developer experience.
  • Support performance and cost optimization across compute, storage, and orchestration systems.

Qualifications:

  • 3–6 years of experience as a Data Engineer or ETL Developer in a production environment. 
  • Proficiency in Python and SQL; strong familiarity with Databricks, Spark, or equivalent big-data frameworks. 
  • Experience with workflow orchestration tools such as Airflow, Dagster, Luigi or Prefect. 
  • Deep understanding of data modeling, data warehousing, and distributed data processing. 
  • Knowledge of modern data lakehouse architectures (Delta, Parquet, Iceberg). 
  • Familiarity with CI/CD, GitHub Actions, and data pipeline testing frameworks. 
  • Comfort working in a cross-functional environment with ML, product, and analytics teams. 

Nice to Have:

  • Experience with sports, telemetry, or sensor data pipelines. 
  • Familiarity with streaming frameworks (Kafka, Spark Structured Streaming, Flink). 
  • General knowledge of American football, the NFL, and college football
  • Background in data governance, lineage, and observability tools (Monte Carlo, Great Expectations, Unity Catalog, OpenLineage). 
  • Experience with cloud infrastructure (AWS, GCP, or Azure) and containerization (Docker, Kubernetes). 
  • Exposure to best practices in machine-learning model management and MLOps 

Benefits:

  • Competitive Salary and Bonus Plan
  • Comprehensive health insurance plan
  • Retirement savings plan (401k) with company match
  • Remote working environment
  • A flexible, unlimited time off policy
  • Generous paid holiday schedule - 13 in total including Monday after the Super Bowl

Top Skills

Airflow
AWS
Azure
Dagster
Databricks
Delta
Docker
GCP
Iceberg
Kafka
Kubernetes
Parquet
Python
Spark
SQL

Similar Jobs

Yesterday
In-Office or Remote
Minnetonka, MN, USA
Senior level
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Design and optimize database architectures and data models; build, automate, and maintain ETL/ELT pipelines and large-scale data platforms on Azure; ensure data quality, security, and compliance; collaborate with data scientists to deliver AI/ML solutions; lead projects, mentor junior engineers, and manage vendor and stakeholder relationships.
Top Skills: Sql,Postgres,Mysql,Azure,Ci/Cd,Devops,Mlops,Etl,Elt,Pyspark,Scala Spark,Hive,Hadoop,Nosql,Python,Scala,Data Warehousing,Paas
Yesterday
Remote or Hybrid
Framingham, MA, USA
69K-129K Annually
Mid level
69K-129K Annually
Mid level
Big Data • Healthtech • Software
Design, build, and maintain scalable ETL/ELT pipelines using Python, Spark, Databricks, Airflow and SSIS. Integrate and cleanse diverse healthcare datasets, implement Unity Catalog for metadata and governance, optimize Spark performance and JVM tuning, support Medallion architecture, and collaborate with cross-functional teams to automate CI/CD, observability, and data quality processes.
Top Skills: Python,Scala,Sql,Apache Spark,Databricks,Aws,Ssis,Apache Airflow,Unity Catalog,Jenkins,Gitlab Ci,Parquet,Delta,Csv,Xml,Nosql,Jvm,Medallion Architecture
5 Days Ago
Easy Apply
Remote
United States
Easy Apply
150K-175K Annually
Mid level
150K-175K Annually
Mid level
Healthtech • HR Tech • Software
As a Data Engineer at Vivian Health, you will build and maintain data pipelines, APIs, and contribute to product development using various technologies.
Top Skills: AirflowAPIsAWSBigQueryCloud Data WarehousesCloudFormationDbtDynamoDBElastic BeanstalkInfrastructure-As-CodeJavaScriptLambdaPythonReactRedshiftSnowflakeSQLTerraform

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account