BJAK Logo

BJAK

Lead Engineer, Machine Learning

Posted 18 Days Ago
Remote or Hybrid
Hiring Remotely in United States
Expert/Leader
Remote or Hybrid
Hiring Remotely in United States
Expert/Leader
Leads the end-to-end execution layer for machine learning, including data pipelines, large-model training and fine-tuning, evaluation, inference, deployment, monitoring, and reliability. Designs scalable GPU-based systems, optimizes performance and cost, partners with research and application engineering, and translates model capabilities into measurable production improvements. Also provides technical leadership, judgment, and support for high-impact ML delivery.
The summary above was generated by AI
About the Role

There are over 5 billion users using basic applications today such email, notes, tasks, calendar and they're not AI-native. Our mission is to build proactive applications for anyone in the world, who are not used to complex prompting. We aim to bring intelligence to conversations, errands, organising and workflows, with minimal to no prompting.

Our product focuses on achieving high reliability for long-running workflows, persistent context, and real-world task completion. We believe products will greatly reduce hallucinations

Our objective is to organise anyone's life, allowing us all to spend time on valuable and meaningful things

As Lead Engineer, Machine Learning, you own the execution layer of our intelligence, turning research and model capabilities into reliable, scalable production systems.

You will work across the model lifecycle: data, training, evaluation, inference, and deployment. This is a hands-on leadership role for someone who wants to operate at the intersection of research, systems, and product.

 
What You'll Own
  • Own the end-to-end ML systems powering our company, from data and training to evaluation, inference, and deployment.

  • Build and evolve training and fine-tuning pipelines for large models.

  • Design evaluation systems that measure capability, robustness, safety, and real-world product performance.

  • Architect high-performance inference systems, optimizing latency, GPU utilization, memory, cost, and reliability.

  • Build data pipelines and systems for high-quality real-world and synthetic training data.

  • Establish reliable production infrastructure for deploying, monitoring, and continuously improving models.

  • Partner closely with research and application engineering to turn model capabilities into product improvements.

  • Make pragmatic technical trade-offs and rapidly iterate based on real-world performance.

 
What We're Looking For
  • Experience building and shipping ML systems used in production, not just research prototypes.

  • Strong understanding of modern large-model training, fine-tuning, evaluation, and inference.

  • Strong software engineering and systems fundamentals.

  • Experience operating ML workloads at meaningful scale, particularly GPU-based systems.

  • Strong technical judgment and the ability to navigate ambiguous problems independently.

  • A bias toward experimentation, measurement, and shipping.

  • High standards for correctness, reliability, and production quality.

 
Outcomes
  • Research and models reliably translate into production-ready solutions with clear performance and quality targets.

  • ML pipelines, training loops, and inference systems are stable, efficient, and maintainable.

  • Production issues are detected, debugged, and resolved quickly, minimizing user impact.

  • Team members are supported, aligned, and able to deliver high-impact ML work with minimal friction.

  • Iterations on models and systems are measurable, safe, and improve user experience over time.

 
Tech Stack
  • Python

  • PyTorch / JAX

  • GPU-based training and inference system

 
Ideal Experience
  • You have built or shipped real ML systems used by people, not just demos.

  • You are comfortable working with large models and understanding their failure modes.

  • You write strong, production-grade code and care about system correctness.

 
How We Work

We are a small, high-talent-density, hands-on team. Engineers have broad ownership and are expected to exercise strong judgment and execute independently.

We make decisions quickly, work closely together, and balance speed with engineering fundamentals. We care less about process and more about building something exceptional.

 
Interview process

If there appears to be a fit, we'll reach to schedule 3, but no more than 4 interviews.

Applications are evaluated by our technical team members. Interviews will be conducted via virtual meetings and/or onsite.

We value transparency and efficiency, so expect a prompt decision. If you've demonstrated the exceptional skills and mindset we're looking for, we'll extend an offer to join us. This isn't just a job offer; it's an invitation to be part of a team that's bringing AI to have practical benefits to billions globally.

Similar Jobs

Yesterday
In-Office or Remote
146K-250K Annually
Expert/Leader
146K-250K Annually
Expert/Leader
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Leads enterprise AI/ML architecture for LLM, agentic, multi-agent, and RAG applications. Designs scalable, secure, compliant frameworks; remains hands-on with coding and integrations; drives optimization, observability, reliability, governance, and reusable components. Partners with product, engineering, and compliance teams, evaluates emerging AI technologies, and mentors senior engineers. Requires extensive software engineering and AI architecture experience, strong Python skills, cloud and containerization expertise, and experience deploying AI solutions in regulated environments.
Top Skills: Agentic WorkflowsArizeAWSAzureCi/CdDockerKubernetesLlmsMulti-Agent SystemsNext.JsPhoenixPostgresPythonReactRetrieval-Augmented Generation (Rag)
3 Days Ago
In-Office or Remote
146K-250K Annually
Expert/Leader
146K-250K Annually
Expert/Leader
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Leads enterprise AI/ML architecture and engineering strategy, developing scalable AI-powered applications, agentic and generative AI solutions, streaming data pipelines, and cloud-native platforms. The role uses Python, Java, Spring Boot, Next.js, Kafka, and PostgreSQL while establishing architecture standards, governance, security, and reliability practices. Responsibilities include technical leadership across programs, stakeholder influence, mentoring engineering teams, evaluating emerging technologies, and driving responsible AI adoption.
Top Skills: Agentic AiAgnoAngularApache KafkaCloud-Native ArchitecturesGenerative AiJavaKnowledge GraphsLanggraphMlopsNestjsNext.JsPostgresPythonReactSpring BootSpring WebfluxVercel Ai Sdk
17 Days Ago
In-Office or Remote
146K-250K Annually
Senior level
146K-250K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Leads the design, development, and deployment of AI-powered healthcare solutions using Java, Spring Boot microservices, HL7 FHIR, Da Vinci guides, Kafka, relational databases, Docker, Kubernetes, and cloud platforms. Builds REST APIs and event-driven systems across claims, eligibility, prior authorization, and provider domains. Establishes CI/CD automation, observability, security, and responsible AI practices while evaluating emerging technologies and improving operational workflows.
Top Skills: Apache KafkaAWSAzureAzure DevopsDa Vinci Implementation GuidesDockerElkGitGitGithub ActionsGrafanaHibernateHl7 FhirJava 17/21JenkinsJwtKubernetesMySQLOauth 2.0Oauth Client CredentialsOraclePostgresPrometheusSplunkSpring BootSpring Data Jpa

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account