Cox Exponential Logo

Cox Exponential

Research Scientist/Engineer, Efficient ML Systems

Posted 15 Days Ago
Remote or Hybrid
Hiring Remotely in CA, USA
Mid level
Remote or Hybrid
Hiring Remotely in CA, USA
Mid level
Research and build efficient ML systems for large-scale LLMs and agentic RL: design algorithms and system techniques, prototype in training/inference stacks, run large-scale experiments, and translate findings into production or publications.
The summary above was generated by AI
About Goaly

At Goaly, our mission is to make custom AI affordable for every business. Our founding team comes from the front lines of top AI labs and tech giants (Meta MSL, TikTok AI, Google DeepMind, xAI, Microsoft Research, etc.), where we built large-scale training infrastructure powering trillion-parameter models and scaled GenAI models to a global user base. Now, we are building something we wish we had before: a platform that makes training and adapting custom AI affordable for all modern companies, not just Big Tech. Our north star is ambitious: for a domain-specific task, reach 90% of SOTA performance at less than 10% of the cost. To get a taste of what we are doing, see our first tech blog.


About the Role

As an AI Research Scientist (Efficient ML Systems) at Goaly, you will research and build the systems that make frontier-scale models practical. This role sits at the intersection of algorithms, systems, and hardware efficiency.You will design and evaluate new training and inference techniques, prototype them in real systems, and push them to production-scale workloads.

Your work will either ship directly into our core platform or lead to publications at top venues such as NeurIPS, ICML, ICLR, or CVPR.This is not a paper-only role. You will write real systems code, run large-scale experiments, and directly shape how modern LLMs and RL systems are trained and deployed.


Core Responsibilities

  • Research efficient ML systems: Invent and evaluate algorithms and system techniques that improve LLM and agentic RL training and inference efficiency (memory, compute, communication, and stability).
  • Scale agentic RL: Design and optimize large-scale agentic RL pipelines, including asynchronous training, experience management, reward modeling, and long-horizon stability.
  • End-to-end experimentation: Design large-scale experiments spanning model architecture, training algorithms, distributed systems, and hardware-aware optimization.
  • System-aware research: Prototype research ideas directly in training and inference stacks (e.g., parallelism strategies, attention kernels, RL training pipelines) and validate them at scale.
  • Production & publication: Translate successful ideas into production-ready systems and/or publish them at top-tier conferences with full internal support.

Requirements

  • Ph.D. or Master's degree in CS, AI, Systems, or related fields (Exceptional undergraduates with strong research capabilities may be considered).
  • Strong foundation in LLM or large-scale ML training, including Transformers, attention mechanisms, distributed training, and optimization methods.
  • Experience or strong interest in agentic RL or large-scale reinforcement learning systems, including stability, scalability, or long-horizon training challenges.
  • Demonstrated interest in efficiency-focused research, such as training acceleration, memory optimization, parallelism, kernels, or RL system robustness.
  • Proficient in PyTorch or JAX. Clean coding style and strong command of Python.
  • Adaptability: A fast learner with a strong sense of responsibility, capable of wearing multiple hats and handling cross-stack challenges.


Bonus Points

  • First-author publications at top conferences (NeurIPS, ICML, ICLR, CVPR, ACL).
  • High-star open-source projects on Hugging Face or Gold/Silver medals in Kaggle competitions.


Similar Jobs

Yesterday
Easy Apply
Remote or Hybrid
Easy Apply
199K-257K Annually
Mid level
199K-257K Annually
Mid level
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Lead and coach a team of ~8 account executives selling full-cycle to SMBs: hire, onboard, train, review pipeline, support deals, create training content, and collaborate cross-functionally to meet quarterly targets.
Top Skills: Salesforce (Sfdc)
Yesterday
In-Office or Remote
USA
112K-207K Annually
Senior level
112K-207K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Lead the CFC CRM product for HCP/HCO-facing colleagues, define roadmap, translate stakeholder needs into prioritized enhancements, write requirements and user stories, partner with engineering, UX, vendors and compliance to deliver scalable global CRM solutions and measure business impact.
Top Skills: Crm PlatformsLife Sciences CloudOceSalesforceVeeva
Yesterday
Remote or Hybrid
Charlotte, NC, USA
124K-280K Annually
Senior level
124K-280K Annually
Senior level
Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Lead large Oracle Field Service implementation projects as a Solutions Architect senior manager. Design and deploy Oracle Fusion Service and Field Service Cloud solutions, guide assessment and future-state planning, interact with senior clients, coach teams, apply delivery methodologies and accelerators, and drive continuous improvement.
Top Skills: Oracle Customer ExperienceOracle EpmOracle Field Service CloudOracle FinOracle Fusion ServiceOracle HcmOracle Lead ManagementOracle Marketing AutomationOracle Sales AutomationOracle Scm

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account