Neurons Lab Logo

Neurons Lab

AI Architect / Tech Lead (mahjong game)

Posted 25 Days Ago
Be an Early Applicant
Remote
Hiring Remotely in Greece
Senior level
Remote
Hiring Remotely in Greece
Senior level
Lead end-to-end development of an AI mahjong companion, including the gameplay decision model, LLM explanation layer, valid-action integration, win detection, training pipeline, evaluation harness, and low-latency serving. Own architecture, AWS deployment, Langfuse observability, and the two-second response budget while leading a small AI pod and communicating technical decisions to the client’s CTO and engineering team.
The summary above was generated by AI
About the project (description, duration, stage)

Hands-on Tech Lead for an AI Companion in an online mahjong game. The client is a social gaming company (web3 element) that scales its product and team. We deliver the AI side of their game as their embedded AI partner.

The AI Companion plays mahjong at a strong level and explains its moves. The core of the role is to build the mahjong-playing algorithm: a dedicated decision-making model (RL, imitation learning, or search-based — trained on the client's hand-history data) with an LLM reasoning layer on top. Key design constraints: a valid-action contract with the game engine (the bridge supplies legal moves), win detection, and a 2-second response budget per move. Explanations run async. Support for more than one rule set (riichi and regional variants) is on the roadmap.

Duration: 3 months, 0.5 FTE.

What you'll actually do (example tasks)
  • Design and build the mahjong-playing algorithm: choose and defend the approach (imitation learning on hand histories, RL / self-play, search with MCTS, or a hybrid), then train, evaluate, and ship it.

  • Own the technical architecture end to end: game model + LLM reasoning layer, valid-action mask, win detection, and the API contract with the client's game bridge.

  • Hit the 2-second response budget: design and measure the inference path, batching, and caching; keep a latency buffer for the client-facing number.

  • Define what data and event names we need from the client (hand histories, event streams); build the training and calibration pipeline on that data.

  • Build and run the evaluation harness: measure play strength against the client's reference points, and validate explanation quality.

  • Stand up LLM observability with Langfuse (async logging, N+1 batch) as an early sprint quick win.

  • Take over context from the team and lead the sprint work with the AI Engineer; work with the client's Product Owner in a scrum process.

  • Front the client's CTO and engineers on technical decisions; explain trade-offs in plain language and in depth when asked.

  • Watch the risks the account team flagged: licensing on new training data, engine-bridge capabilities, and multi-rule-set scope.

Skills (hands-on first)
  • Game AI / sequential decision-making: hands-on RL, imitation learning, or search-based agents (MCTS, self-play) — ideally for imperfect-information games (mahjong, poker, card games)

  • Expert Python for ML systems; strong software engineering (APIs, testing, CI)

  • Model training on gameplay data end to end: data → training → evaluation → serving

  • LLM application engineering: reasoning layers, prompt and context design, structured outputs, guardrails

  • Low-latency inference: profiling, batching, caching, model-size trade-offs against a hard time budget

  • LLM observability and evaluation (Langfuse or similar)

  • AWS deployment for ML workloads

  • Technical leadership of a small pod; clear written and spoken communication with client engineers and executives

Knowledge
  • Game theory for imperfect-information games; evaluation of play strength (win rates, Elo-style ratings, baseline agents)

  • Game-engine integration patterns (event streams, action masks, state bridges)

  • Web3 / gaming product context — plus, not required

  • AWS Well-Architected for ML workloads

Experience

Key characteristics (ideally 4/4):

  • Hands-on ML/AI engineering at production scale

  • Shipped an AI system inside a live product with hard latency limits

  • Cloud hyperscaler experience (AWS preferred)

  • Technology consulting / client-facing delivery background

Role-specific characteristics:

  • 6+ years hands-on ML/AI engineering, with real game AI or sequential decision-making work (RL / MCTS / self-play — not only LLM apps)

  • Trained models on user or gameplay data end-to-end (data → training → evaluation → serving)

  • Led small delivery teams while still coding personally

  • Comfortable owning an architecture in front of a technical client CTO

Questions for Applicants
  • Imperfect information: mahjong hides most tiles from each player. How does hidden information change your algorithm choice compared to a perfect-information game like chess?

  • Latency budget: tell us about a system you shipped with a hard response-time limit. How did you design, measure, and defend the budget?

  • LLM + model hybrid: how would you combine a trained game model with an LLM explanation layer so the explanation never contradicts the move?

  • Hands-on + lead: how do you balance personally coding the hard parts with leading an engineer and fronting the client?

Similar Jobs

3 Days Ago
In-Office or Remote
USA
215K-358K Annually
Senior level
215K-358K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Leads measurement and product marketing for eight enterprise AI platforms. Builds shared KPI, taxonomy, scorecard, data-quality, and value-realization frameworks; translates usage and outcome data into executive insights and investment guidance. Oversees positioning, internal launches, campaigns, enablement, adoption, and audience segmentation. Partners across product, engineering, data, communications, and business teams while building and managing teams responsible for analytics, marketing, and communications.
Top Skills: Ai PlatformsBusiness IntelligenceDashboardsData ContractsData FabricData VisualizationEvent TaxonomyExperimentationKnowledge GraphsKpi FrameworksOkr PlatformsProduct AnalyticsTelemetry
3 Days Ago
In-Office or Remote
USA
163K-272K Annually
Senior level
163K-272K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Leads product strategy, roadmap, lifecycle ownership, adoption, and measurable outcomes for a greenfield enterprise AI platform. Defines platform boundaries, governance controls, evaluation, registration, and cost-tracking capabilities while translating evolving risk, privacy, security, and GxP requirements into usable product features. Partners with engineering, design, legal, compliance, risk, security, and global agent-building teams to prioritize investments, guide delivery, communicate direction, and drive platform adoption.
Top Skills: AgileAIAi AgentsAi GovernanceGxpLeanModel Evaluation
4 Days Ago
In-Office or Remote
Site of Old Bullion, NV, USA
177K-294K Annually
Senior level
177K-294K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Build, test, and ship production-grade full-stack software and AI-powered systems for Pfizer’s Medical Affairs portfolio. Integrate LLMs, agentic AI, and RAG pipelines; develop scalable systems from prototypes through production; optimize prompts for cost, latency, and quality; integrate enterprise platforms; and monitor AI/ML performance. Collaborate with architecture, product, UX, and data engineering teams while maintaining strong code quality, documentation, testing, and verification practices.
Top Skills: Agentic AiAgileApi-First DesignAWSAzureCi/CdGdprGxpHipaaLlmopsLlmsMlopsRagSalesforce Life Sciences CloudSalesforce Marketing CloudScrumVector DatabasesVeeva Crm

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account