Pika Logo

Pika

Software Engineer, AI Infra

Reposted 4 Hours Ago
In-Office
Palo Alto, CA
250K-350K Annually
Senior level
In-Office
Palo Alto, CA
250K-350K Annually
Senior level
Design and build scalable backend and infrastructure for autonomous AI agents: real-time messaging, agent runtimes, LLM integrations, vector retrieval, APIs, databases, and production reliability while mentoring engineers and driving architectural decisions.
The summary above was generated by AI

About the Role

 

We are looking for a Staff/Lead Software Engineer, AI Infrastructure, to play a critical role in building and scaling the core infrastructure that powers Pika’s AI capabilities. In this position, you will lead the design and implementation of GPU infrastructure, AI model serving APIs, and general AI infrastructure execution—enabling cutting-edge machine learning features that drive our products.

 

You will be responsible for architecting robust, distributed systems optimized for high-performance AI workloads, large-scale GPU orchestration, and low-latency, reliable API serving. Your work will directly impact the way users experience and interact with generative AI at scale. As a senior technical leader, you’ll also mentor engineers, drive best practices, and set the technical vision for AI infrastructure at Pika.

 

What You’ll Do

 
  • Design, develop, and maintain scalable GPU infrastructure for training and serving state-of-the-art AI models

  • Architect and optimize high-throughput, low-latency APIs for AI model serving and inference

  • Lead the orchestration, scheduling, and efficient utilization of heterogeneous GPU resources across clusters

  • Build and support robust systems for model deployment, monitoring, scaling, and reliability in production environments

  • Collaborate with ML, backend, and platform engineering teams to deliver seamless AI-powered product features

  • Drive technical direction, code reviews, and mentorship across the AI Infrastructure team

 

What We’re Looking For

 
  • Strong experience (5+ years) as a software engineer working on systems infrastructure, including hands-on work with ML serving and GPU orchestration

  • Deep knowledge of distributed systems, Kubernetes (or similar orchestration frameworks), and cloud-native infrastructure (AWS/GCP/Azure)

  • Proven expertise in building and optimizing APIs for large-scale AI model serving (TensorFlow Serving, Triton, TorchServe, or similar)

  • Familiarity with the challenges of high-throughput, scalable GPU fleet management, scheduling, and efficient model execution

  • Proficiency in backend languages such as Python, Go, or C++, and experience optimizing for performance and reliability

  • Ownership mentality and the drive to solve complex problems independently in ambiguous, high-growth environments

  • Excellent communication, collaborative, and mentorship skills

 

Nice to Have

 
  • Experience with multi-modal AI model infrastructure (LLMs, generative models, video/image/speech models)

  • Background in building infra for multi-tenant SaaS, enterprise AI/ML platforms, or operational automation at scale

  • Previous startup experience or experience leading high-impact projects through ambiguity and rapid iteration

  • Experience with competitive coding or large-scale distributed computing environments

 

What We Offer

 
  • Competitive salary in the AI industry

  • Equity in a rapidly growing team shaping the future of AI

  • Comprehensive health benefits, monthly stipends, and company retreats

  • A supportive and collaborative office culture—everyone builds, ships, and learns together

 

About Pika

 

At Pika, we’re building the infrastructure that empowers everyone to create videos and express ideas through advanced AI. Our team is passionate about removing technical barriers to creativity, and we thrive on working together to solve hard problems. We’re based in Palo Alto, CA, with a collaborative team working in-office 3–5 days a week.

Similar Jobs

3 Days Ago
In-Office
220K-400K Annually
Mid level
220K-400K Annually
Mid level
Artificial Intelligence • Software • Automation
Build and operate core AI agent infrastructure: orchestration, context management, filesystem and sandboxed execution, model routing and serving, databases, Kubernetes, networking, and cloud systems to make long-running agent workloads fast, reliable, safe, and cost-efficient.
Top Skills: Cloud InfrastructureContainersDatabasesDistributed SystemsFilesystemsKubernetesModel RoutingModel ServingNetworkingQueuesSandboxed Execution
4 Hours Ago
Remote or Hybrid
United States
Entry level
Entry level
Fintech • Machine Learning • Software • Financial Services
This position is an opportunity to fill a form for future job matching after the ASPLOS Conference with IMC Trading.
4 Hours Ago
Hybrid
178K-313K Annually
Senior level
178K-313K Annually
Senior level
Artificial Intelligence • Cloud • Machine Learning • Mobile • Software • Virtual Reality • App development
Design and implement generative ML systems (image, video, audio, multimodal LLMs) and deliver on-device and server-side inference. Build GenAI pipelines and AR experiences, prototype with cross-functional teams, and optimize efficient models for real-time mobile and wearable applications.
Top Skills: Audio GenerationAugmented RealityC++ClassificationDiffusion ModelsGenerative ModelsImage GenerationLlmsMobile Real-Time SoftwareObject DetectionOn-Device InferencePythonPyTorchSegmentationTensorFlowTrackingVideo Generation

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account