Skip to content
AI Engineer Jobs
FriendliAI

Forward Deployed Engineer - AI Inference

FriendliAI

Location
Onsite (San Francisco, California)
Employment
Full-time
Level
Mid Level
Posted 1 day ago

About the Role

FriendliAI is building the fastest inference cloud for agents, delivering high throughput and low latency for frontier open-weight models. This role involves designing large-scale deployment architectures for LLM inference and collaborating directly with customer engineering teams to solve production-grade AI challenges.

Skills

Kubernetes Docker Terraform Helm Cloud infrastructure DevOps Reliability engineering Distributed systems Networking Performance tuning GPU-based computing Generative AI Model serving CI/CD Observability Backend systems

Benefits

  • Health check-up

Perks

  • Flexible working hours
  • Free meals

Full job details