Skip to content
AI Engineer Jobs
General Compute

Founding Inference Engineer

General Compute

Location
Onsite (San Francisco, California)
Employment
Full-time
Level
Senior Level
Posted 6 days ago

About the Role

General Compute is building an ASIC-first AI neocloud to deliver 5-7x faster token generation than GPU-based competitors. As a Founding Inference Engineer, you will own the end-to-end inference serving stack, optimizing performance and reliability for high-speed AI workloads.

Skills

LLM Inference Request Batching KV-Cache Management Autoscaling ASIC Fleet Management Systems Architecture Concurrency Networking Scheduling Production Systems Infrastructure Engineering Monitoring Alerting Performance Tuning Hardware Utilization

Full job details