Skip to content
AI Engineer Jobs
Groq

Sr. Staff Inference Serving Engineer

Groq

Location
Hybrid (San Francisco Bay Area, California)
Compensation
$341k - $401k/yr
Employment
Full-time
Level
Senior Level
Posted 1 day ago

About the Role

Groq is building a high-performance software stack to convert bare metal compute into a token-producing engine for large language models. This role involves optimizing inference serving systems to ensure low latency and high throughput in production environments.

Skills

LLM Inference serving Transformer architectures Model compilation Graph optimization Batching Caching Scheduling Memory management Quantization Performance profiling Systems programming GPU acceleration Kernel tuning

Perks

  • Hybrid Work
  • LTI Program

Full job details