A technical owner working with large enterprise and digital-native customers at Fireworks, turning complex AI problems into production systems fast while building executive-level relationships.
Responsibilities
- ▹Build end-to-end POCs and MVPs with customer engineering teams inside their own codebases and infrastructure
- ▹Architect inference foundations for customers whose core product depends on GenAI, sizing deployments to scale
- ▹Run load tests and establish latency, throughput, and cost baselines against realistic customer traffic profiles
- ▹Deploy and validate new model families on inference frameworks (vLLM, SGLang)
- ▹Guide customers on model selection and fine-tuning strategy (SFT, DPO, RFT)
- ▹Build and run fine-tuning pipelines directly with customers
- ▹Lead structured discovery conversations to unpack customer pain points and success criteria
- ▹Own the technical relationship from first engagement through production deployment
- ▹Spend time on-site with customers, building trust and momentum in person
- ▹Translate recurring customer pain points into concrete product proposals
Requirements
- ▹5+ years in a hands-on, customer-facing technical role (Forward Deployed Engineer, Applied AI Engineer, Solutions Architect, ML Engineer with field exposure, or technical founder)
- ▹Demonstrated ability to build production software with customers
- ▹Strong Python skills, comfortable with production code
- ▹Familiarity with Kubernetes and infrastructure engineering
- ▹Working knowledge of the LLM stack: inference tradeoffs, model serving, fine-tuning workflows (SFT at minimum, DPO/RFT a strong plus)
- ▹Experience with cloud infrastructure (AWS, Azure, GCP) and deploying models on GPU infrastructure
- ▹Exceptional communication across executive and engineering levels
Nice to have
- ▹10+ years in technical field or engineering roles
- ▹Experience with inference serving frameworks (vLLM, SGLang, TensorRT-LLM)
- ▹Experience operating as a technical authority inside a customer's environment
- ▹Track record taking GenAI POCs from prototype to production-scale deployments
- ▹Experience with hyperscaler AI platforms (Azure AI Foundry, AWS Bedrock/SageMaker, GCP Vertex)
- ▹Experience building or integrating agentic systems or tool-use chains
Soft skills
Executive and engineering-level communication in the same dayStakeholder management across large, longer-cycle organizationsBuilding trust through in-person presenceIndependent, structured discovery leadership
What we offer
- ▹$200,000-$260,000 on-target earnings plus equity
- ▹Meaningful equity in a fast-growing startup
- ▹Competitive salary and comprehensive benefits package
About the company
Fireworks is building the future of generative AI infrastructure, delivering one of the industry's fastest and most scalable inference platforms. A Series C company valued at $4 billion, backed by investors including Benchmark, Sequoia, Lightspeed, and Index, and founded by veterans of Meta PyTorch and Google Vertex AI.
Similar jobs

Job
Customer Reliability Engineer (Airflow)
Astronomer
DatabricksDbt
+8
💰 Salary: not specified
🌍 Remote
🗣️ EN

Job
Cyber AI Deployment Engineer
Accenture Federal Services
AI/MLBitbucket
+7
💰 Salary: not specified
🏢 On-site
Washington
🗣️ EN

Job
Professional Services Engineer
BigID
AI/ML
+12
$90,000–$110,000/yr
gross
🌍 Remote
🗣️ EN

Job
AI Security Associate Manager
Accenture Federal Services
Llm
💰 Salary: not specified
🏢 On-site
Arlington
🗣️ EN

Job
Enterprise Engineer
Branch
+5
💰 Salary: not specified
🌍 Remote
🗣️ EN

Job
Solutions Integration Engineer II
Samsara
+7
$73,780–$111,600/yr
gross
🌍 Remote
🗣️ EN
