Forward Deployed Engineer (Generative AI)
Tiger Analytics Inc. · United States; Canada
The role
This role involves deploying, integrating, and scaling enterprise Generative AI solutions directly within customer engineering teams. The engineer will operationalize Large Language Models and retrieval systems on Google Cloud Platform, bridging AI research and production infrastructure. The position requires deep expertise in GCP, Vertex AI, and LLM orchestration.
- Location
- United States; Canada
- Work mode
- Onsite
- Level
- Senior
What we know that the posting doesn’t say
- Seen todaystill listed on the employer’s careers page
- Posted 25 days agothe first time we saw it
About Tiger Analytics Inc.
Tiger Analytics is an advanced analytics consulting firm that partners with Fortune 500 companies to generate business value from data, with expertise in Machine Learning, Data Science, and AI.
What you would do
- Deploy, fine-tune, and optimize large-scale Gen AI models in customer cloud environments.
- Architect scalable infrastructure for AI workloads using GPU/TPU orchestration and high-performance storage.
- Design and implement high-throughput data ingestion pipelines and Vector Database architectures for RAG.
- Act as primary technical consultant guiding clients on AI safety, prompt engineering, and cost optimization.
- Feed edge-case deployment insights back to core AI research and platform engineering teams.
- Manage customer expectations around LLM non-determinism, hallucinations, and performance trade-offs.
Must have
- Advanced knowledge of Vertex AI primitives including Studio, Model Registry, and Endpoint deployment.
- Hands-on experience with LLM orchestration tools like LangChain, LlamaIndex, and AutoGen.
- Experience with deep learning frameworks such as PyTorch and Hugging Face on GCP.
- Production experience with Vertex AI Vector Search or managed vector stores like Milvus or Pinecone.
- Proficiency in model serving frameworks like vLLM, TGI, or Triton Inference Server.
- Deep expertise in Google Kubernetes Engine for managing GPU/TPU workloads and autoscaling.
- Mastery of Terraform to provision secure GCP environments, IAM roles, and Vertex AI resources.
- Strong coding skills in Python or Go with emphasis on clean, concurrent code.
- Experience with Google Cloud SDK.
What you get
- Opportunity for significant career development in a fast-growing environment.
- High degree of individual responsibility.
- Equal employment opportunities.
Key skills
- gcp
- vertex ai
- vertex ai studio
- model registry
- endpoint deployment
- vertex ai pipelines
- kubeflow
- vertex ai vector search
- langchain
- llamaindex
- autogen
- pytorch
- hugging face
- milvus
- pinecone
- pgvector
- cloud sql
- spanner
- vllm
- tgi
- triton inference server
- gke
- terraform
- python
Tiger Analytics is looking for experienced Forward Deployed Engineer (Generative AI) with Gen AI experience to join our fast-growing advanced analytics consulting firm. Our employees bring deep expertise in Machine Learning, Data Science, and AI. We are the trusted analytics partner for multiple Fo…
Extracted from the employer’s posting. Read it in full on Tiger Analytics Inc.’s careers pageApply without filling the form
Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.
Other roles at Tiger Analytics Inc.
- Senior Data Analyst- BankingMcLean, Virginia, United States; Richmond, Virginia, United States
- Forward Deployed AI Engineer -Neo4j / Knowledge GraphRemote
- Senior Data Product Manager - BankingMcLean, Virginia, United States
- Senior/Lead Data Scientist - Agentic AINew York, New York, United States; San Francisco, California, United States; Dallas, Texas, United States
- Forward Deployed Engineer (Generative AI)Austin, Texas, United States; Dallas, Texas, United States; Chicago, Illinois, United States; New York, New York, United States
- MLOps Lead EngineerSt. Louis, Missouri, United States