General informationRef Id: HREZ-12478
Published Date: Jun 4, 2026
Department:
Department
Job location:
Job location
Work arrangement: Onsite
Job ID: 3516516
Employment type:
Employment type
Number of openings:
Number of openings
Location: CA, USA
About the Role:
We are seeking a talented AI Engineer specializing in Large Language Model (LLM) applications to join our growing AI team at a pioneering AI startup in Silicon Valley. You will help develop and deploy AI systems that leverage foundation models to solve real-world business problems.
Responsibilities:
- Build LLM-powered applications including RAG systems and AI agents using frameworks like LangChain or LlamaIndex
- Assist with fine-tuning and evaluating pre-trained models for domain-specific use cases
- Help build and maintain ML inference infrastructure on cloud platforms
- Develop evaluation frameworks to measure model quality and track performance over time
- Implement prompt engineering strategies and context management techniques
- Collaborate with product managers and engineers to integrate AI features into products via APIs
- Stay current with AI research trends and prototype new techniques for potential adoption
- Support responsible AI practices including basic guardrails and model monitoring
Education:
Bachelor's or Master's degree in Computer Science, Artificial Intelligence, Machine Learning, or a related technical field.
Experience:
3+ years of software engineering experience with at least 1 year working on machine learning or AI-related projects.
Required Skills:
1. Proficiency in Python with experience using PyTorch or Hugging Face Transformers; familiarity with experiment tracking tools like MLflow or Weights & Biases
2. Hands-on experience building LLM applications; familiarity with vector databases (Pinecone, Weaviate, or pgvector) and embedding models
3. Basic understanding of transformer architecture and common LLM concepts such as tokenization, context windows, and quantization
4. Familiarity with MLOps practices including model versioning, CI/CD for ML, and cloud-based deployment
5. Experience with at least one cloud platform (AWS, GCP, or Azure) for deploying ML workloads
6. Solid Python software engineering skills: API development with FastAPI or Flask, writing clean and testable code
7. Awareness of AI safety considerations and responsible AI principles
Compensation:
Salary: $170,000 - $230,000 per year
Benefits: Top-tier health benefits, equity package, 401(k) with 6% company match, $5,000 annual learning budget, conference sponsorship, and GPU workstation
Please input Company info text. This section will be hidden when left empty.
< Back to job list