Back to all roles
AI Engineering

LLM Engineer

We're seeking an exceptional LLM Engineer to build and scale production AI systems for our portfolio of Series A–C companies. You'll work on cutting-edge AI infrastructure, ship features that directly impact product velocity, and help define what great AI engineering looks like in a growth-stage context.

📧 Email your resume tohr@thehirehare.com

What you'll do.

  • 🚀 Design, build, and deploy production LLM applications using state-of-the-art models (GPT-4, Claude, Gemini, Llama 3.1, Mistral)
  • 🎯 Implement advanced prompt engineering, few-shot learning, and RLHF pipelines to optimize model performance
  • 📊 Build comprehensive evaluation frameworks to measure model quality, latency, cost, and hallucination rates
  • 🏗️ Architect context management systems, RAG pipelines, and hybrid retrieval strategies
  • ⚡ Optimize inference infrastructure for cost and latency (prompt caching, request batching, quantization, speculative decoding)
  • 🔄 Own the full lifecycle from prototyping to production deployment, monitoring, and continuous improvement

What we're looking for.

Requirements

  • 3+ years of production experience building and shipping AI/ML systems (or 5+ for senior roles)
  • Strong programming skills in Python and experience with modern AI frameworks (PyTorch, TensorFlow, JAX)
  • Deep understanding of LLMs, transformers, and the latest AI research
  • Production experience with cloud infrastructure (AWS, GCP, Azure) and containerization (Docker, Kubernetes)
  • Strong systems thinking — you understand trade-offs between accuracy, latency, cost, and reliability
  • Growth-stage experience preferred — you thrive in fast-moving, ambiguous environments

Nice to have

  • Open-source contributions or published research in AI/ML
  • Experience working at an AI-first company or research lab
  • Familiarity with the latest AI tools, frameworks, and models
  • Strong writing and documentation skills
  • Experience mentoring or leading junior team members

🛠️ Tech stack.

Python, TypeScript
OpenAI API, Anthropic Claude, Google Gemini, Llama 3.1, Mistral
LangChain, LlamaIndex, Haystack, DSPy
Vector DBs: Pinecone, Weaviate, Qdrant, pgvector
PyTorch, Transformers, vLLM, Text Generation Inference
Docker, Kubernetes, AWS/GCP/Azure
FastAPI, Redis, PostgreSQL

Interested in this role?

📧 Send your resume tohr@thehirehare.com