Further
Senior AI Engineer
Cleveland, Ohio; Dallas, Texas · Senior
Sponsorship not specifiedDetected 12 days ago
TypeScriptPythonFastAPIBackend DevelopmentCode ReviewGitPostgreSQLVector DatabasesGCPCloud PlatformsGraphQLRESTgRPCMachine LearningData ScienceLLMsRAGAgentic AILLMOpsLangGraph
About the role
- If you love data and are looking for unlimited growth opportunities, we want to talk with you about joining Further.
- We have an award winning culture of extraordinary people.
Responsibilities
- Lead the implementation of rigorous evaluation frameworks to monitor model performance, drift, and cost in real-time.
- Architect and develop high-performance backend services and APIs using Python (FastAPI) to serve large language models at scale.
- Design advanced Retrieval-Augmented Generation (RAG) systems, selecting and managing vector databases and optimizing embedding strategies for accuracy and speed.
- Establish comprehensive model observability and guardrail systems to monitor real-time performance, detect distribution drift, and implement automated safety filters that mitigate hallucinations, bias, and toxic outputs in production environments.
- Build robust integration layers that connect AI agents securely to external enterprise systems, CRMs, and legacy databases.
- Collaborate with infrastructure teams to define deployment strategies, ensuring solutions scale dynamically under load.
Requirements
- 3 years of experience working directly with GenAI/RAG/LLM architecture
- Expert proficiency in Python AI application development and modern API architecture (REST, GraphQL, gRPC) using enterprise standards like static type checking and data validation.
- Hands-on expertise with vector databases (Pinecone, Weaviate, PostgreSQL) and search algorithms.
- Strong understanding of LLMOps principles, including model registry, versioning, and serving infrastructure specifically in Google Cloud.
Nice to have
- Experience in Typescript development for prototyping and integrations
- Proficiency with git workflows and understanding of standard application development processes
- Knowledge of advanced prompt engineering and fine-tuning techniques (LoRA, PEFT).
- Experience optimizing inference costs and latency for large-scale deployments.
- Previous experience in a client-facing consulting role, managing diverse stakeholders and navigating complex organizational structures.
- Any Google Cloud Professional Certification
- What you'll be doing in this role:
- Define the end-to-end architecture for AI products on cloud platforms (preferably Google Cloud Platform), ensuring high availability, security, and cost-effectiveness.
Skills
- Further is a data, cloud, and AI company whose focus is helping companies turn raw data into the right decisions.
- Our purpose is to enable people to thrive so that businesses can thrive.
Benefits
- Conduct code reviews, provide technical guidance, and foster a culture of continuous learning and innovation within the engineering team.
Apply directly at Further →Create a free account for alerts like thisView Further immigration profile
This listing is sourced directly from Further's careers page and normalized into a canonical job model.