Custom RAG Application & AI Support Chatbot Development

Employer not named by the sourceRemote

AI/MLFull StackFrontendAI Tooling

Apply on the company’s site

Frontier is not the employer and does not collect applications.

About this role

Python, Frontend Development, FastAPI, REST API, Large Language Model, LangChain, AI Chatbot Development, AI Model Development, Vector Databases, AI Integration · We are seeking an experienced AI/LLM engineer to develop an end-to-end Retrieval-Augmented Generation (RAG) system for automated customer support chat.

Scope of Work • Data Ingestion & Preprocessing: Parse and chunk internal documentation (PDFs, Markdown, FAQs, and ticket logs) with automated re-indexing.

• Vector Storage & Semantic Search: Set up and optimize a vector database (e.g., Chroma, Pinecone, Weaviate, or pgvector) for hybrid/semantic retrieval.

• LLM Pipeline: Integrate LLMs (OpenAI GPT-4/3.5, Claude, or open-source models) with prompt chaining that provides accurate answers and explicit source citations.

• Backend & API: Build secure, low-latency REST/WebSocket endpoints using Python (FastAPI).

• Frontend Interface: Lightweight chat widget or React component ready to embed into our platform.

Key Acceptance Criteria Sub-3 second latency for end-to-end question retrieval and response generation.

Grounded responses strictly based on retrieved context to prevent hallucinations.

Clean, modular codebase provided in a GitHub repository with Docker setup scripts and concise documentation.

To Apply: Please share brief links/examples of RAG pipelines or LLM applications