NextMantra AI
Website:
beremarkable.ai
Company:
https://www.linkedin.com/company/next-mantra-ai
Seniority: Entry level
Industries: IT Services and IT Consulting
Job details:
Company Description NextMantra AI is a technology-driven talent intelligence company that helps organizations build scalable digital capabilities and high-impact technology teams. The company combines AI-powered hiring technology with expert-led talent services to accelerate digital transformation and deliver complex technology initiatives faster. Its proprietary AI Adaptive Cross-Questioning Video Assessment Platform conducts dynamic, real-time technical interviews, ensuring only rigorously evaluated, project-ready candidates reach hiring managers. In addition to the platform, NextMantra AI offers IT consulting and specialized talent solutions with domain expert evaluation and hands-on technical screening. The company delivers strategic talent and technology consulting across high-growth industries such as financial services, SaaS, fintech, healthtech, e-commerce, deep tech, gaming, and more, with coverage across SEA, MENA, APAC, and the USA.
Job Description – Generative AI Engineer (RAG)
📍 Location: Gurgaon
💼 Experience: 2–5 Years
🕒 Employment Type: Full-Time
About the Role
We are looking for a passionate Generative AI Engineer with hands-on experience in Retrieval-Augmented Generation (RAG), Speech-to-Text (STT), and Text-to-Speech (TTS) technologies. The ideal candidate will design and develop AI-powered conversational applications, intelligent assistants, and enterprise GenAI solutions using modern LLM frameworks and cloud AI services.
Key Responsibilities
Design, develop, and deploy Generative AI applications using Large Language Models (LLMs).
Build and optimize RAG (Retrieval-Augmented Generation) pipelines for enterprise knowledge retrieval.
Develop Speech-to-Text (STT) and Text-to-Speech (TTS) solutions for voice-enabled AI applications.
Integrate LLMs such as OpenAI GPT, Claude, Gemini, or Llama into production applications.
Develop AI chatbots, AI agents, and virtual assistants.
Build vector search solutions using Pinecone, Weaviate, ChromaDB, FAISS, or similar vector databases.
Optimize prompt engineering, context management, and response quality.
Work with structured and unstructured data sources, including PDFs, documents, APIs, and databases.
Collaborate with product and engineering teams to deliver scalable AI solutions.
Monitor, evaluate, and continuously improve AI model performance.
Required Skills
✅ 2–5 years of software development experience with Python.
✅ Strong hands-on experience with:
Retrieval-Augmented Generation (RAG)
Large Language Models (LLMs)
Prompt Engineering
LangChain or LlamaIndex
OpenAI, Claude, Gemini, or open-source LLMs
Vector Databases (Pinecone, FAISS, ChromaDB, Weaviate)
✅ Experience with:
Speech-to-Text (Whisper, Deepgram, Azure Speech, Google Speech API, etc.)
Text-to-Speech (ElevenLabs, Azure TTS, Google TTS, Amazon Polly, etc.)
✅ Strong knowledge of:
REST APIs
FastAPI or Flask
Python
Git
SQL/NoSQL databases
✅ Familiarity with cloud platforms such as AWS, Azure, or GCP.
Preferred Skills
Experience with AI Agents and Agentic AI frameworks.
Knowledge of MCP (Model Context Protocol) and AI tool integration.
Experience deploying AI applications using Docker and Kubernetes.
Familiarity with Hugging Face Transformers and open-source LLMs.
Experience with OCR, document processing, and multimodal AI.
Knowledge of monitoring and evaluation frameworks for LLM applications.
What We Offer
Opportunity to work on cutting-edge Generative AI products.
Exposure to enterprise-scale AI solutions.
Collaborative and innovative work environment.
Competitive salary and career growth opportunities.
Click on Apply to know more.