AI Engineer — LLM / VLM
New
J
JobgetherArtificial Intelligence
Remote work from India, with the role based in Kolkata and structured as a remote position.Full-Time
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Required Skills
- DockerPythonPyTorchFastAPILLM
Requirements
- Strong Python programming skills and software engineering fundamentals.
- Hands-on professional experience with LLMs and/or VLMs and generative AI systems.
- Solid understanding of Transformers, attention mechanisms, tokenization, and embeddings.
- Practical experience with PyTorch and Hugging Face Transformers.
- Hands-on experience designing and implementing RAG systems and vector-search solutions.
- Knowledge of prompt engineering and LLM evaluation methodologies.
- Experience developing APIs and REST services with Git, Docker, and CI/CD.
- Familiarity with vector databases such as FAISS, Milvus, Pinecone, Weaviate, or pgvector.
- Understanding of cloud-based AI infrastructure (AWS, Azure, or GCP).
- Experience with multimodal models like Qwen-VL, LLaVA, Gemini, or GPT vision models.
- Strong analytical and problem-solving abilities.
Responsibilities
- Design, develop, and deploy production-ready AI applications powered by Large Language Models and Vision-Language Models.
- Build end-to-end RAG pipelines covering document ingestion, chunking, embedding generation, retrieval, and reranking.
- Apply prompt engineering, supervised fine-tuning, and LoRA/QLoRA to improve model performance.
- Develop multimodal AI solutions that process text, images, PDFs, charts, and tables.
- Build AI agents and tool-calling workflows that combine reasoning, retrieval, and external tools.
- Optimize model inference and serving for latency, throughput, and cost efficiency.
- Develop production APIs using Python, FastAPI, and Docker while applying CI/CD practices.
- Create evaluation frameworks to measure AI system accuracy, hallucination rates, and safety.
View Full Description & ApplyYou'll be redirected to the employer's site