Search by job, company or skills

  • Posted 22 days ago
  • Be among the first 10 applicants

Job Description

Company Overview:

At TechX, we are pioneers in delivering cutting-edge solutions that empower businesses to thrive in today's digital landscape. With a strong focus on Cloud Transformation (AWS), Data Modernization, and Generative AI, we bring unparalleled expertise to drive innovation, efficiency, and growth for our clients.

We specialize in banking and financial services, retail, manufacturing, and transportation, delivering impactful solutions that address industry-specific challenges and significant market shifts.

Together, we'll build a future where innovation knows no limits.

Job Overview:

We are looking for an AI Engineer focused on Retrieval-Augmented Generation (RAG) and knowledge pipelines. You will build the ingestion, indexing, and retrieval layer that lets our AI agents answer with cited, traceable, domain-specific knowledge. This layer is foundational to the accuracy and trust of our AI products.

Key Responsibilities:

  • Design and implement RAG pipelines: document ingestion, chunking, embedding, indexing, retrieval, and citation.
  • Build domain knowledge packs with source citation and traceable answers.
  • Implement data classification, access control, and quality checks for sensitive / regulated corpora.
  • Tune retrieval quality — embeddings, re-ranking, hybrid search — and support evaluation of answer accuracy and grounding.
  • Integrate the knowledge layer with the agent runtime and orchestration APIs.
  • Collaborate with Solution Engineers and customers to onboard real document corpora into demo- and production-quality packs.

Key Requirements:

  • Minimum 2+ years of experience in GenAI or NLP.
  • Strong hands-on experience with RAG and Agentic frameworks (LangChain, LlamaIndex), and a solid understanding of RAG, Agentic, and DeepSearch.
  • Experience integrating with vector databases including Qdrant, ChromaDB, OpenSearch, and Pinecone.
  • Understanding of the Model Context Protocol (MCP).
  • Familiar with AI platforms such as Azure AI, Vertex AI; bonus if familiar with AWS Bedrock / SageMaker and chatbot development.
  • Familiar with the Hugging Face transformers library, PyTorch, and TensorFlow.
  • Strong foundation in Computer Science.
  • Experience implementing REST API, gRPC, WebSocket, microservices, and web technologies.
  • Experience with vLLM, llama.cpp, and GGUF-type models.
  • Understanding of and proficiency in prompt engineering.

What We Offer:

  • Innovative Environment: A dynamic and collaborative work environment where innovation is encouraged.
  • Professional Growth: Opportunities for professional growth and development through continuous learning and exposure to cutting-edge technologies.
  • Competitive Package: Competitive salary and benefits package
  • Impactful Work: The chance to work on cutting-edge projects with industry-leading clients, making a tangible impact on their business success.

Apply now!

Due to the high volume of applicants, only shortlisted candidates will be contacted. We apologize for this inconvenience!

More Info

Job Type:
Industry:
Employment Type:

About Company

Job ID: 150703905

Similar Jobs

Ho Chi Minh, Vietnam

Skills:

DatabasesDockerMicrosoft AzureKubernetesPythonData warehouse solutionsAgentic AI frameworksAI ML applicationsMLOps tools

Ho Chi Minh, Vietnam

Skills:

agentic workflowsLLM APIsLLMsAI toolingvector databasesClaudebackend frameworksAI orchestration tooling

Ho Chi Minh, Vietnam

Skills:

Salesforce CrmMLopsPythonpgvectorAI Agentsvector databasesAnthropicPineconeLLM ecosystemChromaLlamaOpenAIWeaviateGemini

Ho Chi Minh, Vietnam

Skills:

TensorflowNlpMachine LearningPytorchPythonComputer VisionDeep LearningOpenAI’s GPTRetrieval-Augmented Generation

Vietnam, Ho Chi Minh

Skills:

KubernetesPythonDockerLangChainMilvusLangfusePydanticVector databasesaiohttpLangSmithW BPineconeLangGraphasyncioQdrant