

Search by job, company or skills

Company Description
VinSmart Future (VSF) is the leading technology company within the Vingroup Corporation, formed by the merger of the group's entire technology ecosystem, including VinApp, VinIT, VinBigdata, and other tech units. As a core driver of Vingroup's future growth, VSF is at the forefront of technological development, with artificial intelligence (AI) as its foundation. With a talented team of nearly 4,000 local and international technology experts, VSF focuses on creating high-utility technologies that enhance lives and connect data, models, and infrastructure to unlock new possibilities.
Job Description
• Design, train, and deploy AI models for production use, covering CV, NLP, generative AI, and speech systems (ASR, TTS, speech-to-speech).
• Build and optimize pipelines for pretraining, fine-tuning, inference, and evaluation, applying
techniques such as transformers, RAG, RLHF/DPO, and model optimization (distillation, quantization).
• Adapt foundation models (e.g., Whisper, Qwen, XTTS) for Vietnamese and multi-domain use cases, improve performance (accuracy, latency, naturalness), and develop conversational AI systems like voice assistants.
• Collaborate with cross-functional teams to deliver end-to-end AI solutions, build data pipelines and evaluation frameworks, and continuously apply new AI research into real-world applications.
Requirements
Strong foundation (mandatory):
• Python (advanced), solid coding practices (clean, modular, production-ready)
• Deep Learning frameworks: PyTorch (preferred) or TensorFlow
• Strong fundamentals: ML/DL, NLP, linear algebra, probability, optimization
Model development & deployment:
• Hands-on experience training, fine-tuning, and deploying models to production
• Experience with transformer-based models (LLMs, seq2seq)
• Familiar with RAG, LLM applications OR model inference pipelines
At least 3+ YOE in ONE domain expertise:
• NLP/LLM: Transformers, RAG, prompt engineering, conversational AI
• Speech AI: ASR or TTS, working with models like Whisper, XTTS, speech pipelines
• Computer Vision: Detection, tracking, segmentation
• Agentic AI: Multi-agent systems, tool-using agents, workflows
Model training & optimization:
• Experience with fine-tuning (SFT), domain adaptation, evaluation metrics
• Familiar with at least one: RLHF / DPO / PPO OR model optimization (quantization, distillation)
Engineering & deployment:
• Experience with API development (FastAPI/Flask) or microservices
• Familiar with Docker, CI/CD, cloud (AWS/GCP)
• Working knowledge of ML lifecycle / MLOps (MLflow, W&B is a plus)
Nice to Have
• Experience in Vietnamese Speech AI (ASR/TTS) or multi-accent systems
• Experience in Smart City / large-scale systems / real-time AI
• Knowledge of streaming systems, low-latency inference, speech-to-speech AI
• Background in research, publications, or open-source AI projects
Contact:
Ms. Nguyệt Zalo/Call: 0769 288 088
Email: [Confidential Information]
Office:
Hanoi: Technopark, Gia Lam
Job ID: 152370677
Skills:
Java, Microsoft Office 365, C, Cassandra, Node.js, Sql, Javascript, Gcp, Docker, Elasticsearch, Microsoft Azure, MongoDB, Kubernetes, Python, AWS, no-SQL, LXD, LXC
Skills:
pruning , Nlp, Pytorch, Python, MLops, Memory, Agentic systems, multi-agent orchestration, HuggingFace, TensorRT, knowledge distillation, ONNX, Llm, TFLite, Transformer LLM architectures, on-device inference, prompt context engineering, quantization, fine-tuning, LLMOps, model compression, vLLM, tool use planning, RAG, function calling
Skills:
Tensorflow, Nlp, MLops, Pytorch, FastAPI, Ocr, Python, Computer Vision, LangChain, LangGraph, Conversational AI
Skills:
AWS, Pytorch, Tensorflow, Python, Azure, Gcp, LangChain, AI ML frameworks, MLOps practices, HuggingFace
Skills:
container , Data Structures And Algorithms, Message Queue, Hadoop, Sql, Microservices, Nosql, Spark, Python, Tensor, Deep learning models, Text data, Design Patterns for Agentic AI, Large Language Model APIs, IT system architecture, AI programming libraries