Full-time · Hanoi
-
Job Type: Full-time, Onsite
-
Location: H10 Building, 475 Nguyen Trai, Thanh Xuan, Hanoi
-
Working Hours: Monday - Friday (08:00 AM - 05:30 PM)
-
Website: https://wigin.ai/
ABOUT WIGIN AI:
Wigin is a product-led AI company powered by a team of leading engineers across the artificial intelligence and technology sectors. The company specializes in building production-grade solutions that transform frontier AI research into reliable, scalable systems engineered to deliver real, measurable value for businesses. Rather than just creating theoretical models, Wigin focuses on engineering practical products that win at scale and optimize how businesses actually run.
The company operates across three core domains:
(1) AI Services
(2) AI Investment Intelligence
(3) AI Products
ABOUT THE ROLE:
We are looking for a Full-time
AI Engineer
to join our team and work on projects involving
Large Language Models (LLMs) and Generative AI
.
You will work closely with a technical team to design, fine-tune, evaluate, optimize, and deploy advanced language models for real-world applications. This is a great opportunity for candidates who want to work deeply with frontier LLMs, Retrieval-Augmented Generation (RAG) architecture, AI Agent systems, and large-scale GPU infrastructure.
WHAT YOU'LL DO:
-
Research, develop, and apply state-of-the-art Large Language Models (LLMs) and Generative Text/Multimodal AI models.
-
Build, optimize, and evaluate advanced LLM applications, including RAG (Retrieval-Augmented Generation), AI Agents, and workflow automation systems.
-
Design, pre-train, fine-tune (SFT, LoRA/QLoRA), and align (RLHF/DPO) open-source LLMs for domain-specific tasks.
-
Optimize model latency, throughput, and inference cost using modern frameworks and techniques.
-
Read research papers, experiment with new methods (Prompt Engineering, Context Extension, Function Calling), and apply them to practical AI products.
-
Collaborate with backend and product engineering teams to integrate LLM pipelines into scalable production systems.
-
Monitor model safety, hallucinations, performance, and continuously improve model quality.
WHAT WE'RE LOOKING FOR:
-
Strong foundation in Machine Learning, Deep Learning, and Natural Language Processing (NLP). This position is open to Vietnamese candidates only.
-
Solid understanding of Transformer architectures, Attention mechanisms, and modern LLM paradigms.
-
Hands-on experience working with LLMs (commercial APIs like OpenAI/Anthropic or open-source models like Llama/Qwen).
-
Practical experience in building
RAG architectures
, Vector Databases (e.g., Milvus, Qdrant, Pinecone), or
AI Agent frameworks
(e.g., LangChain, LlamaIndex, AutoGen/CrewAI).
-
Ability to read and implement research papers and technical documentation effectively.
-
Strong problem-solving skills, clean coding practices, and a proactive engineering mindset.
-
Ability to work independently and collaborate effectively with a multi-disciplinary team.
Nice to have
-
Experience in fine-tuning LLMs (LoRA, QLoRA, DeepSpeed, Unsloth, PEFT).
-
Experience with LLM inference optimization engines (e.g., vLLM, TensorRT-LLM, Ollama, TGI) and techniques (Quantization, Speculative Decoding).
WHAT WE OFFER
-
Competitive salary package and benefits:
Up to 40M.
-
Flexible working time.
-
Young, dynamic, and growth-oriented working environment.
-
Access to premium technical resources and AI tools (Claude Max and Codex).
-
Opportunity to work with large GPU infrastructure and frontier LLM frameworks.
To apply, send your resume and a short note about what you've built to
hr@wigin.ai
or
Contact 035.602.1236
(Ms.Thu Trang)