Middle AI Research Engineer (Focus LLM)
Hạn nộp hồ sơ: 23/09/2026 (Còn 28 ngày)
Ứng tuyển sớm để được ưu tiên
Kết nối với Nhà tuyển dụng để tìm hiểu thông tin và gia tăng cơ hội trúng tuyển
Nhà tuyển dụng đang online
Mô tả công việc
ABOUT WIGIN AI:
Wigin is a product-led AI company powered by a team of leading engineers across the artificial intelligence and technology sectors. The company specializes in building production-grade solutions that transform frontier AI research into reliable, scalable systems engineered to deliver real, measurable value for businesses. Rather than just creating theoretical models, Wigin focuses on engineering practical products that win at scale and optimize how businesses actually run.
The company operates across three core domains:
(1) AI Services
(2) AI Investment Intelligence
(3) AI Products
ABOUT THE ROLE:
We are looking for a Full-time AI Engineer to join our team and work on projects involving Large Language Models (LLMs) and Generative AI.
You will work closely with a technical team to design, fine-tune, evaluate, optimize, and deploy advanced language models for real-world applications. This is a great opportunity for candidates who want to work deeply with frontier LLMs, Retrieval-Augmented Generation (RAG) architecture, AI Agent systems, and large-scale GPU infrastructure
WHAT YOU'LL DO:
Research, develop, and apply state-of-the-art Large Language Models (LLMs) and Generative Text/Multimodal AI models.
Build, optimize, and evaluate advanced LLM applications, including RAG (Retrieval-Augmented Generation), AI Agents, and workflow automation systems.
Design, pre-train, fine-tune (SFT, LoRA/QLoRA), and align (RLHF/DPO) open-source LLMs for domain-specific tasks.
Optimize model latency, throughput, and inference cost using modern frameworks and techniques.
Read research papers, experiment with new methods (Prompt Engineering, Context Extension, Function Calling), and apply them to practical AI products.
Collaborate with backend and product engineering teams to integrate LLM pipelines into scalable production systems.
Monitor model safety, hallucinations, performance, and continuously improve model quality.
Wigin is a product-led AI company powered by a team of leading engineers across the artificial intelligence and technology sectors. The company specializes in building production-grade solutions that transform frontier AI research into reliable, scalable systems engineered to deliver real, measurable value for businesses. Rather than just creating theoretical models, Wigin focuses on engineering practical products that win at scale and optimize how businesses actually run.
The company operates across three core domains:
(1) AI Services
(2) AI Investment Intelligence
(3) AI Products
ABOUT THE ROLE:
We are looking for a Full-time AI Engineer to join our team and work on projects involving Large Language Models (LLMs) and Generative AI.
You will work closely with a technical team to design, fine-tune, evaluate, optimize, and deploy advanced language models for real-world applications. This is a great opportunity for candidates who want to work deeply with frontier LLMs, Retrieval-Augmented Generation (RAG) architecture, AI Agent systems, and large-scale GPU infrastructure
WHAT YOU'LL DO:
Research, develop, and apply state-of-the-art Large Language Models (LLMs) and Generative Text/Multimodal AI models.
Build, optimize, and evaluate advanced LLM applications, including RAG (Retrieval-Augmented Generation), AI Agents, and workflow automation systems.
Design, pre-train, fine-tune (SFT, LoRA/QLoRA), and align (RLHF/DPO) open-source LLMs for domain-specific tasks.
Optimize model latency, throughput, and inference cost using modern frameworks and techniques.
Read research papers, experiment with new methods (Prompt Engineering, Context Extension, Function Calling), and apply them to practical AI products.
Collaborate with backend and product engineering teams to integrate LLM pipelines into scalable production systems.
Monitor model safety, hallucinations, performance, and continuously improve model quality.
Yêu cầu
Strong fundamentals in Machine Learning, Deep Learning, and NLP; deep understanding of Transformer architecture and Attention mechanisms.
Research background (thesis/dissertation, publications, university research labs, ML/Kaggle competitions, reproducing paper results) with a desire to transition into applied AI engineering to build real-world products.
2-3 years of AI/ML experience working with LLMs, using commercial APIs (OpenAI, Anthropic) or open-source models (Llama, Qwen). Preference for candidates born between 1999-2002 who graduated from top tech universities (HUST, UET, etc.).
Hands-on experience building RAG systems, Vector DBs (Milvus, Qdrant, Pinecone), or Agent Frameworks (LangChain, LlamaIndex, AutoGen/CrewAI).
Ability to read, understand, and independently implement research papers.
Research mindset: driven to understand why a solution works rather than just making the code run.
Strong problem-solving skills and the ability to work independently.
Nice to Have:
Fine-tuning experience: LoRA, QLoRA, DeepSpeed, Unsloth, PEFT.
Inference optimization: vLLM, TensorRT-LLM, Ollama, TGI, Quantization, Speculative Decoding.
Research background (thesis/dissertation, publications, university research labs, ML/Kaggle competitions, reproducing paper results) with a desire to transition into applied AI engineering to build real-world products.
2-3 years of AI/ML experience working with LLMs, using commercial APIs (OpenAI, Anthropic) or open-source models (Llama, Qwen). Preference for candidates born between 1999-2002 who graduated from top tech universities (HUST, UET, etc.).
Hands-on experience building RAG systems, Vector DBs (Milvus, Qdrant, Pinecone), or Agent Frameworks (LangChain, LlamaIndex, AutoGen/CrewAI).
Ability to read, understand, and independently implement research papers.
Research mindset: driven to understand why a solution works rather than just making the code run.
Strong problem-solving skills and the ability to work independently.
Nice to Have:
Fine-tuning experience: LoRA, QLoRA, DeepSpeed, Unsloth, PEFT.
Inference optimization: vLLM, TensorRT-LLM, Ollama, TGI, Quantization, Speculative Decoding.
Quyền lợi
1. Competitive Compensation & Comprehensive Benefits
Competitive salary package, up to 40M VND/month
Attractive project bonus for outstanding contributions and successful project delivery.
13th-month salary and full benefits in accordance with Vietnamese labor law
Annual leave, social insurance, health insurance and other statutory benefits
Regular salary review twice a year.
2. Flexible & Modern Working
Monday-Friday, with flexible check-in time from [protected info] WFH day/week
Focus on results, ownership, and sustainable work-life balance
3. Work on Challenging AI Projects
Work on diverse, challenging, and high-impact AI projects
Directly tackle real-world problems across LLM, Generative AI, AI Agents, RAG and AI infrastructure
Work with large-scale GPU infrastructure and modern AI frameworks
Freedom to experiment, research, and turn ideas into production-ready products
4. High Ownership & Direct Impact
Join a small, highly technical team with a high level of autonomy
Work directly with the CEO and core technical team
Take ownership of meaningful problems and see your work go from idea to production
Fast decision-making, minimal bureaucracy, and plenty of room to make an impact
5. Premium AI Tools & Resources
Unlimited access to Claude Max and Codex
Access to large-scale GPU infrastructure for research, experimentation, and model development
Premium technical resources and the latest AI tools to accelerate your work
Competitive salary package, up to 40M VND/month
Attractive project bonus for outstanding contributions and successful project delivery.
13th-month salary and full benefits in accordance with Vietnamese labor law
Annual leave, social insurance, health insurance and other statutory benefits
Regular salary review twice a year.
2. Flexible & Modern Working
Monday-Friday, with flexible check-in time from [protected info] WFH day/week
Focus on results, ownership, and sustainable work-life balance
3. Work on Challenging AI Projects
Work on diverse, challenging, and high-impact AI projects
Directly tackle real-world problems across LLM, Generative AI, AI Agents, RAG and AI infrastructure
Work with large-scale GPU infrastructure and modern AI frameworks
Freedom to experiment, research, and turn ideas into production-ready products
4. High Ownership & Direct Impact
Join a small, highly technical team with a high level of autonomy
Work directly with the CEO and core technical team
Take ownership of meaningful problems and see your work go from idea to production
Fast decision-making, minimal bureaucracy, and plenty of room to make an impact
5. Premium AI Tools & Resources
Unlimited access to Claude Max and Codex
Access to large-scale GPU infrastructure for research, experimentation, and model development
Premium technical resources and the latest AI tools to accelerate your work
Thông tin khác
Thời gian làm việc
Thứ 2 - Thứ 6 (từ 08:00 đến 17:30)
Thứ 2 - Thứ 6 (từ 08:00 đến 17:30)
Thông tin chung
- Thu nhập: 25 - 40 triệu
Nơi làm việc
- Hà Nội: Tòa H10, Ngõ 475 Nguyễn Trãi, Phường Thanh Liệt (huyện Thanh Trì cũ)
Việc làm tương tự khác
Tổng Công ty Mạng lưới Viettel - Chi nhánh Tập đoàn Công nghiệp - Viễn thông Quân đội
Hà Nội, Thái Bình
15 - 25 triệu
NGÂN HÀNG TMCP ĐẠI CHÚNG VIỆT NAM - PVcomBank
Hà Nội, Hải Phòng
Thương lượng
Đội ngũ hỗ trợ của JobOKO sẵn sàng đồng hành, tư vấn và giới thiệu những cơ hội việc làm phù hợp, giúp Ứng viên tự tin phát triển sự nghiệp và chinh phục mục tiêu nghề nghiệp bền vững.
Hotline CSKH
1900.63.63.84
Công ty Cổ phần JobOKO Toàn cầu
Đội ngũ hỗ trợ của JobOKO luôn chủ động tư vấn các giải pháp tuyển dụng tối ưu, cam kết đồng hành và hỗ trợ Quý Nhà tuyển dụng đạt được hiệu quả tuyển dụng bền vững.
Hotline CSKH
0962.107.888
Công ty Cổ phần JobOKO Toàn cầu