Urgent

AI Engineer - Video Analysis Core

Closed

Computer Vision AI/ML AI/Artificial Intelligence

勤務地

Ho Chi Minh

総空席数

1 人

福利厚生

フル社会保険

年間給与の見直し

旅行/会社の旅行

仕事用のノートパソコン/デスクトップ

業績ボーナス

追加の健康保険

13ヶ月目の給与

その他の福利厚生

職務概要

Established for over 20 years in Vietnam with more than 350 members, GIANTY is proud of providing numerous successful services to millions of users worldwide OUR CORE SERVICES: Development Services (Outsource & Offshore), In-House Development (Apps, Games, Anime, Arts, Design, Big Data/Analytics, IoT, VR, AI) Your Mission You will be the founding AI engineer to architect and build the core video analysis engine. The system will combine multi-modal perception (video, audio, text), LLMs, and multi-agent reasoning to understand human behavior, actions, and outcomes from raw video streams. You will work closely with product leaders and domain experts to turn ideas into working features, ensuring our platform adapts across domains while staying fast, accurate, and scalable. - Design and implement the AI pipeline for video understanding, including: - Frame extraction; object and behavior detection - Temporal event segmentation and context tagging - Multimodal fusion (vision, audio, transcripts, sensor data) - Integrate LLMs, RAG, and multi-agent orchestration to interpret, summarize, and reason over video events. - Build real-time or near-real-time inference workflows with a focus on performance and reliability. - Own end-to-end delivery from research prototypes to production-ready APIs and services. - Collaborate with frontend engineers and PMs to define features and deliver “show, not tell” demos weekly. - Optimize models for both edge devices and cloud deployment. - Keep the system domain-agnostic, ensuring adaptability across Education, Retail, Operations, Security, and other verticals.

必要なスキルと経験

- Builder mindset: You turn abstract ideas into fast, tangible results. - Hands-on problem solver: You don’t wait for perfect specs—you ship MVPs, iterate quickly, and improve continuously. - Strong AI/ML foundations with at least 3 years of proven experience in: - Video analytics & computer vision (YOLO, OpenCV, Mediapipe, etc.) - Multimodal models & LLM integration (OpenAI, Gemini, or similar) - Retrieval-Augmented Generation (RAG) pipelines - Multi-agent frameworks (LangChain, CrewAI, AutoGen, etc.) - Technical skills: Python, PyTorch/TensorFlow, API development, and cloud/edge deployment (AWS, GCP, Jetson, etc.). - Data technologies: Vector databases, embeddings, and scalable data pipelines.

この求人に応募する理由

- A core role shaping a multi-domain AI product platform from zero to scale. - Opportunity to own the architecture and roadmap of a critical system. - A builder culture: fast iterations, weekly user demos, “done > perfect”. - Exposure to multi-domain problems (Education, Retail, Industrial). - Competitive salary + stock options for early team members. - Performance review (end of year) and salary review (in June) every year. - SHUI and Health Insurance - Working hour: 8:00-17:00, Mon to Fri

これはプレミアムコンテンツです。すべて表示するにはログインしてください。. ログイン