4 часа назад
Principal Engineer - Agentic AI Architect
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Principal Engineer - Agentic AI Architect (Agentic AI/FPGA): Designing scalable agentic AI systems that integrate LLMs, RAG pipelines, autonomous agents, and FPGA/ASIC platforms with an accent on hardware-aware system architecture and AI inference optimization. Focus on building multi-agent orchestration frameworks, optimizing LLM deployment through quantization and compilation, and bridging software intelligence with heterogeneous hardware acceleration.
Location: Shanghai, China
Company
is hiring for a role focused on agentic AI architecture and FPGA/ASIC system integration.
What you will do
- Design and implement agentic AI systems using LLMs, tools, multi-agent orchestration, and RAG pipelines.
- Define end-to-end AI architectures spanning data ingestion, reasoning, action, and scalable deployment.
- Co-design AI systems with FPGA, ASIC, and heterogeneous CPU/GPU/FPGA platforms.
- Optimize LLM inference through quantization, compilation, partitioning, and hardware acceleration.
- Build reusable platforms and components for memory management, tool integration, and knowledge retrieval using LangChain, LlamaIndex, and custom orchestration frameworks.
- Lead architecture reviews, mentor senior engineers, and guide cross-functional technical strategy.
Requirements
- Master’s or Bachelor’s degree and 10+ years of experience in AI/ML systems, distributed systems, or hardware-software co-design.
- Recognized expertise in AI and LLMs, including LLM architectures, inference, RAG systems, vector databases, and autonomous agents.
- Experience building production-grade LLM applications, multi-step reasoning workflows, and reliable AI systems.
- Strong background in FPGA, ASIC, EDA systems, heterogeneous computing, and AI workload optimization.
- Programming experience with Python; C++ is preferred.
- Experience with distributed systems, APIs, microservices, cloud or edge AI deployment, model fine-tuning, quantization, distillation, and on-device AI.
Nice to have
- Knowledge of compiler stacks such as MLIR or TVM.
- Experience with robotics, autonomous systems, or real-time AI pipelines.
Culture & Benefits
- Regular full-time employment.
- Work focused on AI reasoning systems, FPGA acceleration, and next-generation computing platforms.
- Opportunity to influence AI and semiconductor integration through intelligent hardware orchestration and real-time decision-making pipelines.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →