Staff AI Compiler Engineer
Full Job Description
About EnCharge AI
EnCharge AI is a leader in advanced AI hardware and software systems for edge-to-cloud computing. Our robust, scalable next-generation in-memory computing technology provides orders-of-magnitude higher compute efficiency compared to today's best-in-class solutions.
The Role
We are seeking an experienced Staff/Principal AI Compiler Engineer to spearhead efforts in developing and optimizing graph compilers tailored to cutting-edge AI and ML workloads. You will collaborate with hardware architects, software developers, and researchers to enhance performance, optimize computation graphs, and enable efficient model deployment on EnCharge's Inference Accelerators.
Key Responsibilities
- Architect, design, and implement optimizations for AI model execution on graph compilers to improve performance, reduce latency, and maximize hardware utilization.
- Work closely with ML researchers and engineers to understand and address hardware-specific challenges in deploying AI models.
- Perform performance optimizations for neural network models, including layer fusion, operator fusion, and graph-level transformations.
- Develop compiler optimizations that convert high-level AI models (e.g., TensorFlow, PyTorch) into intermediate representations (IR).
- Implement parsing, semantic analysis, and IR generation for deep learning frameworks.
- Research and integrate the latest advancements in compiler design, ML model optimizations, and hardware acceleration.
- Provide leadership, mentorship, and technical guidance to a team of engineers focused on graph compiler optimizations.
Company
Mulya Technologies
Mulya Technologies, formerly known as The Human Capital and merged with Mulya Consulting, is a premium corporate HR services company specializing in aligning business goals with human resources for th...