Lead the implementation, optimization, and delivery of machine learning, computer vision, and numeric libraries across heterogeneous hardware architectures. Own technical execution, quality, timelines, and customer outcomes while mentoring engineers in low-level optimization, parallel programming, vectorization, profiling, debugging, benchmarking, and performance tuning. Collaborate with customers and stakeholders to identify performance requirements and bottlenecks, make technical decisions, and deliver high-performance software solutions.
We are seeking a highly skilled and motivated technical lead to drive the implementation and
optimization of machine learning, computer vision, and numeric libraries across target
hardware architectures, including CPUs, GPUs, DSPs, NPUs, and other accelerators. The role
involves technical ownership and accountability for project execution and delivery, driving high
performance software solutions, providing technical leadership to engineering teams, and
working closely with customers to successfully deliver optimized libraries.
Key Responsibilities
• Lead the implementation, optimization, and delivery of machine learning, computer
vision, and numeric libraries for heterogeneous hardware architectures, including CPUs,
GPUs, DSPs, NPUs, and accelerators.
• Own project execution with accountability for technical delivery, quality, timelines, and
successful customer outcomes.
• Provide technical leadership, guidance, and mentorship to engineers in areas such as
low-level optimization, parallel programming, vectorization, memory hierarchy
optimization, profiling, performance analysis, debugging, benchmarking, and
performance tuning to enable successful execution and delivery of projects.
• Work directly with customers and internal stakeholders to understand performance
requirements, optimization goals, bottlenecks, and expected outcomes
.
• Conduct technical reviews, problem-solving, and technical decision-making to support
successful execution of projects.
• Stay current with advancements in machine learning, computer vision, high
performance computing, and heterogeneous compute architectures.
Requirements
Qualifications
• BTech/BE/MTech/ME/MS/PhD in CSE/IT/ECE or related field.
• 5+ years of experience in software development, performance engineering, algorithm
optimization, porting, or low-level software development.
• Proven experience in leading technical execution of projects, mentoring engineers, and
driving successful project delivery.
• Strong interest and passion for technical leadership, team guidance, customer
engagement, and ownership of project execution and delivery.
• Strong expertise in C/C++, CUDA, OpenCL, HIP, SYCL, or similar programming models.
• Hands-on experience optimizing software for CPUs, GPUs, DSPs, NPUs, or other
accelerator architectures.
• Strong understanding of parallel computing concepts, SIMD/vectorization, threading
models, cache/memory optimization, profiling, benchmarking, and performance tuning.
• Strong debugging, analytical, problem-solving, and technical decision-making skills.
• Excellent communication, collaboration, and stakeholder management skills.
Who This Role Would Suit
This role would suit strong technical professionals with hands-on expertise in performance
optimization, low-level software development, or heterogeneous computing who are looking to
grow into a larger technical leadership role.
We are looking for individuals who are passionate about taking ownership of project execution
and delivery, mentoring and guiding teams, enabling successful outcomes through technical
leadership, and contributing to the growth and success of the organization.
MulticoreWare Chennai, Tamil Nadu, IND Office
DLF IT Park Block 3, Ground Floor, Mount Poonamallee Road, Manapakkam, , Chennai, TamilNadu , India, 600089
Similar Jobs
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Lead the delivery of AI products across international markets by coordinating internal engineers, contractors, and vendors. Establish engineering standards, make technical decisions, review code, ensure testing and release readiness, and guide solutions from proof of concept through enterprise production. Oversee data, AI, generative AI, agentic AI, and full-stack applications while managing dependencies, technical debt, and delivery quality. Collaborate with product, engineering, data science, security, and global stakeholders, and evaluate emerging technologies.
Top Skills:
Agentic AiAirflowAWSBackend EngineeringCi/CdConfluenceData EngineeringDbtFull-Stack DevelopmentGenerative AiJIRALlmopsMachine LearningMlopsPythonSnowflakeSQL
Greentech • Other
Lead design, development, testing, and delivery of full-stack web applications and microservices. Establish engineering standards, build scalable AWS-based solutions, design secure REST APIs, write unit tests, lead distributed feature teams, mentor engineers, and collaborate with stakeholders using agile practices.
Top Skills:
AWSCSSHTML5JavaJavaScriptJqueryMicroservicesReact JsRest ApisSpring Boot
Greentech • Other
Lead cross-functional teams to design and build scalable full-stack digital products leveraging AI/ML. Conduct AI/ML research, architect solutions, implement ReactJS/Python stacks, APIs and databases, enforce testing, security and DevSecOps practices, and mentor junior engineers in an Agile environment.
Top Skills:
AgileAi LibrariesApplication SecurityAWSBddCloudDevsecopsKanbanLinuxMocking FrameworksNoSQLPythonReactRestful ApisSQLTddUnit Testing
What you need to know about the Chennai Tech Scene
To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.

