A complete application in a minute — tailored resume and cover letter, ready to send.
Flairdeck in Bengaluru, Karnataka invites an engineer to deploy and optimize Large Language Models on edge devices, preferably NVIDIA Jetson platforms. You will design practical LLM-based solutions using prompt engineering, preprocessing, caching, data creation, and occasional model fine-tuning to meet real-world constraints.
The role emphasizes reliable and efficient LLM applications under limited compute, memory, latency, and power on edge devices, with opportunities to work on end-to-end
We are looking for an engineer experienced in deploying and optimizing Large Language Models on edge devices, preferably NVIDIA Jetson platforms. The role is not limited to model inference; the candidate should be able to design practical LLM-based solutions for real-world scenarios using prompt engineering, input preprocessing, caching strategies, data creation, and model fine-tuning when required. The ideal candidate should understand how to make LLM applications reliable, efficient, and context-aware under edge-device constraints such as limited compute, memory, latency, and power.
Design practical LLM-based solutions for real-world scenarios using prompt engineering, input preprocessing, caching strategies, data creation, and model fine-tuning when required. Understand how to make LLM applications reliable, efficient, and context-aware under edge-device constraints such as limited compute, memory, latency, and power.
Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.