Strong programming skills in C++ Strong experience with Llama.cpp and ggml inference engines, which facilitates the deployment of models to specific GPU architectures Experience with any GPU framework between Cuda, Vulkan, Metal, OpenCL Good understanding of deep learning concepts and model architectures Experience with transformers, LLMs, Diffusion models Demonstrated ability to rapidly assimilate new technologies and techniques A degree in Computer Science, AI, Machine Learning, or a related field, complemented by a solid track record in AI R&D Bonus points if: You know how to train/fine-tune a LLM You have productionized models You have research experience in new model architectures You have experience with distributed systems Javascript experience Important information for candidates Recruitment scams have become increasingly common. To protect yourself, please keep the following in mind when applying for roles: Apply only through our official channels. We do not use third-party platforms or agencies for recruitment unless clearly stated. All open roles are listed on our official careers page: https://tether.recruitee.com/ Verify the recruiter’s identity. All our recruiters have verified LinkedIn profiles. If you’re unsure, you can confirm their identity by checking their profile or contacting us through our website. Be cautious of unusual communication methods. We do not conduct interviews over WhatsApp, Telegram, or SMS. All communication is done through official company emails and platforms. Double-check email addresses. All communication from us will come from emails ending in @tether.to or @tether.io We will never request payment or financial details. If someone asks for personal financial information or payment at any point during the hiring process, it is a scam. Please report it immediately. When in doubt, feel free to reach out through our official website.