Jobs
No opportunities available
There are currently no job opportunities available. Please check back later.
LLM Inference Frameworks and Optimization Engineer
LLM Inference Frameworks and Optimization Engineer
Together AIAbout the role
Together.ai is seeking an Inference Frameworks and Optimization Engineer to design, develop, and optimize distributed inference engines for large language and multimodal models. This role focuses on achieving low-latency, high-throughput inference through GPU/accelerator optimizations and software-hardware co-design. The ideal candidate will have 3+ years of experience in deep learning inference frameworks, distributed systems, or high-performance computing, with proficiency in Python and C++/CUDA. This is a unique opportunity to shape the future of LLM inference infrastructure and push the boundaries of AI performance and scalability.
Similar jobs
Browse Jobs by Role
Chief of StaffProduct MarketingForward Deployed EngineerForward DeployedSoftware EngineerProduct ManagerData ScientistDesignSalesMarketingOperationsEngineering ManagerFinanceCustomer SuccessHR & People OpsHardware EngineerStrategyAccountingFounding EngineerGrowthDevOps & InfrastructureMachine Learning Engineer
