Zorky CRMZorky CRM
EN|RU
@ekaterinovikova

Jobs

Live IT jobs from 1000+ sources across CIS and Europe. Updated daily. About →
Stack filter:triton
Team Lead в команду ML-инфраструктуры R&D
leadofficeRU
c++llmmlopspythonpytorchtriton
❣️ Team Lead в команду ML-инфраструктуры R&D Мы создаём инфраструктуру для обучения и дообучения больших языковых (LLM) и визуально-языковых (VLM) моделей. Наша задача — сделать с
Apply
ML Engineer (Internship and Full-time)
San Francisco, US
pytorchtriton
Tilde Research is a moonshot AI lab advancing mechanistic interpretability, new architectures, and pretraining science. We build foundational understanding of models to advance the
Apply
Kernel Engineer (Internship and Full-time)
San Francisco, US
pytorchtriton
Tilde Research is a moonshot AI lab advancing mechanistic interpretability, new architectures, and pretraining science. We build foundational understanding of models to advance the
Apply
Member of Technical Staff - Foundations
principalremoteSan Francisco / Tel Aviv / Zurich, US
fine-tuningpythonpytorchtriton
Tzafon is a foundation model lab building scalable compute systems and advancing machine intelligence, with offices in San Francisco, Zurich & Tel Aviv. We’ve raised over $12m in f
Apply
Software Engineer - Triton
leadoffice~$6.9K /moUK, GB
pytorchtritonusability
About the job Build the framework support that helps AI developers unlock Graphcore hardware. You will help Graphcore accelerators work seamlessly with state-of-the-art ML framewor
Apply
Software Engineer, Inference
officeSan Francisco, US
c++llmpythontritonvllm
Overview Pulse is tackling one of the most persistent challenges in data infrastructure: extracting accurate, structured information from complex documents at scale. We have a brea
Apply
MLOps / ML Infrastructure Engineer
middleremoteМосква, RU
ci/cdgrafanakuberneteslinuxmlflowmlops+2
Мы ищем инженера, который будет отвечать за стабильную работу ML-инфраструктуры и сервисов инференса в production. В этой роли предстоит настраивать и поддерживать Triton Inference
Apply
AI/ML Engineer в направление «Интеллектуальные решения»
middleofficeАстана, KZ
devopsdockerfine-tuninggitkuberneteslangchain+6
Ключевые задачи: Разбирать запросы заказчиков и предлагать, каким классом решений их закрывать; Разрабатывать и улучшать RAG-пайплайны: поиск, реранжирование, сборка контекста, цит
Apply
GPU Performance Engineer
San Francisco HQ, US
c++pythontriton
We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI
Apply
Вакансия MLE
leadremoteRU
ci/cddockerfastapillmpythonsql+2
Вакансия MLE 🟢Удаленно в РФ🟢Полная занятость 🟢ЗП обсуждается на собеседовании 🟢В компанию Ростелеком #ds_ml 🟦Задачи – Планирование архитектуры приложений с использованием LL
Apply
Inference Performance Engineer
remoteNew York, NY, US
c++llmpythontransformerstritonvllm
About the role Serving frontier models at scale requires solving novel systems problems at every layer of the stack. As an Inference Performance Engineer, you'll own the runtime th
Apply
Machine Learning Research Engineer (MLRE) - GPUs
remoteSan Francisco Office, US
pytorchtriton
Why Achira At Achira, we are building a team of world-class scientists, ML researchers, and engineers to work together to move beyond the beaten path in drug discovery. We are acti
Apply
CUDA-разработчик (инфраструктура и ускорение вычислений) в Автономный транспорт
officeRU
c++triton
CUDA-разработчик (инфраструктура и ускорение вычислений) в Автономный транспорт Стек: CUDA, C++, GPU architecture, ONNX, Machine Learning, Performance optimization, TensorRT, Trit
Apply
Member of Technical Staff — Inference-Multi-Hardware
principalofficePalo Alto, CA, US
c++pythontriton
About the Role RadixArk is seeking a Member of Technical Staff - Inference-Multi-Hardware to push the limits of performance for frontier AI systems. Most performance engineering as
Apply
Member of Technical Staff – AI Inference platform
principalZürich, CH
grafanaopentelemetryprometheuspythonsolidtriton+1
Your mission You will make Lyceum's AI inference platform reliable, secure, and scalable - ensuring it performs under pressure as we grow to thousands of concurrent users. While ot
Apply
Test Engineer (XSW Tester) - Storage Server & Video Management Systems
Taipei City
agiletritonusability
Keenfinity is a global leader in smart security and professional communications—born from Bosch Building Technologies and now an agile, independent company under Triton Partners.Ou
Apply
视频生成模型 · 训练 Infra 工程师
remoteSan Francisco Bay Area, US
pytorchtriton
我们在训练自研视频生成基础模型(DiT / Flow Matching),需要一位既能搭起训练平台、又能把研究代码变成数百卡集群上稳定结果的工程师。你不只是用平台的人,更是建平台的人。 你会做 • 训练平台搭建:从作业调度、断点续训、监控告警到数据 / 权重流水线,把分散的脚本沉淀为团队可复用的训练基础设施。 • 数百卡规模的分布式训练:FSDP、张量并行、
Apply
Linux System Engineer
taipei
agilelinuxtriton
Keenfinity is a global leader in smart security and professional communications—born from Bosch Building Technologies and now an agile, independent company under Triton Partners.Ou
Apply
1 / 1