More on Temu
IT و برنامهنویسی
کار از راه دور: Machine Learning Engineer - Inference Optimization
More on Temu
Featherless AI
این پیوند شما را به آگهی اصلی میبرد. Donator درخواست نمیگیرد.
آگهیهای تازه در تلگرام
هر آگهی دورکاری تازه، همان لحظه که منتشر میشود.
@donator_jobs_fa
کار از راه دور: Machine Learning Engineer - Inference Optimization، دسته «IT و برنامهنویسی»
این آگهی را Featherless AI برای موقعیت Machine Learning Engineer - Inference Optimization منتشر کرده است. این موقعیت در دسته «IT و برنامهنویسی» جای میگیرد و بهطور کامل دورکاری است. شرکت هیچ محدودیتی برای محل زندگی داوطلب نمیگذارد، پس میتوانید از همان جایی که زندگی میکنید درخواست بدهید.
آگهی روی یک ابزار تأکید کرده است، Machine Learning، و تجربه کار با همین یکی تعیینکننده خواهد بود.
کارفرما رقمی اعلام نکرده است و این موضوع در مصاحبه روشن میشود. آگهی درباره ساعت کاری هیچ شرطی نگذاشته است.
این موقعیت از یک بررسی خودکار گذشته است: آگهیهایی که مجوز کار خارجی، اسپانسر ویزا، تابعیت مشخص یا سکونت در کشوری معین میخواهند، به این فهرست نمیرسند.
ابزارها
در یک نگاه
- شرکت
- Featherless AI
- دسته
- IT و برنامهنویسی
- چه کسی میتواند درخواست بدهد
- از هر کشور جهان
- شکل همکاری
- بهطور کامل دورکاری
- ابزارها
- Machine Learning
- نوع همکاری
- تماموقت
- زمان انتشار
- 24 سپتامبر 2026 (امروز)
- فعال
- تا 23 نوامبر 2026
- منبع
- Himalayas
موقعیتهای تازه در ایمیل: IT و برنامهنویسی
تابلو روزی چند بار بهروز میشود. فقط وقتی مینویسیم که موقعیت تازه و بررسیشدهای منتشر شده باشد. نه اسپم، نه نیاز به حساب کاربری.
قبلاً عضو شدهاید؟ مدیریت تنظیمات
توضیح کارفرما
About the Role We’re looking for a Machine Learning Engineer to own and push the limits of model inference performance at scale . You’ll work at the intersection of research and production-turning cutting-edge models into fast, reliable, and cost-efficient systems that serve real users. This role is ideal for someone who enjoys deep technical work, profiling systems down to the kernel/GPU level, and translating research ideas into production-grade performance gains. What You’ll Do
* Optimize inference latency, throughput, and cost for large-scale ML models in production
* Profile and bottleneck GPU/CPU inference pipelines (memory, kernels, batching, IO)
* Implement and tune techniques such as:
* Quantization (fp16, bf16, int8, fp8)
* KV-cache optimization & reuse
* Speculative decoding, batching, and streaming
* Model pruning or architectural simplifications for inference
* Collaborate with research engineers to productionize new model architectures
* Build and maintain inference-serving systems (e.g. Triton, custom runtimes, or bespoke stacks)
* Benchmark performance across hardware (NVIDIA / AMD GPUs, CPUs) and cloud setups
* Improve system reliability, observability, and cost efficiency under real workloads
What We’re Looking For
* Strong experience in ML inference optimization or high-performance ML systems
* Solid understanding of deep learning internals (attention, memory layout, compute graphs)
* Hands-on experience with PyTorch (or similar) and model deployment
* Familiarity with GPU performance tuning (CUDA, ROCm, Triton, or kernel-level optimizations)
* Experience scaling inference for real users (not just research benchmarks)
* Comfortable working in fast-moving startup environments with ownership and ambiguity
Nice to Have
* Experience with LLM or long-context model inference
* Knowledge of inference frameworks (TensorRT, ONNX Runtime, vLLM, Triton)
* Experience optimizing across different hardware vendors
* Open-source contributions in ML systems or inference tooling
* Background in distributed systems or low-latency services
Why Join Us
* Real ownership over performance-critical systems
* Direct impact on product reliability and unit economics
* Close collaboration with research, infra, and product
* Competitive compensation + meaningful equity at Series A
* A team that cares about engineering quality, not hype
Originally posted on Himalayas
متن به زبان اصلی کارفرما نگه داشته شده است، چون درخواست را هم به همان زبان میفرستید.
پرسشهای پرتکرار درباره این موقعیت
میتوانم از جایی که زندگی میکنم برای «Machine Learning Engineer - Inference Optimization» درخواست بدهم؟
بله. برای این آگهی، Featherless AI داوطلب از هر کشور جهان میپذیرد، پس مجوز کار کشوری دیگر لازم ندارید. آگهی از یک بررسی خودکار گذشته است: اگر کارفرما مجوز کار، اسپانسر ویزا یا سکونت در کشوری معین میخواست، روی این تابلو نمیآمد.
چه دستمزدی اعلام شده است؟
برای این آگهی، Featherless AI هیچ دستمزدی منتشر نکرده است. بیشتر آگهیهای دورکاری عددی نمیدهند و این موضوع در مصاحبه روشن میشود.
چطور درخواست بدهم؟
شما مستقیماً به کارفرما درخواست میدهید، از راه آگهی اصلی که در Himalayas منتشر شده است. Donator درخواست نمیگیرد، کمیسیون نمیخواهد و رزومهای نگه نمیدارد.
این چه نوع آگهیای است؟
یک موقعیت بهطور کامل دورکاری در دسته «IT و برنامهنویسی». آگهیهای ترکیبی و هر چیزی که حضور در دفتر بخواهد روی این تابلو منتشر نمیشود.
گواهینامههای رایگان برای این موقعیت
این آگهی Machine Learning را میخواهد. گواهینامههای رایگان زیر دقیقاً همین ابزارها را پوشش میدهند.
OCI AI Foundations Associate
exam is free tooOracle · ~10 h
A free certification on the AI track, exam included, taken inside MyLearn. An Agentic AI Foundations version also exists.
strong brandValid: 18-24 monthsAI Fundamentals and Generative AI Essentials
digital badgeIBM SkillsBuild · ~12 h
Credly badges an employer recognises. The tests require 80% to pass.
strong brandAnthropic Academy: 21 courses
certificateAnthropic · ~20 h
Courses on the Claude API, MCP and Claude Code. Each carries a free certificate, and Georgia is a supported country.
moderate weightDeep Reinforcement Learning course
certificateHugging Face · ~30 h
Entirely free with no deadline. 80% earns a Certificate of Completion, 100% a Certificate of Honors.
moderate weight
موقعیتهای مشابه
AI Researcher - Multilingual Data
Featherless AI · امروز
تازهاز هر جاPythonMachine LearningSenior Software Engineer - React / Node.js
Adalo · امروز
تازهاز هر جاReactNode.jsEngineering Manager - MLOps & Analytics
Canonical · امروز
تازهاز هر جاPythonMachine LearningJunior Linux Kernel Engineer - Ubuntu
Canonical · امروز
تازهاز هر جا