donator JOBS

IT 与编程

Senior Research Engineer - Voice(远程职位)

Synthesia

限欧洲远程Europe, UK
申请

链接会跳转到原始招聘信息。Donator 不代收申请。

可分享的海报
发布时间: (6 天前)有效期: 截至 2026年9月28日

Senior Research Engineer - Voice:「IT 与编程」类别,完全远程

这条招聘信息来自 Synthesia,岗位是 Senior Research Engineer - Voice:这是高级职位,公司期待由你自己拿主意。职位属于「IT 与编程」类别,完全远程。公司招的是居住在欧洲的候选人。

公司没有公布具体数字,这一项会在面试时谈定。对工作时间,招聘信息没有提出任何条件。

这个职位通过了自动核查:凡是要求外国工作许可、签证担保、特定国籍,或者必须居住在指定国家的招聘信息,都不会进入列表。

要点

公司
Synthesia
类别
IT 与编程
谁可以申请
来自欧洲任何国家
工作方式
完全远程
级别
高级
用工形式
全职
发布时间
2026年8月19日 (6 天前)
有效期
截至 2026年9月28日
来源
Jobicy

新职位邮件提醒:IT 与编程

职位板每天更新数次。只有出现新的、已核查的职位时才会写信。不发垃圾邮件,也不需要注册账号。

已经订阅过了? 管理订阅设置

公司发布的职位描述

Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US.

As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations.

Following our recent Series E funding round, where we raised $200 million, our valuation stands at $4 billion. Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow.

What you'll do at Synthesia

As a Research Engineer you will join a team of 40+ Researchers and Engineers within the R&D Department working on cutting-edge challenges in the Generative AI space, with a focus on creating high-quality, expressive and real-time synthetic voices. Within the team you’ll have the opportunity to work on the applied side of our research efforts and directly impact our solutions that are used worldwide by over 60,000 businesses.

If you are an expert in ML, LLMs, speech generation, conversational models , this is your chance to make a global impact. You will join our Audio Post-Training Team , which works on generative speech and voice synthesis , ensuring our in-house voice models reach production-level quality, speed, and robustness. Typical projects include:

* Develop and evaluate streaming and speech-to-speech systems, enabling low-latency, interactive voice synthesis.

* Adapt models for new conditioning inputs (emotion, speed, prosody, speaker control, etc.).

* Implement post-training optimization techniques (quantization, pruning, distillation) to improve efficiency and latency in real-time speech generation.

* Integrate and test novel architectures, such as neural codecs, diffusion, or flow-matching models, to enhance realism and responsiveness.

* Contribute to defining new evaluation metrics for conversational speech, including latency-aware and online MOS prediction systems.

* Stay updated with the latest research in audio diffusion, autoregressive models, neural codecs, and multimodal LLMs.

* Apply DPO (Direct Preference Optimization) and distillation to fine-tune large-scale speech models.

What we're looking for:

* Strong understanding of generative modeling, ideally applied to sequential or multimodal data.

* Hands-on experience with large language models (LLMs) or similar transformer-based architectures.

* High proficiency in PyTorch, including experience with distributed training and model optimization.

* Solid grasp of time-series modeling and tokenization, preferably in the context of audio or speech.

* Demonstrated ability to prototype quickly, test hypotheses, and iterate efficiently.

* Proven experience in training deep learning models end-to-end, from data preparation to evaluation.

* Strong general software engineering skills, enabling contributions to a large, shared research infrastructure.

Nice to have experience:

* Experience with real-time or streaming architectures is a big plus.

* Familiarity with state-of-the-art architectures in audio and speech generation (e.g., diffusion models, neural codecs, flow-matching models, autoregressive decoders).

* Experience with speech-to-speech or text-to-speech (TTS) systems.

* Evidence of original research contributions, such as publications or open-source work in top-tier venues (e.g., ICASSP, Interspeech, NeurIPS, ICML).

职位正文保留公司发布时的原文,因为投递的时候用的也是同一种语言。

该职位发布于 Jobicy. 原始招聘信息 (Jobicy)

关于这个职位的常见问题

我可以在自己居住的地方申请「Senior Research Engineer - Voice」这个职位吗?

可以。对于这个职位,Synthesia 接受来自欧洲任何国家的候选人,所以你不需要其他国家的工作许可。这条招聘信息通过了自动核查:如果雇主要求工作许可、签证担保,或者必须居住在某个特定国家,它就不会出现在本站。

标明的报酬是多少?

对于这个职位,Synthesia 没有公布报酬。多数远程招聘信息不给出具体数字,这件事会在面试时谈定。

怎么申请?

你通过发布在 Jobicy 上的原始招聘信息,直接向雇主提交申请。Donator 不代收申请,不收取佣金,也不保存简历。

这是什么类型的职位?

这是「IT 与编程」类别中的一个完全远程职位。混合办公的招聘信息,以及任何需要到办公室的职位,本站都不会发布。

适合这个职位的免费证书

这个领域里分量最重的几张免费证书。每一张都在发证方自己的页面上核实过。

全部免费证书:「IT and cloud」

相似职位

远程职位:Senior Research Engineer - Voice | Donator