ModelBest
Posted 3mo ago

全双工语音算法专家

ModelBest
Beijing, Beijing, China
OnsiteFull Time
Responsibilities
  • leading research
  • developing models
  • improving models
Requirements
  • Master's in computer science or related,3+ years voice algorithm development,experience with ASR/TTS/LLM/end-to-end speech models,proficiency with TensorFlow or PyTorch and Python,strong engineering and communication skills
Technical tools mentioned
TensorFlowPyTorchPython

Job description

1. 主导语音大模型的前沿算法研究及产业落地;
2. 研发具有语音理解、生成、对话能力的端到端模型,开发流式全模态(语音/文本/视觉)大模型;
3. 持续跟踪前沿技术动态,能够对领域最新技术进行及时吸纳和改进,通过技术创新和工程实践,支撑模型能力提升;
4. 通过论文和技术报告等形式提升团队的技术影响力。

硕士及以上学历,计算机相关专业;
1. 三年以上语音算法开发经验,深入参与过语音交互类全链路模型及产品建设,熟悉 ASR/TTS/LLM/端到端语音大模型 等语音 AI 相关技术原理,并对相关技术有深入理解和思考;
2. 掌握 AI 产品开发的开源工具和框架(如TensorFlow、PyTorch)。具备出色的编程能力和工程能力,熟练掌握Python或其他相关编程语言,具有良好的代码编写习惯和程序开发经验;
3. 具备良好的团队合作精神、沟通能力以及分析和解决问题的能力。

About ModelBest

Develops efficient, on-device large language models and edge AI.

Year founded
2022
Employees
200
Organization type
Private
Latest investment
Raised $5.00B Funding Round (2026) — led by China Telecom, Shenzhen Capital Group, Inovance Capital, Primavera Capital Group
Subsidiaries
Headquarters
CN

Similar jobs

Speech Algorithm Engineer roles near Beijing, Beijing
6d
Save
Mark Applied
Hide
语音算法工程师
Beijing, Beijing, China
OnsiteFull Time
Duxiaoman
Duxiaoman: A financial technology providing consumer credit and wealth solutions.
Experience in speech or multimodal large models; understanding of ASR, TTS, VAD, dialogue systems, data processing, and model architectures; Python, Linux, and deep learning framework proficiency.
Python, Linux, ASR, TTS, VAD, Fun-Audio-Chat, Moshi, Qwen3Omni
2w
Save
Mark Applied
Hide
混元大模型语音算法工程师(北京/上海)
Beijing or Shanghai or Shenzhen
OnsiteFull Time
Tencent
TencentHKEX: 0700: Provides integrated internet services, digital entertainment, and cloud technology.
4+ YOEExperience in speech/audio large-model development, strong coding and algorithm skills, proficiency in Python/C/C++, familiarity with model training frameworks, and solid math/signal processing background.
Python, C, C++, PyTorch, Megatron, DeepSpeed
3w
Save
Mark Applied
Hide
语音算法岗位-2027届
Beijing or Chongqing or Chengdu
OnsiteFull Time
Mashang Consumer Finance
Mashang Consumer Finance: A licensed financial institution providing digital consumer lending services.
Master's degree in computer science/electronic information/artificial intelligence, speech algorithm project experience, proficient in Python and basic AI coding, publications a plus; targeted to 2027 graduates.
Python
1mo
Save
Mark Applied
Hide
语音算法工程师实习生
Shenzhen or Beijing
OnsiteInternship
SenseTime
SenseTimeHKEX: 0020: Develops artificial intelligence software and computer vision technology.
Graduate-level in AI/ML/signal processing/computer science, familiarity with speech models and toolkits, strong research and PyTorch implementation skills, proficiency in Python/C/C++/Java/Shell, and experience with speech synthesis/recognition techniques.
kaldi, K2, wenet, espnet, whisper, FunASR, PyTorch, Python, C, C++, Java, Shell, RNN-T, conformer, CTC, SSL, LLM, diffusion, SpearTTS, ChatTTS, IndexTTS, CosyVoice
1mo
Save
Mark Applied
Hide
语音算法工程师(多模态理解大模型-说话人方向)-Data语音
Beijing, Beijing, China
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
Design and improve speaker-related capabilities for multimodal audio understanding models; experience with speaker clustering/diarization/recognition, deep learning, large models, and production-scale data pipelines; strong coding skills.
AHC, ResNet, EcapaTDNN, ALM, LLM, Codex, TRAE, C++, Python
1mo
Save
Mark Applied
Hide
语音算法工程师-语音感知-北京
Beijing, Beijing, China
OnsiteFull Time
Li Auto
Li AutoNASDAQ / HKEX: LI / 2015: Designing and manufacturing premium smart electric vehicles.
Master's degree or above in AI/electronic information/computer science; expertise in wake-word, speech recognition, multimodal large models; proficiency in C/C++ and Python; research publications preferred.
C, C++, Python
3mo
Save
Mark Applied
Hide
语音算法工程师
Shenzhen or Beijing
OnsiteFull Time
SenseTime
SenseTimeHong Kong Stock Exchange: 0020: AI software provider of computer vision and deep learning.
Graduate degree in AI/ML/signal processing/computer science, strong ML and signal processing foundation, experience with speech recognition/synthesis models and toolkits, proficiency in PyTorch and programming in Python/C/C++/Java, publication or competition track record preferred.
RNN-T, conformer, CTC, Kaldi, K2, WeNet, ESPnet, Whisper, FunASR, PyTorch, SSL, LLM, diffusion, Python, C, C++, Java, Shell, SpearTTS, ChatTTS, IndexTTS, CosyVoice
3mo
Save
Mark Applied
Hide
实习-语音大模型算法工程师(TTS)
Beijing or Shanghai
OnsiteInternship
Nio
NioNYSE: NIO: Designs and manufactures smart premium electric vehicles.
Conduct TTS research and engineering including data processing, pretraining, fine-tuning, RL, and model evaluation; bachelor\u0002s degree in relevant fields required; strong communication and learning ability.