Arrastra un PDF aquí, pega una URL o toca para seleccionar un archivo
ClickUp
USA
At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. Role Overview You'll own and evolve the AI systems behind ClickUp's voice platform: real-time streaming transcription, intelligent reformatting, context-aware mention detection, and voice-to-action pipelines. This is a high-impact, hands-on role where you'll push the boundaries of what voice interfaces can do inside a productivity tool used by millions.
Design, build, and optimize real-time speech-to-text pipelines (streaming ASR, VAD, audio processing) Improve transcription accuracy through context injection (user names, teams, custom vocabulary, language detection) Develop and maintain LLM-powered post-processing (grammar correction, filler removal, mention resolution, formatting) Build voice-to-action systems that parse natural language into structured workspace commands Evaluate, benchmark, and integrate ASR models (Whisper, AssemblyAI, Fireworks, etc.) for cost, latency, and accuracy Collaborate with product and platform teams to ship voice features across MAX Desktop, Mobile, Web, and Browser Extension Explore multimodal AI capabilities (screen + voice + text) for next-gen
Ref. 2F516
Arrastra un PDF aquí, pega una URL o toca para seleccionar un archivo
Tu Anuncio para Redes Sociales
×Esta imagen está optimizada para Instagram y LinkedIn. ¡Compártela para atraer talento!
Se envía un enlace que abre esta oferta directamente.
Tu Anuncio para TikTok / Reels
Esta imagen vertical está optimizada para TikTok, Instagram Reels y YouTube Shorts.