Full-Time

Applied AI Engineer

AssemblyAI

AssemblyAI

51-200 employees

API-based speech transcription and analytics

Compensation Overview

$150k - $250k/yr

+ OTE

London, UK + 2 more

More locations: San Francisco, CA, USA | New York, NY, USA

Hybrid

Hybrid role; occasional travel to NYC/SF and international client sites.

Category
AI & Machine Learning (1)
Required Skills
Claude
Python
JavaScript
Ruby
Java
TypeScript
Go

Get referred to AssemblyAI

See people who can refer or advise you

Requirements
  • Located in or willing to relocate to New York City or San Francisco
  • 4+ years of customer-facing experience in highly technical environments, requiring strong software engineering skills, including roles such as Solutions Architect, Developer Advocate, Technical Solutions Engineer, or equivalent
  • Strong programming skills in Python and/or TypeScript/JavaScript, with the ability to read and debug code in other languages
  • Excellent written and verbal communication skills
  • Experience with web APIs, WebSockets, and real-time data streaming
  • Understanding of audio fundamentals (sample rates, encoding formats, codecs) or willingness to learn quickly
  • Ability to leverage AI tools creatively to move quickly—proficiency with tools like Claude, Cursor, Windsurf, or similar AI-assisted development tools
  • Problem-solving mindset—able to debug unfamiliar codebases, identify root causes, and propose creative solutions
  • Comfort with ambiguity and fast-paced work—first line of defense on novel problems
  • Ability and willingness to travel domestically and internationally as needed (approximately 10-25%)
Responsibilities
  • Guide customers through their entire product journey—from initial implementation and technical onboarding, through development and testing, all the way to production deployment, and continue to provide support post-launch to ensure their implementation is performing optimally and meeting their business objectives
  • Build custom demos and prototypes that showcase how AssemblyAI's models can solve specific customer use cases
  • Lead technical discovery sessions with prospects and customers to understand their requirements and design optimal implementations
  • Assist with competitive evaluations by providing benchmarks, designing fair comparison tests, and helping customers evaluate our models against alternatives; work closely with customers to ensure they have the data and support needed to make informed decisions
  • Troubleshoot production issues for customers using our streaming and async APIs, sometimes requiring you to debug code in languages you may not use daily (Java, Go, Ruby, etc.)
  • Provide technical guidance on architecture, scaling, data privacy, and best practices for integrating our APIs
  • Create proofs-of-concept that demonstrate the "art of the possible" with Speech AI
  • Travel to customer sites for technical workshops, implementation support, and relationship building with key accounts
  • Serve as the voice of the customer to our Product and Research teams—test and dogfood new products and features before they reach customers, providing feedback from a practitioner's perspective
  • Advocate for improvements and new capabilities based on real-world use cases you encounter, ensuring our roadmap aligns with customer needs
  • Contribute to documentation, guides, and example code that help developers integrate our APIs more effectively
  • Identify patterns in customer issues and work with Engineering to address root causes (SDK improvements, better error messages, new features)
  • Collaborate with Sales and Success teams to ensure smooth handoffs and consistent customer experience
  • Participate in team off-sites to align on strategy, share learnings, and build team cohesion
  • Stay current on AI/ML developments, particularly in speech recognition, NLP, and audio processing
  • Experiment with new use cases and applications of our models
  • Build and maintain an open-source repository of demos, integrations, and reference implementations
  • Represent AssemblyAI at conferences and industry events, delivering talks or running workshops when appropriate
Desired Qualifications
  • Experience in a startup or high-growth environment
  • Understanding of cloud infrastructure (AWS, GCP, Azure) and deployment patterns
  • Experience building prototypes or MVPs quickly
  • Public speaking or conference presentation experience

AssemblyAI provides an API-driven Speech AI platform that converts audio and video into text and insights. Its models handle automatic transcription and add-ons like speaker diarization, sentiment analysis, and PII redaction, all delivered through scalable API endpoints so customers can integrate transcription and audio analytics into their own apps. Unlike many competitors, AssemblyAI emphasizes a combination of high accuracy, broad capabilities (including privacy-focused features), and ongoing model improvements from a research-driven team, with a subscription-based pricing model for access and data processed. The company aims to help businesses unlock voice data by turning speech into actionable information that can power products and services across tech, media, and enterprise sectors.

Company Size

51-200

Company Stage

Series C

Total Funding

$113.1M

Headquarters

San Francisco, California

Founded

2017

Get referred to AssemblyAI

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • G2 named AssemblyAI a Leader in Spring 2026 Voice Recognition on March 23, 2026.
  • AssemblyAI's July 2026 releases expanded language support and improved speaker diarization.
  • Its October 2025 platform suite added Guardrails, Speech Understanding, and LLM Gateway, broadening customer stickiness.

What critics are saying

  • Tesseract Systems sued AssemblyAI on May 27, 2026, creating patent-defense distraction and costs.
  • Deepgram raised $130 million in January 2026 and IBM integrated Deepgram in February 2026.
  • OpenAI Whisper and Deepgram compress pricing, weakening AssemblyAI's API moat by 2026.

What makes AssemblyAI unique

  • AssemblyAI launched Universal-3.5 Pro on July 7, 2026 with native code-switching and diarization.
  • Its April 29, 2026 Voice Agent API unifies STT, reasoning, and TTS in one WebSocket.
  • Medical Mode, launched March 25, 2026, targets clinical terminology across multilingual healthcare transcription.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Competitive Salary + Bonus

Equity

401k

100% Remote team

Unlimited PTO

Premium medical, vision, & dental care

$1K budget for your home office setup

New Macbook Pro (or PC if you prefer)

2x/year company paid team retreat

Growth & Insights and Company News

Headcount

6 month growth

-4%

1 year growth

-6%

2 year growth

-2%
AssemblyAI
Jan 9th, 2025
Dev.to x AssemblyAI: Winter Speech-to-Text Challenge Winners

Last month, AssemblyAI, Inc. partnered with developer community hub Dev.to to host a winter hacking challenge!

Blockchain News
Nov 25th, 2024
Implementing Speech-to-Text with JavaScript and Node.js

AssemblyAI has released a comprehensive tutorial on utilizing its API to convert audio and video files into text using JavaScript and Node.js.

TechCrunch
Oct 15th, 2024
Gladia believes real-time processing is the next frontier of audio transcription APIs

Gladia competes with other well-funded companies in the space, such as AssemblyAI, Deepgram and Speechmatics.

AssemblyAI
Oct 3rd, 2024
AI-powered meeting company Supernormal launches customizable Voice Agents

AssemblyAI, Inc. recently launched Voice Agents, a new addition to its AI meeting platform that allows users to create customizable conversational agents to handle routine conversations like inbound sales, customer support, candidate screens, and employee surveys-saving users time and providing an instant and engaging experience for end-users.

Megadumpload
Sep 27th, 2024
AssemblyAI Launches Postman Collection for Enhanced API Testing

AssemblyAI launches Postman collection for enhanced API testing.