Tech Stack
Job Description, Responsibilities & Requirements
About the Position
We are seeking a skilled ML/AI Engineer to join our team and deliver cutting-edge ML/AI models in the ASR domain and legal insights more broadly. The core of this role is building, deploying, and serving ML models in production.
A key part of this role is speech engineering - prior ASR/speech experience is not a must, but a willingness to learn and grow into this domain is essential. Experience with LLM-based solutions and agentic flows is an advantage, not a requirement.
Responsibilities
- Build, deploy, and serve ML models in production - the core of the role - optimizing inference performance, latency, and efficient GPU utilization.
- Support and extend our speech solutions, including 3rd-party ASR integrations and self-hosted ML models.
- Evaluate and optimize key ASR/ML metrics such as WER, latency, F1 score, EER, and DER.
- Secondary: Integrate LLMs and agentic flows into production solutions where they add value, including fine-tuning existing models.
Requirements
Core (required):
- Solid background in machine/deep learning, with 4+ years of experience.
- Strong Python skills and experience working in Linux environments.
- Familiarity with deep learning libraries (e.g., PyTorch, TensorFlow) and training workflows.
- Experience with MLOps pipelines and serving/deploying models in production.
- Experience optimizing models for inference, including GPU acceleration.
- Experience with Docker, gRPC, and serverless / micro-services architectures on cloud infrastructure (AWS/GCP/Azure).
- Experience with real-time / streaming systems and low-latency processing (e.g. streaming pipelines, live sessions).
Nice to have (secondary / optional):
- Experience building and productizing agentic flows and AI-driven solutions (1+ year).
- Familiarity with LLM / agent orchestration frameworks such as CrewAI, AWS Bedrock, Temporal, AgentCore, or LiteLLM.
- Prior ASR / speech experience - not required; willingness to learn this domain is what matters most.
We Offer
- Competitive salary
- Opportunity to work in a fast-paced, startup environment
- Continuous learning and growth opportunities
- Collaborative and tech-savvy team environment
About the Company
Our team of 250 professionals builds AI-driven solutions that help the world’s leading organizations get real value from their audio and video content. We serve high-end customers across the Legal, Media and Education markets with a next-generation verbal-intelligence platform that brings AI into the core of their workflows.
Our in-house technology combines advanced LLMs, GenAI capabilities and speech AI to power everything from legal-tech automation and real-time insights to intelligent media and learning workflows. Beyond transcription or accessibility, Verbit provides actionable intelligence: smart summaries, analytics, search, knowledge extraction, legal-deposition insights and more.
With more than 3,000 global customers relying on Verbit-including top law firms, universities, enterprises and media organizations-we’re building the future of how speech data is processed, understood and used.
We’re a group of:
- Tech-savvy individuals who are always open to more growth and learning opportunities
- Adaptable and flexible people who thrive in a fast-paced, startup environment
- Creative minds who rethink and question how to outperform past results
- Effective communicators who can promote and represent Verbit’s tech and brand