AI/ML Applied Engineer

karumi (yc f25) United State
Visa Sponsorship
Apply
AI Summary

Join our AI engineering team to build core intelligence behind our platform, combining cutting-edge AI with practical systems work. Design voice experiences, build browser agents, and optimize LLM behavior for production reliability. Ship working AI features that solve real problems.

Key Highlights
Build and optimize voice AI systems
Design browser agents that navigate web applications
Implement browser automation with computer vision and DOM understanding
Key Responsibilities
Build and optimize voice AI systems
Design browser agents that navigate web applications
Implement browser automation with computer vision and DOM understanding
Engineer prompt systems and LLM workflows
Create evaluation frameworks
Integrate multimodal AI
Manage AI Infrastructure
Monitor and improve AI system performance
Technical Skills Required
Production experience with LLMs Speech AI (STT/TTS systems) Browser automation (Playwright, Puppeteer, Selenium) Python Async programming and real-time systems Prompt engineering Retrieval systems Agent frameworks
Benefits & Perks
Meaningful equity stake
Visa sponsorship available
Work on cutting-edge voice AI and browser agents
Nice to Have
Experience building autonomous agents or multi-step AI workflows
Knowledge of computer vision for UI understanding and visual grounding
Fine-tuning or training language models for specialized tasks

Job Description


AI/ML Applied Engineer
The Opportunity

Join our AI engineering team in the US to build the core intelligence behind our platform. You'll work at the intersection of voice AI, browser automation, and large language models - creating agents that can listen, speak, navigate interfaces, and interact naturally with users in real-time.

This role combines cutting-edge AI with practical systems work. You'll design voice experiences, build browser agents that understand and control web applications, and optimize LLM behavior for production reliability. We ship working AI features that solve real problems, balancing innovation with pragmatic constraints.


We sponsor visas for qualified candidates.


Core Responsibilities

- Build and optimize voice AI systems using speech-to-text and text-to-speech models

- Design browser agents that navigate, understand, and interact with web applications

- Implement browser automation with computer vision and DOM understanding

- Engineer prompt systems and LLM workflows for consistent, intelligent behavior

- Create evaluation frameworks to measure voice quality, agent accuracy, and user experience

- Integrate multimodal AI - combining voice, vision, and language understanding

- Build real-time AI pipelines where latency and reliability are critical

- Manage the AI Infrastructure and take care of it

- Monitor and improve AI system performance in production environments


Technical Requirements

- Production experience with LLMs (OpenAI, Anthropic, or open-source models)

- Hands-on work with speech AI (STT/TTS systems like Deepgram, ElevenLabs, Whisper)

- Experience with browser automation (Playwright, Puppeteer, Selenium) or computer vision

- Strong Python skills with async programming and real-time systems

- Understanding of prompt engineering, retrieval systems, and agent frameworks

- Ability to debug complex AI behaviors and build observability tools

- Software engineering fundamentals for production AI systems


Nice to Have

- Experience building autonomous agents or multi-step AI workflows

- Knowledge of computer vision for UI understanding and visual grounding

- Fine-tuning or training language models for specialized tasks

- Real-time audio processing and streaming architectures

- Background in NLP, machine learning research, or AI systems


Why Karumi
  • Meaningful equity stake in a backed, fast-growing company
  • Work on cutting-edge voice AI and browser agents in production
  • Shape how AI systems interact with users and software interfaces
  • Small team with direct impact on core product capabilities
  • Visa sponsorship available


About Karumi
  • Karumi lets SaaS companies deliver personalized product demos at scale, 24/7, in any language. Our AI agent learns from any software product, joins video calls with prospects, and navigates the product in a real browser while talking, listening, and taking actions. By delivering the “aha moment” instantly, Karumi eliminates demo waiting times and boosts conversion.

Similar Jobs

Explore other opportunities that match your interests

Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Not Applicable

employia

United State

Full Stack Product Engineer

Programming
6h ago

Premium Job

Sign up is free! Login or Sign up to view full details.

•••••• •••••• ••••••
Job Type ••••••
Experience Level ••••••

employia

United State
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Mid-Senior level

clera

United State

Subscribe our newsletter

New Things Will Always Update Regularly