Cutshort logo
For Employers
VMax eSolutions India Pvt Ltd
VMax eSolutions India Pvt Ltd cover picture
VMax eSolutions India Pvt Ltd logo

VMax eSolutions India Pvt Ltd

https://vmaxindia.com
Founded :
2008
Type :
Services
Size :
100-1000
Stage :
Profitable

About

VMAX is an ISO 90012015 and ISO 27001:2013 certified organisation based in Hyderabad, India. It builds custom, tailor-made, and scalable products that can be easily integrated with third-party systems. VMAX provides cutting-edge solutions to its clients and has helped organizations emerge as winners in their respective industries.
Read more

Company social profiles

twitter

Jobs at VMax eSolutions India Pvt Ltd

VMax eSolutions India Pvt Ltd
at VMax eSolutions India Pvt Ltd
Bachu Sai Nikheel
Posted by Bachu Sai Nikheel
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Hyderabad
10 - 15 yrs
₹35L - ₹45L / yr
Generative AI
PEFT (Parameter-Efficient Fine-Tuning)
Voice processing
Artificial Intelligence (AI)
GPU computing
+3 more

We are seeking an experienced AI Architect to design, build, and scale production-ready AI voice conversation agents deployed locally (on-prem / edge / private cloud) and optimized for GPU-accelerated, high-throughput environments.

You will own the end-to-end architecture of real-time voice systems, including speech recognition, LLM orchestration, dialog management, speech synthesis, and low-latency streaming pipelines—designed for reliability, scalability, and cost efficiency.

This role is highly hands-on and strategic, bridging research, engineering, and production infrastructure.


Key Responsibilities

Architecture & System Design

  • Design low-latency, real-time voice agent architectures for local/on-prem deployment
  • Define scalable architectures for ASR → LLM → TTS pipelines
  • Optimize systems for GPU utilization, concurrency, and throughput
  • Architect fault-tolerant, production-grade voice systems (HA, monitoring, recovery)

Voice & Conversational AI

  • Design and integrate:
  • Automatic Speech Recognition (ASR)
  • Natural Language Understanding / LLMs
  • Dialogue management & conversation state
  • Text-to-Speech (TTS)
  • Build streaming voice pipelines with sub-second response times
  • Enable multi-turn, interruptible, natural conversations

Model & Inference Engineering

  • Deploy and optimize local LLMs and speech models (quantization, batching, caching)
  • Select and fine-tune open-source models for voice use cases
  • Implement efficient inference using TensorRT, ONNX, CUDA, vLLM, Triton, or similar

Infrastructure & Production

  • Design GPU-based inference clusters (bare metal or Kubernetes)
  • Implement autoscaling, load balancing, and GPU scheduling
  • Establish monitoring, logging, and performance metrics for voice agents
  • Ensure security, privacy, and data isolation for local deployments

Leadership & Collaboration

  • Set architectural standards and best practices
  • Mentor ML and platform engineers
  • Collaborate with product, infra, and applied research teams
  • Drive decisions from prototype → production → scale

Required Qualifications

Technical Skills

  • 7+ years in software / ML systems engineering
  • 3+ years designing production AI systems
  • Strong experience with real-time voice or conversational AI systems
  • Deep understanding of LLMs, ASR, and TTS pipelines
  • Hands-on experience with GPU inference optimization
  • Strong Python and/or C++ background
  • Experience with Linux, Docker, Kubernetes

AI & ML Expertise

  • Experience deploying open-source LLMs locally
  • Knowledge of model optimization:
  • Quantization
  • Batching
  • Streaming inference
  • Familiarity with voice models (e.g., Whisper-like ASR, neural TTS)

Systems & Scaling

  • Experience with high-QPS, low-latency systems
  • Knowledge of distributed systems and microservices
  • Understanding of edge or on-prem AI deployments

Preferred Qualifications

  • Experience building AI voice agents or call automation systems
  • Background in speech processing or audio ML
  • Experience with telephony, WebRTC, SIP, or streaming audio
  • Familiarity with Triton Inference Server / vLLM
  • Prior experience as Tech Lead or Principal Engineer

What We Offer

  • Opportunity to architect state-of-the-art AI voice systems
  • Work on real-world, high-scale production deployments
  • Competitive compensation and equity (if applicable)
  • High ownership and technical influence
  • Collaboration with top-tier AI and infrastructure talent
Read more
VMax eSolutions India Pvt Ltd
at VMax eSolutions India Pvt Ltd
Bachu Sai Nikheel
Posted by Bachu Sai Nikheel
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Hyderabad
3 - 5 yrs
₹20L - ₹25L / yr
Artificial Intelligence (AI)
AI Agents
Voice recognition
Generative AI
skill iconMachine Learning (ML)

Company Description


VMax e-Solutions India Private Limited, based in Hyderabad, is a dynamic organization specializing in Open Source ERP Product Development and Mobility Solutions. As an ISO 9001:2015 and ISO 27001:2013 certified company, VMax is dedicated to delivering tailor-made and scalable products, with a strong focus on e-Governance projects across multiple states in India. The company's innovative technologies aim to solve real-life problems and enhance the daily services accessed by millions of citizens. With a culture of continuous learning and growth, VMax provides its team members opportunities to develop expertise, take ownership, and grow their careers through challenging and impactful work.


About the Role


We’re hiring a Senior Data Scientist with deep real-time voice AI experience and strong backend engineering skills.


1. You’ll own and scale our end-to-end voice agent pipeline that powers AI SDRs, customer support 2. agents, and internal automation agents on calls. This is a hands-on, highly technical role where you’ll design and optimize low-latency, high-reliability voice systems.


3. You’ll work closely with our founders, product, and platform teams, with significant ownership over architecture, benchmarks.


What You’ll Do


1. Own the voice stack end-to-end – from telephony / WebRTC entrypoints to STT, turn-taking, LLM reasoning, and TTS back to the caller.


2. Design for real-time – architect and optimize streaming pipelines for sub-second latency, barge-in, interruptions, and graceful recovery on bad networks.


3. Integrate and tune models – evaluate, select, and integrate STT/TTS/LLM/VAD providers (and self-hosted models) for different use-cases, balancing quality, speed, and cost.


4. Build orchestration & tooling – implement agent orchestration logic, evaluation frameworks, call simulators, and dashboards for latency, quality, and reliability.


5. Harden for production – ensure high availability, observability, and robust fault-tolerance for thousands of concurrent calls in customer VPCs.


6. Shape the voice roadmap – influence how voice fits into our broader Agentic OS vision (simulation, analytics, multi-agent collaboration, etc.).


You’re a Great Fit If You Have


1. 6+ years of software engineering experience (backend or full-stack) in production systems.


2. Strong experience building real-time voice agents or similar systems using:


STT / ASR (e.g. Whisper, Deepgram, Assembly, AWS Transcribe, GCP Speech)


TTS (e.g. ElevenLabs, PlayHT, AWS Polly, Azure Neural TTS)


VAD / turn-taking and streaming audio pipelines


LLMs (e.g. OpenAI, Anthropic, Gemini, local models)


3. Proven track record designing and operating low-latency, high-throughput streaming systems (WebRTC, gRPC, websockets, Kafka, etc.).


4. Hands-on experience integrating ML models into live, user-facing applications with real-time inference & monitoring.


5. Solid backend skills with Python and TypeScript/Node.js; strong fundamentals in distributed systems, concurrency, and performance optimization.


6. Experience with cloud infrastructure – especially AWS (EKS, ECS, Lambda, SQS/Kafka, API Gateway, load balancers).


7. Comfortable working in Kubernetes / Docker environments, including logging, metrics, and alerting.


8. Startup DNA – at least 2 years in an early or mid-stage startup where you shipped fast, owned outcomes, and worked close to the customer.


Nice to Have


1. Experience self-hosting AI models (ASR / TTS / LLMs) and optimizing them for latency, cost, and reliability.


2. Telephony integration experience (e.g. Twilio, Vonage, Aircall, SignalWire, or similar).


3. Experience with evaluation frameworks for conversational agents (call quality scoring, hallucination checks, compliance rules, etc.).


4. Background in speech processing, signal processing, or dialog systems.


5. Experience deploying into enterprise VPC / on-prem environments and working with security/compliance constraints.

Read more
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo

Similar companies

Incubyte cover picture
Incubyte's logo

Incubyte

https://incubyte.co
Founded
2020
Type
Services
Size
20-100
Stage
Bootstrapped

About the company

About Us

Incubyte is an AI-first software development agency built on the principles of software craftsmanship—where how we build is just as important as what we build. We partner with organizations across stages, from enterprises looking to scale and modernize to early-stage founders bringing new ideas to life.


At Incubyte, AI is deeply integrated across the software development lifecycle to drive speed, efficiency, and smarter outcomes. Guided by Software Craftsmanship values and Extreme Programming practices, we combine high velocity with disciplined engineering to deliver reliable, high-impact solutions.

We don’t just build software—we incubate dedicated engineering teams. From designing systems to shaping team structures and organizational strategy, we enable our clients to launch and scale products that are relevant today and resilient for the future.


Whether you’re scaling an existing product, building from scratch, or optimizing manual processes, we help you move faster with confidence:

  • Scale and modernize your product
  • Launch quickly and iterate continuously
  • Automate processes for non-linear growth
  • Build systems that are stable, predictable, and measurable


Our approach is rooted in ownership. As a DevOps-driven organization, our engineers take responsibility for the entire lifecycle—from development to release—ensuring quality at every step.


Founded by product professionals, we bring a strong product mindset into services. We’re driven by curiosity, continuous learning, and a passion for building great software the right way.


We’re always looking for people who care deeply about code, craftsmanship, and growth. Join us if you’re excited to build, learn, and make an impact.

Jobs

13

TalentXO cover picture
TalentXO's logo

TalentXO

https://app.talentxo.com
Founded
2018
Type
Product
Size
Stage
Profitable

About the company

Jobs

106

Hiraya Digital cover picture
Hiraya Digital's logo

Hiraya Digital

https://hiraya.digital
Founded
2023
Type
Services
Size
0-20
Stage
Profitable

About the company

We are hiring for multiple clients

Jobs

4

TheCodersHub cover picture
TheCodersHub's logo

TheCodersHub

https://thecodershub.co.in
Founded
2024
Type
Products & Services
Size
0-20
Stage
Bootstrapped

About the company

TheCodersHub is a startup that offers services in #androiddevelopment, #webdevelopment, #softwaredevelopment, #mobiledevelopment, and #iosdevelopment. While still in the early stages, we are focused on growth and innovation within the tech industry.

Jobs

6

Alpha cover picture
Alpha's logo

Alpha

https://get-alpha.ai/careers
Founded
2025
Type
Products & Services
Size
0-20
Stage
Bootstrapped

About the company

Alpha brings a new AI native environment to build effective and reliable AI agents with no code.

Jobs

3

Dino Ventures cover picture
Dino Ventures's logo

Dino Ventures

https://indidino.com
Founded
2024
Type
Product
Size
20-100
Stage
Bootstrapped

About the company

At IndiDino, we are building Bharat-focused products for a billion people with the goal of creating a net positive impact. We are driven by a simple mission — to create meaningful products that solve real problems at scale and improve the lives of people across India. By combining innovation, technology, and a deep understanding of Bharat, we aim to build solutions that empower communities and create lasting value. 

Jobs

3

AmpereHour energy cover picture
AmpereHour energy's logo

AmpereHour energy

https://amperehourenergy.co.
Founded
2017
Type
Products & Services
Size
100-1000
Stage
Raised funding

About the company

Jobs

3

NSM Consultant cover picture
NSM Consultant's logo

NSM Consultant

https://nsm-consultant.com
Founded
2012
Type
Products & Services
Size
20-100
Stage
Bootstrapped

About the company

NSM Consultant offers expert software development, business consulting, and IT solutions. Led by Ningaiah Mahesh with 27+ years experience and $300M sales track record.

Jobs

2

Founded
Type
Size
Stage

About the company

Jobs

3

QuantumProp AI cover picture
QuantumProp AI's logo

QuantumProp AI

https://quantumpropai.com
Founded
2024
Type
Product
Size
20-100
Stage
Profitable

About the company

QuantumPropAI is a revolutionary financial technology company that makes AI-powered trading accessible to everyone. Our platform empowers individuals to earn a passive income by leveraging the power of advanced artificial intelligence, without requiring any prior trading experience or technical knowledge. We handle the complexities of market analysis and strategy execution, allowing our members to earn with a simple, daily, five-minute commitment.

Jobs

1

Want to work at VMax eSolutions India Pvt Ltd?
VMax eSolutions India Pvt Ltd's logo
Why apply via Cutshort?
Connect with actual hiring teams and get their fast response. No spam.
Find more jobs