Cutshort logo
VMax eSolutions India Pvt Ltd
VMax eSolutions India Pvt Ltd cover picture
VMax eSolutions India Pvt Ltd logo

VMax eSolutions India Pvt Ltd

https://vmaxindia.com
Founded :
2008
Type :
Services
Size :
100-1000
Stage :
Profitable

About

VMAX is an ISO 90012015 and ISO 27001:2013 certified organisation based in Hyderabad, India. It builds custom, tailor-made, and scalable products that can be easily integrated with third-party systems. VMAX provides cutting-edge solutions to its clients and has helped organizations emerge as winners in their respective industries.
Read more

Company social profiles

twitter

Jobs at VMax eSolutions India Pvt Ltd

VMax eSolutions India Pvt Ltd
at VMax eSolutions India Pvt Ltd
Bachu Sai Nikheel
Posted by Bachu Sai Nikheel
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Hyderabad
10 - 15 yrs
₹35L - ₹45L / yr
Generative AI
PEFT (Parameter-Efficient Fine-Tuning)
Voice processing
Artificial Intelligence (AI)
GPU computing
+3 more

We are seeking an experienced AI Architect to design, build, and scale production-ready AI voice conversation agents deployed locally (on-prem / edge / private cloud) and optimized for GPU-accelerated, high-throughput environments.

You will own the end-to-end architecture of real-time voice systems, including speech recognition, LLM orchestration, dialog management, speech synthesis, and low-latency streaming pipelines—designed for reliability, scalability, and cost efficiency.

This role is highly hands-on and strategic, bridging research, engineering, and production infrastructure.


Key Responsibilities

Architecture & System Design

  • Design low-latency, real-time voice agent architectures for local/on-prem deployment
  • Define scalable architectures for ASR → LLM → TTS pipelines
  • Optimize systems for GPU utilization, concurrency, and throughput
  • Architect fault-tolerant, production-grade voice systems (HA, monitoring, recovery)

Voice & Conversational AI

  • Design and integrate:
  • Automatic Speech Recognition (ASR)
  • Natural Language Understanding / LLMs
  • Dialogue management & conversation state
  • Text-to-Speech (TTS)
  • Build streaming voice pipelines with sub-second response times
  • Enable multi-turn, interruptible, natural conversations

Model & Inference Engineering

  • Deploy and optimize local LLMs and speech models (quantization, batching, caching)
  • Select and fine-tune open-source models for voice use cases
  • Implement efficient inference using TensorRT, ONNX, CUDA, vLLM, Triton, or similar

Infrastructure & Production

  • Design GPU-based inference clusters (bare metal or Kubernetes)
  • Implement autoscaling, load balancing, and GPU scheduling
  • Establish monitoring, logging, and performance metrics for voice agents
  • Ensure security, privacy, and data isolation for local deployments

Leadership & Collaboration

  • Set architectural standards and best practices
  • Mentor ML and platform engineers
  • Collaborate with product, infra, and applied research teams
  • Drive decisions from prototype → production → scale

Required Qualifications

Technical Skills

  • 7+ years in software / ML systems engineering
  • 3+ years designing production AI systems
  • Strong experience with real-time voice or conversational AI systems
  • Deep understanding of LLMs, ASR, and TTS pipelines
  • Hands-on experience with GPU inference optimization
  • Strong Python and/or C++ background
  • Experience with Linux, Docker, Kubernetes

AI & ML Expertise

  • Experience deploying open-source LLMs locally
  • Knowledge of model optimization:
  • Quantization
  • Batching
  • Streaming inference
  • Familiarity with voice models (e.g., Whisper-like ASR, neural TTS)

Systems & Scaling

  • Experience with high-QPS, low-latency systems
  • Knowledge of distributed systems and microservices
  • Understanding of edge or on-prem AI deployments

Preferred Qualifications

  • Experience building AI voice agents or call automation systems
  • Background in speech processing or audio ML
  • Experience with telephony, WebRTC, SIP, or streaming audio
  • Familiarity with Triton Inference Server / vLLM
  • Prior experience as Tech Lead or Principal Engineer

What We Offer

  • Opportunity to architect state-of-the-art AI voice systems
  • Work on real-world, high-scale production deployments
  • Competitive compensation and equity (if applicable)
  • High ownership and technical influence
  • Collaboration with top-tier AI and infrastructure talent
Read more
VMax eSolutions India Pvt Ltd
at VMax eSolutions India Pvt Ltd
Bachu Sai Nikheel
Posted by Bachu Sai Nikheel
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Hyderabad
3 - 5 yrs
₹20L - ₹25L / yr
Artificial Intelligence (AI)
AI Agents
Voice recognition
Generative AI
skill iconMachine Learning (ML)

Company Description


VMax e-Solutions India Private Limited, based in Hyderabad, is a dynamic organization specializing in Open Source ERP Product Development and Mobility Solutions. As an ISO 9001:2015 and ISO 27001:2013 certified company, VMax is dedicated to delivering tailor-made and scalable products, with a strong focus on e-Governance projects across multiple states in India. The company's innovative technologies aim to solve real-life problems and enhance the daily services accessed by millions of citizens. With a culture of continuous learning and growth, VMax provides its team members opportunities to develop expertise, take ownership, and grow their careers through challenging and impactful work.


About the Role


We’re hiring a Senior Data Scientist with deep real-time voice AI experience and strong backend engineering skills.


1. You’ll own and scale our end-to-end voice agent pipeline that powers AI SDRs, customer support 2. agents, and internal automation agents on calls. This is a hands-on, highly technical role where you’ll design and optimize low-latency, high-reliability voice systems.


3. You’ll work closely with our founders, product, and platform teams, with significant ownership over architecture, benchmarks.


What You’ll Do


1. Own the voice stack end-to-end – from telephony / WebRTC entrypoints to STT, turn-taking, LLM reasoning, and TTS back to the caller.


2. Design for real-time – architect and optimize streaming pipelines for sub-second latency, barge-in, interruptions, and graceful recovery on bad networks.


3. Integrate and tune models – evaluate, select, and integrate STT/TTS/LLM/VAD providers (and self-hosted models) for different use-cases, balancing quality, speed, and cost.


4. Build orchestration & tooling – implement agent orchestration logic, evaluation frameworks, call simulators, and dashboards for latency, quality, and reliability.


5. Harden for production – ensure high availability, observability, and robust fault-tolerance for thousands of concurrent calls in customer VPCs.


6. Shape the voice roadmap – influence how voice fits into our broader Agentic OS vision (simulation, analytics, multi-agent collaboration, etc.).


You’re a Great Fit If You Have


1. 6+ years of software engineering experience (backend or full-stack) in production systems.


2. Strong experience building real-time voice agents or similar systems using:


STT / ASR (e.g. Whisper, Deepgram, Assembly, AWS Transcribe, GCP Speech)


TTS (e.g. ElevenLabs, PlayHT, AWS Polly, Azure Neural TTS)


VAD / turn-taking and streaming audio pipelines


LLMs (e.g. OpenAI, Anthropic, Gemini, local models)


3. Proven track record designing and operating low-latency, high-throughput streaming systems (WebRTC, gRPC, websockets, Kafka, etc.).


4. Hands-on experience integrating ML models into live, user-facing applications with real-time inference & monitoring.


5. Solid backend skills with Python and TypeScript/Node.js; strong fundamentals in distributed systems, concurrency, and performance optimization.


6. Experience with cloud infrastructure – especially AWS (EKS, ECS, Lambda, SQS/Kafka, API Gateway, load balancers).


7. Comfortable working in Kubernetes / Docker environments, including logging, metrics, and alerting.


8. Startup DNA – at least 2 years in an early or mid-stage startup where you shipped fast, owned outcomes, and worked close to the customer.


Nice to Have


1. Experience self-hosting AI models (ASR / TTS / LLMs) and optimizing them for latency, cost, and reliability.


2. Telephony integration experience (e.g. Twilio, Vonage, Aircall, SignalWire, or similar).


3. Experience with evaluation frameworks for conversational agents (call quality scoring, hallucination checks, compliance rules, etc.).


4. Background in speech processing, signal processing, or dialog systems.


5. Experience deploying into enterprise VPC / on-prem environments and working with security/compliance constraints.

Read more
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo

Similar companies

Quantiphi cover picture
Quantiphi's logo

Quantiphi

https://quantiphi.com
Founded
2013
Type
Products & Services
Size
1000-5000
Stage
Profitable

About the company

Quantiphi is an award-winning AI-first digital engineering company driven by the desire to reimagine and realize transformational opportunities at the heart of the business. Since its inception in 2013, Quantiphi has solved the toughest and most complex business problems by combining deep industry experience, disciplined cloud, and data-engineering practices, and cutting-edge artificial intelligence research to achieve accelerated and quantifiable business results.

Jobs

8

The OmniJobs cover picture
The OmniJobs's logo

The OmniJobs

https://theomnijobs.com
Founded
2021
Type
Services
Size
10-50
Stage
Bootstrapped

About the company

Jobs

12

Blitzy cover picture
Blitzy's logo

Blitzy

https://blitzy.com
Founded
2024
Type
Product
Size
0-20
Stage
Raised funding

About the company

Blitzy is a Boston, MA based Generative AI Start-up with an established office in Pune, India. We are on a mission to automate custom software creation to unlock the next industrial revolution. We're backed by multiple tier 1 investors, have success as founders at the last start-up, and dozens of Generative AI patents to our names.


Our Culture

Our Co-Founder and CTO is a Serial Gen AI Inventor who grew up in Pune, India, is a BITS Pilani graduate, and worked at NVIDIA's Pune office for 6 years. There, he was promoted 5 times in 6 years and was transferred to the NVIDIA Headquarters in Santa Clara, California. After making significant contributions to NVIDIA, he proceeded to attend Harvard for his dual Masters in Engineering and MBA from HBS. Our other Co-Founder/CEO is a successful Serial Entrepreneur who has built multiple companies. As a team, we work very hard, have a curious mind-set, and believe in a low-ego high output approach.


Funding Journey

In September 2024, Blitzy secured $4.4M in seed funding from prominent investors including Link Ventures, Asymmetric Capital Partners, Flybridge, and four other strategic investors, demonstrating strong market confidence in their autonomous software development platform.


Our Values

  1. We move Blitzy Fast: Time is both our company's and our client's most precious asset. We move fast and fearlessly to innovate internally and deliver exceptional software externally to our clients.
  2. We have a Championship Mindset: We operate like a professional sports team. We win as a team by holding ourselves and each other to high standards, collaborating in-person, and remaining focused on the mission.
  3. We have a Passion for Invention: We are inventors at heart. We value starting with best practices and open source, but we are pushing the frontier of what is possible.
  4. We Work for the Customer: We focus on delivering outsized value to the customers we work with and expanding those relationships to deep, meaningful partnerships.


What We Ask of Candidates

Please ask yourself if you are ready for a challenge before applying. Even in optimal conditions, Start-Ups are hard, and are always a lot of work. What you do week to week will change. If this feels exciting, not concerning, that's a good sign.

Jobs

3

FreeADS cover picture
FreeADS's logo

FreeADS

https://freeads.no
Founded
2025
Type
Products & Services
Size
0-20
Stage
Bootstrapped

About the company

Jobs

2

Honeybee Digital cover picture
Honeybee Digital's logo

Honeybee Digital

https://honeybeedigital.com
Founded
2014
Type
Services
Size
0-20
Stage
Bootstrapped

About the company

Honeybee Digital - is a leading digital marketing and website design Service Provides services like SEO, web designing, Wordpress, influencer marketing from Gandhinagar, Gujarat, India.

Jobs

11

AGI Ready cover picture
AGI Ready's logo

AGI Ready

https://agiready.io
Founded
2025
Type
Products & Services
Size
0-20
Stage
Bootstrapped

About the company

Build intelligent GTM automation that gets smarter as AI advances. Lead enrichment, workflow automation, and unified data systems for Series A+ companies scaling revenue without scaling headcount.

Jobs

1

Gunpowder Innovations cover picture
Gunpowder Innovations's logo

Gunpowder Innovations

https://gunpowderinnovations.com
Founded
2023
Type
Products & Services
Size
0-20
Stage
Raised funding

About the company

Gunpowder Digital offers AI-driven financial advice, custom fintech solutions, and mobile app development for wealth management and startups in the UK

Jobs

3

Power tech Digital Limited cover picture
Power tech Digital Limited's logo

Power tech Digital Limited

https://powertech.digital
Founded
2024
Type
Products & Services
Size
0-20
Stage
Bootstrapped

About the company

UK-based software consultancy specializing in AI-driven custom development, Microsoft Power Platform solutions, and innovative SaaS products. Transform your business with PowerTech Digital.

Jobs

2

Founded
2023
Type
Product
Size
0-20
Stage
Raised funding

About the company

Ande is an AI-native, full-stack TypeScript platform built on React, Node.js, GraphQL, and Postgres, running on AWS and powering web, mobile, internal operations, deep integrations, and agentic workflows.

Our product sits at the intersection of enterprise workflows, hospitality operations, payments, compliance, procurement, and AI — giving engineers the opportunity to solve problems that combine polished user experiences with complex real-world systems.


Engineering at Ande is deeply product-oriented and systems-heavy. We care about:

  • Type safety and shared abstractions
  • Fast iteration and observable production systems
  • High-quality user experiences
  • Building durable foundations for a category-defining platform

PMs and engineers work closely with the business domain, contributing directly to:

  • Booking experiences
  • Client entertainment policies
  • Venue operations
  • Spend visibility and approvals
  • Payments and procurement workflows
  • Enterprise integrations
  • AI-driven workflows that reduce manual coordination across enterprises and hospitality partners

Founders

  • Lohit Sarma
  • Ashish Bidadi
  • Michael McDermott


Jobs

0

Kris@Work cover picture
Kris@Work's logo

Kris@Work

https://krisatwork.com
Founded
2024
Type
Product
Size
20-100
Stage
Raised funding

About the company

Kris replaces fragmented outbound with one intelligent sales window. Signal-based prospecting, prioritization & outreach. 15× more qualified meetings at 1/3 the cost.

Jobs

1

Want to work at VMax eSolutions India Pvt Ltd?
VMax eSolutions India Pvt Ltd's logo
Why apply via Cutshort?
Connect with actual hiring teams and get their fast response. No spam.
Find more jobs