Design, train, deploy, monitor, and optimize AI/ML models powering a cloud-based voice agent platform. Responsibilities include STT, NLU, and TTS development; MLOps pipelines; AI API services; inference and latency optimization; frontend prototypes and dashboards; research, experimentation, vendor evaluation, documentation, and code quality. The role requires collaboration with backend, frontend, and infrastructure teams.
This is a remote position.
Job timings: Mon - Fri US EST Time zone
Job Location: Pakistan (Remote)
Experience: 5+ years
Job Location: Pakistan (Remote)
Experience: 5+ years
PRODUCT CONTEXT
CloudPSO's first standardized AI Agent product enables enterprises to deploy intelligent voice assistants for customer support, sales, and internal operations. Its core capabilities include:
- Speech-to-Text (STT): Real-time voice transcription.
- Natural Language Understanding (NLU): Intent recognition, context awareness, and conversation control.
- Text-to-Speech (TTS): Natural, human-like voice responses.
- Enterprise Platform: Cloud-native deployment with scalability, low latency, security, observability, and enterprise-grade reliability.
We are looking for an experienced AI Engineer / Developer to join our product development team and drive the AI capabilities of this platform.
ROLE OVERVIEW
The AI Engineer will be responsible for designing, training, deploying, and optimizing AI/ML models that power the Voice Agent platform. This role spans the full AI lifecycle—from research and experimentation to production deployment and monitoring. The ideal candidate has deep expertise in speech and NLP technologies, strong engineering practices, and the ability to collaborate across backend, frontend, and infrastructure teams. Frontend development skills are mandatory, as this role includes building prototypes, dashboards, and internal demos.
KEY RESPONSIBILITIES
- Model Development: Design, train, and fine-tune models and pipelines for STT, NLU, and TTS use cases.
- MLOps: Build and maintain MLOps pipelines for model deployment, monitoring, and retraining.
- API Development: Develop APIs and services that expose AI capabilities to backend and frontend systems.
- Performance Optimization: Collaborate with backend engineers to optimize inference performance and end-to-end latency.
- Frontend Prototyping: Create frontend prototypes, dashboards, and demonstrations using a modern frontend framework.
- Research & Experimentation: Conduct research and controlled experiments to improve model accuracy, quality, and performance.
- Vendor Evaluation: Evaluate external models and providers against product requirements and measurable benchmarks.
- Documentation: Document model architectures, experiments, evaluation results, and deployment processes.
- Code Quality: Participate in code reviews and maintain high engineering and reproducibility standards.
Requirements
PREFERRED SKILLS
- Experience with real-time streaming using WebSocket or WebRTC.
- Knowledge of model quantization, pruning, and inference optimization.
- Familiarity with SIP, WebRTC, or PSTN integration.
- Open-source contributions or research publications in AI, speech, or voice domains.
REQUIRED QUALIFICATIONS
- Experience: 7–9 years of AI/ML development experience.
- Programming: Strong Python expertise with experience in PyTorch, TensorFlow, Hugging Face, or equivalent frameworks.
- Speech & NLP: Hands-on experience with speech/audio processing and natural language processing techniques.
- Model Deployment: Hands-on deployment experience using ONNX, TensorRT, Triton, or equivalent tooling.
- Frontend (Mandatory): Experience with React, Vue.js, or Angular.
- Cloud AI Platforms: Experience with AWS SageMaker, GCP Vertex AI, Azure ML, or an equivalent cloud AI platform.
- Containerization: Knowledge of Docker and Kubernetes.
- Version Control & CI/CD: Experience with Git and continuous integration and delivery pipelines.
- Analytical Skills: Strong analytical and problem-solving skills.
Benefits
- Medical insurance
- Company gadgets
- Paid time off
- Stock options (ESOP)
- Competitive salary and benefits package.
- Opportunities for professional development and growth.
- Collaborative and innovative work environment.
- Chance to work on cutting-edge cloud projects.
- Supportive and inclusive company culture
Similar Jobs
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Designs and deploys enterprise AI/ML solutions, scalable models, data pipelines, and MLOps architectures in cloud production environments. Responsibilities include model monitoring and optimization, responsible AI implementation, workflow automation, emerging technology evaluation, cross-functional architecture collaboration, code reviews, and mentoring. The role focuses on healthcare challenges and supports production-grade machine learning systems, including real-time inference and continuous integration and deployment.
Top Skills:
AirflowAWSAzureCi/CdDockerFhirGCPGenerative AiHl7KafkaKubeflowKubernetesLarge Language ModelsMlflowMlopsPythonPyTorchRetrieval-Augmented GenerationScikit-LearnSparkSQLTensorFlow
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Design, build, deploy, and scale production AI agents and platforms for legal workflows. Develop and evaluate agent solutions, optimize performance and cost, redesign business processes, and promote responsible AI adoption. Collaborate with Legal, Enterprise AI, Data, Security, Privacy, and Engineering teams to deliver secure, reliable, scalable solutions. Contribute frameworks, training, governance guidance, and innovative applications using modern LLM and AI techniques.
Top Skills:
Advanced RetrievalAgent FrameworksAi AgentsAi GatewaysAi/MlAWSFine-TuningGCPKnowledge GraphsLarge Language ModelsModel Context ProtocolModel CustomizationMulti-Agent SystemsMultimodal Document ProcessingPythonSecure Tool IntegrationsSynthetic Data
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Design, deploy, and scale production AI/ML solutions for healthcare, including predictive models, generative AI, statistical analysis, microservices, APIs, and MLOps pipelines. Analyze complex clinical datasets, implement model monitoring and governance, evaluate emerging AI technologies, and collaborate with engineering, data, product, and clinical teams to improve care delivery and operational efficiency.
Top Skills:
SparkAWSAzureAzure MlCi/CdDatabricksDockerFhirGCPGenerative AiHipaaHitrustHl7KubernetesLarge Language ModelsMlopsNumpyPandasPrompt EngineeringPythonPyTorchRetrieval-Augmented GenerationSagemakerScikit-LearnScipySnowflakeTensorFlowVertex Ai
What you need to know about the NYC Tech Scene
As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.
Key Facts About NYC Tech
- Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
- Key Industries: Artificial intelligence, Fintech
- Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
- Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory


