Loading memory…
Loading memory…
You will own the backbone that runs thousands of concurrent AI-powered phone calls inside bank-grade on-premise environments. Not just keeping pods alive, you architect distributed systems that handle real-time voice, scale STT / LLM / TTS inference across customer GPU clusters, integrate with enterprise telephony (Cisco CUBE, Genesys, Asterisk), and deploy behind the firewalls of largest financial institutions. Your work decides whether our platform answers a bank's rush-hour traffic or leaves customers on dead air. • Own on-prem deployments into OpenShift clusters inside banks. Helm charts, image registries, GPU allocation, CyberArk integration, SAML 2.0 / OIDC SSO. • Scale GPU inference infrastructure for our STT, TTS, and LLM models across multiple customer environments (H100 / H200, NVLink, Triton or vLLM). • Integrate with telephony : Asterisk, SIP trunks, Cisco CUBE, Genesys, WebRTC. SIP header parsing (X-Genesys-*), direction routing, warm transfers, DTMF. • Own reliability : Splunk SIEM forwarding, Langfuse and Grafana observability, incident playbooks for bank-grade 24/7 SLAs. • Security and compliance : RBAC, pentest remediation, KVKK and BDDK compliance patterns, pod se