Job Description
The role focuses on creating advanced routing capabilities, AI-driven traffic optimization, and intelligent decision-making systems designed specifically for next-generation AI applications. Working closely with product, platform, and AI teams, you'll play a key role in defining architecture, driving innovation, and leading large-scale technical initiatives from concept through production.
Responsibilities
Architect and develop AI Gateway services that intelligently manage and route AI and LLM-related traffic.
Design sophisticated load-balancing mechanisms that leverage model behavior, token consumption, latency metrics, and application context to optimize routing decisions.
Enhance ADC and GSLB platforms to better support AI-centric workloads through capabilities such as:
Contribute to advanced networking capabilities such as context-aware traffic orchestration, adaptive optimization, and AI-assisted anomaly detection.
Lead architectural decisions spanning data plane services, control plane functionality, configuration management, and distributed systems design.
Provide technical leadership and mentorship around AI application traffic patterns, scalability considerations, and performance tuning.
Partner with AI research, product management, and platform engineering teams to align technical strategy with business objectives.
Establish best practices for performance testing, benchmarking, capacity planning, and global-scale deployment
We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment regardless of their race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please send a request to HR@insightglobal.com.To learn more about how we collect, keep, and process your private information, please review Insight Global's Workforce Privacy Policy: https://insightglobal.com/workforce-privacy-policy/.
Required Skills & Experience
• Advanced degree in Computer Science, Networking, or related field (MS/PhD preferred).
• Deep knowledge of L4–L7 protocols (TCP/TLS/HTTP/2/3, QUIC), load balancing algorithms and GSLB strategies (DNS‑based, HTTP redirect, anycast).
• Strong programming skills in C/C++ (data plane), Go/Python (control/ML), and performance profiling (perf, eBPF, flamegraphs, VTune/nvprof).
• Applied ML: anomaly detection, time‑series forecasting, classification; experience with PyTorch/TensorFlow.
• Hands‑on experience with AI/ML systems, including:
Model inference pipelines, LLM API integrations
Benefit packages for this role will start on the 1st day of employment and include medical, dental, and vision insurance, as well as HSA, FSA, and DCFSA account options, and 401k retirement account access with employer matching. Employees in this role are also entitled to paid sick leave and/or other paid time off as provided by applicable law.