Ramneek Kalra
Sr. Consultant, Improving Inc.
Hyderābād, India
Actions
I am an Improver working as a Sr. Consultant I at Improving Enterprises Inc, specialized in resolving complex technical issues across containerized workloads in EKS, ECS, AWS Batch, and related infrastructure. My focus: reducing downtime, accelerating resolution, and restoring customer trust—at scale.
With a strong foundation in cloud and edge computing, DevOps, and multi-cloud architectures, I’ve supported high-priority customers through incident diagnostics, root cause analysis, and deep-level troubleshooting in production environments.
What makes me effective?
- Calm under pressure during high-severity events
- Obsessed with digging deep to find the real “why” behind problems
- Trusted by customers and engineers alike for clear, actionable solutions
Beyond my technical-space, I’ve led 1000+ tech education sessions, contributed to open-source, authored 20+ research papers, and helped grow communities as an IEEE Impact Creator. I bring both hands-on engineering skill and a broader commitment to innovation and mentorship.
Area of Expertise
Topics
On-Call Nightmare of Non-Deterministic Stacks: Managing Cascading Failures in Agentic K8s Clusters
Moving autonomous AI agents from a local IDE demo to a high-scale production Kubernetes environment introduces an entirely new class of infrastructure volatility. Unlike traditional deterministic microservices, agentic workloads exhibit unpredictable runtime behaviors, spontaneous burst capacity demands, and non-deterministic memory footprints that easily bypass classic horizontal pod autoscaling.
This session steps away from the AI hype to address the raw infrastructure reality that engineering leaders and platform teams face when agents go rogue. We will explore real-world patterns for isolating probabilistic AI workloads from core business logic using advanced Kubernetes scheduling, custom metric-driven autoscaling, and intelligent resource quotas. Attendees will analyze a blueprint for building resilient platform guardrails that prevent cascading cluster failures, manage vector database state sync issues during outages, and maintain systemic stability when AI logic behaves unpredictably.
AI Agent Security 101: Hands-on with OAuth 2.1 and Network Policies
Almost all agents in production are protected by a static API key as an Environment variable. Here in this workshop, we will achieve security by a vulnerable demo with MCP server.
Adding OAuth 2.1 with Keycloak as the IdP, along with blocking destructive tools via Kyverno policies and lock down pod to pod traffic with Kubernetes Network Policies.
To demonstrate, we will have a live attack and defend round, making this a quite interesting workshop where attendees attempt prompt-injection and credential exfiltration against each other clusters to see what each control catches.
When the Cluster Bleeds: High-Severity Incident Diagnostics for EKS Workloads at Scale
When production Kubernetes clusters go down, clock cycles turn into dollar signs. Moving beyond basic kubectl logs and describe commands, this session dives deep into the trenches of real-world, high-severity EKS and AWS Batch incidents.
Led by a Senior Consultant who specializes in rapid incident restoration, we will break down a repeatable, calm-under-pressure diagnostic framework for containerized workloads. Attendees will walk away with architectural blueprints and operational runbooks covering:
- Tracing silent failures in multi-tenant EKS control planes and VPC CNI bottlenecks.
- Decoupling application bugs from underlying AWS infrastructure drift.
- Real-world post-mortems of root-cause analyses that saved customer trust and minimized downtime.
Cloud Native Benefit: Shifts the focus from theoretical Day-1 deployment to the harsh realities of Day-2 operations, helping platform engineers build resilient, highly observable infrastructure.
Unlocking Liquid Compute: Scaling AI & LLM Inference via Kubernetes Dynamic Resource Allocation
As generative AI and Large Language Models (LLMs) transition from experimental batch jobs to business-critical production infrastructure, orchestrating expensive hardware like GPUs/TPUs has become an operational hurdle. The traditional Device Plugin framework often falls short, forcing platform engineers into manual node pinning, restrictive "all-or-nothing" allocations, and static node selectors.
Enter Dynamic Resource Allocation (DRA)—the game-changing evolution in Kubernetes device management.
In this technical session, we will break down how DRA shifts the paradigm from rigid hardware assignments to a highly fluid, "liquid" resource pool. We will explore how DRA decouples hardware inventory from workload requirements using ResourceSlices and ResourceClaims, allowing the Kube-scheduler to make topology-aware, capability-based scheduling decisions.
Attendees will learn production-ready best practices for running high-volume, low-latency LLM inference workloads. We will cover:
- How to define custom DeviceClasses as abstract blueprints for developers.
- Strategies for fine-grained resource sharing (like VRAM requirements and interconnect topology).
- Essential cluster administration safeguards, including DRA driver deployment, liveness probing, and graceful node draining to minimize scheduling delays.
Whether you are scaling open-source foundational models or building internal developer platforms for data science teams, this talk will give you the architectural blueprint to optimize accelerator utilization, eliminate resource drift, and slash your AI infrastructure TCO.
IEEE DSSYWLC
I spoke on "From Campus to Career: Stress-Free Guide to Industry-Readiness" at The DSSYWLC represents a premier flagship event of the IEEE Delhi Section. It serves as a unified platform bringing together three vital pillars of the engineering community: Young Professionals & Students, Women in Engineering (WIE), and Life Members.
IEEE YP Delhi Summit 1.0
I along with my IEEE YP Delhi Section Team organized IEEE YP Delhi Summit 1.0 which was for the benefit of students' & professionals' community to get guidance from Industrialists and retired engineers to excel in their career-ahead. Alongside, live-hiring & mock interviews booth were setup to make Final-Year students prepare for industry-interviews. This event catered around 90+ participants from across North India Region.
Please note that Sessionize is not responsible for the accuracy or validity of the data provided by speakers. If you suspect this profile to be fake or spam, please let us know.
Jump to top