BEGIN:VCALENDAR
VERSION:2.0
PRODID:-//HasGeek//NONSGML Funnel//EN
DESCRIPTION:Real systems. Real engineers. Real lessons.
X-WR-CALDESC:Real systems. Real engineers. Real lessons.
NAME:Platform Engineering meet-up - July 25
X-WR-CALNAME:Platform Engineering meet-up - July 25
REFRESH-INTERVAL;VALUE=DURATION:PT12H
SUMMARY:Platform Engineering meet-up - July 25
TIMEZONE-ID:Asia/Kolkata
X-PUBLISHED-TTL:PT12H
X-WR-TIMEZONE:Asia/Kolkata
BEGIN:VEVENT
SUMMARY:Druid at 700 MBps: how we stopped babysitting our observability st
 ack
DTSTART:20260725T093000Z
DTEND:20260725T101500Z
DTSTAMP:20260726T005243Z
UID:session/4G1HMH8QtS6fPML8uDnjcp@hasgeek.com
SEQUENCE:2
CATEGORIES:Talk (30 mins),Bangalore meet-ups
CREATED:20260725T101938Z
DESCRIPTION:{Describe your session in 2 paragraphs}\nConfluent's observabi
 lity platform handles around 7 million events per second. Apache Druid is 
 what makes that work. It's where our metrics land and where every dashboar
 d\, alert\, and on-call investigation eventually pulls from. This session 
 is about how we actually operate Druid in production: how we lay out inges
 tion\, the data modeling choices that bit us before they helped us\, and t
 he kinds of failures you only really learn about by hitting them.\nI'll sp
 end most of the time on the two changes that made the biggest difference f
 or us. The first is how we split our clusters into a scraping tier for ing
 estion and recent queries\, and a historical tier for older data. The seco
 nd is the set of automations we built to handle the stuff Druid operators 
 see all the time: stuck tasks\, segments that won't balance\, queries that
  run forever\, coordinator weirdness. None of it is fancy\, but together i
 t's the difference between getting paged at 3am and not.\n\n{Mention 1-2 t
 akeaways from your session}\n1. How separating workloads\, rather than jus
 t scaling a single system\, can solve a class of reliability problems that
  more capacity won't.\n\n2. A practical way to think about which operation
 al pain points are worth automating away\, and which ones aren't.\n\n{Whic
 h audiences is your session going to beneficial for?}\nSREs and platform e
 ngineers running observability or large-scale data systems\, and anyone th
 inking about how to scale operations without scaling the on-call rotation.
  The patterns will land hardest if you operate a real-time analytics or me
 trics store\, but most of the ideas apply to any stateful data infrastruct
 ure under heavy load.\n\n{Add your bio - who you are\; where you work}\nPa
 rth Agrawal\, Senior Software Engineer at Confluent. I have been working o
 n the observability platform behind Confluent Cloud for the last 3 years. 
LAST-MODIFIED:20260725T102119Z
LOCATION:Bangalore
ORGANIZER;CN=Rootconf:MAILTO:no-reply@hasgeek.com
URL:https://hasgeek.com/rootconf/platform-engineering-meet-up-july-25/sche
 dule/druid-at-700-mbps-how-we-stopped-babysitting-our-observability-stack-
 4G1HMH8QtS6fPML8uDnjcp
BEGIN:VALARM
ACTION:display
DESCRIPTION:Druid at 700 MBps: how we stopped babysitting our observabilit
 y stack in 5 minutes
TRIGGER:-PT5M
END:VALARM
END:VEVENT
BEGIN:VEVENT
SUMMARY:From image chaos to control: introducing AVM (Agentic Vulnerabilit
 y Management)
DTSTART:20260725T102000Z
DTEND:20260725T110500Z
DTSTAMP:20260726T005243Z
UID:session/6cZS6Cngz2Q2UTGCMN4LU8@hasgeek.com
SEQUENCE:3
CATEGORIES:Talk (30 mins),Bangalore meet-ups
CREATED:20260725T101952Z
DESCRIPTION:__Session Description__\nAt Acceldata\, managing container ima
 ge sprawl\, addressing vulnerabilities\, and maintaining compliance across
  scaling cloud-native environments became a complex and fragmented challen
 ge. To solve this\, we built AVM (Agentic Vulnerability Management) - an e
 nterprise-grade platform designed to centralize and secure container image
  management. By acting as a unified control plane for container security\,
  AVM transforms security from a reactive bottleneck into an automated\, co
 llaborative workflow.\n\nDesigned specifically for DevOps\, platform\, and
  development teams\, AVM enables continuous vulnerability scanning\, polic
 y-driven governance and automated remediation via custom AI agents develop
 ed solely for vulnerability tracking and fixing. It automates the entire i
 mage lifecycle - from discovery and analysis to fix and promotion - ensuri
 ng that only secure\, compliant\, and pre-approved images reach production
 . By integrating security natively into the deployment workflow\, AVM sign
 ificantly reduces operational overhead while simultaneously boosting devel
 oper velocity and strengthening your overall security posture.\n\n__Key Ta
 keaways__\n- Security as a Platform Primitive (Secure-by-Default): Learn h
 ow platform teams can easily provision trusted\, pre-approved base images\
 , establishing a robust and uncompromising security baseline across all Ku
 bernetes environments.\n\n- Drastically Reduced DevOps Overhead: Discover 
 how to effectively automate vulnerability scanning\, strict policy enforce
 ment\, and seamless image promotion to minimize manual toil and operationa
 l complexity.\n\n__Target Audience__\nThis session is highly beneficial fo
 r:\n- DevOps Engineers and Site Reliability Engineers (SREs)\n- Platform E
 ngineering Teams\n- Release Management Teams\n- Cloud Security Practitione
 rs\n\n__Speaker Bios__\nDhanish Siddharth Bandaru: Dhanish is a Software E
 ngineer on the DevOps team at Acceldata\, specializing in container securi
 ty and platform engineering within cloud-native environments. He builds an
 d operates robust systems for vulnerability management\, CI/CD automation\
 , and infrastructure workflows. His work focuses on delivering scalable\, 
 automated solutions that simultaneously enhance developer productivity and
  ensure secure software delivery.\n\nHarshdeep Dhiman: Harsh is a Senior S
 oftware Engineer specializing in DevOps\, cloud-native systems\, and datab
 ase management. He designs and builds scalable\, automated infrastructure 
 and backend systems with a strong focus on CI/CD pipelines\, containerised
  environments\, and database performance and reliability.He works on impro
 ving system efficiency\, ensuring secure and seamless software delivery\, 
 and optimizing data workflows to support high-performance applications in 
 production environments.\n\n__Additional Resources__\nAVM Blog: https://en
 gineering.acceldata.io/from-image-chaos-to-control-how-avm-brings-automati
 on-and-lineage-to-container-security/
LAST-MODIFIED:20260725T102137Z
LOCATION:Bangalore
ORGANIZER;CN=Rootconf:MAILTO:no-reply@hasgeek.com
URL:https://hasgeek.com/rootconf/platform-engineering-meet-up-july-25/sche
 dule/from-image-chaos-to-control-introducing-avm-agentic-vulnerability-man
 agement-6cZS6Cngz2Q2UTGCMN4LU8
BEGIN:VALARM
ACTION:display
DESCRIPTION:From image chaos to control: introducing AVM (Agentic Vulnerab
 ility Management) in 5 minutes
TRIGGER:-PT5M
END:VALARM
END:VEVENT
BEGIN:VEVENT
SUMMARY:Keep calm and fail over: engineering chaos on CloudNative Postgres
  clusters
DTSTART:20260725T111000Z
DTEND:20260725T115500Z
DTSTAMP:20260726T005243Z
UID:session/9Rjz9u98E662gf9dQNUPJr@hasgeek.com
SEQUENCE:2
CATEGORIES:Talk (30 mins),Bangalore meet-ups
CREATED:20260725T102000Z
DESCRIPTION:**Session Description**  \n\nWhile Kubernetes is brilliant for
  stateless workloads\, running stateful workloads like relational database
 s on it often introduces operational fragility. Many teams reach for manag
 ed cloud databases or hand-built operators\, resulting in inconsistent pat
 terns and manual failovers. In this session\, we will share Nutanix's jour
 ney of migrating our diverse PostgreSQL footprint to a unified\, Kubernete
 s-native solution using CloudNativePG (CNPG). We will walk through the com
 plexity gap of database challenges in Kubernetes\, why we evaluated and ad
 opted CNPG over other operators\, and how its declarative state management
  and built-in replication capabilities helped us achieve seamless self-hea
 ling.\n\nTo validate this resilience\, we battle-tested CNPG using Chaos E
 ngineering with Chaos Mesh. We didn't just break things and hope for the b
 est\; we employed a hypothesis-driven approach to inject network partition
 s\, pod failures\, and resource exhaustion into our primary and replica no
 des. We'll share our methodology for modeling SLAs\, our chaos experiments
  (such as primary network partitioning and kube-apiserver disruptions)\, a
 nd how these tests proved CNPG's robustness. Attendees will see how we ens
 ured fast\, clean failovers and prioritized data correctness over availabi
 lity during split-brain scenarios\, ultimately managing over 500 applicati
 ons across 30+ clusters with confidence.\n\n**Takeaways**\n\n- **Operation
 al Efficiency & Resilience:** Discover how to efficiently run Postgres wor
 kloads in cloud-native environments by leveraging a declarative operator f
 or automated reconciliation\, day-2 operations\, and self-healing.\n- **Va
 lidating High Availability with Chaos Engineering:** Learn how to implemen
 t a hypothesis-driven chaos testing methodology to prove built-in failover
  mechanisms\, validate RPO/RTO thresholds\, and ensure data consistency un
 der infrastructure stress.\n\n**Target Audience**  \n\nThis session is hig
 hly beneficial for Platform Engineers\, Site Reliability Engineers (SREs)\
 , DevOps practitioners\, and Database Administrators who are looking to ru
 n or manage highly available stateful workloads (specifically PostgreSQL) 
 on Kubernetes at scale.\n\n**Speaker Bios**\n\n**Vishal Sharma** is a Memb
 er of Technical Staff at Nutanix focused on distributed systems and cloud-
 native infrastructure. Over the past three years\, he has tackled complex 
 platform engineering challenges\, including scaling observability pipeline
 s with OpenTelemetry and building robust developer tooling and platform au
 tomation. He is currently working to ensure high availability for stateful
  workloads on Kubernetes using CloudNativePG.\n\n**Rutwij Nerkar** is a Me
 mber of Technical Staff at Nutanix and has over 4 years of experience work
 ing with distributed systems and platforms at scale. With previous experie
 nce of building systems at scale in product companies\, he's currently tac
 kling platform engineering challenges at Nutanix and working on building a
  robust Postgres experience using CloudnativePG.
LAST-MODIFIED:20260725T102035Z
LOCATION:Bangalore
ORGANIZER;CN=Rootconf:MAILTO:no-reply@hasgeek.com
URL:https://hasgeek.com/rootconf/platform-engineering-meet-up-july-25/sche
 dule/keep-calm-and-fail-over-engineering-chaos-on-cloudnative-postgres-clu
 sters-9Rjz9u98E662gf9dQNUPJr
BEGIN:VALARM
ACTION:display
DESCRIPTION:Keep calm and fail over: engineering chaos on CloudNative Post
 gres clusters in 5 minutes
TRIGGER:-PT5M
END:VALARM
END:VEVENT
END:VCALENDAR
