TGViewer
Channel Public Channel
DevOps & SRE notes

DevOps & SRE notes

@devops_sre_notes

Helpful articles and tools for DevOps&SRE

youtube: https://www.youtube.com/@DevOpsAndSRENotes
Subscribers
13.3K
Photos
57
Videos
1
Links
2.6K

Showing posts older than #2245 · Back to latest

Older Posts 20 shown
Post #2240 13.7K
Migrating from MetalLB to Cilium streamlines Kubernetes networking by consolidating load balancer, IP address management, and network advertisement features into a single tool. This article details how Cilium—starting with version 1.13—natively supports LoadBalancer IP management, BGP (Layer 3) announcements, and Layer 2 (ARP) announcements, eliminating the need for MetalLB in most self-managed clusters. Through practical YAML examples, it demonstrates configuring Cilium IP pools, service selectors, specific IP assignments, and both IPv4 and IPv6 support, as well as advertising service IPs to the network using BGP or ARP, offering a more integrated and simplified approach to Kubernetes networking.

https://isovalent.com/blog/post/migrating-from-metallb-to-cilium/
Isovalent Migrating from MetalLB to Cilium In this blog post, you will learn how to migrate from MetalLB to Cilium for local service advertisement over Layer 2.
  • 👍 5
  • ❤ 2
Post #2237 13.4K
This newsletter explains the challenges of the "hot shard" problem—when a disproportionate amount of traffic targets a single shard, causing resource saturation and degraded performance. The blogpost outlines practical strategies to address this, such as vertical scaling, adding read replicas or caches, distributing hot keys across more shards, choosing better sharding keys and algorithms, implementing load balancing and queueing, controlling traffic with backpressure, and monitoring the cluster for early detection of issues.

https://newsletter.scalablethread.com/p/how-to-handle-hot-shard-problem
Scalablethread How to Handle Hot Shard Problem? Understanding Different Approaches to Address Hot Key/Partition Problem
  • 👍 3
  • ❤ 1
Post #2236 14.3K
Figma’s migration onto Kubernetes is a compelling case study in how a high-growth company can modernize its infrastructure for scalability, reliability, and developer productivity. This article recounts Figma’s decision to move from AWS ECS to Kubernetes (EKS), the challenges they faced with ECS—such as lack of support for StatefulSets, Helm charts, and advanced autoscaling—and the benefits they unlocked by embracing the broader CNCF ecosystem and Kubernetes’ popularity within the industry.

https://www.figma.com/blog/migrating-onto-kubernetes/
Figma How We Migrated onto K8s in Less Than 12 months | Figma Blog Migrating onto Kubernetes can take years. Here’s why we decided it was worth undertaking, and how we moved a majority of our core services.
  • 👍 1
Post #2231 14K
Understanding logical replication in PostgreSQL is crucial for anyone managing data across multiple Postgres instances. This blogpost from EnterpriseDB introduces the basics of logical replication, explaining how it enables selective data replication—such as inserts, updates, and deletes—between databases, even across different Postgres versions, and outlines the practical steps to set up publications and subscriptions for real-time data synchronization.

https://www.enterprisedb.com/blog/logical-replication-postgres-basics
EDB Logical replication in Postgres: Basics In this post we'll explore the basics of logical replication between two Postgres databases as both a user and a developer. Postgres first implemented physical replication where it shipped bytes on disk from one database A to another database B. Database…
  • ❤ 1
  • 👍 1
Post #2230 15.6K
Efficient, disruption-free application updates are essential for modern cloud-native operations. This article on Semaphore explains how Kubernetes’ rolling update deployment strategy enables teams to maintain service continuity while incrementally rolling out new versions.

https://semaphore.io/blog/kubernetes-rolling-update-deployment
Semaphore Kubernetes Deployments: A Guide to the Rolling Update Deployment Strategy - Semaphore The article elaborates on Kubernetes' rolling update deployment strategy, emphasizing incremental changes, adjustable speed, and pause/resume options.
  • ❤ 2
Post #2227 14.3K
This analysis explores how DeepSeek has reimagined the Transformer architecture to achieve greater efficiency and performance in large language models. The piece highlights innovations like Multi-Head Latent Attention and advanced Mixture-of-Experts routing that set DeepSeek apart from conventional approaches.

https://epoch.ai/gradient-updates/how-has-deepseek-improved-the-transformer-architecture
Epoch AI How has DeepSeek improved the Transformer architecture? This Gradient Updates issue goes over the major changes that went into DeepSeek's most recent model.
  • ❤ 5
Post #2226 15K
Railway’s latest narrative details their transition from relying on Google Cloud Platform to building their own physical infrastructure, highlighting the challenges and lessons learned in constructing a custom data center cage. This entry offers a behind-the-scenes look at selecting colocation options, managing power and cooling, and orchestrating the intricate cabling and network setup required for a resilient, high-performance platform.

https://blog.railway.com/p/data-center-build-part-one
Railway Blog So You Want to Build Your Own Data Center When it comes to infrastructure engineering, building a data center is probably closer to building a house than to deploying a Terraform stack.
  • 👍 3
Post #2223 14.5K
Kubernetes network policies are essential for controlling how traffic flows between pods, namespaces, and external endpoints in your cluster, helping you enforce security and compliance requirements. This guide by Scott Rigby explains the differences between Layer 4 (L4) and Layer 7 (L7) policies, their pros and cons, and how combining both approaches—using tools like Linkerd—can help you achieve a robust, zero-trust security model tailored to modern cloud-native environments.

https://www.buoyant.io/blog/a-guide-to-modern-kubernetes-network-policies
www.buoyant.io A guide to modern Kubernetes network policies In the world of Kubernetes, network policies are essential for controlling traffic within your cluster. But what are they really? And why, when and how should you implement them?
  • 👍 2
  • ❤ 1
Older posts →
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →