DevOps&SRE Library
19.7K subscribers
431 photos
2 videos
2 files
5.41K links
Библиотека статей по теме DevOps и SRE.

Реклама: @ostinostin
Контент: @mxssl

РКН: https://www.gosuslugi.ru/snet/67704b536aa9672b963777b3
Download Telegram
Kubernetes v1.35: Extended Toleration Operators to Support Numeric Comparisons (Alpha)

In Kubernetes v1.35, we're introducing Extended Toleration Operators as an alpha feature. This enhancement adds Gt (Greater Than) and Lt (Less Than) operators to spec.tolerations, enabling threshold-based scheduling decisions that unlock new possibilities for SLA-based placement, cost optimization, and performance-aware workload distribution.


https://kubernetes.io/blog/2026/01/05/kubernetes-v1-35-numeric-toleration-operators
📹Вебинар: Выбор между Serverless и Kubernetes для AI-ворклоадов: как определить оптимальную платформу под задачу

На открытом уроке рассмотрим:
- В чем различаются Serverless-подходы и Kubernetes при работе с AI-ворклоадами;
- Какие преимущества и ограничения есть у каждого подхода с точки зрения масштабируемости, стоимости и сложности эксплуатации;
- Какие трейдоффы нужно учитывать при выборе платформы: холодный старт, управление состоянием, поддержка GPU;
- Как обосновывать выбор архитектуры для разных AI-сценариев на практическом воркшопе.

После занятия вы будете знать:
- Как сравнивать Serverless и Kubernetes для различных AI-задач;
- Как выбирать платформу оркестрации в зависимости от требований к нагрузке, бюджету и архитектуре решения;
- Как учитывать ключевые технические ограничения при проектировании AI-инфраструктуры;
- Как аргументированно обосновывать выбор платформы для задач масштабирования, потоковой обработки данных и построения гибридных сред.

⚠️ Открытый урок проходит в преддверии старта курса «ИИ-архитектор».

👉 Для участия зарегистрируйтесь: https://vk.cc/cZywHe

Реклама. ООО «Отус онлайн-образование», ОГРН 1177746618576, www.otus.ru, erid: 2VtzqxE1hEa
Migrating from Bitnami PostgreSQL to CloudNative-PG on Kubernetes

If you're running PostgreSQL on Kubernetes, chances are you've used Bitnami's popular Helm charts. They've been a go-to for many, but a significant change is on the horizon. As outlined in this GitHub issue, Bitnami is moving its production-ready charts and images to a commercial offering. For those of us who rely on and advocate for open-source solutions, this means it's time to find a robust alternative.


https://k8scockpit.tech/posts/cloudnative-pg
Ingress-nginx уходит в прошлое. С марта 2026 поддержка прекратилась. А что вместо него? Предлагаем посмотреть на Gateway API, новый стандарт Kubernetes SIG.

23 июля в 17:00 старший SRE-инженер MWS Cloud Евгений Макеев на практике покажет:

чем маршрутизация Gateway API отличается от Ingress
как выбрать контроллер
как установить Gateway API в Managed Kubernetes
как настроить безопасное подключение через TLS-сертификат

Будет полезно для DevOps, платформенным инженерам и разработчикам.

Регистрируйтесь по ссылке
Please open Telegram to view this post
VIEW IN TELEGRAM
Case Study: Reducing Complexity By Migrating from K8S to ECS Fargate for NetworkLessons

The Kubernetes story is one I hear often. Teams adopt K8s expecting operational simplicity, only to discover they've traded application complexity for infrastructure complexity. For a solo founder focused on content creation, maintaining a Kubernetes cluster was simply the wrong trade-off.


https://dev.to/aws-builders/case-study-reducing-complexity-by-migrating-from-k8s-to-ecs-fargate-for-networklessons-3271
Database State Management in Kubernetes: Running SQL Server on AKS with GitOps

This article is about the patterns that actually work when you're managing real databases in Kubernetes, not hello-world demos: when your boss says "everything should be containerized" but your databases laugh in the face of ephemeral pods, and when GitOps meets a 200GB production database that absolutely cannot lose a single transaction.


https://medium.com/@firaassboui/database-state-management-in-kubernetes-running-sql-server-on-aks-with-gitops-69286a87f8de
Deploy LLM Models on OpenShift

Operators make life easier, but they are not always an option. In this post, I'll walk through a practical way to deploy large language models on OpenShift without relying on the OpenShift AI or NVIDIA operators. The approach uses llama.cpp as a lightweight runtime engine and runs a quantized GGUF model to enable efficient inference with minimal dependencies.


https://medium.com/@ahmeddraz/deploy-llm-models-on-openshift-84ecb014f09a
Enforcing Signed Container Images in Kubernetes Using Cosign & Kyverno (Helm-based Setup)

This article explains how Cosign, Kyverno, and Harbor can work together to enforce image signature verification in Kubernetes. The approach is well-suited for enterprise environments with private registries, custom TLS certificates, and restricted access to public transparency logs.


https://medium.com/@hansakabiyon99/enforcing-signed-container-images-in-kubernetes-using-cosign-kyverno-helm-based-setup-646209ecb8ce
Modernizing Jenkins: From Static Agents to Kubernetes Dynamic Pods

Jenkins might feel like legacy tech, but it's still powering CI/CD at thousands of companies. Instead of ripping it out, here's how to modernize it with Kubernetes — solving cost tracking, resource isolation, and scalability problems along the way.


https://blog.stackademic.com/modernizing-jenkins-from-static-agents-to-kubernetes-dynamic-pods-fbda3f897018
Building a Local Data Platform with Kubernetes and Terraform

This blog aims to provide a practical example of how core tools like Kubernetes, Terraform, and DevContainers can be combined to build a local data platform in a structured and maintainable way.


https://blog.dataengineerthings.org/building-a-local-data-platform-with-kubernetes-and-terraform-9547a4256a7f
netfence

Netfence runs as a daemon on your VM/container hosts and automatically injects eBPF filter programs into cgroups and network interfaces, with a built-in DNS server that resolves allowed domains and populates the IP allowlist.


https://github.com/danthegoodman1/netfence
endpoint-monitoring-operator

A lightweight, extensible Kubernetes Operator that probes any endpoint—HTTP/JSON, TCP, DNS, ICMP, Trino, OpenSearch, and more—and routes alerts to Slack or e-mail with a simple Custom Resource.


https://github.com/iam404/endpoint-monitoring-operator
argocd-diff-preview

Argo CD Diff Preview is a tool that renders the diff between two branches in a Git repository. It is designed to render manifests generated by Argo CD, providing a clear and concise view of the changes between two branches. It operates similarly to Atlantis for Terraform, creating a plan that outlines the proposed changes.


https://github.com/dag-andersen/argocd-diff-preview
From Push to Production: Our Deployment Pipeline with Argo CD

So we built a pipeline that’s intentionally “boring”: GitHub pull requests, GitHub Actions, Kubernetes, and Argo CD — split across staging and production.

This post is a walkthrough of how a feature goes from a developer’s machine to real users, and why we made the choices we did.


https://medium.com/openmirai/from-push-to-production-our-deployment-pipeline-with-argo-cd-00e55b3feee9
From Minutes to Seconds: How I Eliminated Kubernetes Image Pull Delays

How I reduced pod startup times from minutes to seconds with intelligent image preloading


https://medium.com/@yyadid7/from-minutes-to-seconds-how-i-eliminated-kubernetes-image-pull-delays-16f166327576
Nomad on OpenShift: The case for the control plane

If Red Hat trusts OpenShift to run the control plane for their largest infrastructure orchestrator, the same pattern should apply to your smallest.


https://hashicorpengineering.substack.com/p/nomad-on-openshift-the-control-plane
Deep Dive: How linkerd-destination works in the Linkerd Service Mesh

Recently, in our daily operations, we took a deep dive into the inner workings of linkerd-destination, one of the most critical components of the Linkerd control plane.


https://medium.com/@bezarsnba/deep-dive-the-linkerd-destination-service-en-19f6efd1b308
🔍Тестовое собеседование с Head of DevOps уже завтра

28 июля(уже завтра!) в 19:00 по мск приходи онлайн на открытое собеседование, чтобы посмотреть на настоящее интервью на Middle DevOps-разработчика.

Как это будет:
📂 Александр Хренников, Head of DevOps в KTS с опытом 14+ лет, будет задавать реальные вопросы и задачи разработчику-добровольцу
📂 Александр будет комментировать каждый ответ респондента, чтобы дать понять, чего от вас ожидает собеседующий на интервью
📂 В конце можно будет задать любой вопрос Александру

Это бесплатно. Эфир проходит в рамках менторской программы от ШОРТКАТ для DevOps-разработчиков, которые хотят повысить свой грейд, ЗП и прокачать скиллы.

Переходи в нашего бота, чтобы получить ссылку на эфир → @shortcut_devops_bot

Реклама.
О рекламодателе.
Please open Telegram to view this post
VIEW IN TELEGRAM
Uniform API server access using clientcmd

If you've ever wanted to develop a command line client for a Kubernetes API, especially if you've considered making your client usable as a kubectl plugin, you might have wondered how to make your client feel familiar to users of kubectl. In fact, the Kubernetes project provides two libraries to help you handle kubectl-style command line arguments in Go programs: clientcmd and cli-runtime (which uses clientcmd). This article will show how to use the former.


https://kubernetes.io/blog/2026/01/19/clientcmd-apiserver-access
CloudNativePG (CNPG) - install (2.18) and first test: simulate transient failure

I'm starting a series of blog posts to explore CloudNativePG (CNPG), a Kubernetes operator for PostgreSQL that automates high availability in containerized environments.


https://dev.to/franckpachot/cloudnativepg-install-218-and-first-test-transient-failure-4ml