kubernetes

Author	SHA1	Message	Date
Kubernetes Prow Robot	c7cc7886e2	Merge pull request #116702 from vinaykul/restart-free-pod-vertical-scaling-podmutation-fix Fix pod object update that may cause data race	2023-03-21 19:26:36 -07:00
vinay kulkarni	f41702b8d2	Return updatedPod if resize upon successful checkpointing of allocated resources	2023-03-22 00:24:00 +00:00
vinay kulkarni	d753893260	Do not modify original pod object when processing pod resource resize	2023-03-18 17:57:25 +00:00
vinay kulkarni	358474b71d	Explicitly return from checkpoint update failures. SyncPod will retry	2023-03-17 18:00:04 +00:00
vinay kulkarni	f66e8848ee	Fix pod object update that may cause data race	2023-03-17 08:50:52 +00:00
Paco Xu	7afcfe1826	kubelet: use filepath.Clean before init, validate it in setupDataDirs	2023-03-17 15:45:39 +08:00
Michal Wozniak	3d68f362c3	Give terminal phase correctly to all pods that will not be restarted	2023-03-16 21:25:29 +01:00
Kubernetes Prow Robot	28fa3cbbf1	Merge pull request #115847 from moshe010/pod-resource-api-dra-upstream Extend the PodResources API to include resources allocated by DRA	2023-03-14 14:12:26 -07:00
Kubernetes Prow Robot	89a9c0c8bb	Merge pull request #96120 from LorbusChris/kubelet-journal-logs KEP 2258: add node log query	2023-03-14 14:12:14 -07:00
Kubernetes Prow Robot	6a111bebe2	Merge pull request #116377 from kinvolk/rata/userns KEP-127: user namespace support for stateless pods	2023-03-14 10:40:43 -07:00
Moshe Levi	2a568bcfc8	kubelet podresources: extend List to support Dynamic Resources and implement Get API Signed-off-by: Moshe Levi <moshele@nvidia.com>	2023-03-14 19:33:04 +02:00
Francesco Romani	5e03998991	kubelet: podresources: pack parameters in a struct To enable rate limiting, needed for GA graduation, we need to pass more parameters to the already crowded `ListenAndServePodresources` function. To tidy up a bit, pack the parameters in a helper struct, with no intended changes in behavior. Signed-off-by: Francesco Romani <fromani@redhat.com>	2023-03-14 19:33:01 +02:00
Aravindh Puthiyaparambil	d12696c20f	kubelet: Expose simple journald and Get-WinEvent shims on the logs endpoint Provide an administrator a streaming view of journal logs on Linux systems using journalctl, and event logs on Windows systems using the Get-WinEvent PowerShell cmdlet without them having to implement a client side reader. Only available to cluster admins. The implementation for journald on Linux was originally done by Clayton Coleman. Introduce a heuristics approach to query logs The logs query for node objects will follow a heuristics approach when asked to query for logs from a service. If asked to get the logs from a service foobar, it will first check if foobar logs to the native OS service log provider. If unable to get logs from these, it will attempt to get logs from /var/foobar, /var/log/foobar.log or /var/log/foobar/foobar.log in that order. The logs sub-command can also directly serve a file if the query looks like a file. Co-authored-by: Clayton Coleman <ccoleman@redhat.com> Co-authored-by: Christian Glombek <cglombek@redhat.com>	2023-03-14 08:54:36 -07:00
Kubernetes Prow Robot	204a9a1f17	Merge pull request #116459 from ffromani/podresources-ratelimit-minimal add podresources DOS prevention using rate limit	2023-03-14 08:36:45 -07:00
Kubernetes Prow Robot	c8f001d798	Merge pull request #114504 from vrutkovs/tracing-kubelet-toplevel kubelet: create top-level traces for pod sync and GC	2023-03-14 03:12:16 -07:00
Rodrigo Campos	ec0410a266	kubelet: Move userns manager to its own package To that end, we need to add one kubelet getter listPodsFromDisk(). Other than that, it is a pretty trivial move. Signed-off-by: Rodrigo Campos <rodrigoca@microsoft.com>	2023-03-13 22:28:04 +01:00
Vadim Rutkovsky	556d774945	kubelet: create top-level traces for pod sync and GC This starts new top level OpenTelemetry spans every time syncPod or image / container GC is invoked	2023-03-11 10:42:14 +01:00
vinay kulkarni	01b96e7704	Rename ContainerStatus.ResourcesAllocated to ContainerStatus.AllocatedResources	2023-03-10 14:49:26 +00:00
Francesco Romani	09517c27c4	kubelet: podresources: pack parameters in a struct To enable rate limiting, needed for GA graduation, we need to pass more parameters to the already crowded `ListenAndServePodresources` function. To tidy up a bit, pack the parameters in a helper struct, with no intended changes in behavior. Signed-off-by: Francesco Romani <fromani@redhat.com>	2023-03-10 10:28:52 +01:00
Kubernetes Prow Robot	e57d968323	Merge pull request #116015 from SataQiu/clean-kubelet-20230223 kubelet: remove the deprecated --master-service-namespace flag	2023-03-09 22:43:34 -08:00
Kubernetes Prow Robot	45b96eae98	Merge pull request #113145 from smarterclayton/zombie_terminating_pods kubelet: Force deleted pods can fail to move out of terminating	2023-03-09 15:32:30 -08:00
Clayton Coleman	6b9a381185	kubelet: Force deleted pods can fail to move out of terminating If a CRI error occurs during the terminating phase after a pod is force deleted (API or static) then the housekeeping loop will not deliver updates to the pod worker which prevents the pod's state machine from progressing. The pod will remain in the terminating phase but no further attempts to terminate or cleanup will occur until the kubelet is restarted. The pod worker now maintains a store of the pods state that it is attempting to reconcile and uses that to resync unknown pods when SyncKnownPods() is invoked, so that failures in sync methods for unknown pods no longer hang forever. The pod worker's store tracks desired updates and the last update applied on podSyncStatuses. Each goroutine now synchronizes to acquire the next work item, context, and whether the pod can start. This synchronization moves the pending update to the stored last update, which will ensure third parties accessing pod worker state don't see updates before the pod worker begins synchronizing them. As a consequence, the update channel becomes a simple notifier (struct{}) so that SyncKnownPods can coordinate with the pod worker to create a synthetic pending update for unknown pods (i.e. no one besides the pod worker has data about those pods). Otherwise the pending update info would be hidden inside the channel. In order to properly track pending updates, we have to be very careful not to mix RunningPods (which are calculated from the container runtime and are missing all spec info) and config- sourced pods. Update the pod worker to avoid using ToAPIPod() and instead require the pod worker to directly use update.Options.Pod or update.Options.RunningPod for the correct methods. Add a new SyncTerminatingRuntimePod to prevent accidental invocations of runtime only pod data. Finally, fix SyncKnownPods to replay the last valid update for undesired pods which drives the pod state machine towards termination, and alter HandlePodCleanups to: - terminate runtime pods that aren't known to the pod worker - launch admitted pods that aren't known to the pod worker Any started pods receive a replay until they reach the finished state, and then are removed from the pod worker. When a desired pod is detected as not being in the worker, the usual cause is that the pod was deleted and recreated with the same UID (almost always a static pod since API UID reuse is statistically unlikely). This simplifies the previous restartable pod support. We are careful to filter for active pods (those not already terminal or those which have been previously rejected by admission). We also force a refresh of the runtime cache to ensure we don't see an older version of the state. Future changes will allow other components that need to view the pod worker's actual state (not the desired state the podManager represents) to retrieve that info from the pod worker. Several bugs in pod lifecycle have been undetectable at runtime because the kubelet does not clearly describe the number of pods in use. To better report, add the following metrics: kubelet_desired_pods: Pods the pod manager sees kubelet_active_pods: "Admitted" pods that gate new pods kubelet_mirror_pods: Mirror pods the kubelet is tracking kubelet_working_pods: Breakdown of pods from the last sync in each phase, orphaned state, and static or not kubelet_restarted_pods_total: A counter for pods that saw a CREATE before the previous pod with the same UID was finished kubelet_orphaned_runtime_pods_total: A counter for pods detected at runtime that were not known to the kubelet. Will be populated at Kubelet startup and should never be incremented after. Add a metric check to our e2e tests that verifies the values are captured correctly during a serial test, and then verify them in detail in unit tests. Adds 23 series to the kubelet /metrics endpoint.	2023-03-08 22:03:51 -06:00
vinay kulkarni	b0dce923f1	Add Get interfaces for container's checkpointed ResourcesAllocated and Resize values, remove error logging for valid standalone kubelet scenario	2023-03-06 09:50:12 +00:00
vinay kulkarni	12435b26fc	Fix nil pointer access panic in kubelet from uninitialized pod allocation checkpoint manager in standalone kubelet scenario	2023-03-04 08:07:40 +00:00
Patrick Ohly	dad95e1be6	update lease controller Passing in a context instead of a stop channel has several advantages: - ensures that client-go calls return as soon as the controller is asked to stop - contextual logging can be used By passing that context down to its own functions and checking it while waiting, the lease controller also doesn't get stuck in backoffEnsureLease anymore (https://github.com/kubernetes/kubernetes/issues/116196).	2023-03-02 15:06:00 +01:00
ruiwen-zhao	572e6e0ffb	Add MaxParallelImagePulls support Signed-off-by: ruiwen-zhao <ruiwen@google.com>	2023-03-02 03:57:59 +00:00
SataQiu	91089ce65b	kubelet: remove the deprecated --master-service-namespace flag	2023-03-01 18:44:59 +08:00
Chen Wang	7db339dba2	This commit contains the following: 1. Scheduler bug-fix + scheduler-focussed E2E tests 2. Add cgroup v2 support for in-place pod resize 3. Enable full E2E pod resize test for containerd>=1.6.9 and EventedPLEG related changes. Co-Authored-By: Vinay Kulkarni <vskibum@gmail.com>	2023-02-24 18:21:21 +00:00
Vinay Kulkarni	f2bd94a0de	In-place Pod Vertical Scaling - core implementation 1. Core Kubelet changes to implement In-place Pod Vertical Scaling. 2. E2E tests for In-place Pod Vertical Scaling. 3. Refactor kubelet code and add missing tests (Derek's kubelet review) 4. Add a new hash over container fields without Resources field to allow feature gate toggling without restarting containers not using the feature. 5. Fix corner-case where resize A->B->A gets ignored 6. Add cgroup v2 support to pod resize E2E test. KEP: /enhancements/keps/sig-node/1287-in-place-update-pod-resources Co-authored-by: Chen Wang <Chen.Wang1@ibm.com>	2023-02-24 18:21:21 +00:00
Ed Bartosh	4f88332ab4	kubelet: prepare DRA resources before CNI setup	2023-02-06 20:40:11 +02:00
Kubernetes Prow Robot	b532f2b3e7	Merge pull request #112136 from pacoxu/migrate-runtime-endpoint-flags kubelet: migrate container runtime endpoint flag to config	2023-01-03 09:29:31 -08:00
Jordan Liggitt	78cb3862f1	Fix indentation/spacing in comments to render correctly in godoc	2022-12-17 23:27:38 -05:00
Paco Xu	f28f40e521	remove a flag check that was introduced in #112542 ; address several comments Signed-off-by: Paco Xu <paco.xu@daocloud.io>	2022-12-13 14:00:29 +08:00
Aditi Sharma	214a0ee7b8	Migrate container runtime endpoint flag to config Signed-off-by: Aditi Sharma <adi.sky17@gmail.com> Signed-off-by: Paco Xu <paco.xu@daocloud.io>	2022-12-13 14:00:29 +08:00
Kubernetes Prow Robot	a668924cb6	Merge pull request #113255 from claudiubelu/path-filepath-update-kubelet Replaces path.Operation with filepath.Operation (kubelet)	2022-12-09 22:27:41 -08:00
arrowfeng	6a57404e28	kubelet: cleanup secretManager and configManager in podManager Signed-off-by: arrowfeng <289716347@qq.com>	2022-11-14 23:05:32 +08:00
Ed Bartosh	ae0f38437c	kubelet: add support for dynamic resource allocation Dependencies need to be updated to use github.com/container-orchestrated-devices/container-device-interface. It's not decided yet whether we will implement Topology support for DRA or not. Not having any toppology-related code will help to avoid wrong impression that DRA is used as a hint provider for the Topology Manager.	2022-11-11 21:58:03 +01:00
Jingyuan Liang	9f5c5b82a9	kubelet: Keep trying fast status update at startup until node is ready	2022-11-09 15:55:20 +00:00
Kubernetes Prow Robot	70263d55b2	Merge pull request #113501 from pacoxu/fix-startReflector kubelet: fix nil pointer in startReflector for standalone mode	2022-11-09 03:50:12 -08:00
Peter Hunt	6298ce68e2	kubelet: wire ListPodSandboxMetrics Signed-off-by: Peter Hunt <pehunt@redhat.com>	2022-11-08 14:47:08 -05:00
Claudiu Belu	b9bf3e5c49	Replaces path.Operation with filepath.Operation (kubelet) The path module has a few different functions: Clean, Split, Join, Ext, Dir, Base, IsAbs. These functions do not take into account the OS-specific path separator, meaning that they won't behave as intended on Windows. For example, Dir is supposed to return all but the last element of the path. For the path "C:\some\dir\somewhere", it is supposed to return "C:\some\dir\", however, it returns ".". Instead of these functions, the ones in filepath should be used instead.	2022-11-08 16:05:48 +00:00
Harshal Patil	86284d42f8	Add support for Evented PLEG Signed-off-by: Harshal Patil <harpatil@redhat.com> Co-authored-by: Swarup Ghosh <swghosh@redhat.com>	2022-11-08 20:06:16 +05:30
David Ashpole	64af1adace	Second attempt: Plumb context to Kubelet CRI calls (#113591 ) * plumb context from CRI calls through kubelet * clean up extra timeouts * try fixing incorrectly cancelled context	2022-11-05 06:02:13 -07:00
Kubernetes Prow Robot	c8a3657bde	Merge pull request #113307 from andrewsykim/apiserver-identity-hostname apiserver identity: use persistent names for lease objects	2022-11-04 07:28:25 -07:00
Kubernetes Prow Robot	1bf4af4584	Merge pull request #111930 from azylinski/new-histogram-pod_start_sli_duration_seconds New histogram: Pod start SLI duration	2022-11-04 07:28:14 -07:00
Paco Xu	89e4836dde	add ut for kubelet standalone mode	2022-11-04 18:17:51 +08:00
Andrew Sy Kim	72f2e1cc0d	lease controller: update NewController to accept leaseName as a parameter, remove NewControllerWithLeaseName Signed-off-by: Andrew Sy Kim <andrewsy@google.com>	2022-11-04 00:44:13 -04:00
Paco Xu	57a3af1f87	kubelet: don't set secret and configmap manager if running in standalone mode	2022-11-03 17:46:52 +08:00
Kubernetes Prow Robot	98742f9d77	Merge pull request #110747 from harshanarayana/cleanup/GIT-110737/logging-improvements structured-logging: replace KObjs with KObjSlice for logging	2022-11-03 00:49:34 -07:00
Antonio Ojea	9c2b333925	Revert "plumb context from CRI calls through kubelet" This reverts commit `f43b4f1b95`.	2022-11-02 13:37:23 +00:00

1 2 3 4 5 ...

2201 Commits