podman

Commit Graph

Author	SHA1	Message	Date
Daniel J Walsh	81eb71fe36	Fix permissions on initially created named volumes Permission of volume should match the directory it is being mounted on. Fixes: https://github.com/containers/podman/issues/10188 Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2021-06-14 11:56:48 -04:00
Adrian Reber	240bbc3bfa	Fix pre-checkpointing Unfortunately --pre-checkpointing never worked as intended and recent changes to runc have shown that it is broken. To create a pre-checkpoint CRIU expects the paths between the pre-checkpoints to be a relative path. If having a previous checkpoint it needs the be referenced like this: --prev-images-dir ../parent Unfortunately Podman was giving runc (and CRIU) an absolute path. Unfortunately, again, until March 2021 CRIU silently ignored if the path was not relative and switch back to normal checkpointing. This has been now fixed in CRIU and runc and running pre-checkpoint with the latest runc fails, because runc already sees that the path is absolute and returns an error. This commit fixes this by giving runc a relative path. This commit also fixes a second pre-checkpointing error which was just recently introduced. So summarizing: pre-checkpointing never worked correctly because CRIU ignored wrong parameters and recent changes broke it even more. Now both errors should be fixed. [NO TESTS NEEDED] Signed-off-by: Adrian Reber <areber@redhat.com> Signed-off-by: Adrian Reber <adrian@lisas.de>	2021-06-10 15:29:24 +02:00
Paul Holzinger	18fa124dfc	Improve systemd-resolved detection When 127.0.0.53 is the only nameserver in /etc/resolv.conf assume systemd-resolved is used. This is better because /etc/resolv.conf does not have to be symlinked to /run/systemd/resolve/stub-resolv.conf in order to use systemd-resolved. [NO TESTS NEEDED] Fixes: #10570 Signed-off-by: Paul Holzinger <pholzing@redhat.com>	2021-06-08 18:14:00 +02:00
Adrian Reber	8aa5340ade	Add parameter to specify checkpoint archive compression The checkpoint archive compression was hardcoded to `archive.Gzip`. There have been requests to make the used compression algorithm selectable. There was especially the request to not compress the checkpoint archive to be able to create faster checkpoints when not compressing it. This also changes the default from `gzip` to `zstd`. This change should not break anything as the restore code path automatically handles whatever compression the user provides during restore. Signed-off-by: Adrian Reber <areber@redhat.com>	2021-06-07 08:07:15 +02:00
Paul Holzinger	df2e7e00fc	add ipv6 nameservers only when the container has ipv6 enabled The containers /etc/resolv.conf allways preserved the ipv6 nameserves from the host even when the container did not supported ipv6. Check if the cni result contains an ipv6 address or slirp4netns has ipv6 support enabled and only add the ipv6 nameservers when this is the case. The test needs to have an ipv6 nameserver in the hosts /etc/hosts but we should never mess with this file on the host. Therefore the test is skipped when no ipv6 is detected. Fixes #10158 Signed-off-by: Paul Holzinger <pholzing@redhat.com>	2021-06-03 10:19:36 +02:00
OpenShift Merge Robot	a7fa0da4a5	Merge pull request #10334 from mheon/add_relabel_vol_plugin Ensure that :Z/:z/:U can be used with named volumes	2021-05-17 16:28:21 -04:00
OpenShift Merge Robot	9a9118b831	Merge pull request #10366 from ashley-cui/secretoptions Support uid,gid,mode options for secrets	2021-05-17 16:24:20 -04:00
Ashley Cui	cf30f160ad	Support uid,gid,mode options for secrets Support UID, GID, Mode options for mount type secrets. Also, change default secret permissions to 444 so all users can read secret. Signed-off-by: Ashley Cui <acui@redhat.com>	2021-05-17 14:35:55 -04:00
Baron Lenardson	c8dfcce6db	Add host.containers.internal entry into container's etc/hosts This change adds the entry `host.containers.internal` to the `/etc/hosts` file within a new containers filesystem. The ip address is determined by the containers networking configuration and points to the gateway address for the containers networking namespace. Closes #5651 Signed-off-by: Baron Lenardson <lenardson.baron@gmail.com>	2021-05-17 08:21:22 -05:00
Matthew Heon	6efca0bbac	Ensure that :Z/:z/:U can be used with named volumes Docker allows relabeling of any volume passed in via -v, even including named volumes. This normally isn't an issue at all, given named volumes get the right label for container access automatically, but this becomes an issue when volume plugins are involved - these aren't managed by Podman, and may well be unaware of SELinux labelling. We could automatically relabel these volumes on creation, but I'm still reluctant to do that (feels like it could break things). Instead, let's allow :z and :Z to be used with named volumes, so users can explicitly request relabel of a volume plugin-backed volume. We also get :U at the same time. I don't see any real need for it but it also doesn't seem to hurt, so I didn't bother disabling it. Fixes #10273 Signed-off-by: Matthew Heon <mheon@redhat.com>	2021-05-17 09:10:59 -04:00
OpenShift Merge Robot	141ba94f97	Merge pull request #10221 from ashley-cui/envsec Add support for environment variable secrets	2021-05-07 05:34:26 -04:00
Daniel J Walsh	f528511bf6	Revert Patch to relabel if selinux not enabled Revert : https://github.com/containers/podman/pull/9895 Turns out that if Docker is in --selinux-enabeled, it still relabels if the user tells the system to, even if running a --privileged container or if the selinux separation is disabled --security-opt label=disable. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2021-05-06 18:00:16 -04:00
Ashley Cui	2634cb234f	Add support for environment variable secrets Env var secrets are env vars that are set inside the container but not commited to and image. Also support reading from env var when creating a secret. Signed-off-by: Ashley Cui <acui@redhat.com>	2021-05-06 14:00:57 -04:00
Giuseppe Scrivano	27ac750c7d	cgroup: fix rootless --cgroup-parent with pods extend to pods the existing check whether the cgroup is usable when running as rootless with cgroupfs. commit `17ce567c68` introduced the regression. Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2021-05-06 08:33:28 +02:00
Valentin Rothberg	0f7d54b026	migrate Podman to containers/common/libimage Migrate the Podman code base over to `common/libimage` which replaces `libpod/image` and a lot of glue code entirely. Note that I tried to leave bread crumbs for changed tests. Miscellaneous changes: * Some errors yield different messages which required to alter some tests. * I fixed some pre-existing issues in the code. Others were marked as `//TODO`s to prevent the PR from exploding. * The `NamesHistory` of an image is returned as is from the storage. Previously, we did some filtering which I think is undesirable. Instead we should return the data as stored in the storage. * Touched handlers use the ABI interfaces where possible. * Local image resolution: previously Podman would match "foo" on "myfoo". This behaviour has been changed and Podman will now only match on repository boundaries such that "foo" would match "my/foo" but not "myfoo". I consider the old behaviour to be a bug, at the very least an exotic corner case. * Futhermore, "foo:none" does not resolve to a local image "foo" without tag anymore. It's a hill I am (almost) willing to die on. * `image prune` prints the IDs of pruned images. Previously, in some cases, the names were printed instead. The API clearly states ID, so we should stick to it. * Compat endpoint image removal with _force_ deletes the entire not only the specified tag. Signed-off-by: Valentin Rothberg <rothberg@redhat.com>	2021-05-05 11:30:12 +02:00
Giuseppe Scrivano	17ce567c68	cgroup: always honor --cgroup-parent with cgroupfs if --cgroup-parent is specified, always honor it without doing any detection whether cgroups are supported or not. Closes: https://github.com/containers/podman/issues/10173 Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2021-05-03 12:36:03 +02:00
Sebastian Jug	db7cff8c86	Add support for CDI device configuration - Persist CDIDevices in container config - Add e2e test - Log HasDevice error and add additional condition for safety Signed-off-by: Sebastian Jug <seb@stianj.ug>	2021-04-20 09:18:52 -04:00
Giuseppe Scrivano	2fad29ccb2	cgroup: do not set cgroup parent when rootless and cgroupfs do not set the cgroup parent when running as rootless with cgroupfs, even if cgroup v2 is used. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1947999 Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2021-04-12 16:55:55 +02:00
Daniel J Walsh	6831c72f6a	Don't relabel volumes if running in a privileged container Docker does not relabel this content, and openstack is running containers in this manner. There is a penalty for doing this on each container, that is not worth taking on a disable SELinux container. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2021-04-05 13:07:36 -04:00
Paul Holzinger	54b588c07d	rootless cni without infra container Instead of creating an extra container create a network and mount namespace inside the podman user namespace. This ns is used to for rootless cni operations. This helps to align the rootless and rootful network code path. If we run as rootless we just have to set up a extra net ns and initialize slirp4netns in it. The ocicni lib will be called in that net ns. This design allows allows easier maintenance, no extra container with pause processes, support for rootless cni with --uidmap and possibly more. The biggest problem is backwards compatibility. I don't think live migration can be possible. If the user reboots or restart all cni containers everything should work as expected again. The user is left with the rootless-cni-infa container and image but this can safely be removed. To make the existing cni configs work we need execute the cni plugins in a extra mount namespace. This ensures that we can safely mount over /run and /var which have to be writeable for the cni plugins without removing access to these files by the main podman process. One caveat is that we need to keep the netns files at `XDG_RUNTIME_DIR/netns` accessible. `XDG_RUNTIME_DIR/rootless-cni/{run,var}` will be mounted to `/{run,var}`. To ensure that we keep the netns directory we bind mount this relative to the new root location, e.g. XDG_RUNTIME_DIR/rootless-cni/run/user/1000/netns before we mount the run directory. The run directory is mounted recursive, this makes the netns directory at the same path accessible as before. This also allows iptables-legacy to work because /run/xtables.lock is now writeable. Signed-off-by: Paul Holzinger <paul.holzinger@web.de>	2021-04-01 17:27:03 +02:00
なつき	a2e834d0d9	[NO TESTS NEEDED] Fix for kernel without CONFIG_USER_NS Signed-off-by: Natsuki <i@ntk.me>	2021-03-26 21:03:24 -07:00
TomSweeneyRedHat	5b2e71dc5b	Validate passed in timezone from tz option Erik Sjolund reported an issue where a badly formated file could be passed into the `--tz` option and then the date in the container would be badly messed up: ``` erik@laptop:~$ echo Hello > file.txt erik@laptop:~$ podman run --tz=../../../home/erik/file.txt --rm -ti docker.io/library/alpine cat /etc/localtime Hello erik@laptop:~$ podman --version podman version 3.0.0-rc1 erik@laptop:~$ ``` This fix checks to make sure the TZ passed in is a valid value and then proceeds with the rest of the processing. This was first reported as a potential security issue, but it was thought not to be. However, I thought closing the hole sooner rather than later would be good. Signed-off-by: TomSweeneyRedHat <tsweeney@redhat.com>	2021-03-21 17:25:35 -04:00
Valentin Rothberg	d0d084dd8c	turn hidden --trace into a NOP The --trace has helped in early stages analyze Podman code. However, it's contributing to dependency and binary bloat. The standard go tooling can also help in profiling, so let's turn `--trace` into a NOP. [NO TESTS NEEDED] Signed-off-by: Valentin Rothberg <rothberg@redhat.com>	2021-03-08 09:22:42 +01:00
Adrian Reber	91b2f07d5b	Use functions and defines from checkpointctl No functional changes. [NO TESTS NEEDED] - only moving code around Signed-off-by: Adrian Reber <areber@redhat.com>	2021-03-02 17:00:06 +00:00
Adrian Reber	bf92e21113	Move checkpoint/restore code to pkg/checkpoint/crutils To be able to reuse common checkpoint/restore functions this commit moves code to pkg/checkpoint/crutils. This commit has not functional changes. It only moves code around. [NO TESTS NEEDED] - only moving code around Signed-off-by: Adrian Reber <areber@redhat.com>	2021-03-02 17:00:06 +00:00
Paul Holzinger	90050671b7	Add dns search domains from cni response to resolv.conf This fixes slow local host name lookups. see containers/dnsname#57 Signed-off-by: Paul Holzinger <paul.holzinger@web.de>	2021-02-24 10:41:56 +01:00
Eduardo Vega	874f2327e6	Add U volume flag to chown source volumes Signed-off-by: Eduardo Vega <edvegavalerio@gmail.com>	2021-02-22 22:55:19 -06:00
Valentin Rothberg	5dded6fae7	bump go module to v3 We missed bumping the go module, so let's do it now :) * Automated go code with github.com/sirkon/go-imports-rename * Manually via `vgrep podman/v2` the rest Signed-off-by: Valentin Rothberg <rothberg@redhat.com>	2021-02-22 09:03:51 +01:00
Daniel J Walsh	5d1ec2960d	Do not reset storage when running inside of a container Currently if the host shares container storage with a container running podman, the podman inside of the container resets the storage on the host. This can cause issues on the host, as well as causes the podman command running the container, to fail to unmount /dev/shm. podman run -ti --rm --privileged -v /var/lib/containers:/var/lib/containers quay.io/podman/stable podman run alpine echo hello * unlinkat /var/lib/containers/storage/overlay-containers/a7f3c9deb0656f8de1d107e7ddff2d3c3c279c11c1635f233a0bffb16051fb2c/userdata/shm: device or resource busy * unlinkat /var/lib/containers/storage/overlay-containers/a7f3c9deb0656f8de1d107e7ddff2d3c3c279c11c1635f233a0bffb16051fb2c/userdata/shm: device or resource busy Since podman is volume mounting in the graphroot, it will add a flag to /run/.containerenv to tell podman inside of container whether to reset storage or not. Since the inner podman is running inside of the container, no reason to assume this is a fresh reboot, so if "container" environment variable is set then skip reset of storage. Also added tests to make sure /run/.containerenv is runnig correctly. Fixes: https://github.com/containers/podman/issues/9191 Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2021-02-16 14:18:53 -05:00
OpenShift Merge Robot	7fb347a3d4	Merge pull request #9399 from vrothberg/home-sweet-home do not set empty $HOME	2021-02-16 11:39:27 -05:00
Valentin Rothberg	2ec0e3b650	do not set empty $HOME Make sure to not set an empty $HOME for containers and let it default to "/". https://github.com/containers/crun/pull/599 is required to fully address #9378. Partially-Fixes: #9378 Signed-off-by: Valentin Rothberg <rothberg@redhat.com>	2021-02-16 14:21:45 +01:00
Daniel J Walsh	3d50393f09	Don't chown workdir if it already exists Currently podman is always chowning the WORKDIR to root:root This PR will return if the WORKDIR already exists. Fixes: https://github.com/containers/podman/issues/9387 Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2021-02-16 04:52:02 -05:00
Paul Holzinger	78c8a87362	Enable whitespace linter Use the whitespace linter and fix the reported problems. [NO TESTS NEEDED] Signed-off-by: Paul Holzinger <paul.holzinger@web.de>	2021-02-11 23:01:56 +01:00
Paul Holzinger	ef2fc90f2d	Enable stylecheck linter Use the stylecheck linter and fix the reported problems. [NO TESTS NEEDED] Signed-off-by: Paul Holzinger <paul.holzinger@web.de>	2021-02-11 23:01:29 +01:00
Matthew Heon	ea910fc535	Rewrite copy-up to use buildah Copier The old copy-up implementation was very unhappy with symlinks, which could cause containers to fail to start for unclear reasons when a directory we wanted to copy-up contained one. Rewrite to use the Buildah Copier, which is more recent and should be both safer and less likely to blow up over links. At the same time, fix a deadlock in copy-up for volumes requiring mounting - the Mountpoint() function tried to take the already-acquired volume lock. Fixes #6003 Signed-off-by: Matthew Heon <mheon@redhat.com>	2021-02-10 14:21:37 -05:00
Ashley Cui	832a69b0be	Implement Secrets Implement podman secret create, inspect, ls, rm Implement podman run/create --secret Secrets are blobs of data that are sensitive. Currently, the only secret driver supported is filedriver, which means creating a secret stores it in base64 unencrypted in a file. After creating a secret, a user can use the --secret flag to expose the secret inside the container at /run/secrets/[secretname] This secret will not be commited to an image on a podman commit Signed-off-by: Ashley Cui <acui@redhat.com>	2021-02-09 09:13:21 -05:00
Valentin Rothberg	821ef6486a	fix logic when not creating a workdir When resolving the workdir of a container, we may need to create unless the user set it explicitly on the command line. Otherwise, we just do a presence check. Unfortunately, there was a missing return that lead us to fall through into attempting to create and chown the workdir. That caused a regression when running on a read-only root fs. Fixes: #9230 Signed-off-by: Valentin Rothberg <rothberg@redhat.com>	2021-02-05 09:50:07 +01:00
Valentin Rothberg	0f668aa085	workdir presence checks A container's workdir can be specified via the CLI via `--workdir` and via an image config with the CLI having precedence. Since images have a tendency to specify workdirs without necessarily shipping the paths with the root FS, make sure that Podman creates the workdir. When specified via the CLI, do not create the path, but check for its existence and return a human-friendly error. NOTE: `crun` is performing a similar check that would yield exit code 127. With this change, however, Podman performs the check and yields exit code 126. Since this is specific to `crun`, I do not consider it to be a breaking change of Podman. Fixes: #9040 Signed-off-by: Valentin Rothberg <rothberg@redhat.com>	2021-01-26 09:02:21 +01:00
Giuseppe Scrivano	ef654941d1	libpod: move slirp magic IPs to consts Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2021-01-22 08:08:27 +01:00
Matthew Heon	b53cb57680	Initial implementation of volume plugins This implements support for mounting and unmounting volumes backed by volume plugins. Support for actually retrieving plugins requires a pull request to land in containers.conf and then that to be vendored, and as such is not yet ready. Given this, this code is only compile tested. However, the code for everything past retrieving the plugin has been written - there is support for creating, removing, mounting, and unmounting volumes, which should allow full functionality once the c/common PR is merged. A major change is the signature of the MountPoint function for volumes, which now, by necessity, returns an error. Named volumes managed by a plugin do not have a mountpoint we control; instead, it is managed entirely by the plugin. As such, we need to cache the path in the DB, and calls to retrieve it now need to access the DB (and may fail as such). Notably absent is support for SELinux relabelling and chowning these volumes. Given that we don't manage the mountpoint for these volumes, I am extremely reluctant to try and modify it - we could easily break the plugin trying to chown or relabel it. Also, we had no less than 5 separate implementations of inspecting a volume floating around in pkg/infra/abi and pkg/api/handlers/libpod. And none of them used volume.Inspect(), the only correct way of inspecting volumes. Remove them all and consolidate to using the correct way. Compat API is likely still doing things the wrong way, but that is an issue for another day. Fixes #4304 Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2021-01-14 15:35:33 -05:00
zhangguanzhang	0cff5ad0a3	Fxes /etc/hosts duplicated every time after container restarted in a pod Signed-off-by: zhangguanzhang <zhangguanzhang@qq.com>	2021-01-13 19:03:35 +08:00
OpenShift Merge Robot	db5e7ec4c4	Merge pull request #8947 from Luap99/cleanup-code Fix problems reported by staticcheck	2021-01-12 13:15:35 -05:00
Paul Holzinger	8452b768ec	Fix problems reported by staticcheck `staticcheck` is a golang code analysis tool. https://staticcheck.io/ This commit fixes a lot of problems found in our code. Common problems are: - unnecessary use of fmt.Sprintf - duplicated imports with different names - unnecessary check that a key exists before a delete call There are still a lot of reported problems in the test files but I have not looked at those. Signed-off-by: Paul Holzinger <paul.holzinger@web.de>	2021-01-12 16:11:09 +01:00
unknown	2aa381f2d0	add pre checkpoint Signed-off-by: Zhuohan Chen <chen_zhuohan@163.com>	2021-01-10 21:38:28 +08:00
OpenShift Merge Robot	49db79e735	Merge pull request #8781 from rst0git/cr-volumes Add support for checkpoint/restore of containers with volumes	2021-01-08 10:41:05 -05:00
Giuseppe Scrivano	ecedda63a6	rootless: automatically split userns ranges writing to the id map fails when an extent overlaps multiple mappings in the parent user namespace: $ cat /proc/self/uid_map 0 1000 1 1 100000 65536 $ unshare -U sleep 100 & [1] 1029703 $ printf "0 0 100\n" \| tee /proc/$!/uid_map 0 0 100 tee: /proc/1029703/uid_map: Operation not permitted This limitation is particularly annoying when working with rootless containers as each container runs in the rootless user namespace, so a command like: $ podman run --uidmap 0:0:2 --rm fedora echo hi Error: writing file `/proc/664087/gid_map`: Operation not permitted: OCI permission denied would fail since the specified mapping overlaps the first mapping (where the user id is mapped to root) and the second extent with the additional IDs available. Detect such cases and automatically split the specified mapping with the equivalent of: $ podman run --uidmap 0:0:1 --uidmap 1:1:1 --rm fedora echo hi hi A fix has already been proposed for the kernel[1], but even if it accepted it will take time until it is available in a released kernel, so fix it also in pkg/rootless. [1] https://lkml.kernel.org/lkml/20201203150252.1229077-1-gscrivan@redhat.com/ Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2021-01-07 09:42:27 +01:00
Radostin Stoyanov	288ccc4c84	Include named volumes in container migration When migrating a container with associated volumes, the content of these volumes should be made available on the destination machine. This patch enables container checkpoint/restore with named volumes by including the content of volumes in checkpoint file. On restore, volumes associated with container are created and their content is restored. The --ignore-volumes option is introduced to disable this feature. Example: # podman container checkpoint --export checkpoint.tar.gz <container> The content of all volumes associated with the container are included in `checkpoint.tar.gz` # podman container checkpoint --export checkpoint.tar.gz --ignore-volumes <container> The content of volumes is not included in `checkpoint.tar.gz`. This is useful, for example, when the checkpoint/restore is performed on the same machine. # podman container restore --import checkpoint.tar.gz The associated volumes will be created and their content will be restored. Podman will exit with an error if volumes with the same name already exist on the system or the content of volumes is not included in checkpoint.tar.gz # podman container restore --ignore-volumes --import checkpoint.tar.gz Volumes associated with container must already exist. Podman will not create them or restore their content. Signed-off-by: Radostin Stoyanov <rstoyanov@fedoraproject.org>	2021-01-07 07:51:19 +00:00
Radostin Stoyanov	17f50fb4bf	Use Options as exportCheckpoint() argument Instead of individual values from ContainerCheckpointOptions, provide the options object. This is a preparation for the next patch where one more value of the options object is required in exportCheckpoint(). Signed-off-by: Radostin Stoyanov <rstoyanov@fedoraproject.org>	2021-01-07 07:48:41 +00:00
Matthew Heon	8f844a66d5	Ensure that user-specified HOSTNAME is honored When adding the HOSTNAME environment variable, only do so if it is not already present in the spec. If it is already present, it was likely added by the user, and we should honor their requested value. Fixes #8886 Signed-off-by: Matthew Heon <mheon@redhat.com>	2021-01-06 09:46:21 -05:00
Josh Soref	4fa1fce930	Spelling Signed-off-by: Josh Soref <jsoref@users.noreply.github.com>	2020-12-22 13:34:31 -05:00
Giuseppe Scrivano	f711f5a68d	podman: drop checking valid rootless UID do not check whether the specified ID is valid in the user namespace. crun handles this case[1], so the check in Podman prevents to get to the OCI runtime at all. $ podman run --user 10:0 --uidmap 0:0:1 --rm -ti fedora:33 sh -c 'id; cat /proc/self/uid_map' uid=10(10) gid=0(root) groups=0(root),65534(nobody) 10 0 1 [1] https://github.com/containers/crun/pull/556 Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2020-12-11 15:43:33 +01:00
OpenShift Merge Robot	9b3a81a002	Merge pull request #8571 from Luap99/podman-network-reload Implement pod-network-reload	2020-12-08 06:15:40 -05:00
Matthew Heon	b0286d6b43	Implement pod-network-reload This adds a new command, 'podman network reload', to reload the networks of existing containers, forcing recreation of firewall rules after e.g. `firewall-cmd --reload` wipes them out. Under the hood, this works by calling CNI to tear down the existing network, then recreate it using identical settings. We request that CNI preserve the old IP and MAC address in most cases (where the container only had 1 IP/MAC), but there will be some downtime inherent to the teardown/bring-up approach. The architecture of CNI doesn't really make doing this without downtime easy (or maybe even possible...). At present, this only works for root Podman, and only locally. I don't think there is much of a point to adding remote support (this is very much a local debugging command), but I think adding rootless support (to kill/recreate slirp4netns) could be valuable. Signed-off-by: Matthew Heon <matthew.heon@pm.me> Signed-off-by: Paul Holzinger <paul.holzinger@web.de>	2020-12-07 19:26:23 +01:00
OpenShift Merge Robot	f01630acf3	Merge pull request #8476 from rhatdan/containerenv Add containerenv information to /run/.containerenv	2020-12-04 11:56:24 -05:00
Daniel J Walsh	d9154e97eb	Add containerenv information to /run/.containerenv We have been asked to leak some information into the container to indicate: * The name and id of the container * The version of podman used to launch the container * The image name and ID the container is based on. * Whether the container engine is running in rootless mode. Fixes: https://github.com/containers/podman/issues/6192 Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-12-03 13:32:15 -05:00
Daniel J Walsh	f00cc25a7c	Drop default log-level from error to warn Our users are missing certain warning messages that would make debugging issues with Podman easier. For example if you do a podman build with a Containerfile that contains the SHELL directive, the Derective is silently ignored. If you run with the log-level warn you get a warning message explainging what happened. $ podman build --no-cache -f /tmp/Containerfile1 /tmp/ STEP 1: FROM ubi8 STEP 2: SHELL ["/bin/bash", "-c"] STEP 3: COMMIT --> 7a207be102a 7a207be102aa8993eceb32802e6ceb9d2603ceed9dee0fee341df63e6300882e $ podman --log-level=warn build --no-cache -f /tmp/Containerfile1 /tmp/ STEP 1: FROM ubi8 STEP 2: SHELL ["/bin/bash", "-c"] STEP 3: COMMIT WARN[0000] SHELL is not supported for OCI image format, [/bin/bash -c] will be ignored. Must use `docker` format --> 7bd96fd25b9 7bd96fd25b9f755d8a045e31187e406cf889dcf3799357ec906e90767613e95f These messages will no longer be lost, when we default to WARNing level. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-12-03 06:28:09 -05:00
Daniel J Walsh	20160af018	Switch from pkg/secrets to pkg/subscriptions The buildah/pkg/secrts package was move to containers/common/pkg/subscriptions. Switch to using this by default. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-11-26 07:30:18 -05:00
OpenShift Merge Robot	42ec4cf87f	Merge pull request #8290 from vrothberg/fix-8265 use container cgroups path	2020-11-17 14:00:09 +01:00
Valentin Rothberg	39bf07694c	use container cgroups path When looking up a container's cgroup path, parse /proc/[PID]/cgroup. This will work across all cgroup managers and configurations and is supported on cgroups v1 and v2. Fixes: #8265 Signed-off-by: Valentin Rothberg <rothberg@redhat.com>	2020-11-17 12:29:50 +01:00
Daniel J Walsh	4ca4234af1	Make sure /etc/hosts populated correctly with networks The --hostname and containername should always be added to containers. Added some tests to make sure you can always ping the hostname and container name from within the container. Fixes: https://github.com/containers/podman/issues/8095 Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-11-16 16:40:50 -05:00
Daniel J Walsh	3ee44d942e	Add better support for unbindable volume mounts Allow users to specify unbindable on volume command line Switch internal mounts to rprivate to help prevent leaks. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-11-02 07:19:12 -05:00
OpenShift Merge Robot	1fe79dd677	Merge pull request #8177 from rhatdan/wrap Stop excessive wrapping of errors	2020-10-30 19:52:17 +01:00
Andy Librian	6779c1cfc2	Improve setupSystemd, grab mount options from the host fixes #7661 Signed-off-by: Andy Librian <andylibrian@gmail.com>	2020-10-30 20:51:34 +07:00
Daniel J Walsh	831d7fb0d7	Stop excessive wrapping of errors Most of the builtin golang functions like os.Stat and os.Open report errors including the file system object path. We should not wrap these errors and put the file path in a second time, causing stuttering of errors when they get presented to the user. This patch tries to cleanup a bunch of these errors. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-10-30 05:34:04 -04:00
Valentin Rothberg	65a618886e	new "image" mount type Add a new "image" mount type to `--mount`. The source of the mount is the name or ID of an image. The destination is the path inside the container. Image mounts further support an optional `rw,readwrite` parameter which if set to "true" will yield the mount writable inside the container. Note that no changes are propagated to the image mount on the host (which in any case is read only). Mounts are overlay mounts. To support read-only overlay mounts, vendor a non-release version of Buildah. Signed-off-by: Valentin Rothberg <rothberg@redhat.com>	2020-10-29 15:06:22 +01:00
Paul Holzinger	f391849c22	Don't error if resolv.conf does not exists If the resolv.conf file is empty we provide default dns servers. If the file does not exists we error and don't create the container. We should also provide the default entries in this case. This is also what docker does. Fixes #8089 Signed-off-by: Paul Holzinger <paul.holzinger@web.de>	2020-10-22 19:21:07 +02:00
Matthew Heon	0864d82cb5	Add hostname to /etc/hosts for --net=none This does not match Docker, which does not add hostname in this case, but it seems harmless enough. Fixes #8095 Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2020-10-21 13:45:41 -04:00
Matthew Heon	1b288a35ba	Ensure that hostname is added to hosts with net=host When a container uses --net=host the default hostname is set to the host's hostname. However, we were not creating any entries in `/etc/hosts` despite having a hostname, which is incorrect. This hostname, for Docker compat, will always be the hostname of the host system, not the container, and will be assigned to IP 127.0.1.1 (not the standard localhost address). Also, when `--hostname` and `--net=host` are both passed, still use the hostname from `--hostname`, not the host's hostname (we still use the host's hostname by default in this case if the `--hostname` flag is not passed). Fixes #8054 Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2020-10-20 10:31:15 -04:00
Paul Holzinger	2e65497dea	Fix possible panic in libpod container restore We need to do a length check before we can access the networkStatus slice by index to prevent a runtime panic. Fixes #8026 Signed-off-by: Paul Holzinger <paul.holzinger@web.de>	2020-10-15 11:50:29 +02:00
Daniel J Walsh	6ca8067956	Setup HOME environment when using --userns=keep-id Currently the HOME environment is set to /root if the user does not override it. Also walk the parent directories of users homedir to see if it is volume mounted into the container, if yes, then set it correctly. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-10-14 16:45:24 -04:00
Matthew Heon	4d800a5f45	Store cgroup manager on a per-container basis When we create a container, we assign a cgroup parent based on the current cgroup manager in use. This parent is only usable with the cgroup manager the container is created with, so if the default cgroup manager is later changed or overridden, the container will not be able to start. To solve this, store the cgroup manager that created the container in container configuration, so we can guarantee a container with a systemd cgroup parent will always be started with systemd cgroups. Unfortunately, this is very difficult to test in CI, due to the fact that we hard-code cgroup manager on all invocations of Podman in CI. Fixes #7830 Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2020-10-08 15:25:06 -04:00
Daniel J Walsh	3ae47f7d2b	Populate /etc/hosts file when run in a user namespace We do not populate the hostname field with the IP Address when running within a user namespace. Fixes https://github.com/containers/podman/issues/7490 Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-10-07 08:39:44 -04:00
Kir Kolyshkin	684d0079d2	Lowercase some errors This commit is courtesy of ``` for f in $(git ls-files .go \| grep -v ^vendor/); do \ sed -i 's/$errors\..$"Error /\1"error /' $f; done for f in $(git ls-files .go \| grep -v ^vendor/); do \ sed -i 's/$errors\..$"Failed to /\1"failed to /' $f; done ``` etc. Self-reviewed using `git diff --word-diff`, found no issues. Signed-off-by: Kir Kolyshkin <kolyshkin@gmail.com>	2020-10-05 15:56:44 -07:00
Kir Kolyshkin	4878dff3e2	Remove excessive error wrapping In case os.Open[File], os.Mkdir[All], ioutil.ReadFile and the like fails, the error message already contains the file name and the operation that fails, so there is no need to wrap the error with something like "open %s failed". While at it - replace a few places with os.Open, ioutil.ReadAll with ioutil.ReadFile. - replace errors.Wrapf with errors.Wrap for cases where there are no %-style arguments. Signed-off-by: Kir Kolyshkin <kolyshkin@gmail.com>	2020-10-05 15:30:37 -07:00
Giuseppe Scrivano	d30121969f	libpod: check the gid is present before adding it check there are enough gids in the user namespace before adding supplementary gids from /etc/group. Follow-up for `baede7cd27` Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2020-10-02 13:13:44 +02:00
Daniel J Walsh	baede7cd27	Add additionalGIDs from users in rootless mode There is a risk here, that if the GID does not exists within the User Namespace the container will fail to start. This is only likely to happen in HPC Envioronments, and I think we should add a field to disable it for this environment, Added a FIXME for this issue. We currently have this problem with running a rootfull container within a user namespace, it will fail if the GID is not available. I looked at potentially checking the usernamespace that you are assigned to, but I believe this will be very difficult to code up and to figure out. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-10-01 10:54:28 -04:00
Daniel J Walsh	912f952c1f	Fix --systemd=always regression The kernel will not allow you to modify existing mount flags on a volume when bind mounting it to another place. Since /sys/fs/cgroup/systemd is mounted noexec on the host, it needs to be mounted with the same flags in the rootless container. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-09-15 13:29:44 -04:00
Eduardo Vega	6a1233597a	Determine if resolv.conf points to systemd-resolved Signed-off-by: Eduardo Vega <edvegavalerio@gmail.com>	2020-09-11 23:31:07 -06:00
OpenShift Merge Robot	861451a462	Merge pull request #7541 from mheon/modify_group Make an entry in /etc/group when we modify /etc/passwd	2020-09-10 17:05:02 -04:00
Matthew Heon	f57c39fc7c	Make an entry in /etc/group when we modify /etc/passwd To ensure that the user running in the container ahs a valid entry in /etc/passwd so lookup functions for the current user will not error, Podman previously began adding entries to the passwd file. We did not, however, add entries to the group file, and this created problems - our passwd entries included the group the user is in, but said group might not exist. The solution is to mirror our logic for /etc/passwd modifications to also edit /etc/group in the container. Unfortunately, this is not a catch-all solution. Our logic here is only advanced enough to add to the group file - so if the group already exists but we add a user not a part of it, we will not modify that existing entry, and things remain inconsistent. We can look into adding this later if we absolutely need to, but it would involve adding significant complexity to this already massively complicated function. While we're here, address an edge case where Podman could add a user or group whose UID overlapped with an existing user or group. Also, let's make users able to log into users we added. Instead of generating user entries with an 'x' in the password field, indicating they have an entry in /etc/shadow, generate a '*' indicating the user has no password but can be logged into by other means e.g. ssh key, su. Fixes #7503 Fixes #7389 Fixes #7499 Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2020-09-10 13:02:31 -04:00
Akihiro Suda	f82abc774a	rootless: support `podman network create` (CNI-in-slirp4netns) Usage: ``` $ podman network create foo $ podman run -d --name web --hostname web --network foo nginx:alpine $ podman run --rm --network foo alpine wget -O - http://web.dns.podman Connecting to web.dns.podman (10.88.4.6:80) ... <h1>Welcome to nginx!</h1> ... ``` See contrib/rootless-cni-infra for the design. Signed-off-by: Akihiro Suda <akihiro.suda.cz@hco.ntt.co.jp>	2020-09-09 15:47:38 +09:00
Daniel J Walsh	d68a6b52ec	We should not be mounting /run as noexec when run with --systemd The system defaults /run to "exec" mode, and we default --read-only mounts on /run to "exec", so --systemd should follow suit. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-09-02 08:00:22 -04:00
Matthew Heon	3875040f13	Ensure rootless containers without a passwd can start We want to modify /etc/passwd to add an entry for the user in question, but at the same time we don't want to require the container provide a /etc/passwd (a container with a single, statically linked binary and nothing else is perfectly fine and should be allowed, for example). We could create the passwd file if it does not exist, but if the container doesn't provide one, it's probably better not to make one at all. Gate changes to /etc/passwd behind a stat() of the file in the container returning cleanly. Fixes #7515 Signed-off-by: Matthew Heon <mheon@redhat.com>	2020-08-31 18:15:43 -04:00
Giuseppe Scrivano	3967c46544	vendor: update opencontainers/runtime-spec Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2020-08-21 19:06:04 +02:00
Daniel J Walsh	bd63a252f3	Don't limit the size on /run for systemd based containers We had a customer incident where they ran out of space on /run. If you don't specify size, it will be still limited to 50% or memory available in the cgroup the container is running in. If the cgroup is unlimited then the /run will be limited to 50% of the total memory on the system. Also /run is mounted on the host as exec, so no reason for us to mount it noexec. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-08-18 14:31:00 -04:00
Matthew Heon	7b3cf0c085	Change /sys/fs/cgroup/systemd mount to rprivate I used the wrong propagation first time around because I forgot that rprivate is the default propagation. Oops. Switch to rprivate so we're using the default. Signed-off-by: Matthew Heon <mheon@redhat.com>	2020-08-12 09:15:02 -04:00
Matthew Heon	a064cfc99b	Ensure correct propagation for cgroupsv1 systemd cgroup On cgroups v1 systems, we need to mount /sys/fs/cgroup/systemd into the container. We were doing this with no explicit mount propagation tag, which means that, under some circumstances, the shared mount propagation could be chosen - which, combined with the fact that we need a mount to mask /sys/fs/cgroup/systemd/release_agent in the container, means we would leak a never-ending set of mounts under /sys/fs/cgroup/systemd/ on container restart. Fortunately, the fix is very simple - hardcode mount propagation to something that won't leak. Signed-off-by: Matthew Heon <mheon@redhat.com>	2020-08-11 09:53:36 -04:00
Matthew Heon	333d9af77a	Ensure WORKDIR from images is created A recent crun change stopped the creation of the container's working directory if it does not exist. This is arguably correct for user-specified directories, to protect against typos; it is definitely not correct for image WORKDIR, where the image author definitely intended for the directory to be used. This makes Podman create the working directory and chown it to container root, if it does not already exist, and only if it was specified by an image, not the user. Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2020-08-03 14:44:52 -04:00
OpenShift Merge Robot	044a7cb100	Merge pull request #6991 from mheon/change_passwd_ondisk Make changes to /etc/passwd on disk for non-read only	2020-07-29 14:27:50 -04:00
Daniel J Walsh	a5e37ad280	Switch all references to github.com/containers/libpod -> podman Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-07-28 08:23:45 -04:00
Matthew Heon	bae6853906	Make changes to /etc/passwd on disk for non-read only Bind-mounting /etc/passwd into the container is problematic becuase of how system utilities like `useradd` work. They want to make a copy and then rename to try to prevent breakage; this is, unfortunately, impossible when the file they want to rename is a bind mount. The current behavior is fine for read-only containers, though, because we expect useradd to fail in those cases. Instead of bind-mounting, we can edit /etc/passwd in the container's rootfs. This is kind of gross, because the change will show up in `podman diff` and similar tools, and will be included in images made by `podman commit`. However, it's a lot better than breaking important system tools. Fixes #6953 Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2020-07-23 14:27:19 -04:00
Daniel J Walsh	4c4a00f63e	Support default profile for apparmor Currently you can not apply an ApparmorProfile if you specify --privileged. This patch will allow both to be specified simultaniosly. By default Apparmor should be disabled if the user specifies --privileged, but if the user specifies --security apparmor:PROFILE, with --privileged, we should do both. Added e2e run_apparmor_test.go Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-07-22 06:27:20 -04:00
Ashley Cui	d4d3fbc155	Add --umask flag for create, run --umask sets the umask inside the container Defaults to 0022 Co-authored-by: Daniel J Walsh <dwalsh@redhat.com> Signed-off-by: Ashley Cui <acui@redhat.com>	2020-07-21 14:22:30 -04:00
Qi Wang	020d81f113	Add support for overlay volume mounts in podman. Add support -v for overlay volume mounts in podman. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com> Signed-off-by: Qi Wang <qiwan@redhat.com>	2020-07-20 09:48:55 -04:00
Matthew Heon	1ad7042a34	Preserve passwd on container restart We added code to create a `/etc/passwd` file that we bind-mount into the container in some cases (most notably, `--userns=keep-id` containers). This, unfortunately, was not persistent, so user-added users would be dropped on container restart. Changing where we store the file should fix this. Further, we want to ensure that lookups of users in the container use the right /etc/passwd if we replaced it. There was already logic to do this, but it only worked for user-added mounts; it's easy enough to alter it to use our mounts as well. Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2020-07-15 10:25:46 -04:00
Matthew Heon	4b784b377c	Remove all instances of named return "err" from Libpod This was inspired by https://github.com/cri-o/cri-o/pull/3934 and much of the logic for it is contained there. However, in brief, a named return called "err" can cause lots of code confusion and encourages using the wrong err variable in defer statements, which can make them work incorrectly. Using a separate name which is not used elsewhere makes it very clear what the defer should be doing. As part of this, remove a large number of named returns that were not used anywhere. Most of them were once needed, but are no longer necessary after previous refactors (but were accidentally retained). Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2020-07-09 13:54:47 -04:00
Daniel J Walsh	6c6670f12a	Add username to /etc/passwd inside of container if --userns keep-id If I enter a continer with --userns keep-id, my UID will be present inside of the container, but most likely my user will not be defined. This patch will take information about the user and stick it into the container. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-07-07 08:34:31 -04:00
OpenShift Merge Robot	9532509c50	Merge pull request #6836 from ashley-cui/tzlibpod Add --tz flag to create, run	2020-07-06 13:28:20 -04:00
Valentin Rothberg	8489dc4345	move go module to v2 With the advent of Podman 2.0.0 we crossed the magical barrier of go modules. While we were able to continue importing all packages inside of the project, the project could not be vendored anymore from the outside. Move the go module to new major version and change all imports to `github.com/containers/libpod/v2`. The renaming of the imports was done via `gomove` [1]. [1] https://github.com/KSubedi/gomove Signed-off-by: Valentin Rothberg <rothberg@redhat.com>	2020-07-06 15:50:12 +02:00
Ashley Cui	9a1543caec	Add --tz flag to create, run --tz flag sets timezone inside container Can be set to IANA timezone as well as `local` to match host machine Signed-off-by: Ashley Cui <acui@redhat.com>	2020-07-02 13:30:59 -04:00
Giuseppe Scrivano	6ee5f740a4	podman: add new cgroup mode split When running under systemd there is no need to create yet another cgroup for the container. With conmon-delegated the current cgroup will be split in two sub cgroups: - supervisor - container The supervisor cgroup will hold conmon and the podman process, while the container cgroup is used by the OCI runtime (using the cgroupfs backend). Closes: https://github.com/containers/libpod/issues/6400 Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2020-06-25 17:16:12 +02:00
Daniel J Walsh	5b3503c0a1	Add container name to the /etc/hosts within the container This will allow containers that connect to the network namespace be able to use the container name directly. For example you can do something like podman run -ti --name foobar fedora ping foobar While we can do this with hostname now, this seems more natural. Also if another container connects on the network to this container it can do podman run --network container:foobar fedora ping foobar And connect to the original container,without having to discover the name. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-06-20 06:20:46 -04:00
Daniel J Walsh	200cfa41a4	Turn on More linters - misspell - prealloc - unparam - nakedret Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-06-15 07:05:56 -04:00
Giuseppe Scrivano	8ef1b461ae	libpod: fix check for slirp4netns netns fix the check for c.state.NetNS == nil. Its value is changed in the first code block, so the condition is always true in the second one and we end up running slirp4netns twice. Closes: https://github.com/containers/libpod/issues/6538 Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2020-06-11 13:06:26 +02:00
Giuseppe Scrivano	6c27e27b8c	container: do not set hostname when joining uts do not set the hostname when joining an UTS namespace, as it could be owned by a different userns. Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2020-06-10 14:52:10 +02:00
Giuseppe Scrivano	a389eab8d1	container: make resolv.conf and hosts accessible in userns when running in a new userns, make sure the resolv.conf and hosts files bind mounted from another container are accessible to root in the userns. Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2020-06-10 14:46:48 +02:00
Qi Wang	77e4b077b9	check --user range for rootless containers Check --user range if it's a uid for rootless containers. Returns error if it is out of the range. From https://github.com/containers/libpod/issues/6431#issuecomment-636124686 Signed-off-by: Qi Wang <qiwan@redhat.com>	2020-06-02 11:28:58 -04:00
Daniel J Walsh	35829854a2	Fix mountpont in SecretMountsWithUIDGID In FIPS Mode we expect to work off of the Mountpath not the Rundir path. This is causing FIPS Mode checks to fail. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-05-19 16:33:24 -04:00
Giuseppe Scrivano	b69ba30b14	libpod: set hostname from joined container when joining a UTS namespace, take the hostname from the destination container. Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2020-04-27 17:08:53 +02:00
Daniel J Walsh	e62d081770	Update podman to use containers.conf Add more default options parsing Switch to using --time as opposed to --timeout to better match Docker. Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-04-20 16:11:36 -04:00
Giuseppe Scrivano	3a0a727110	userns: support --userns=auto automatically pick an empty range and create an user namespace for the container. Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2020-04-06 16:32:36 +02:00
Daniel J Walsh	4352d58549	Add support for containers.conf vendor in c/common config pkg for containers.conf Signed-off-by: Qi Wang qiwan@redhat.com Signed-off-by: Daniel J Walsh <dwalsh@redhat.com>	2020-03-27 14:36:03 -04:00
Matthew Heon	b6954758bb	Attempt manual removal of CNI IP allocations on refresh We previously attempted to work within CNI to do this, without success. So let's do it manually, instead. We know where the files should live, so we can remove them ourselves instead. This solves issues around sudden reboots where containers do not have time to fully tear themselves down, and leave IP address allocations which, for various reasons, are not stored in tmpfs and persist through reboot. Fixes #5433 Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2020-03-19 17:20:31 -04:00
Matthew Heon	b41c864d56	Ensure that exec sessions inherit supplemental groups This corrects a regression from Podman 1.4.x where container exec sessions inherited supplemental groups from the container, iff the exec session did not specify a user. Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2020-02-28 11:32:56 -05:00
Matthew Heon	97323808ed	Add network options to podman pod create Enables most of the network-related functionality from `podman run` in `podman pod create`. Custom CNI networks can be specified, host networking is supported, DNS options can be configured. Also enables host networking in `podman play kube`. Fixes #2808 Fixes #3837 Fixes #4432 Fixes #4718 Fixes #4770 Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2020-02-19 11:29:30 -05:00
Valentin Rothberg	67165b7675	make lint: enable gocritic `gocritic` is a powerful linter that helps in preventing certain kinds of errors as well as enforcing a coding style. Signed-off-by: Valentin Rothberg <rothberg@redhat.com>	2020-01-13 14:27:02 +01:00
Adrian Reber	225c7ae6c9	Correctly export the root file-system changes When doing a checkpoint with --export the root file-system diff was not working as expected. Instead of getting the changes from the running container to the highest storage layer it got the changes from the highest layer to that parent's layer. For a one layer container this could mean that the complete root file-system is part of the checkpoint. With this commit this changes to use the same functionality as 'podman diff'. This actually enables to correctly diff the root file-system including tracking deleted files. This also removes the non-working helper functions from libpod/diff.go. Signed-off-by: Adrian Reber <areber@redhat.com>	2019-12-09 13:29:36 +01:00
Matthew Heon	b0b9103cca	Allow chained network namespace containers The code currently assumes that the container we delegate network namespace to will never further delegate to another container, so when looking up things like /etc/hosts and /etc/resolv.conf we won't pull the correct files from the chained dependency. The changes to resolve this are relatively simple - just need to keep looking until we find a container without NetNsCtr set. Fixes #4626 Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2019-12-03 10:27:15 -05:00
OpenShift Merge Robot	e4275b3453	Merge pull request #4493 from mheon/add_removing_state Add ContainerStateRemoving	2019-12-02 16:31:11 +01:00
Adrian Reber	5e43c7cde1	Disable checkpointing of containers started with --rm Trying to checkpoint a container started with --rm works, but it makes no sense as the container, including the checkpoint, will be deleted after writing the checkpoint. This commit inhibits checkpointing containers started with '--rm' unless '--export' is used. If the checkpoint is exported it can easily be restored from the exported checkpoint, even if '--rm' is used. To restore a container from a checkpoint it is even necessary to manually run 'podman rm' if the container is not started with '--rm'. Signed-off-by: Adrian Reber <areber@redhat.com>	2019-11-28 20:25:45 +01:00
Matthew Heon	25cc43c376	Add ContainerStateRemoving When Libpod removes a container, there is the possibility that removal will not fully succeed. The most notable problems are storage issues, where the container cannot be removed from c/storage. When this occurs, we were faced with a choice. We can keep the container in the state, appearing in `podman ps` and available for other API operations, but likely unable to do any of them as it's been partially removed. Or we can remove it very early and clean up after it's already gone. We have, until now, used the second approach. The problem that arises is intermittent problems removing storage. We end up removing a container, failing to remove its storage, and ending up with a container permanently stuck in c/storage that we can't remove with the normal Podman CLI, can't use the name of, and generally can't interact with. A notable cause is when Podman is hit by a SIGKILL midway through removal, which can consistently cause `podman rm` to fail to remove storage. We now add a new state for containers that are in the process of being removed, ContainerStateRemoving. We set this at the beginning of the removal process. It notifies Podman that the container cannot be used anymore, but preserves it in the DB until it is fully removed. This will allow Remove to be run on these containers again, which should successfully remove storage if it fails. Fixes #3906 Signed-off-by: Matthew Heon <mheon@redhat.com>	2019-11-19 15:38:03 -05:00
Radostin Stoyanov	368d2ecfb6	container-restore: Fix restore with user namespace When restoring a container with user namespace, the user namespace is created by the OCI runtime, and the network namespace is created after the user namespace to ensure correct ownership. In this case PostConfigureNetNS will be set and the value of c.state.NetNS would be nil. Hence, the following error occurs: $ sudo podman run --name cr \ --uidmap 0:1000:500 \ -d docker.io/library/alpine \ /bin/sh -c 'i=0; while true; do echo $i; i=$(expr $i + 1); sleep 1; done' $ sudo podman container checkpoint cr $ sudo podman container restore cr ... panic: runtime error: invalid memory address or nil pointer dereference [signal SIGSEGV: segmentation violation code=0x1 addr=0x30 pc=0x13a5e3c] Signed-off-by: Radostin Stoyanov <rstoyanov1@gmail.com>	2019-11-17 00:34:02 +00:00
Jakub Filak	2497b6c77b	podman: add support for specifying MAC I basically copied and adapted the statements for setting IP. Closes #1136 Signed-off-by: Jakub Filak <jakub.filak@sap.com>	2019-11-06 16:22:19 +01:00
Urvashi Mohnani	2a149ad90a	Vendor in latest containers/buildah Pull in changes to pkg/secrets/secrets.go that adds the logic to disable fips mode if a pod/container has a label set. Signed-off-by: Urvashi Mohnani <umohnani@redhat.com>	2019-11-01 09:41:09 -04:00
Valentin Rothberg	11c282ab02	add libpod/config Refactor the `RuntimeConfig` along with related code from libpod into libpod/config. Note that this is a first step of consolidating code into more coherent packages to make the code more maintainable and less prone to regressions on the long runs. Some libpod definitions were moved to `libpod/define` to resolve circular dependencies. Signed-off-by: Valentin Rothberg <rothberg@redhat.com>	2019-10-31 17:42:37 +01:00
Giuseppe Scrivano	0d5d6dab57	systemd: mask /sys/fs/cgroup/systemd/release_agent when running in systemd mode on cgroups v1, make sure the /sys/fs/cgroup/systemd/release_agent is masked otherwise the container is able to modify it and execute scripts on the host. Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2019-10-25 21:50:29 +02:00
Matthew Heon	b6a7d88397	When restoring containers, reset cgroup path Previously, `podman checkport restore` with exported containers, when told to create a new container based on the exported checkpoint, would create a new container, with a new container ID, but not reset CGroup path - which contained the ID of the original container. If this was done multiple times, the result was two containers with the same cgroup paths. Operations on these containers would this have a chance of crossing over to affect the other one; the most notable was `podman rm` once it was changed to use the --all flag when stopping the container; all processes in the cgroup, including the ones in the other container, would be stopped. Reset cgroups on restore to ensure that the path matches the ID of the container actually being run. Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2019-10-10 14:53:29 -04:00
Matthew Heon	6f630bc09b	Move OCI runtime implementation behind an interface For future work, we need multiple implementations of the OCI runtime, not just a Conmon-wrapped runtime matching the runc CLI. As part of this, do some refactoring on the interface for exec (move to a struct, not a massive list of arguments). Also, add 'all' support to Kill and Stop (supported by runc and used a bit internally for removing containers). Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2019-10-10 10:19:32 -04:00
OpenShift Merge Robot	e4835f6b01	Merge pull request #4086 from mheon/cni_del_on_refresh Force a CNI Delete on refreshing containers	2019-09-25 09:35:40 +02:00
Matthew Heon	b57d2f4cc7	Force a CNI Delete on refreshing containers CNI expects that a DELETE be run before re-creating container networks. If a reboot occurs quickly enough that containers can't stop and clean up, that DELETE never happens, and Podman currently wipes the old network info and thinks the state has been entirely cleared. Unfortunately, that may not be the case on the CNI side. Some things - like IP address reservations - may not have been cleared. To solve this, manually re-run CNI Delete on refresh. If the container has already been deleted this seems harmless. If not, it should clear lingering state. Fixes: #3759 Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2019-09-24 09:52:11 -04:00
Gabi Beyer	5813c8246e	rootless: Rearrange setup of rootless containers In order to run Podman with VM-based runtimes unprivileged, the network must be set up prior to the container creation. Therefore this commit modifies Podman to run rootless containers by: 1. create a network namespace 2. pass the netns persistent mount path to the slirp4netns to create the tap inferface 3. pass the netns path to the OCI spec, so the runtime can enter the netns Closes #2897 Signed-off-by: Gabi Beyer <gabrielle.n.beyer@intel.com>	2019-09-24 11:01:28 +02:00
Giuseppe Scrivano	fb353f6f42	execuser: look at the source for /etc/{passwd,group} overrides look if there are bind mounts that can shadow the /etc/passwd and /etc/group files. In that case, look at the bind mount source. Closes: https://github.com/containers/libpod/pull/4068#issuecomment-533782941 Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2019-09-21 22:11:09 +02:00
Giuseppe Scrivano	e42e1c45ae	container: make sure $HOME is always set If the HOME environment variable is not set, make sure it is set to the configuration found in the container /etc/passwd file. It was previously depending on a runc behavior that always set HOME when it is not set. The OCI runtime specifications do not require HOME to be set so move the logic to libpod. Closes: https://github.com/debarshiray/toolbox/issues/266 Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2019-09-20 16:01:38 +02:00
Giuseppe Scrivano	a249c98db8	linux: fix systemd with --cgroupns=private When --cgroupns=private is used we need to mount a new cgroup file system so that it points to the correct namespace. Needs: https://github.com/containers/crun/pull/88 Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2019-09-12 14:33:26 +02:00
OpenShift Merge Robot	9cf852c305	Merge pull request #3927 from openSUSE/manager-annotations Add `ContainerManager` annotation to created containers	2019-09-11 09:34:14 +02:00
OpenShift Merge Robot	7ac6ed3b4b	Merge pull request #3581 from mheon/no_cgroups Support running containers without CGroups	2019-09-11 00:58:46 +02:00
Matthew Heon	c2284962c7	Add support for launching containers without CGroups This is mostly used with Systemd, which really wants to manage CGroups itself when managing containers via unit file. Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2019-09-10 10:52:37 -04:00
Sascha Grunert	df036f9f8e	Add `ContainerManager` annotation to created containers This change adds the following annotation to every container created by podman: ```json "Annotations": { "io.containers.manager": "libpod" } ``` Target of this annotaions is to indicate which project in the containers ecosystem is the major manager of a container when applications share the same storage paths. This way projects can decide if they want to manipulate the container or not. For example, since CRI-O and podman are not using the same container library (libpod), CRI-O can skip podman containers and provide the end user more useful information. A corresponding end-to-end test has been adapted as well. Relates to: https://github.com/cri-o/cri-o/pull/2761 Signed-off-by: Sascha Grunert <sgrunert@suse.com>	2019-09-10 09:37:14 +02:00
Matthew Heon	b6106341fb	When first mounting any named volume, copy up Previously, we only did this for volumes created at the same time as the container. However, this is not correct behavior - Docker does so for all named volumes, even those made with 'podman volume create' and mounted into a container later. Fixes #3945 Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2019-09-09 17:17:39 -04:00
Matthew Heon	77f9234513	Ignore ENOENT on umount of SHM Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2019-09-06 10:25:53 -04:00
Matthew Heon	de9a394fcf	Correctly report errors on unmounting SHM When we fail to remove a container's SHM, that's an error, and we need to report it as such. This may be part of our lingering storage woes. Also, remove MNT_DETACH. It may be another cause of the storage removal failures. Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2019-09-05 17:12:27 -04:00
Matthew Heon	a760e325f3	Add ability for volumes with options to mount/umount When volume options and the local volume driver are specified, the volume is intended to be mounted using the 'mount' command. Supported options will be used to volume the volume before the first container using it starts, and unmount the volume after the last container using it dies. This should work for any local filesystem, though at present I've only tested with tmpfs and btrfs. Signed-off-by: Matthew Heon <matthew.heon@pm.me>	2019-09-05 17:12:27 -04:00
baude	8818e358bf	handle dns response from cni when cni returns a list of dns servers, we should add them under the right conditions. the defined conditions are as follows: - if the user provides dns, it and only it are added. - if not above and you get a cni name server, it is added and a forwarding dns instance is created for what was in resolv.conf. - if not either above, the entries from the host's resolv.conf are used. Signed-off-by: baude <bbaude@redhat.com> Signed-off-by: baude <bbaude@redhat.com>	2019-09-03 10:10:05 -05:00
Giuseppe Scrivano	69727abdf6	cgroup: fix regression when running systemd commit `223fe64dc0` introduced the regression. When running on cgroups v1, bind mount only /sys/fs/cgroup/systemd as rw, as the code did earlier. Also, simplify the rootless code as it doesn't require any special handling when using --systemd. Closes: https://bugzilla.redhat.com/show_bug.cgi?id=1737554 Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2019-08-06 19:46:34 +02:00
baude	97b84dedf3	Revert "rootless: Rearrange setup of rootless containers" This reverts commit `80dcd4bebc`. Signed-off-by: baude <bbaude@redhat.com>	2019-08-06 09:51:38 -05:00
OpenShift Merge Robot	337358ae63	Merge pull request #3690 from adrianreber/ignore-static-ip restore: added --ignore-static-ip option	2019-08-05 16:11:50 +02:00
OpenShift Merge Robot	e2f38cdaa4	Merge pull request #3310 from gabibeyer/rootlessKata rootless: Rearrange setup of rootless containers *CIRRUS: TEST IMAGES*	2019-08-05 14:26:04 +02:00
Adrian Reber	c23b92b409	restore: added --ignore-static-ip option If a container is restored multiple times from an exported checkpoint with the help of '--import --name', the restore will fail if during 'podman run' a static container IP was set with '--ip'. The user can tell the restore process to ignore the static IP with '--ignore-static-ip'. Signed-off-by: Adrian Reber <areber@redhat.com>	2019-08-02 10:10:54 +02:00
OpenShift Merge Robot	5056964d09	Merge pull request #3677 from giuseppe/systemd-cgroupsv2 systemd, cgroupsv2: not bind mount /sys/fs/cgroup/systemd	2019-08-01 11:35:20 +02:00
Giuseppe Scrivano	223fe64dc0	systemd, cgroupsv2: not bind mount /sys/fs/cgroup/systemd when running on a cgroups v2 system, do not bind mount the named hierarchy /sys/fs/cgroup/systemd as it doesn't exist anymore. Instead bind mount the entire /sys/fs/cgroup. Signed-off-by: Giuseppe Scrivano <gscrivan@redhat.com>	2019-08-01 07:31:06 +02:00

1 2 3 4 5 ...

349 Commits