systemd

mirror of https://github.com/systemd/systemd.git synced 2024-12-23 21:35:11 +03:00

Author	SHA1	Message	Date
Topi Miettinen	c0548df0a2	core: firewall integration with ControlGroupNFTSet= New directive `ControlGroupNFTSet=` provides a method for integrating services into firewall rules with NFT sets. Example: ``` table inet filter { ... set timesyncd { type cgroupsv2 } chain ntp_output { socket cgroupv2 != @timesyncd counter drop accept } ... } ``` /etc/systemd/system/systemd-timesyncd.service.d/override.conf ``` [Service] ControlGroupNFTSet=inet:filter:timesyncd ``` ``` $ sudo nft list set inet filter timesyncd table inet filter { set timesyncd { type cgroupsv2 elements = { "system.slice/systemd-timesyncd.service" } } } ```	2022-06-08 16:12:25 +00:00
Benjamin Franzke	a25d9395ad	tree-wide: streamline wiki links * Avoid traling slash as most links are defined without. * Always use https:// protocol and www. subdomain Allows for easier tree-wide linkvalidation for our migration to systemd.io.	2022-05-21 14:28:03 +02:00
Zbigniew Jędrzejewski-Szmek	6f83ea60e9	man: beef up the description of systemd-oomd.service The gist of the description is moved from systemd.resource-control to systemd-oomd man page. Cross-references to OOMPolicy, memory.oom.group, oomctl, ManagedOOMSwap and ManagedOOMMemoryPressure are added in all places. The descriptions are also more down-to-earth: instead of talking about "taking action" let's just say "kill". We might add configuration for different actions in the future, but we're not there yet, so let's just describe what we do now.	2022-04-28 15:46:44 +02:00
Sho Iizuka	17cfd6f96f	man: how to unset CPUQuota= This description will help users who are trying to reset the already configured CPUQuota= by trying incorrect ways such as CPUQuota=0 or CPUQUota=infinity.	2021-12-13 19:43:56 +00:00
Lennart Poettering	49e9218ae3	Merge pull request #20768 from pdmorrow/shutdown_cgroup_ctrl cgroups: apply StartupAllowedCPUs= and StartupAllowedMemoryNodes= during shutdown	2021-09-27 13:44:54 +02:00
Zbigniew Jędrzejewski-Szmek	a14e028e86	man: cross-reference DeviceAllow= and PrivateDevices= They are somewhat similar, but not easy to discover, esp. considering that they are described in different pages. For PrivateDevices=, split out the first paragraph that gives the high-level overview. (The giant second paragraph could also use some heavy editing to break it up into more digestible chunks, alas.)	2021-09-27 09:19:02 +02:00
Peter Morrow	058a2d8f13	man: Startup* updates for systemd.resource-control All Startup*= directives now also apply to the shutdown phase as well as boot phase.	2021-09-24 15:09:54 +01:00
Peter Morrow	c93a7d4ad3	docs: update docs with StartupAllowedCPUs and StartupAllowedMemoryNodes details Signed-off-by: Peter Morrow <pemorrow@linux.microsoft.com>	2021-09-15 09:52:12 +01:00
Yu Watanabe	d4e30ad1fb	tree-wide: fix typo	2021-08-22 09:46:22 +01:00
Mauricio Vásquez	795ccb03e0	man: add RestrictNetworkInterfaces= documentation Signed-off-by: Mauricio Vásquez <mauricio@kinvolk.io>	2021-08-18 15:55:54 -05:00
Julia Kartseva	120338ae33	man: document ip proto in SocketBind{Allow\|Deny}=	2021-06-30 00:36:33 -07:00
Lennart Poettering	7dbc38db50	man: explicit say for priority/weight values whether more is more or less Fixes: #17523	2021-05-26 12:42:13 +01:00
Lennart Poettering	f80a206aa4	socket-bind: use lowercase "ipv4"/"ipv6" spelling In most of our codebase when we referenced "ipv4" and "ipv6" on the right-hand-side of an assignment, we lowercases it (on the left-hand-side we used CamelCase, and thus "IPv4" and "IPv6"). In particular all across the networkd codebase the various "per-protocol booleans" use the lower-case spelling. Hence, let's use lower-case for SocketBindAllow=/SocketBindDeny= too, just make sure things feel like they belong together better. (This work is not included in any released version, hence let's fix this now, before any fixes in this area would be API breakage) Follow-up for #17655	2021-05-11 15:37:31 +02:00
Julia Kartseva	6359811021	man: add SocketBind{Allow\|Deny}= documentation	2021-04-26 16:26:28 -07:00
Julia Kartseva	ee08909059	man: add BPFProgram= documentation	2021-04-09 20:28:47 -07:00
Zbigniew Jędrzejewski-Szmek	34507fa9e9	man: remove details of ManagedOOMPreference implementation	2021-02-25 21:14:04 +01:00
Zbigniew Jędrzejewski-Szmek	a8136f1bc0	man: advertise shared drop-ins more systemd.unit(5) is a wall of text. And this particular feature can be very useful in the context of resource control. Let's avertise this cool feature a bit more. Fixes #17900.	2021-02-25 21:14:04 +01:00
Zbigniew Jędrzejewski-Szmek	326152af4d	man: use markup more in description of ManagedOOMPreference= Follow-up for `d8a4d64bc3`.	2021-02-25 21:14:04 +01:00
Anita Zhang	d8a4d64bc3	man: document ManagedOOMPreference=	2021-02-12 12:46:22 -08:00
Anita Zhang	0a9f93443b	oom: rework *MemoryPressureLimit= properties to have 1/10000 precision Requested in https://github.com/systemd/systemd/pull/15206#discussion_r505506657, preserve the full granularity for memory pressure limits (permyriad) instead of capping out at percent.	2021-02-02 17:52:48 -08:00
Pavel Hrdina	16455ee2b1	man: fix small issue in AllowedMemoryNodes description It should not mention "CPU" but "NUMA nodes".	2021-01-30 18:19:17 +01:00
Zbigniew Jędrzejewski-Szmek	75909cc7e4	man: various typos and other small issues Fixes #18397.	2021-01-29 08:42:39 +01:00
Yu Watanabe	db9ecf0501	license: LGPL-2.1+ -> LGPL-2.1-or-later	2020-11-09 13:23:58 +09:00
Anita Zhang	cf3e57884e	man: document systemd-oomd and related items	2020-10-09 02:40:19 -07:00
Lennart Poettering	037857507a	man: fix xml tags	2020-08-20 13:19:01 +02:00
Benjamin Berg	29bb3d7fc4	man: Improve MemoryMin=/MemoryLow= description The description didn't really explain how the distribution mechanism works exactly and the relationship of leaf and slice units. Update the documentation and also explicitly explain the expected behaviour as it is created by the memory_recursiveprot cgroup2 mount option.	2020-08-19 11:17:02 +02:00
Lennart Poettering	6b000af4f2	tree-wide: avoid some loaded terms https://tools.ietf.org/html/draft-knodel-terminology-02 https://lwn.net/Articles/823224/ This gets rid of most but not occasions of these loaded terms: 1. scsi_id and friends are something that is supposed to be removed from our tree (see #7594) 2. The test suite defines an API used by the ubuntu CI. We can remove this too later, but this needs to be done in sync with the ubuntu CI. 3. In some cases the terms are part of APIs we call or where we expose concepts the kernel names the way it names them. (In particular all remaining uses of the word "slave" in our codebase are like this, it's used by the POSIX PTY layer, by the network subsystem, the mount API and the block device subsystem). Getting rid of the term in these contexts would mean doing some major fixes of the kernel ABI first. Regarding the replacements: when whitelist/blacklist is used as noun we replace with with allow list/deny list, and when used as verb with allow-list/deny-list.	2020-06-25 09:00:19 +02:00
Lennart Poettering	92d64d1444	man: s/PROGRAMM/PROGRAM/	2020-06-23 17:13:26 +02:00
Zbigniew Jędrzejewski-Szmek	e1a0423266	man: reword description of IPAddressDeny/Allow a bit	2020-05-26 11:13:06 +02:00
Anita Zhang	5403e15337	man: update list of supported controllers	2020-03-05 13:53:29 +09:00
Lennart Poettering	f27a21d48b	man: document the limits of the block device discovery for IO cgroup options Fixes: #14271	2020-01-17 10:08:13 +01:00
Zbigniew Jędrzejewski-Szmek	246be82bd4	man: link to specific sections of cgroups-v2 document The document is rather huge, and a specific link is easier to consume. The form is a bit strange because troff puts the symlink at the bottom, keyed by title, so we need to use the same link target in all places.	2020-01-09 16:47:34 +01:00
Zbigniew Jędrzejewski-Szmek	bb6d563a50	doc: link to html versions of cgroup docs Also stop linking to some (obsolete) v1 documentation.	2020-01-09 16:47:34 +01:00
Lennart Poettering	3a827125e7	man: stop recommending modprobe -abq in ExecStartPre=	2020-01-07 19:00:56 +01:00
Zbigniew Jędrzejewski-Szmek	f8b68539d0	man: fix a few bogus entries in directives index When wrong element types are used, directives are sometimes placed in the wrong section. Also, strip part of text starting with "'", which is used in a few places and which is displayed improperly in the index.	2019-11-21 22:06:30 +01:00
Chris Down	ba79e19cb2	cgroup: docs: memory.high doc fixups The docs just tautologically call this the "high limit". Just call it throttling as we do in cgroup-v2.rst.	2019-09-30 14:30:14 +01:00
Chris Down	b62087d4d0	cgroup: docs: Mention unbounded protection for memory.{low,min} I got asked why Memory{Low,Min} don't allow "infinity". They do, but the docs don't say that like they already do for Memory{High,Max}.	2019-09-30 14:23:32 +01:00
Pavel Hrdina	047f5d63d7	cgroup: introduce support for cgroup v2 CPUSET controller Introduce support for configuring cpus and mems for processes using cgroup v2 CPUSET controller. This allows users to limit which cpus and memory NUMA nodes can be used by processes to better utilize system resources. The cgroup v2 interfaces to control it are cpuset.cpus and cpuset.mems where the requested configuration is written. However, it doesn't mean that the requested configuration will be actually used as parent cgroup may limit the cpus or mems as well. In order to reflect the real configuration cgroup v2 provides read-only files cpuset.cpus.effective and cpuset.mems.effective which are exported to users as well.	2019-09-24 15:16:07 +02:00
Lennart Poettering	3ff668cb9a	man: reword DeviceAllow= documentation Don't claim we'd use cgroup.deny much. It's just a way to remove stuff from device lists, which is nothing we allow users to explicitly configure. Also, extend documentation when wildcards may be used, and when not.	2019-07-31 16:06:15 +02:00
Lennart Poettering	00d85bbb60	man: document the modprobe hack for DeviceAllow=	2019-07-23 13:30:56 +02:00
Kai Lüke	fab347489f	bpf-firewall: custom BPF programs through IP(Ingress\|Egress)FilterPath= Takes a single /sys/fs/bpf/pinned_prog string as argument, but may be specified multiple times. An empty assignment resets all previous filters. Closes https://github.com/systemd/systemd/issues/10227	2019-06-25 09:56:16 +02:00
Chris Down	acdb4b5236	cgroup: Polish hierarchically aware protection docs a bit I missed adding a section in `systemd.resource-control` about DefaultMemoryMin in #12332. Also, add a NEWS entry going over the general concept.	2019-05-08 12:06:32 +01:00
Chris Down	c52db42b78	cgroup: Implement default propagation of MemoryLow with DefaultMemoryLow In cgroup v2 we have protection tunables -- currently MemoryLow and MemoryMin (there will be more in future for other resources, too). The design of these protection tunables requires not only intermediate cgroups to propagate protections, but also the units at the leaf of that resource's operation to accept it (by setting MemoryLow or MemoryMin). This makes sense from an low-level API design perspective, but it's a good idea to also have a higher-level abstraction that can, by default, propagate these resources to children recursively. In this patch, this happens by having descendants set memory.low to N if their ancestor has DefaultMemoryLow=N -- assuming they don't set a separate MemoryLow value. Any affected unit can opt out of this propagation by manually setting `MemoryLow` to some value in its unit configuration. A unit can also stop further propagation by setting `DefaultMemoryLow=` with no argument. This removes further propagation in the subtree, but has no effect on the unit itself (for that, use `MemoryLow=0`). Our use case in production is simplifying the configuration of machines which heavily rely on memory protection tunables, but currently require tweaking a huge number of unit files to make that a reality. This directive makes that significantly less fragile, and decreases the risk of misconfiguration. After this patch is merged, I will implement DefaultMemoryMin= using the same principles.	2019-04-12 17:23:58 +02:00
Lennart Poettering	ef81ce6e80	man: clarify which addresses are affected by IPAddressAllow=/IPAddressDeny= For ingress traffic it's the source address of IP packets we check, for egress traffic it's the destination address. Mention that.	2019-03-29 16:17:55 +01:00
Zbigniew Jędrzejewski-Szmek	3a54a15760	man: use same header for all files The "include" files had type "book" for some raeason. I don't think this is meaningful. Let's just use the same everywhere. $ perl -i -0pe 's^..DOCTYPE (book\|refentry) PUBLIC "-//OASIS//DTD DocBook XML V4.[25]//EN"\s+"http^<!DOCTYPE refentry PUBLIC "-//OASIS//DTD DocBook XML V4.5//EN"\n "http^gms' man/*.xml	2019-03-14 14:42:05 +01:00
Zbigniew Jędrzejewski-Szmek	0307f79171	man: standarize on one-line license header No need to waste space, and uniformity is good. $ perl -i -0pe 's\|\n+<!--\sSPDX-License-Identifier: LGPL-2.1..\s-->\|\n<!-- SPDX-License-Identifier: LGPL-2.1+ -->\|gms' man/*.xml	2019-03-14 14:29:37 +01:00
Filipe Brandenburger	10f2864111	core: add CPUQuotaPeriodSec= This new setting allows configuration of CFS period on the CPU cgroup, instead of using a hardcoded default of 100ms. Tested: - Legacy cgroup + Unified cgroup - systemctl set-property - systemctl show - Confirmed that the cgroup settings (such as cpu.cfs_period_ns) were set appropriately, including updating the CPU quota (cpu.cfs_quota_ns) when CPUQuotaPeriodSec= is updated. - Checked that clamping works properly when either period or (quota * period) are below the resolution of 1ms, or if period is above the max of 1s.	2019-02-14 11:04:42 -08:00
Yu Watanabe	d1698b82e6	man: add referecne to systemd-system.conf	2019-02-01 12:31:51 +01:00
Chris Down	c72703e26d	cgroup: Add DisableControllers= directive to disable controller in subtree Some controllers (like the CPU controller) have a performance cost that is non-trivial on certain workloads. While this can be mitigated and improved to an extent, there will for some controllers always be some overheads associated with the benefits gained from the controller. Inside Facebook, the fix applied has been to disable the CPU controller forcibly with `cgroup_disable=cpu` on the kernel command line. This presents a problem: to disable or reenable the controller, a reboot is required, but this is quite cumbersome and slow to do for many thousands of machines, especially machines where disabling/enabling a stateful service on a machine is a matter of several minutes. Currently systemd provides some configuration knobs for these in the form of `[Default]CPUAccounting`, `[Default]MemoryAccounting`, and the like. The limitation of these is that Default*Accounting is overrideable by individual services, of which any one could decide to reenable a controller within the hierarchy at any point just by using a controller feature implicitly (eg. `CPUWeight`), even if the use of that CPU feature could just be opportunistic. Since many services are provided by the distribution, or by upstream teams at a particular organisation, it's not a sustainable solution to simply try to find and remove offending directives from these units. This commit presents a more direct solution -- a DisableControllers= directive that forcibly disallows a controller from being enabled within a subtree.	2018-12-03 15:40:31 +00:00
Lennart Poettering	077c40bc52	man: link Delegate= documentation up with the markdown docs	2018-11-26 18:43:23 +01:00
Lennart Poettering	964c4eda5b	man: also use "yes"/"no" rather than "true"/"false" in man pages We usually use yes/no in all our unit files, do the same in the man pages. Triggered by: https://github.com/systemd/systemd/pull/9824#issuecomment-420729987	2018-10-13 12:59:29 +02:00
Tejun Heo	6ae4283cb1	core: add IODeviceLatencyTargetSec This adds support for the following proposed latency based IO control mechanism. https://lkml.org/lkml/2018/6/5/428	2018-08-22 16:46:18 +02:00
Ryutaroh Matsumoto	be60dd3ec8	Various accountings are not implied by their controllers The original manpage says "Implies BBBAccounting" many times but actually that accounting is not implied by the respective resource control in v239 with the unified cgroup hierarchy. This commit removes those false explanations.	2018-07-20 16:44:40 +02:00
Chen Qi	49bdfaba92	man/systemd.resource-control.xml: point user to correct url cpu.cfs_quota_us is actually explained in sched-bwc.txt instead of sched-design-CFS.txt.	2018-07-18 13:17:24 +02:00
Tejun Heo	4842263577	core: add MemoryMin The kernel added support for a new cgroup memory controller knob memory.min in bf8d5d52ffe8 ("memcg: introduce memory.min") which was merged during v4.18 merge window. Add MemoryMin to support memory.min.	2018-07-12 08:21:43 +02:00
Zbigniew Jędrzejewski-Szmek	514094f933	man: drop mode line in file headers This is already included in .dir-locals, so we don't need it in the files themselves.	2018-07-03 01:32:25 +02:00
Lennart Poettering	30ce657e5d	Merge pull request #9301 from keszybz/man-drop-authorgroup man: drop unused <authorgroup> tags from man sources	2018-06-14 15:29:24 +02:00
Zbigniew Jędrzejewski-Szmek	0cd41d4dff	Drop my copyright headers perl -i -0pe 's/\sCopyright © .... Zbigniew Jędrzejewski.?\n/\n/gms' man/xml git grep -e 'Copyright.Jędrzejewski' -l \| xargs perl -i -0pe 's/(#\n)?# +Copyright © [0-9, -]+ Zbigniew Jędrzejewski.?\n//gms' git grep -e 'Copyright.Jędrzejewski' -l \| xargs perl -i -0pe 's/\s\/\\\\s+Copyright © [0-9, -]+ Zbigniew Jędrzejewski[^\n]?\s\\\\/\s/\n\n/gms' git grep -e 'Copyright.Jędrzejewski' -l \| xargs perl -i -0pe 's/\s+Copyright © [0-9, -]+ Zbigniew Jędrzejewski[^\n]//gms'	2018-06-14 13:03:20 +02:00
Zbigniew Jędrzejewski-Szmek	fdbbee37d5	man: drop unused <authorgroup> tags from man sources Docbook styles required those to be present, even though the templates that we use did not show those names anywhere. But something changed semi-recently (I would suspect docbook templates, but there was only a minor version bump in recent years, and the changelog does not suggest anything related), and builds now work without those entries. Let's drop this dead weight. Tested with F26-F29, debian unstable. $ perl -i -0pe 's/\s<authorgroup>.<.authorgroup>//gms' man/*xml	2018-06-14 12:22:18 +02:00
Lennart Poettering	96b2fb93c5	tree-wide: beautify remaining copyright statements Let's unify an beautify our remaining copyright statements, with a unicode ©. This means our copyright statements are now always formatted the same way. Yay.	2018-06-14 10:20:21 +02:00
Lennart Poettering	818bf54632	tree-wide: drop 'This file is part of systemd' blurb This part of the copyright blurb stems from the GPL use recommendations: https://www.gnu.org/licenses/gpl-howto.en.html The concept appears to originate in times where version control was per file, instead of per tree, and was a way to glue the files together. Ultimately, we nowadays don't live in that world anymore, and this information is entirely useless anyway, as people are very welcome to copy these files into any projects they like, and they shouldn't have to change bits that are part of our copyright header for that. hence, let's just get rid of this old cruft, and shorten our codebase a bit.	2018-06-14 10:20:20 +02:00
Zbigniew Jędrzejewski-Szmek	11a1589223	tree-wide: drop license boilerplate Files which are installed as-is (any .service and other unit files, .conf files, .policy files, etc), are left as is. My assumption is that SPDX identifiers are not yet that well known, so it's better to retain the extended header to avoid any doubt. I also kept any copyright lines. We can probably remove them, but it'd nice to obtain explicit acks from all involved authors before doing that.	2018-04-06 18:58:55 +02:00
Zbigniew Jędrzejewski-Szmek	2f75b05c24	man: IPAccounting for slices in now allowed Also split that description into paragraphs by subject.	2018-02-22 14:53:55 +01:00
Lennart Poettering	99f3baa983	man: clarify that the controllers listed on Delegate= might not be the only ones	2017-11-21 11:54:08 +01:00
Zbigniew Jędrzejewski-Szmek	572eb058cf	Add SPDX license identifiers to man pages	2017-11-19 19:08:15 +01:00
Zbigniew Jędrzejewski-Szmek	c12ad58c41	man: remove note about CPU controller being unmerged https://git.kernel.org/pub/scm/linux/kernel/git/torvalds/linux.git/commit/?id=0d5936344f30aba0f6ddb92b030cb6a05168efe6 In principle we shouldn't merge this until after 4.15 is released, but the chances of a revert upstream are low, and in that unlikely scenario we can just revert this patch, it's a trivial documentation update after all.	2017-11-19 14:15:42 +01:00
Yu Watanabe	1bdfc7b951	core/cgroup: assigning empty string to Delegate= resets list of controllers (#7336 ) Before this, assigning empty string to Delegate= makes no change to the controller list. This is inconsistent to the other options that take list of strings. After this, when empty string is assigned to Delegate=, the list of controllers is reset. Such behavior is consistent to other options and useful for drop-in configs. Closes #7334.	2017-11-17 10:04:25 +01:00
Lennart Poettering	a9f01ad1bf	man: document the new Delegate= syntax	2017-11-13 10:49:15 +01:00
Jakub Wilk	dcfaecc70a	man: fix typos (#7029 )	2017-10-10 21:59:03 +02:00
Daniel Mack	8d8631d4c9	man: document the new ip accounting and filting directives	2017-09-22 15:24:55 +02:00
Zbigniew Jędrzejewski-Szmek	1245e4132b	man: use "filename" not "file name" by default We settled on "filename" and "file system", so change a couple of places for consistency. The exception is when there's an adjective before "file" that binds more strongly then "name": "password file name", "output file name", etc. Those cases are left intact.	2017-09-15 17:18:29 +02:00
John Lin	45f09f939b	man: explicitly distinguish "implicit dependencies" and "default dependencies" Fixes: #6793	2017-09-13 11:39:09 +08:00
AsciiWolf	28a0ad81ee	man: use https:// in URLs	2017-02-21 16:28:04 +01:00
Lennart Poettering	c7458f9399	man: avoid abbreviated "cgroups" terminology (#4396 ) Let's avoid the overly abbreviated "cgroups" terminology. Let's instead write: "Linux Control Groups (cgroups)" is the long form wherever the term is introduced in prose. Use "control groups" in the short form wherever the term is used within brief explanations. Follow-up to: #4381	2016-10-17 09:50:26 -04:00
Zbigniew Jędrzejewski-Szmek	74b47bbd5d	man: add crosslink between systemd.resource-control(5) and systemd.exec(5) Fixes #4379.	2016-10-15 18:38:20 -04:00
Tejun Heo	7d862ab8c2	core: make settings for unified cgroup hierarchy supersede the ones for legacy hierarchy (#4269 ) There are overlapping control group resource settings for the unified and legacy hierarchies. To help transition, the settings are translated back and forth. When both versions of a given setting are present, the one matching the cgroup hierarchy type in use is used. Unfortunately, this is more confusing to use and document than necessary because there is no clear static precedence. Update the translation logic so that the settings for the unified hierarchy are always preferred. systemd.resource-control man page is updated to reflect the change and reorganized so that the deprecated settings are at the end in its own section.	2016-10-14 21:07:16 -04:00
WaLyong Cho	96e131ea09	core: introduce MemorySwapMax= Similar to MemoryMax=, MemorySwapMax= limits swap usage. This controls controls "memory.swap.max" attribute in unified cgroup.	2016-08-30 11:11:45 +09:00
Tejun Heo	66ebf6c0a1	core: add cgroup CPU controller support on the unified hierarchy Unfortunately, due to the disagreements in the kernel development community, CPU controller cgroup v2 support has not been merged and enabling it requires applying two small out-of-tree kernel patches. The situation is explained in the following documentation. https://git.kernel.org/cgit/linux/kernel/git/tj/cgroup.git/tree/Documentation/cgroup-v2-cpu.txt?h=cgroup-v2-cpu While it isn't clear what will happen with CPU controller cgroup v2 support, there are critical features which are possible only on cgroup v2 such as buffered write control making cgroup v2 essential for a lot of workloads. This commit implements systemd CPU controller support on the unified hierarchy so that users who choose to deploy CPU controller cgroup v2 support can easily take advantage of it. On the unified hierarchy, "cpu.weight" knob replaces "cpu.shares" and "cpu.max" replaces "cpu.cfs_period_us" and "cpu.cfs_quota_us". [Startup]CPUWeight config options are added with the usual compat translation. CPU quota settings remain unchanged and apply to both legacy and unified hierarchies. v2: - Error in man page corrected. - CPU config application in cgroup_context_apply() refactored. - CPU accounting now works on unified hierarchy.	2016-08-07 09:45:39 -04:00
Zbigniew Jędrzejewski-Szmek	0d5299ef5a	Merge pull request #3843 from maxime1986/minor-systemd.resource-control	2016-07-31 21:15:17 -04:00
Maxime de Roucy	c23b2c70bf	documentation: cgroup-v1 and systemd user instance Explain in the systemd.resource-control man that systemd user instance can't use resource control on cgroup-v1.	2016-07-31 15:00:59 +02:00
Maxime de Roucy	65c1cdb282	documentation: add cgroup-v2.txt link add cgroup-v2.txt link in section "Unified and Legacy Control Group Hierarchies" of systemd.resource-control man.	2016-07-31 14:38:56 +02:00
Lennart Poettering	83f8e80857	core: support percentage specifications on TasksMax= This adds support for a TasksMax=40% syntax for specifying values relative to the system's configured maximum number of processes. This is useful in order to neatly subdivide the available room for tasks within containers.	2016-07-22 15:33:12 +02:00
Lennart Poettering	328583dbc3	man: minor fixes	2016-06-14 19:50:38 +02:00
Lennart Poettering	875ae5661a	core: optionally, accept a percentage value for MemoryLimit= and related settings If a percentage is used, it is taken relative to the installed RAM size. This should make it easier to write generic unit files that adapt to the local system.	2016-06-14 19:50:38 +02:00
Tejun Heo	e57c9ce169	core: always use "infinity" for no upper limit instead of "max" (#3417 ) Recently added cgroup unified hierarchy support uses "max" in configurations for no upper limit. While consistent with what the kernel uses for no upper limit, it is inconsistent with what systemd uses for other controllers such as memory or pids. There's no point in introducing another term. Update cgroup unified hierarchy support so that "infinity" is the only term that systemd uses for no upper limit.	2016-06-03 17:49:05 +02:00
Tejun Heo	da4d897e75	core: add cgroup memory controller support on the unified hierarchy (#3315 ) On the unified hierarchy, memory controller implements three control knobs - low, high and max which enables more useable and versatile control over memory usage. This patch implements support for the three control knobs. * MemoryLow, MemoryHigh and MemoryMax are added for memory.low, memory.high and memory.max, respectively. * As all absolute limits on the unified hierarchy use "max" for no limit, make memory limit parse functions accept "max" in addition to "infinity" and document "max" for the new knobs. * Implement compatibility translation between MemoryMax and MemoryLimit. v2: - Fixed missing else's in config_parse_memory_limit(). - Fixed missing newline when writing out drop-ins. - Coding style updates to use "val > 0" instead of "val". - Minor updates to documentation.	2016-05-27 18:10:18 +02:00
Tejun Heo	538b48524c	core: translate between IO and BlockIO settings to ease transition Due to the substantial interface changes in cgroup unified hierarchy, new IO settings are introduced. Currently, IO settings apply only to unified hierarchy and BlockIO to legacy. While the transition is necessary, it's painful for users to have to provide configs for both. This patch implements translation from one config set to another for configs which make sense. * The translation takes place during application of the configs. Users won't see IO or BlockIO settings appearing without being explicitly created. * The translation takes place only if there is no config for the matching cgroup hierarchy type at all. While this doesn't provide comprehensive compatibility, it should considerably ease transition to the new IO settings which are a superset of BlockIO settings. v2: - Update test-cgroup-mask.c so that it accounts for the fact that CGROUP_MASK_IO and CGROUP_MASK_BLKIO move together. Also, test/parent.slice now sets IOWeight instead of BlockIOWeight.	2016-05-18 17:35:12 -07:00
Tejun Heo	ac06a0cf8a	core: add support for IOReadIOPSMax and IOWriteIOPSMax cgroup IO controller supports maximum limits for both bandwidth and IOPS but systemd resource control currently only supports bandwidth limits. This patch adds support for IOReadIOPSMax and IOWriteIOPSMax when unified cgroup hierarchy is in use. It isn't difficult to also add BlockIOReadIOPS and BlockIOWriteIOPS for legacy hierarchies but IO control on legacy hierarchies is half-broken anyway, so let's leave it alone for now.	2016-05-18 13:50:56 -07:00
Lennart Poettering	0069a0dd14	man: clarify that IOXyz= only applies to the unified hierarchy, and BlockIOXyz= to the legacy hierarchy With this change for each setting we say which hierarachy it applies to briefly in the first sentence of the description, plus in longer form in an extra pargraph at the end, with a recommendation for the counterpart of the option in the other hierarchy. Also adds markup and the "=" suffix to all mentioned settings.	2016-05-16 22:48:45 +02:00
Tejun Heo	13c31542cc	core: add io controller support on the unified hierarchy On the unified hierarchy, blkio controller is renamed to io and the interface is changed significantly. * blkio.weight and blkio.weight_device are consolidated into io.weight which uses the standardized weight range [1, 10000] with 100 as the default value. * blkio.throttle.{read\|write}_{bps\|iops}_device are consolidated into io.max. Expansion of throttling features is being worked on to support work-conserving absolute limits (io.low and io.high). * All stats are consolidated into io.stats. This patchset adds support for the new interface. As the interface has been revamped and new features are expected to be added, it seems best to treat it as a separate controller rather than trying to expand the blkio settings although we might add automatic translation if only blkio settings are specified. * io.weight handling is mostly identical to blkio.weight[_device] handling except that the weight range is different. * Both read and write bandwidth settings are consolidated into CGroupIODeviceLimit which describes all limits applicable to the device. This makes it less painful to add new limits. * "max" can be used to specify the maximum limit which is equivalent to no config for max limits and treated as such. If a given CGroupIODeviceLimit doesn't contain any non-default configs, the config struct is discarded once the no limit config is applied to cgroup. * lookup_blkio_device() is renamed to lookup_block_device(). Signed-off-by: Tejun Heo <htejun@fb.com>	2016-05-05 16:43:06 -04:00
Martin Pitt	5e939dd6a4	man: fix cgroup attributes for device throttling	2016-04-05 15:28:47 +02:00
Martin Pitt	c51fa94772	man: update links to kernel.org cgroup documentation This recently moved from /cgroups/ to /cgroup-v1/. Fixes #2958	2016-04-05 10:48:06 +02:00
Daniel Mack	50f48ad37a	cgroup: remove support for NetClass= directive Support for net_cls.class_id through the NetClass= configuration directive has been added in v227 in preparation for a per-unit packet filter mechanism. However, it turns out the kernel people have decided to deprecate the net_cls and net_prio controllers in v2. Tejun provides a comprehensive justification for this in his commit, which has landed during the merge window for kernel v4.5: https://git.kernel.org/cgit/linux/kernel/git/torvalds/linux.git/commit/?id=bd1060a1d671 As we're aiming for full support for the v2 cgroup hierarchy, we can no longer support this feature. Userspace tool such as nftables are moving over to setting rules that are specific to the full cgroup path of a task, which obsoletes these controllers anyway. This commit removes support for tweaking details in the net_cls controller, but keeps the NetClass= directive around for legacy compatibility reasons.	2016-02-10 16:38:56 +01:00
Lennart Poettering	ae0a5fb1e1	man: document special considerations when mixing templated service units and DefaultDependencies=no Fixes #2189.	2016-01-29 16:50:50 +01:00
Lennart Poettering	0af20ea2ee	core: add new DefaultTasksMax= setting for system.conf This allows initializing the TasksMax= setting of all units by default to some fixed value, instead of leaving it at infinity as before.	2015-11-13 19:50:52 +01:00
Lennart Poettering	c129bd5df3	man: document automatic dependencies For all units ensure there's an "Automatic Dependencies" section in the man page, and explain which dependencies are automatically added in all cases, and which ones are added on top if DefaultDependencies=yes is set. This is also done for systemd.exec(5), systemd.resource-control(5) and systemd.unit(5) as these pages describe common behaviour of various unit types.	2015-11-11 20:47:07 +01:00
Jan Engelhardt	b938cb902c	doc: correct punctuation and improve typography in documentation	2015-11-06 13:00:02 +01:00
Lennart Poettering	d817000dea	man: move documentation about NetClass from systemd.unit(5) to systemd.resource-control(5) This is after all where we expose all the other cgroup props, especially those that can be adjusted dynamically.	2015-10-19 23:07:18 +02:00
Lennart Poettering	d53d94743c	core: refactor cpu shares/blockio weight cgroup logic Let's stop using the "unsigned long" type for weights/shares, and let's just use uint64_t for this, as that's what we expose on the bus. Unify parsers, and always validate the range for these fields. Correct the default blockio weight to 500, since that's what the kernel actually uses. When parsing the weight/shares settings from unit files accept the empty string as a way to reset the weight/shares value. When getting it via the bus, uniformly map (uint64_t) -1 to unset. Open up StartupCPUShares= and StartupBlockIOWeight= to transient units.	2015-09-11 18:31:49 +02:00
Lennart Poettering	03a7b521e3	core: add support for the "pids" cgroup controller This adds support for the new "pids" cgroup controller of 4.3 kernels. It allows accounting the number of tasks in a cgroup and enforcing limits on it. This adds two new setting TasksAccounting= and TasksMax= to each unit, as well as a gloabl option DefaultTasksAccounting=. This also updated "cgtop" to optionally make use of the new kernel-provided accounting. systemctl has been updated to show the number of tasks for each service if it is available. This patch also adds correct support for undoing memory limits for units using a MemoryLimit=infinity syntax. We do the same for TasksMax= now and hence keep things in sync here.	2015-09-10 18:41:06 +02:00

1 2 3 4

170 Commits