cadvisor

Author	SHA1	Message	Date
David Ashpole	3166cdae87	add utils/clock dependency	2017-11-21 16:19:57 -08:00
David Ashpole	3d6ad6dd86	on demand metrics	2017-11-20 14:51:04 -08:00
Rohit Agarwal	4a35130019	Collect container-level GPU metrics using NVML. When cAdvisor starts up, it would read the `vendor` files in `/sys/bus/pci/devices/*` to see if any NVIDIA devices (vendor ID: 0x10de) are attached to the node. If no NVIDIA devices are found, this code path would become dormant for the rest of cAdvisor lifetime. If NVIDIA devices are found, we would start a goroutine that would check for the presence of NVML by trying to dynamically load it at regular intervals. We need to do this regular checking instead of doing it just once because it may happen that cAdvisor is started before the NVIDIA drivers and NVML are installed. Once the NVML dynamic loading succeeds, we would use NVML’s query methods to find out how many devices exist on the node and create a map from their minor numbers to their handles and cache that map. The goroutine would exit at this point. If we detected the presence of NVML in the previous step, whenever a new container is detected by cAdvisor, cAdvisor would read the `devices.list` file from the container's devices cgroup. The `devices.list` file lists the major:minor number of all the devices that the container is allowed to access. If we find any device with major number 195 (which is the major number assigned to NVIDIA devices), we would cache the list of corresponding minor numbers for that container. During every housekeeping operation, in addition to collecting all the existing metrics, we will use the cached NVIDIA device minor numbers and the map from minor numbers to device handles to get metrics for GPU devices attached to the container.	2017-11-06 11:54:59 -08:00
Seth Jennings	3ba4699c12	skip subcontainer update on v2 calls	2017-08-17 12:36:38 -05:00
Tim St. Clair	35f7bc5ee7	Move mocks to testing package to remove +build tags Some go tools (e.g. godef, gorename) don't handle +build tags well, so I refactored some packages to remove the test tags from cAdvisor.	2016-07-06 14:15:43 -07:00
Thomas Orozco	2e1f0e2a08	Use a dedicated CpuLoadReader per container This ensures each goroutine is given its own Netlink connection, and presumably avoids having a message destined for one goroutine read by another goroutine.	2016-05-18 09:34:13 +02:00
Lei Xue	dbbe38dfed	re-order the import package	2015-11-30 16:43:22 +08:00
Piotr Szczesniak	90ca5f9286	Moved max_housekeeping and allow_dynamic_housekeeping flags to cadvisor.go	2015-07-21 20:26:57 +02:00
anushree-n	e2e193c1fd	Add metrics caching	2015-07-20 11:24:20 -07:00
Rohit Jnagal	1a2781819e	Separate in-memory cache from storage drivers.	2015-06-02 16:06:01 +00:00
Victor Marmol	bce54ce3f5	Run custom collectors in container housekeeping. This will allow us to register and run custom collectors for each container.	2015-05-04 15:57:18 -07:00
Rohit Jnagal	872546ba3a	Bulk move current info api to info/v1. Making room for info/v2.	2015-03-04 00:47:28 +00:00
Victor Marmol	63f5714db8	Add Start and End to ContainerInfoRequests. This will allow querying a certain time range.	2015-02-22 19:07:40 -08:00
Rohit Jnagal	2406b6c55b	Add derived stats tracking to containers.	2015-02-16 20:44:34 +00:00
Rohit Jnagal	cbdd96a554	Add task load stats to containers. The stats are only populated when cAdvisor is running outside network namespaces. We'll add a different backend to retrieve the same data from within namespaces.	2015-01-16 23:25:22 +00:00
Victor Marmol	e3ab15417c	Ignore update errors for dead containers. This adds an Exists() interface to detect when the container is dead. Before reporting an update error we check is Exists() is true. Some documentation was added as well.	2014-11-22 05:31:11 +08:00
Victor Marmol	62b37eca1d	Adding --log_cadvisor_usage flag to log the resource usage of cAdvisor. This helps during performance analysis.	2014-10-15 08:18:38 -07:00
Satnam Singh	f3c92c381a	More if undos.	2014-09-24 11:07:48 -07:00
Satnam Singh	8551d376d8	Undo changes to if statements as requested by vmarmol. Fix typos in my changes.	2014-09-24 10:59:18 -07:00
Satnam Singh	bae82a583d	A few minor Go style suggestions.	2014-09-24 10:53:52 -07:00
Victor Marmol	e759059a09	Refactor and dependency inject containerData deps	2014-09-22 10:20:54 -07:00
Victor Marmol	1636c3e759	Change ContainerHandlerFactories to decide what containers they support. This allows a ContainerHandlerFactory to register a CanHandle() function which is called to determine whether the factory can handle a particular container. This commit disables being able to run cAdvisor without lmctfy. This should be enabled again with a "no-op" global factory which I would like to do in a separate PR.	2014-07-16 16:48:45 -07:00
Nan Deng	f8542e2a33	check subcontainers	2014-07-07 16:51:49 -07:00
Nan Deng	7df7989849	test get info	2014-07-03 22:25:03 -07:00
Nan Deng	4df4b0ee2e	unit tests for containerData	2014-07-03 22:08:24 -07:00

25 Commits