gvisor - Container Runtime Sandbox

Age	Commit message (Collapse)	Author
2019-06-24	Implement /proc/net/tcp.	Rahat Mahmood
	PiperOrigin-RevId: 254854346
2019-06-24	platform/ptrace: specify PTRACE_O_TRACEEXIT for stub-processes	Andrei Vagin
	The tracee is stopped early during process exit, when registers are still available, allowing the tracer to see where the exit occurred, whereas the normal exit notifi? cation is done after the process is finished exiting. Without this option, dumpAndPanic fails to get registers. PiperOrigin-RevId: 254852917
2019-06-24	Use correct statx syscall number for amd64.	Nicolas Lacasse
	The previous number was for the arm architecture. Also change the statx tests to force them to run on gVisor, which would have caught this issue. PiperOrigin-RevId: 254846831
2019-06-24	Allow to change logging options using 'runsc debug'	Fabricio Voznika
	New options are: runsc debug --strace=off\|all\|function1,function2 runsc debug --log-level=warning\|info\|debug runsc debug --log-packets=true\|false Updates #407 PiperOrigin-RevId: 254843128
2019-06-22	Implement statx.	Nicolas Lacasse
	We don't have the plumbing for btime yet, so that field is left off. The returned mask indicates that btime is absent. Fixes #343 PiperOrigin-RevId: 254575752
2019-06-21	Fix the logic for sending zero window updates.	Bhasker Hariharan
	Today we have the logic split in two places between endpoint Read() and the worker goroutine which actually sends a zero window. This change makes it so that when a zero window ACK is sent we set a flag in the endpoint which can be read by the endpoint to decide if it should notify the worker to send a nonZeroWindow update. The worker now does not do the check again but instead sends an ACK and flips the flag right away. Similarly today when SO_RECVBUF is set the SetSockOpt call has logic to decide if a zero window update is required. Rather than do that we move the logic to the worker goroutine and it can check the zeroWindow flag and send an update if required. PiperOrigin-RevId: 254505447
2019-06-21	gvisor/fs: getdents returns 0 if offset is equal to FileMaxOffset	Andrei Vagin
	FileMaxOffset is a special case when lseek(d, 0, SEEK_END) has been called. PiperOrigin-RevId: 254498777
2019-06-21	Remove O(n) lookup on unlink/rename	Michael Pratt
	Currently, the path tracking in the gofer involves an O(n) lookup of child fidRefs. This causes a significant overhead on unlinks in directories with lots of child fidRefs (<4k). In this transition, pathNode moves from sync.Map to normal synchronized maps. There is a small chance of contention in walk, but the lock is held for a very short time (and sync.Map also had a chance of requiring locking). OTOH, sync.Map makes it very difficult to add a fidRef reverse map. PiperOrigin-RevId: 254489952
2019-06-21	Deflake TestSimpleReceive failures due to timeouts	Brad Burlage
	This test will occasionally fail waiting to read a packet. From repeated runs, I've seen it up to 1.5s for waitForPackets to complete. PiperOrigin-RevId: 254484627
2019-06-21	ext4 block group descriptor implementation in disk layout package.	Ayush Ranjan
	PiperOrigin-RevId: 254482180
2019-06-21	Add //pkg/flipcall.	Jamie Liu
	Flipcall is a (conceptually) simple local-only RPC mechanism. Compared to unet, Flipcall does not support passing FDs (support for which will be provided out of band by another package), requires users to establish connections manually, and requires user management of concurrency since each connected Endpoint pair supports only a single RPC at a time; however, it improves performance by using shared memory for data (reducing memory copies) and using futexes for control signaling (which is much cheaper than sendto/recvfrom/sendmsg/recvmsg). PiperOrigin-RevId: 254471986
2019-06-21	Add list of stuck tasks to panic message	Fabricio Voznika
	PiperOrigin-RevId: 254450309
2019-06-21	Update pathNode documentation to reflect reality	Michael Pratt
	Neither fidRefs or children are (directly) synchronized by mu. Remove the preconditions that say so. That said, the surrounding does enforce some synchronization guarantees (e.g., fidRef.renameChildTo does not atomically replace the child in the maps). I've tried to note the need for callers to do this synchronization. I've also renamed the maps to what are (IMO) clearer names. As is, it is not obvious that pathNode.fidRefs is a map of child fidRefs rather than self fidRefs. PiperOrigin-RevId: 254446965
2019-06-21	kernel: call t.mu.Unlock() explicitly in WithMuLocked	Andrei Vagin
	defer here doesn't improve readability, but we know it slower that the explicit call. PiperOrigin-RevId: 254441473
2019-06-21	Delete dangling comment line.	Nicolas Lacasse
	This was from an old comment, which was superseded by the existing comment which is correct. PiperOrigin-RevId: 254434535
2019-06-21	Update comment	Fabricio Voznika
	PiperOrigin-RevId: 254428866
2019-06-20	Close FD on TcpSocketTest loop failure.	Ian Gudger
	This helps prevent the blocking call from getting stuck and causing a test timeout. PiperOrigin-RevId: 254325926
2019-06-20	Deflake TestSIGALRMToMainThread.	Neel Natu
	Bump up the threshold on number of SIGALRMs received by worker threads from 50 to 200. Even with the new threshold we still expect that the majority of SIGALRMs are received by the thread group leader. PiperOrigin-RevId: 254289787
2019-06-20	Preallocate auth.NewAnonymousCredentials() in contexttest.TestContext.	Jamie Liu
	Otherwise every call to, say, fs.ContextCanAccessFile() in a benchmark using contexttest allocates new auth.Credentials, a new auth.UserNamespace, ... PiperOrigin-RevId: 254261051
2019-06-20	Add package docs to seqfile and ramfs	Michael Pratt
	These are the only packages missing docs: https://godoc.org/gvisor.dev/gvisor PiperOrigin-RevId: 254261022
2019-06-20	Unmark amutex_test as flaky.	Rahat Mahmood
	PiperOrigin-RevId: 254254058
2019-06-20	Implement madvise(MADV_DONTFORK)	Neel Natu
	PiperOrigin-RevId: 254253777
2019-06-20	Drop extra character	Michael Pratt
	PiperOrigin-RevId: 254237530
2019-06-19	Deflake SendFileTest_Shutdown.	Ian Gudger
	The sendfile syscall's backing doSplice contained a race with regard to blocking. If the first attempt failed with syserror.ErrWouldBlock and then the blocking file became ready before registering a waiter, we would just return the ErrWouldBlock (even if we were supposed to block). PiperOrigin-RevId: 254114432
2019-06-19	Mark tcp_socket test flaky (for real)	Michael Pratt
	The tag on the binary has no effect. It must be on the test. PiperOrigin-RevId: 254103480
2019-06-19	Deflake mount_test.	Nicolas Lacasse
	Inode ids are only stable across Save/Restore if we have an open FD on the inode. All tests that compare inode ids must therefor hold an FD open. PiperOrigin-RevId: 254086603
2019-06-19	Abort loop on failure	Michael Pratt
	As-is, on failure these will infinite loop, resulting in test timeout instead of failure. PiperOrigin-RevId: 254074989
2019-06-19	Add renamed children pathNodes to target parent	Michael Pratt
	Otherwise future renames may miss Renamed calls. PiperOrigin-RevId: 254060946
2019-06-19	fileOp{On,At} should pass the remaning symlink traversal count.	Nicolas Lacasse
	And methods that do more traversals should use the remaining count rather than resetting. PiperOrigin-RevId: 254041720
2019-06-19	Add MountNamespace to task.	Nicolas Lacasse
	This allows tasks to have distinct mount namespace, instead of all sharing the kernel's root mount namespace. Currently, the only way for a task to get a different mount namespace than the kernel's root is by explicitly setting a different MountNamespace in CreateProcessArgs, and nothing does this (yet). In a follow-up CL, we will set CreateProcessArgs.MountNamespace when creating a new container inside runsc. Note that "MountNamespace" is a poor term for this thing. It's more like a distinct VFS tree. When we get around to adding real mount namespaces, this will need a better naem. PiperOrigin-RevId: 254009310
2019-06-19	Mark tcp_socket test flaky.	Neel Natu
	PiperOrigin-RevId: 253997465
2019-06-18	Attempt to fix TestPipeWritesAccumulate	Fabricio Voznika
	Test fails because it's reading 4KB instead of the expected 64KB. Changed the test to read pipe buffer size instead of hardcode and added some logging in case the reason for failure was not pipe buffer size. PiperOrigin-RevId: 253916040
2019-06-18	Use return values from syscalls in eventfd tests.	Rahat Mahmood
	PiperOrigin-RevId: 253890611
2019-06-18	Kill sandbox process when 'runsc do' exits	Fabricio Voznika
	PiperOrigin-RevId: 253882115
2019-06-18	Add Container/Sandbox args struct for creation	Fabricio Voznika
	There were 3 string arguments that could be easily misplaced and it makes it easier to add new arguments, especially for Container that has dozens of callers. PiperOrigin-RevId: 253872074
2019-06-18	Replace usage of deprecated strtoul/strtoull	Brad Burlage
	PiperOrigin-RevId: 253864770
2019-06-18	Fix PipeTest_Streaming timeout	Fabricio Voznika
	Test was calling Size() inside read and write loops. Size() makes 2 syscalls to return the pipe size, making the test do a lot more work than it should. PiperOrigin-RevId: 253824690
2019-06-18	gvisor/fs: don't update file.offset for sockets, pipes, etc	Andrei Vagin
	sockets, pipes and other non-seekable file descriptors don't use file.offset, so we don't need to update it. With this change, we will be able to call file operations without locking the file.mu mutex. This is already used for pipes in the splice system call. PiperOrigin-RevId: 253746644
2019-06-18	gvisor/kokoro: don't modify tests names in the BUILD file	Andrei Vagin
	PiperOrigin-RevId: 253746380
2019-06-17	gvisor/bazel: use python2 to build runsc-debian	Andrei Vagin
	$ bazel build runsc:runsc-debian File ".../bazel_tools/tools/build_defs/pkg/make_deb.py", line 311, in GetFlagValue: flagvalue = flagvalue.decode('utf-8') AttributeError: 'str' object has no attribute 'decode' make_deb.py is incompatible with Python3. https://github.com/bazelbuild/bazel/issues/8443 PiperOrigin-RevId: 253691923
2019-06-17	Internal change.	gVisor bot
	PiperOrigin-RevId: 253559564
2019-06-14	Enable Receive Buffer Auto-Tuning for runsc.	Bhasker Hariharan
	Updates #230 PiperOrigin-RevId: 253225078
2019-06-13	Add support for TCP receive buffer auto tuning.	Bhasker Hariharan
	The implementation is similar to linux where we track the number of bytes consumed by the application to grow the receive buffer of a given TCP endpoint. This ensures that the advertised window grows at a reasonable rate to accomodate for the sender's rate and prevents large amounts of data being held in stack buffers if the application is not actively reading or not reading fast enough. The original paper that was used to implement the linux receive buffer auto- tuning is available @ https://public.lanl.gov/radiant/pubs/drs/lacsi2001.pdf NOTE: Linux does not implement DRS as defined in that paper, it's just a good reference to understand the solution space. Updates #230 PiperOrigin-RevId: 253168283
2019-06-13	Plumb context through more layers of filesytem.	Ian Gudger
	All functions which allocate objects containing AtomicRefCounts will soon need a context. PiperOrigin-RevId: 253147709
2019-06-13	Fix deadlock in fasync.	Ian Gudger
	The deadlock can occur when both ends of a connected Unix socket which has FIOASYNC enabled on at least one end are closed at the same time. One end notifies that it is closing, calling (waiter.Queue).Notify which takes waiter.Queue.mu (as a read lock) and then calls (FileAsync).Callback, which takes FileAsync.mu. The other end tries to unregister for notifications by calling (FileAsync).Unregister, which takes FileAsync.mu and calls (waiter.Queue).EventUnregister which takes waiter.Queue.mu. This is fixed by moving the calls to waiter.Waitable.EventRegister and waiter.Waitable.EventUnregister outside of the protection of any mutex used in (FileAsync).Callback. The new test is related, but does not cover this particular situation. Also fix a data race on FileAsync.e.Callback. (FileAsync).Callback checked FileAsync.e.Callback under the protection of FileAsync.mu, but the waiter calling (*FileAsync).Callback could not and did not. This is fixed by making FileAsync.e.Callback immutable before passing it to the waiter for the first time. Fixes #346 PiperOrigin-RevId: 253138340
2019-06-13	Implement getsockopt() SO_DOMAIN, SO_PROTOCOL and SO_TYPE.	Rahat Mahmood
	SO_TYPE was already implemented for everything but netlink sockets. PiperOrigin-RevId: 253138157
2019-06-13	Merge pull request #306 from amscanne:add_readme	Shentubot
	PiperOrigin-RevId: 253137638
2019-06-13	Update canonical repository.	Adin Scannell
	This can be merged after: https://github.com/google/gvisor-website/pull/77 or https://github.com/google/gvisor-website/pull/78 PiperOrigin-RevId: 253132620
2019-06-13	Add p9 and unet benchmarks.	Jamie Liu
	PiperOrigin-RevId: 253122166
2019-06-13	Set the HOME environment variable (fixes #293)	Ian Lewis
	runsc will now set the HOME environment variable as required by POSIX. The user's home directory is retrieved from the /etc/passwd file located on the container's file system during boot. PiperOrigin-RevId: 253120627