[RFC,v2.1,06/30] x86/sgx: Support VMA permissions more relaxed than enclave permissions

From: Reinette Chatre <reinette.chatre@intel.com>

From: Reinette Chatre <reinette.chatre@intel.com>

=== Summary ===

An SGX VMA can only be created if its permissions are the same or
weaker than the Enclave Page Cache Map (EPCM) permissions. After VMA
creation this same rule is again enforced by the page fault handler:
faulted enclave pages are required to have equal or more relaxed
EPCM permissions than the VMA permissions.

On SGX1 systems the additional enforcement in the page fault handler
is redundant and on SGX2 systems it incorrectly prevents access.
On SGX1 systems it is unnecessary to repeat the enforcement of the
permission rule. The rule used during original VMA creation will
ensure that any access attempt will use correct permissions.
With SGX2 the EPCM permissions of a page can change after VMA
creation resulting in the VMA permissions potentially being more
relaxed than the EPCM permissions and the page fault handler
incorrectly blocking valid access attempts.

Enable the VMA's pages to remain accessible while ensuring that
the PTEs are installed to match the EPCM permissions but not be
more relaxed than the VMA permissions.

=== Full Changelog ===

An SGX enclave is an area of memory where parts of an application
can reside. First an enclave is created and loaded (from
non-enclave memory) with the code and data of an application,
then user space can map (mmap()) the enclave memory to
be able to enter the enclave at its defined entry points for
execution within it.

The hardware maintains a secure structure, the Enclave Page Cache Map
(EPCM), that tracks the contents of the enclave. Of interest here is
its tracking of the enclave page permissions. When a page is loaded
into the enclave its permissions are specified and recorded in the
EPCM. In parallel the kernel maintains permissions within the
page table entries (PTEs) and the rule is that PTE permissions
are not allowed to be more relaxed than the EPCM permissions.

A new mapping (mmap()) of enclave memory can only succeed if the
mapping has the same or weaker permissions than the permissions that
were vetted during enclave creation. This is enforced by
sgx_encl_may_map() that is called on the mmap() as well as mprotect()
paths. This rule remains.

One feature of SGX2 is to support the modification of EPCM permissions
after enclave initialization. Enclave pages may thus already be part
of a VMA at the time their EPCM permissions are changed resulting
in the VMA's permissions potentially being more relaxed than the EPCM
permissions.

Allow permissions of existing VMAs to be more relaxed than EPCM
permissions in preparation for dynamic EPCM permission changes
made possible in SGX2.  New VMAs that attempt to have more relaxed
permissions than EPCM permissions continue to be unsupported.

Reasons why permissions of existing VMAs are allowed to be more relaxed
than EPCM permissions instead of dynamically changing VMA permissions
when EPCM permissions change are:
1) Changing VMA permissions involve splitting VMAs which is an
   operation that can fail. Additionally changing EPCM permissions of
   a range of pages could also fail on any of the pages involved.
   Handling these error cases causes problems. For example, if an
   EPCM permission change fails and the VMA has already been split
   then it is not possible to undo the VMA split nor possible to
   undo the EPCM permission changes that did succeed before the
   failure.
2) The kernel has little insight into the user space where EPCM
   permissions are controlled from. For example, a RW page may
   be made RO just before it is made RX and splitting the VMAs
   while the VMAs may change soon is unnecessary.

Remove the extra permission check called on a page fault
(vm_operations_struct->fault) or during debugging
(vm_operations_struct->access) when loading the enclave page from swap
that ensures that the VMA permissions are not more relaxed than the
EPCM permissions. Since a VMA could only exist if it passed the
original permission checks during mmap() and a VMA may indeed
have more relaxed permissions than the EPCM permissions this extra
permission check is no longer appropriate.

With the permission check removed, ensure that PTEs do
not blindly inherit the VMA permissions but instead the permissions
that the VMA and EPCM agree on. PTEs for writable pages (from VMA
and enclave perspective) are installed with the writable bit set,
reducing the need for this additional flow to the permission mismatch
cases handled next.

Signed-off-by: Reinette Chatre <reinette.chatre@intel.com>
---
 Documentation/x86/sgx.rst      | 10 +++++++++
 arch/x86/kernel/cpu/sgx/encl.c | 38 ++++++++++++++++++----------------
 2 files changed, 30 insertions(+), 18 deletions(-)

Message ID	20220304093524.397485-6-jarkko@kernel.org (mailing list archive)
State	New, archived
Headers	show Return-Path: <linux-sgx-owner@kernel.org> X-Spam-Checker-Version: SpamAssassin 3.4.0 (2014-02-07) on aws-us-west-2-korg-lkml-1.web.codeaurora.org Received: from vger.kernel.org (vger.kernel.org [23.128.96.18]) by smtp.lore.kernel.org (Postfix) with ESMTP id 8C2A5C433EF for <linux-sgx@archiver.kernel.org>; Fri, 4 Mar 2022 09:39:08 +0000 (UTC) Received: (majordomo@vger.kernel.org) by vger.kernel.org via listexpand id S238690AbiCDJjx (ORCPT <rfc822;linux-sgx@archiver.kernel.org>); Fri, 4 Mar 2022 04:39:53 -0500 Received: from lindbergh.monkeyblade.net ([23.128.96.19]:55326 "EHLO lindbergh.monkeyblade.net" rhost-flags-OK-OK-OK-OK) by vger.kernel.org with ESMTP id S239247AbiCDJh6 (ORCPT <rfc822;linux-sgx@vger.kernel.org>); Fri, 4 Mar 2022 04:37:58 -0500 Received: from ams.source.kernel.org (ams.source.kernel.org [145.40.68.75]) by lindbergh.monkeyblade.net (Postfix) with ESMTPS id 4716B1AA070; Fri, 4 Mar 2022 01:36:31 -0800 (PST) Received: from smtp.kernel.org (relay.kernel.org [52.25.139.140]) (using TLSv1.2 with cipher ECDHE-RSA-AES256-GCM-SHA384 (256/256 bits)) (No client certificate requested) by ams.source.kernel.org (Postfix) with ESMTPS id D9D23B827BC; Fri, 4 Mar 2022 09:36:29 +0000 (UTC) Received: by smtp.kernel.org (Postfix) with ESMTPSA id 2FA44C340E9; Fri, 4 Mar 2022 09:36:28 +0000 (UTC) DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/simple; d=kernel.org; s=k20201202; t=1646386588; bh=Owr4D7Nct+1o653wDWG1zRmvQTP1ToeylD32o3q9WB8=; h=From:To:Cc:Subject:Date:In-Reply-To:References:From; b=QAH0mSjPpOP8C2zYMLCL3nEvWbsi4edFwFZ33RIrE8W0rJswsuVg54RtpEysYx6hQ W4J29m94IeW4Fp/v8Gchb4JuKYw4jYnNlaJOnzPSdUCyKXLeo4HfjW+kss+U7eFfBX /b/Y/63hpGiVygbI6HKpe2J6jLdHooPk4XGUdN3WTIOQyVH6fEWgTnfzXGGvE406D7 iNtaQrpTQmpukx0eb5AgBiRvle2vYnZLeKmp3MoNKUEAwNAdI3WLuajy7D52CxxBuB 7jXfwpdbJtQnlX0G26dBfT8z8JilGS5T3+bVCTBFA49ahwKagm7M+jIt6SjUAJP7t7 gD9sK9sB4I2gg== From: Jarkko Sakkinen <jarkko@kernel.org> To: linux-sgx@vger.kernel.org Cc: Nathaniel McCallum <nathaniel@profian.com>, Reinette Chatre <reinette.chatre@intel.com>, Jarkko Sakkinen <jarkko@kernel.org>, Dave Hansen <dave.hansen@linux.intel.com>, Thomas Gleixner <tglx@linutronix.de>, Ingo Molnar <mingo@redhat.com>, Borislav Petkov <bp@alien8.de>, x86@kernel.org (maintainer:X86 ARCHITECTURE (32-BIT AND 64-BIT)), "H. Peter Anvin" <hpa@zytor.com>, Jonathan Corbet <corbet@lwn.net>, linux-kernel@vger.kernel.org (open list:X86 ARCHITECTURE (32-BIT AND 64-BIT)), linux-doc@vger.kernel.org (open list:DOCUMENTATION) Subject: [RFC PATCH v2.1 06/30] x86/sgx: Support VMA permissions more relaxed than enclave permissions Date: Fri, 4 Mar 2022 11:35:00 +0200 Message-Id: <20220304093524.397485-6-jarkko@kernel.org> X-Mailer: git-send-email 2.35.1 In-Reply-To: <20220304093524.397485-1-jarkko@kernel.org> References: <20220304093524.397485-1-jarkko@kernel.org> MIME-Version: 1.0 Content-Transfer-Encoding: 8bit Precedence: bulk List-ID: <linux-sgx.vger.kernel.org> X-Mailing-List: linux-sgx@vger.kernel.org
Series	[RFC,v2.1,01/30] x86/sgx: Add short descriptions to ENCLS wrappers \| expand [RFC,v2.1,01/30] x86/sgx: Add short descriptions to ENCLS wrappers [RFC,v2.1,02/30] x86/sgx: Add wrapper for SGX2 EMODPR function [RFC,v2.1,03/30] x86/sgx: Add wrapper for SGX2 EMODT function [RFC,v2.1,04/30] x86/sgx: Add wrapper for SGX2 EAUG function [RFC,v2.1,05/30] Documentation/x86: Document SGX permission details [RFC,v2.1,06/30] x86/sgx: Support VMA permissions more relaxed than enclave permissions [RFC,v2.1,07/30] x86/sgx: Add pfn_mkwrite() handler for present PTEs [RFC,v2.1,08/30] x86/sgx: Export sgx_encl_ewb_cpumask() [RFC,v2.1,09/30] x86/sgx: Rename sgx_encl_ewb_cpumask() as sgx_encl_cpumask() [RFC,v2.1,10/30] x86/sgx: Move PTE zap code to new sgx_zap_enclave_ptes() [RFC,v2.1,11/30] x86/sgx: Make sgx_ipi_cb() available internally [RFC,v2.1,12/30] x86/sgx: Create utility to validate user provided offset and length [RFC,v2.1,13/30] x86/sgx: Keep record of SGX page type [RFC,v2.1,14/30] x86/sgx: Support restricting of enclave page permissions [RFC,v2.1,15/30] selftests/sgx: Add test for EPCM permission changes [RFC,v2.1,16/30] selftests/sgx: Add test for TCS page permission changes [RFC,v2.1,17/30] x86/sgx: Support adding of pages to an initialized enclave [RFC,v2.1,18/30] x86/sgx: Tighten accessible memory range after enclave initialization [RFC,v2.1,19/30] selftests/sgx: Test two different SGX2 EAUG flows [RFC,v2.1,20/30] x86/sgx: Support modifying SGX page type [RFC,v2.1,21/30] x86/sgx: Support complete page removal [RFC,v2.1,22/30] Documentation/x86: Introduce enclave runtime management section [RFC,v2.1,23/30] selftests/sgx: Introduce dynamic entry point [RFC,v2.1,24/30] selftests/sgx: Introduce TCS initialization enclave operation [RFC,v2.1,25/30] selftests/sgx: Test complete changing of page type flow [RFC,v2.1,26/30] selftests/sgx: Test faulty enclave behavior [RFC,v2.1,27/30] selftests/sgx: Test invalid access to removed enclave page [RFC,v2.1,28/30] selftests/sgx: Test reclaiming of untouched page [RFC,v2.1,29/30] x86/sgx: Free up EPC pages directly to support large page ranges [RFC,v2.1,30/30] selftests/sgx: Page removal stress test

[RFC,v2.1,06/30] x86/sgx: Support VMA permissions more relaxed than enclave permissions

Commit Message

Patch