mirror of
https://git.tartarus.org/simon/putty.git
synced 2025-01-10 09:58:01 +00:00
c1a2114b28
I only recently found out that OpenSSH defined their own protocol IDs for AES-GCM, defined to work the same as the standard ones except that they fixed the semantics for how you select the linked cipher+MAC pair during key exchange. (RFC 5647 defines protocol ids for AES-GCM in both the cipher and MAC namespaces, and requires that you MUST select both or neither - but this contradicts the selection policy set out in the base SSH RFCs, and there's no discussion of how you resolve a conflict between them! OpenSSH's answer is to do it the same way ChaCha20-Poly1305 works, because that will ensure the two suites don't fight.) People do occasionally ask us for this linked cipher/MAC pair, and now I know it's actually feasible, I've implemented it, including a pair of vector implementations for x86 and Arm using their respective architecture extensions for multiplying polynomials over GF(2). Unlike ChaCha20-Poly1305, I've kept the cipher and MAC implementations in separate objects, with an arm's-length link between them that the MAC uses when it needs to encrypt single cipher blocks to use as the inputs to the MAC algorithm. That enables the cipher and the MAC to be independently selected from their hardware-accelerated versions, just in case someone runs on a system that has polynomial multiplication instructions but not AES acceleration, or vice versa. There's a fourth implementation of the GCM MAC, which is a pure software implementation of the same algorithm used in the vectorised versions. It's too slow to use live, but I've kept it in the code for future testing needs, and because it's a convenient place to dump my design comments. The vectorised implementations are fairly crude as far as optimisation goes. I'm sure serious x86 _or_ Arm optimisation engineers would look at them and laugh. But GCM is a fast MAC compared to HMAC-SHA-256 (indeed compared to HMAC-anything-at-all), so it should at least be good enough to use. And we've got a working version with some tests now, so if someone else wants to improve them, they can.
45 lines
1.2 KiB
C
45 lines
1.2 KiB
C
/*
|
|
* Windows implementation of the OS query functions that detect Arm
|
|
* architecture extensions.
|
|
*/
|
|
|
|
#include "putty.h"
|
|
|
|
#if !(defined _M_ARM || defined _M_ARM64)
|
|
/*
|
|
* For non-Arm, stub out these functions. This module shouldn't be
|
|
* _called_ in that situation anyway, but it will still be compiled
|
|
* (because that's easier than getting CMake to identify the
|
|
* architecture in all cases).
|
|
*/
|
|
#define IsProcessorFeaturePresent(...) false
|
|
#endif
|
|
|
|
bool platform_aes_neon_available(void)
|
|
{
|
|
return IsProcessorFeaturePresent(PF_ARM_V8_CRYPTO_INSTRUCTIONS_AVAILABLE);
|
|
}
|
|
|
|
bool platform_pmull_neon_available(void)
|
|
{
|
|
return IsProcessorFeaturePresent(PF_ARM_V8_CRYPTO_INSTRUCTIONS_AVAILABLE);
|
|
}
|
|
|
|
bool platform_sha256_neon_available(void)
|
|
{
|
|
return IsProcessorFeaturePresent(PF_ARM_V8_CRYPTO_INSTRUCTIONS_AVAILABLE);
|
|
}
|
|
|
|
bool platform_sha1_neon_available(void)
|
|
{
|
|
return IsProcessorFeaturePresent(PF_ARM_V8_CRYPTO_INSTRUCTIONS_AVAILABLE);
|
|
}
|
|
|
|
bool platform_sha512_neon_available(void)
|
|
{
|
|
/* As of 2020-12-24, as far as I can tell from docs.microsoft.com,
|
|
* Windows on Arm does not yet provide a PF_ARM_V8_* flag for the
|
|
* SHA-512 architecture extension. */
|
|
return false;
|
|
}
|