Open Source LLM

LLVM Adds Additional Protections For Arm's SLS Speculation Vulnerability Mitigation

Revealed earlier this year was the Arm Straight Line Speculation (SLS) vulnerability . SLS was a Google discovery for modern ARMv8 CPUs where speculative execution past unconditional changes in control flow could lead to information disclosure via side-channel analysis. Arm recommended compiler-based mitigations to insert speculation barriers after vulnerable instructions, which GCC and LLVM began adding opt-in protections right away. This weekend some additional SLS functionality was added for LLVM.

On top of the prior SLS mitigation options for LLVM (the initial hardening pass - AArch

Open Source LLM

POCL 1.6 Released For Portable OpenCL Atop CPUs, Other Accelerators

A new feature release of POCL is now available that is the "Portable Computing Language" offering OpenCL execution atop CPUs and other devices like NVIDIA CUDA that have an LLVM back-end.

POCL 1.6 is out as the latest feature release and continues providing OpenCL 1.2 support and a subset of OpenCL 2.0 functionality. POCL is most well known for OpenCL on CPUs but thanks to LLVM also allows targeting NVIDIA GPUs with CUDA, AMD GPUs with HSA, and other possible accelerator targets. POCL makes use

Open Source LLM

OpenBLAS 0.3.13 Released With A RISC-V Port, POWER10 Optimizations

OpenBLAS 0.3.13 was released today as the newest update to this leading open-source BLAS (and LAPACK) implementation.

With OpenBLAS 0.3.13 some of the release highlights include:

- Detection for the Fujitsu Fortran compiler.

- A RISC-V port for the C910V has been added. This initial OpenBLAS RISC-V target is catering to the RISC-V Vector Extension 0.7.1. This is the first RISC-V port by OpenBLAS.

Open Source LLM

AMD AOMP 11.12 Released For OpenMP Offloading To Radeon GPUs

Last week there was the release of AOCC 2.3 as AMD's LLVM Clang downstream focused on Zen-optimized support. Meanwhile on the graphics side of the house, this week ushered in AOMP 11.12 as their LLVM Clang downstream focused on Radeon OpenMP GPU offloading.

AOMP continues maturing as the company's downstream of LLVM Clang that allows OpenMP offloading to Radeon hardware. AMD has been working to upstream their Radeon OMP patches into LLVM, but at least until that's all perfect, AOMP

Open Source LLM

Classic OSMesa Retires In Mesa 21.0 As The Worst Of The Software Rendering Paths

While working on some core Mesa cleaning/improvements, Eric Anholt has retired the classic OSMesa support in next quarter's Mesa 21.0 .

Those wanting Mesa software rendering in 2020 and beyond should really be using LLVMpipe or otherwise Softpipe should LLVM not be available for your software/hardware platform. LLVMpipe offers much better performance not to mention OpenGL 4.6 and is actually maintained. With classic OSMesa code just rotting and being of minimal use these days for off-screen rendering, the classic code

Open Source LLM

Intel AMX Programming Model Lands In LLVM Compiler

One of the big features to look forward to with Intel's Xeon "Sapphire Rapids" is the introduction of AMX as the Advanced Matrix Extensions. While Sapphire Rapids looks to be at least one year out still, the company's open-source compiler engineers have already been hard at work on the software infrastructure support.

AMX is Intel's new programming paradigm with a focus on better AI performance both for training and inference. AMX is built around the concept of "tiles" as a set of two-

Open Source LLM|Intel

Intel Opens Up "IMF LA" As A GPU Compute Speed Boost To Better Compete With Windows

The open-source Intel Graphics Compiler (IGC) that is currently used by their oneAPI Level Zero and OpenCL implementations but likely to see Intel driver Mesa usage in 2021 has a new feature dubbed "IMF LA" that aims to help with the performance and close the gap with Windows.

Released today was IGC 1.0.5761 . This routine update to the Intel Graphics Compiler has a number of low-level compiler additions and other changes as usual. All quite low level but then

Open Source LLM

POCL 1.6-RC1 Released With Better CUDA Performance

POCL as the "Portable Computing Language" that implements OpenCL and allows it to function atop CPUs as well as CUDA-enabled NVIDIA GPUs, HSA-supported AMD GPUs, and other possible back-ends, is preparing for a new feature release.

On Wednesday marked the release of POCL 1.6-RC1 as the test release for the next update to the Portable Computing Language.

POCL continues leveraging LLVM/Clang for doing much of the heavy lifting and supporting a diverse variety of back-ends for CPU execution

Open Source LLM

Clang LTO Support For The Linux Kernel Spun Up A Seventh Time

Google engineers have sent out their latest patches for allowing the mainline Linux kernel to be built with LLVM Clang link-time optimizations (LTO) for greater performance and possibly size benefits.

Google's team has done a good job not only working on the mainline Clang support for the Linux kernel across the likes of AArch64 and x86_64, but also with other related features of interest to them like the Clang LTO abilities to which internally they already leverage extensively. This upstreaming work has been ongoing

Open Source LLM|Apple Silicon|Intel

(PR) LLNL's New 'Ruby' Supercomputer Taps Intel for COVID-19 Research

Intel today announced that Lawrence Livermore National Laboratory (LLNL) will leverage Intel Xeon Scalable processors in "Ruby," its latest high performance computing cluster. The Ruby system will be used for unclassified programmatic work in support of the National Nuclear Security Administration's (NNSA) stockpile stewardship mission, for researching therapeutic drugs and designer antibodies against SARS-CoV-2, the virus that causes COVID-19, and for other open science work at LLNL.

Ruby was built in collaboration with Intel, LLNL, Supermicro and