Open Source LLM

ROCm 6.3 adds several new features including a Fortran compiler, and SGLang

(Image credit: AMD)

AMD has announced ROCm version 6.3 , which adds many new updates to the ROCm ecosystem. The latest iteration of the open-source driver stack features several additions, including SGLang, FlashAttention-2, and a Fortran Compiler.

SGLang is a new runtime in ROCm 6.3 that purportedly improves latency, throughput, and resource utilization by optimizing "cutting-edge" generative AI models on AMD's homebrewed Instinct GPUs. SGLang purportedly achieves up to 6X higher performance on large

Continue reading...

Open Source LLM

Orange enlists Meta and OpenAI to develop AI language models in Africa - Reuters

Orange enlists Meta and OpenAI to develop AI language models in Africa Reuters

Open Source LLM

AMD Releases ROCm 6.3 with SGLang, Fortran Compiler, Multi-Node FFT, Vision Libraries, and More

AMD has released the new ROCm 6.3 version which introduces several new features and optimizations, including SGLang integration for accelerated AI inferencing, a re-engineered FlashAttention-2 for optimized AI training and inference, the introduction of multi-node Fast Fourier Transform (FFT), new Fortran compiler, and enhanced computer vision libraries like rocDecode, rocJPEG, and rocAL.

According to AMD, the SGLang, a runtime that is now supported by ROCm 6.3, is purpose-built for optimizing inference on models like LLMs

Open Source LLM

Nvidia shows AI model that can modify voices, generate novel sounds - Reuters

Nvidia shows AI model that can modify voices, generate novel sounds Reuters

Open Source LLM

Luma AI Revamps its Dream Machine Video Generator Letting Users Create Clips From Images

Luma Labs has given its Dream Machine AI video generator a makeover by introducing a new text-to-image foundation model called Photon.

[ Read More ]


Open Source LLM

Bloomberg: Apple developing new ‘LLM Siri’ for iOS 19 and macOS 16

While Apple is in the midst of rolling out new Apple Intelligence features as part of iOS 18, it’s also working on yet-to-be-announced features for iOS 19. One of those features is an upgraded version of Siri powered by more advanced large language models, or LLMs. The news was detailed in a new report today .

Mark Gurman reports that this is being referred to as “LLM Siri” inside Apple. The upgraded assistant is currently being tested as part of a separate app


Open Source LLM

Raspberry Pi 5 successfully accelerates LLMs using an eGPU and Vulkan

(Image credit: Jeff Geerling)

A Raspberry Pi 5 hooked up to an AMD Radeon -powered eGPU has been demonstrated using the graphics hardware to accelerate running a Large Language Model (LLM). Of course, it's Pi wizard Jeff Geerling again, and in the video embedded below, he talks us through his experience of leveraging the Vulkan API support to enjoy GPU-accelerated local AI on the Raspberry Pi 5.

GPU accelerated AI—on a Raspberry Pi - YouTube
Continue reading...

Open Source LLM

Khronos Announces Slang Initiative From Open-Source NVIDIA Code

On top of an exciting Vulkan spec update out today, The Khronos Group has announced the Slang Initiative based on NVIDIA's open-source Slang compiler code.

The Khronos Group will be overseeing the development of the open-source Slang shader language and compiler.
"Slang empowers real-time graphics developers with innovative features that complement existing shading languages, including modular code development, portable deployment to multiple target APIs, and neural computation in graphics shaders. Hosting under multi-company governance at Khronos will enable and foster industry-
Continue reading...

Open Source LLM

(PR) Khronos Group Launches Slang Initiative, Hosting Open Source Compiler Contributed by NVIDIA

The Khronos Group, an open consortium of industry leaders in interoperability standards, has announced the launch of the new Slang Initiative. This initiative will oversee and advance the open-source Slang shading language and compiler, building on 15 years of research, development, and deployment experience. Supported by NVIDIA since 2017, Slang has been widely adopted in production projects across the industry.

Slang empowers real-time graphics developers with innovative features that complement existing shading languages, including modular code development, portable deployment to multiple

Open Source LLM

Cerebras video shows AI writing code 75x faster than world's fastest AI GPU cloud — world's largest chip beats AWS's fastest in head-to-head comparison

(Image credit: Cerebras)

Cerebras got Meta’s Llama 3.1 405B large language model to run at 969 tokens per second, 75 times faster than Amazon Web Services' fastest AI service with GPUs could muster.

The LLM was run on Cerebras’s cloud AI service Cerebras Inference, which uses the chip company’s third-generation Wafer Scale Engines rather than GPUs from Nvidia or AMD. Cerebras has always claimed its Inference service is the fastest for generating tokens, the individual

Continue reading...
Menu