Open Source LLM

NVIDIA updates ChatRTX: new models, voice recognition and media search

ChatRTX, the new demo of local LLM powered by NVIDIA GPUs

NVIDIA has updated the tool now supporting more Large Language Models (LLMs).

If you’re wondering why NVIDIA is launching ChatRTX today when a similar tool has already been released, the confusion might arise from the fact that “Chart with RTX” was just a pre-launch tech demo, whereas ChatRTX is the new app, however still officially a demo.

Unlike online tools such as ChatGPT, Google Bard, or Microsoft Copilot, the AI assistant powered

Continue reading...

Open Source LLM

NVIDIA Unveils ChatRTX, Your Own GPU-accelerated AI Assistant with Photo Recognition, Speech Input, Updated Models

NVIDIA today unveiled ChatRTX, the AI assistant that runs locally on your machine, and which is accelerated by your GeForce RTX GPU. NVIDIA had originally launched this as "Chat with RTX" back in February 2024, back then this was regarded more as a public tech demo. We reviewed the application in our feature article. The ChatRTX rebranding is probably aimed at making the name sound more like ChatGPT, which is what the application aims to be—except it runs completely on your machine, and is exhaustively
Continue reading...

Open Source LLM

We Tested NVIDIA's new ChatRTX: Your Own GPU-accelerated AI Assistant with Photo Recognition, Speech Input, Updated Models

NVIDIA today unveiled ChatRTX, the AI assistant that runs locally on your machine, and which is accelerated by your GeForce RTX GPU. NVIDIA had originally launched this as "Chat with RTX" back in February 2024, back then this was regarded more as a public tech demo. We reviewed the application in our feature article. The ChatRTX rebranding is probably aimed at making the name sound more like ChatGPT, which is what the application aims to be—except it runs completely on your machine, and is exhaustively
Continue reading...

Open Source LLM

XZ Backdoor, Nova Driver, Linux 6.9 Features & Ubuntu 24.04 Made For An Exciting April

April 2024 is now in the books after writing 257 original Linux/open-source-related news articles and another 13 featured articles / Linux hardware reviews. Here's a look back at the most exciting (popular) content from April.

As usual, before getting to the monthly highlights, if you enjoy all of the daily and original content on Phoronix please consider supporting operations by disabling any ad-blockers in your web browser when viewing Phoronix. Or join Phoronix Premium to help support

Open Source LLM

OpenAI to use FT content for training AI models in latest media tie-up - Reuters

OpenAI to use FT content for training AI models in latest media tie-up Reuters

Open Source LLM

Llamafile 0.8.1 GPU LLM Offloading Works Now With More AMD GPUs

It was just a few days ago that Llamafile 0.8 released with LLaMA 3 and Grok support along with faster F16 performance. Now this project out of Mozilla for self-contained, easily re-distributable large language model (LLM) deployments is out with a new release.

Most significant with Friday's Llamafile 0.8.1 release is getting GPU support working for more AMD graphics processors / accelerators. Due to some of the AMD offload code within Llamafile only assuming numeric "GFX" graphics IP

Open Source LLM

蘋果公佈裝置上執行的AI模型OpenELM

在微軟、Meta、Google等相繼公佈AI模型後,蘋果終於發聲, 宣佈並開源可在蘋果裝置端執行的AI模型OpenELM家族 以及訓練及推論框架,最小版本僅2.7億參數。

OpenELM全名為開源高效語言模型(Open-source Efficient Language Model),蘋果已在 Hugging Face公開 了4種參

Continue reading...

Open Source LLM

Ubuntu 24.04 LTS lands with Gnome 46, Linux 6.8, and Raspberry Pi 5 support


The biggest Linux distro release of the year, Ubuntu 24.04 brings a new app center and firmware updater based on Flutter, an updated toolchain, and experimental support for TPM-backed full disk encryption.



Read Entire Article


Open Source LLM

Intel Releases OpenVINO 2024.1 With More Gen AI & LLM Features

Intel engineers have just released OpenVINO 2024.1, the newest feature release for this excellent open-source AI toolkit that continues expanding its features and capabilities particularly around Generative AI "GenAI" and Large Language Models (LLMs).

On the generative AI front, OpenVINO 2024.1 adds Mixtral and URLNet models optimized for Intel Xeon CPUs, Stable Diffusion 1.5 / ChatGLM3-6B / Qwen-7B models have been optimized for faster Intel Core Ultra (Meteor Lake)
Continue reading...

Open Source LLM

Llamafile 0.8 Releases With LLaMA3 & Grok Support, Faster F16 Performance

Llamafile has been quite an interesting project out of Mozilla's Ocho group in the era of AI. Llamafile makes it easy to run and distribute large language models (LLMs) that are self-contained within a single file. Llamafile builds off Llama.cpp and makes it easy to ship an entire LLM as a single file with both CPU and GPU execution support. Llamafile 0.8 is out now to join in on the LLaMA3 fun as well as delivering other model support and enhancing the CPU performance.

Llamafile
Continue reading...