Open Source LLM

Mistral AI發表Mistral Large模型及Le Chat聊天機器人,企圖與OpenAI較勁

甫於2023年4月才創立的Mistral AI於本週一(2/26)發表了大型語言模型Mistral Large,以及聊天機器人Le Chat預覽版。

Open Source LLM

Amazon Bedrock將上架Mistral 7B和Mixtral 8x7B兩開源語言模型

AWS 宣佈 將在雲端人工智慧模型平台Bedrock提供Mistral AI的模型,分別是 Mistral 7B Mixtral 8x7B ,而這將使Mistral AI成為繼AI21 Labs、Anthropic、Meta、Stability AI等廠商之後,第7個Bedrock基礎模型供應商,用户將有更多的大型語言模型選擇,以滿足開發各種人工


Open Source LLM

Microsoft inks deal with France’s Mistral AI, an OpenAI rival that has its own chatbot

Microsoft announced an partnership Monday with the French startup Mistral AI that could lessen the software giant’s reliance on ChatGPT-maker OpenAI for supplying the next wave of chatbots and other generative AI products.

Mistral AI emerged less than a year ago but is already what Microsoft described Monday as an “innovator and trailblazer” at the vanguard of building more efficient and cost-effective AI systems.

Microsoft and Mistral didn’t disclose the financial terms of the deal, though Microsoft said it involves a small investment in the Paris-


Open Source LLM

Intel demonstrates PyTorch AI optimizations for accelerating large language models on its Arc Alchemist GPUs

(Image credit: Tom's Hardware)

Intel's Arc Alchemist GPUs can run large language models like Llama 2, thanks to the company's PyTorch extension, as demoed in a recent blog post . The Intel PyTorch Extension, which works on both Windows and Linux, allows LLMs to take advantage of the FP16 performance on Arc GPUs. However, given that Intel says you'll need 14GB of VRAM to use Llama 2 on Intel hardware, it means you'll probably want an

Continue reading...

Open Source LLM

Mesa OpenGL Threading Work Sees Much Reduced Memory Footprint For OpenGL Calls

Longtime AMD open-source Mesa developer Marek Olšák after more than one decade working officially for AMD and years before that as an independent open-source contributor going back to the R300g days still has not run out of new performance optimizations to pursue. The most recent accomplishment for this leading Mesa contributor are some refinements to the OpenGL threading "glthread" code for lowering the memory footprint.

Marek noted in a recent and since merged pull request for Mesa 24.1:
"The biggest
Continue reading...

Open Source LLM

Intel Optimizes PyTorch for Llama 2 for Arc A770 GPU, Uses Higher Precision FP16

Intel just announced optimizations for PyTorch (IPEX) to take advantage of the AI acceleration features of its Arc "Alchemist" GPUs.PyTorch is a popular machine learning library that is often associated with NVIDIA GPUs, but it is actually platform-agnostic. It can be run on a variety of hardware, including CPUs and GPUs. However, performance may not be optimal without specific optimizations. Intel offers such optimizations through the Intel Extension for PyTorch (IPEX), which extends PyTorch with optimizations specifically designed for Intel's compute

Open Source LLM

Intel Optimizes PyTorch for Llama 2 on Arc A770, Higher Precision FP16

Intel just announced optimizations for PyTorch (IPEX) to take advantage of the AI acceleration features of its Arc "Alchemist" GPUs.PyTorch is a popular machine learning library that is often associated with NVIDIA GPUs, but it is actually platform-agnostic. It can be run on a variety of hardware, including CPUs and GPUs. However, performance may not be optimal without specific optimizations. Intel offers such optimizations through the Intel Extension for PyTorch (IPEX), which extends PyTorch with optimizations specifically designed for Intel's compute

Open Source LLM

GLFW 3.4 Brings Better Support For Wayland & Run-Time Platform Selection

GLFW 3.4 has been released as this open-source, multi-platform library used for OpenGL / OpenGL ES / Vulkan development via a platform-independent API. GLFW 3.4 continues supporting Linux, macOS, Windows, and other platforms for offering this nice abstracted solution around graphics and input.

With Friday's release of GLFW 3.4 there is now support for run-time platform selection, improved support for Wayland, enabling both Wayland and X11 support by default, custom heap allocator support
Continue reading...

Open Source LLM

(PR) Google's Gemma Optimized to Run on NVIDIA GPUs, Gemma Coming to Chat with RTX

NVIDIA, in collaboration with Google, today launched optimizations across all NVIDIA AI platforms for Gemma—Google's state-of-the-art new lightweight 2 billion- and 7 billion-parameter open language models that can be run anywhere, reducing costs and speeding innovative work for domain-specific use cases.

Teams from the companies worked closely together to accelerate the performance of Gemma—built from the same research and technology used to create the Gemini models—with NVIDIA TensorRT-LLM, an open-source library for

Open Source LLM

Predibase發布LoRA Land服務,集結25個微調模型之力效能可勝GPT-4

大型語言模型服務公司Predibase推出了一個名為 LoRA Land 的服務,該服務的特點在於集合了25個經微調的大型語言模型,這些模型都是以開源的Mistral-7b模型作為基礎,Predibase針對不同任務對Mistral-7b進行最佳化,使其在不同領域的任務,能