Open Source LLM

Nvidia co-organizes a contest to help build AI dataset to accelerate GPU design

(Image credit: Shutterstock)

Despite their impressive capabilities in generating content, large language models (LLMs) are not so great at designing hardware. Believing this weakness is due to a lack of hardware design data to train the models, Nvidia, Georgia Institute of Technology, and others have organized a contest to help create the needed open-source public dataset.

Nvidia’s director of design automation research, Haoxing (Mark) Ren, recently announced the collaboration on X (formerly known as Twitter). Ren said the

Continue reading...

Open Source LLM

蘋果展示多模態AI訓練框架4M、支援21種模態的AI模型

蘋果本週 公開展示 具備文字、聲音、圖像理解能力的多模態AI模型訓練框架4M,及支援21種模態資料的多模態模型。

4M模型框架全名為極多模態掩碼模型(Massively Multimodal Masked Modeling),為蘋果與瑞士洛桑聯邦理工學院(EPFL)合作開發。研


Open Source LLM

Local Chinese firms rush to fill the AI void after OpenAI abandons China market — Tencent, Baidu, and Alibaba sweeten LLM offers

(Image credit: Shutterstock)

OpenAI made a surprise announcement to its Chinese users that it will leave the market next week, July 9, 2024. Although this move will negatively impact companies relying on access to OpenAI's API for their products, it will also drive innovation and development in local AI models. DigiTimes Asia reported that local tech companies like Tencent, Baidu, and Alibaba are already moving into the market, offering "migration" plans and discounts to affected users.

Another large language model (

Continue reading...

Open Source LLM

Hugging Face第二屆LLM排行榜出爐,中國LLM表現出色

機器學習模型與資料集共享平台 Hugging Face上週公佈 第二屆的開源大型語言模型(LLM)排行榜,表現最佳的是由阿里巴巴所釋出的Qwen 2,且在前十名的LLM中,就有5個來自中國。

Hugging Face主要使用六大測試基準,包括大規模的多工語言理解


Open Source LLM

Servo Web Engine Gets WebGPU Running On OpenGL ES & Other New Features

The Rust-written Servo web layout engine continues progressing for this open-source project now stewarded by the Linux Foundation Europe and seeing code contribution from a range of developers. They have published their June 2024 status update to outline the latest accomplishments for this alternative web engine.

Servo continues moving ahead with a focus on being an embed-friendly web engine as well as on offering up a simple browser thus far for demonstrating its capabilities. Servo has picked up a number of new features in the past month including
Continue reading...

Open Source LLM

Meta發表LLM Compiler,以最佳化程式碼生成及編譯器能力

Meta週四(6/27) 發表了LLM Compiler ,此為奠基於程式碼生成模型 Code Llama 的新模型,額外強化了對編譯器中介語言(IR)、組合語言及最佳化技術的理解,可用來改善所生成的程式碼品質, 目前已可透過Hugging face取得

Meta表示 ,LLM Compiler模型能


Open Source LLM

Chinese AI models storm Hugging Face's LLM chatbot benchmark leaderboard — Alibaba runs the board as major US competitors have worsened

(Image credit: Shutterstock)

Hugging Face has released its second LLM leaderboard to rank the best language models it has tested. The new leaderboard seeks to be a more challenging uniform standard for testing open large language model (LLM) performance across a variety of tasks. Alibaba's Qwen models appear dominant in the leaderboard's inaugural rankings, taking three spots in the top ten.

Pumped to announce the brand new open LLM leaderboard. We burned 300 H100 to re-run new evaluations like

Continue reading...

Open Source LLM

OpenAI signs multi-year content deal with Time magazine - Reuters

OpenAI signs multi-year content deal with Time magazine Reuters

Open Source LLM

AMD's AOMP 19.0-2 Compiler Brings Zero-Copy For CPU-GPU Unified Shared Memory

AMD compiler engineers have released AOMP 19.0-2 as the newest version of their downstream LLVM/Clang compiler that carries all of their latest work around OpenMP/AOCC GPU device offloading to Radeon and Instinct hardware. With this updated AOMP compiler is now run-time support for zero-copy with CPU-GPU unified shared memory and various other new features for this GPU/accelerator-focused compiler.

The big new feature of AOMP 19.0-2 is "significant" run-time feature work
Continue reading...

Open Source LLM

Energy-efficient AI model could be a game changer, 50 times better efficiency with no performance hit

Serving tech enthusiasts for over 25 years.
TechSpot means tech analysis and advice you can trust .

Cutting corners: Researchers from the University of California, Santa Cruz, have devised a way to run a billion-parameter-scale large language model using just 13 watts of power – about as much as a modern LED light bulb. For comparison, a data center-grade GPU used for LLM tasks requires around 700 watts.

AI up to this point has largely been a race to be first, with

Continue reading...