Open Source LLM

蘋果再公佈二款小模型DCLM

繼4月份公佈可在裝置上執行的小語言模型 OpenELM 後,蘋果本週又 公佈 了14億及70億參數的DCLM模型,號稱效能不輸競爭模型如Llama 3、Gemma或Mistral,或是更節省訓練運算資源。

這二款模型是由蘋果DataComp for Language Models(DCLM)團隊開發,並在Hugging Face平


Open Source LLM

(PR) Gigabyte AI TOP Utility Reinventing Your Local AI Fine-tuning

GIGABYTE TECHNOLOGY Co. Ltd, a leading manufacturer of motherboards, graphics cards, and hardware solutions, released the GIGABYTE exclusive groundbreaking AI TOP Utility. With reinvented workflows, user-friendly interface, and real-time progress monitoring, AI TOP Utility provides a reinventing touch of local AI model training and fine-tuning. It features a variety of groundbreaking technologies that can be easily adapted by beginners or experts, for most common open-source LLMs, in anyplace even on your desk.

GIGABYTE AI TOP is the all-
Continue reading...

Open Source LLM

wlroots 0.18 Brings New Wayland Protocols & Support For GPU Reset Recovery

Wlroots 0.18 recently debuted as the newest version of this Wayland library born out of the Sway compositor project. With wlroots 0.18 is support for new Wayland protocols and other exciting features.

New protocols enabled by wlroots 0.18 include linux-drm-syncobj-v1 for explicit synchronization, alpha-modifier-v1 for alpha channel support on surfaces, ext-foreign-toplevel-list-v1 as a protocol for taskbars and app switchers, and ext-transient-seat
Continue reading...

Open Source LLM

微軟開發SpreadsheetLLM以讓LLM更容易理解試算表資料

微軟的AI研究團隊近日發表了一份名為 SpreadsheetLLM的研究報告 ,旨在協助大型語言模型(LLM)更容易理解試算表(Spreadsheet)中的資料,以便正確推論。

試算表由許多儲存格組成,這些儲存格存放著文字、數字、公式或函數,嵌入在許多的行與


Open Source LLM

Former Tesla AI Director reproduces GPT-2 in 24 hours for only $672 — GPT-4 costs $100 million to train

(Image credit: Shutterstock)

OpenAI launched GPT-2 in 2019, reportedly costing $256 per hour to train. However, it’s been five years since then, and we’re already at GPT-4o. Advancements in hardware, software, and data mean that training the same model will take less time and less money, as Andrej Karpathy , the developer behind the project to reproduce GPT-2 in llm.c, has proven.

The primary driver of cost savings

Continue reading...

Open Source LLM

Exclusive: OpenAI working on new reasoning technology under code name ‘Strawberry’ - Reuters

Exclusive: OpenAI working on new reasoning technology under code name ‘Strawberry’ Reuters

Open Source LLM

Mold Linker Gains New Option To Deliver "Massively Faster" Performance

The Mold linker is already a high-speed alternative to the likes of LLVM LLD and GNU Gold. Its performance is very impressive while those using it while carrying out debug builds have the ability to achieve an insane speed-up thanks to a new option.

Mold lead developer Rui Ueyama shared that a new "--separate-debug-file" option has been added to the linker for "massively faster" performance. Linking Clang with debug info included can drop down to less than a half second compared to six

Open Source LLM

Microsoft's AI speech generator achieves human parity but is too dangerous for the public

Serving tech enthusiasts for over 25 years.TechSpot means tech analysis and advice you can trust .

Too Real: Microsoft has developed a new iteration of its neural codec language model, Vall-E, that surpasses previous efforts in terms of naturalness, speech robustness, and speaker similarity. It is the first of its kind to reach human parity in a pair of popular benchmarks, and is apparently so lifelike that Microsoft has no plans to grant access to the public.

Leveraging Vall-E'

Continue reading...

Open Source LLM

LPython 0.22 Released For Ahead-Of-Time Compiler For Python

LPython is an in-development open-source project aiming to be a very fast Python compiler with multiple back-ends . Released this week was LPython 0.22 as the latest step in this crusade.

LPython continues striving to be a great ahead-of-time compiler for Python that is written in C++ and aims for optimal performance across platforms as well as aspiring to be able to transform Python code into other languages like C++ and Fortran.

Continue reading...

Open Source LLM

Meta開發10億以下參數量的小型LLM模型MobileLLM

大廠持續投入終端裝置上的AI模型開發。Llama模型家族獲得眾多開發人員使用後, Meta本週稍早又公佈 可在行動裝置上執行,參數量不到10億的新AI模型家族。

由於在雲端執行上百甚至上千億參數的大型語言模型(LLM)增加雲端運

Continue reading...