Open Source LLM

具備1,760億個參數的語言模型BLOOM開源了

由AI新創Hugging Face主導並協調的 BigScience專案於本週公佈了成果 ,釋出具備1,760億個參數的大型語言模型BLOOM (BigScience Large Open-science Open-access Multilingual Language Model),其規模大過OpenAI GPT-3的1,750億個參數, 而且開放各界下載與存取

BigScience專案是在去年5月


Open Source LLM|Intel

(PR) Intel Releases Open Source AI Reference Kits

Intel has released the first set of open source AI reference kits specifically designed to make AI more accessible to organizations in on-prem, cloud and edge environments. First introduced at Intel Vision, the reference kits include AI model code, end-to-end machine learning pipeline instructions, libraries and Intel oneAPI components for cross-architecture performance. These kits enable data scientists and developers to learn how to deploy AI faster and more easily across healthcare, manufacturing, retail and other industries with higher accuracy, better performance and lower


Open Source LLM

LibreOffice 7.4 RC1 Available For Testing This Latest Open-Source Office Suite

The Document Foundation has made available this morning the LibreOffice 7.4 release candidate as the newest test version of this cross-platform, open-source office suite.

LibreOffice 7.4 is aiming for release in mid-August and this RC1 release is one of the final chances to test this half-year update to this leading free software office suite. LibreOffice 7.4 finally brings WebP image support, a variety of performance improvements, and other evolutionary enhancements.

LibreOffice 7.4 has seen work on various
Continue reading...

Open Source LLM

Tesseract OCR 5.2 Engine Finds Success With AVX-512F

As time goes on more open-source projects are beginning to make better use of AVX-512 support even though it's no longer enabled in the latest Alder Lake processors. After reporting on the big AVX-512 wins for JSON parsing with simdjson , another open-source project finding gains is the Tesseract optical character recognition (OCR) engine.

Tesseract 5.2 was released on Wednesday as the newest feature release to to this open-source OCR engine that has been in development going back
Continue reading...

Open Source LLM

Meta開源能翻譯200種語言的AI模型

Meta昨(6)日宣佈 完成開發能翻譯200種語言的機器翻譯單一AI模型NLLB 200,同時將把該模型及訓練用的資料集一同開源出來。

Meta為實現元宇宙跨語言互動而開發的高品質機器翻譯系統NLLB(No Language Left Behind),該公司宣稱,最新完


Open Source LLM

Yandex開源具備1,000億個參數的YaLM 100B語言模型

俄羅斯最大網路公司Yandex週四(6/23)開源了具備1,000億個參數的YaLM 100B語言模型,宣稱這是全球最大的類生成型已訓練變換模型(Generative Pre-trained Transformer,GPT)的神經網路。

嚴格説來YaLM 100B並不是最大的開源語言模型,


Open Source LLM

NLP AI 訓練更簡單更快速 – Cerebras 宣佈已利用擁有 85 萬核心數量的世界最大晶片進行 NLP 訓練

近日,世界知名加速器晶片公司 Cerebras 宣佈,他們已利用「巨型晶片」為 AI 訓練踏上了重要的一步,訓練出了單晶片最大的 NLP (自然語言處理) AI 模型,而這個模型具有 20 億個參數。

這塊全世界最大的加速器晶片採用 7nm 製程,由

Continue reading...

Open Source LLM

Russia's Yandex opens public access to AI large language model - Reuters.com

  • Summary
  • Companies
  • This content was produced in Russia where the law restricts coverage of Russian military operations in Ukraine

MOSCOW, June 23 (Reuters) - Russian technology company Yandex (YNDX.O) said on Thursday it had made a large language model for artificial intelligence research open to the public, hoping to spawn faster and deeper development of certain AI technologies.

Large language models, which have become a key trend in AI, are powerful programs that can generate paragraphs of text and mimic human conversation.

Yandex, like many Russian


Open Source LLM

Java Benchmarks: OpenJDK 8 Through OpenJDK 19 EA, OpenJ9, GraalVM CE

Stemming from a recent reader request around seeing some fresh OpenJDK performance benchmarks, here are benchmarks of OpenJDK 9 through OpenJDK 18 plus the early access OpenJDK 19 builds. Additionally, OpenJ9 and GraalVM CE were tossed in as alternative implementations.

For satisfying this reader request an Intel Core i5 12600K with Ubuntu 22.04 LTS was used for this round of Java JVM benchmarks. All of the OpenJDK builds tested were obtained from the official OpenJDK binaries and using the latest releases at


Open Source LLM

(PR) Cerebras Systems Sets Record for Largest AI Models Ever Trained on A Single Device

Cerebras Systems, the pioneer in high performance artificial intelligence (AI) computing, today announced, for the first time ever, the ability to train models with up to 20 billion parameters on a single CS-2 system - a feat not possible on any other single device. By enabling a single CS-2 to train these models, Cerebras reduces the system engineering time necessary to run large natural language processing (NLP) models from months to minutes. It also eliminates one of the most painful aspects of NLP—