Open Source LLM

Chinese team wins award for AI booster that may help counter US chip ban

The Mooncake large language learning model is up to five times more efficient than the baseline. Photo: Shutterstock

The team from Beijing-based start-up Moonshot AI and Tsinghua University were handed the Erik Riedel Best Paper Award for their system named Mooncake at the USENIX FAST conference in California last month.

Their large language model (LLM) platform Mooncake helps to reduce dependency on graphics processing units (GPUs) and is up to five times more efficient than the previous baseline,

Continue reading...

Open Source LLM

LLVM 20's Great Fortran Language Support With Flang

With the newly-released LLVM 20.1 compiler stack among the many changes throughout the massive codebase is renaming the "flang-new" compiler just to "flang" . This new Flang compiler front-end has matured quite well over the years to providing robust and reliable Fortran language support within the confines of the LLVM toolchain.

In marking the LLVM 20 stable release and the milestone of flang-new renamed to flang, the LLVM Project blog has put out a blog post outlining all of the accomplishments for

Open Source LLM

OpenAI launches new developer tools as Chinese AI startups gain ground - Reuters

OpenAI launches new developer tools as Chinese AI startups gain ground Reuters

Open Source LLM

ByteDance says new AI technology boosts model training efficiency by 1.7 times

TikTok owner ByteDance said it has achieved a 1.71 times efficiency improvement in large language model (LLM) training, the latest Chinese tech company to achieve a breakthrough that could potentially reduce demand for Nvidia’s high-end graphics processing units (GPUs).

The company’s Doubao development team said they managed to “speed up” LLM training efficiency by “1.71 times” through COMET, an optimised Mixture-of-Experts (MoE) system, according to a recent paper published on

The headquarters of ByteDance is seen in Beijing on September 16, 2020. Photo: AFP

Open Source LLM|Intel

Foxconn launches traditional Chinese large language model for AI-driven manufacturing

Foxconn chairman Young Liu delivers the closing speech at the Hon Hai Tech Day in Taipei, October 8, 2024. Photo: AFP

Foxconn Technology Group, the world’s largest electronics contract manufacturer and major iPhone supplier for Apple, launched its first Chinese large language model (LLM) trained on traditional characters, as the Taiwanese company pushes forward the use of artificial intelligence (AI) in factories.

The new FoxBrain model was trained in a “more efficient and lower-cost” method within

Continue reading...

Open Source LLM

AI之戰︱台半導體巨擘鴻海推首款繁體中文模型 認與DeepSeek有差距

台灣半導體巨擘鴻海星期一宣布,推出首款繁體中文人工智能(AI)大型語言模型,並稱該模型雖與大陸AI公司深度求索(DeepSeek)的蒸餾模型仍有些微差距,但表現已相當接近世界領先水準。

鴻海星期一(10日)在官網發布新聞


Open Source LLM

AI bots now play Mafia with each other on public website, and almost all of them are terrible at it

(Image credit: Shutterstock)

A developer named "Guzus" has created a website where a selection of AI Language Learning Models (LLMs) can play the classic social deduction game Mafia with one another.

Not only can you see the results of who won each match, you can also view a complete transcript of each game played. This culminates in a full ranking for each LLM, to crown who might be the best at fulfilling every role played in Mafia.

To those unfamiliar, the concept of Mafia is simple.

Continue reading...

Open Source LLM

Unofficial ROCm SDK Builder Expanded To Support More GPUs

The community-based ROCm SDK Builder is an unofficial project leveraging the open-source AMD ROCm code and making it easy to build machine learning and GPU compute software across a range of environments and helping ensure proper integration with other machine learning tools and models. The ROCm SDK Builder takes special focus on the consumer Radeon iGPUs and dGPUs that typically aren't as much of a focus for the upstream AMD ROCm stack.

Open-source developer Mika Laitio announced on Thursday the ROCm SDK Builder 6.1.2 release for
Continue reading...

Open Source LLM

(PR) TerraMaster Adds API-in-a-Box Open-Source AI Application Generator to Its F4-424 Max TNAS Series

TerraMaster, in collaboration with its partner Imagery Business Systems, has launched API-in-a-Box, marking a significant milestone in the field of software and hardware development. This pre-configured open-source solution integrates artificial intelligence (AI) and low-code technology with the TerraMaster F4-424 Max TNAS hardware, aiming to accelerate the creation of full-stack applications. It offers a seamless experience from requirement description to production deployment. Below is an analysis of its detailed features, target audience

Open Source LLM

免費、免 VPN 用 OpenAI o3-mini-high 微軟 Copilot 開放無限制使用、兼用深度思考

Microsoft(微軟)今日宣布,其 Copilot 的 Think Deeper 功能已升級,現採用 OpenAI 的 o3-mini-high 模型。最初此功能在去年 10 月推出,只限 Pro 計劃用戶能夠使用,透過 o1 模型協助解決複雜問題。如今所有用戶可免費無限使用此強化版功能,只需點擊