Open Source LLM

Elon Musk teases Grok 3.5 hours after Alibaba’s Qwen3 generates buzz amid US-China AI race

Elon Musk teased the latest version of xAI’s Grok generative AI model hours after Alibaba’s new Qwen3 model started gaining attention from developers. Photo: Reuters

On Tuesday, Alibaba released the third generation of its Qwen family of AI models, with multiple versions, each featuring a different number of parameters. The largest model, with 235 billion parameters, outperformed DeepSeek-R1 and OpenAI’s o1 reasoning models, according to Alibaba, owner of the Post. At 600 million parameters, the most efficient version of the model could be capable of running on a smartphone,


Open Source LLM

Meta introduces Llama application programming interface to attract AI developers - Reuters

Meta introduces Llama application programming interface to attract AI developers Reuters

Open Source LLM

Programmer develops method to run Llama 2 locally on DOS in a weekend

Serving tech enthusiasts for over 25 years.
TechSpot means tech analysis and advice you can trust .

Vintage Hallucinations: A lone developer spent a weekend attempting to run the Llama 2 large language model on old, DOS-based machines. Thanks to the readily available open-source code, the project ultimately succeeded. However, adapting Llama 2 to the archaic DOS environment was no easy feat.

Yeo Kheng Meng, a programmer previously known for creating a DOS client for ChatGPT, has recently embarked on a new AI-related


Open Source LLM

Alibaba unveils Qwen3 AI models that it says outperform DeepSeek R1

Alibaba has released the third generation of its Qwen AI model series. Photo: Reuters

Alibaba Group Holding on Tuesday unveiled the highly anticipated third generation of its open-source artificial intelligence (AI) model series, which promises faster processing and enhanced multilingual capabilities, intensifying competition in an already crowded Chinese market.

The Qwen3 family consists of eight models, ranging from 600 million parameters to 235 billion, with enhancements across all models, according to the Qwen team at Alibaba’s cloud computing unit. Alibaba

Continue reading...

Open Source LLM

OpenAI rolls out new shopping features with ChatGPT search update - Reuters

OpenAI rolls out new shopping features with ChatGPT search update Reuters

Open Source LLM

DeepSeek speculation swirls online over Chinese AI start-up’s much-anticipated R2 model

DeepSeek’s R2 artificial intelligence model was said to have been trained on a server cluster of Huawei Technologies’ Ascend 910B chips, according to Chinese social-media posts. Photo: Shutterstock

The latest speculation about DeepSeek-R2 – the successor to the R1 , reasoning model, which was released in January – that surfaced over the weekend included the product’s imminent launch and the purported new benchmarks it set for cost-efficiency and performance.
That reflects heightened online interest in DeepSeek after it generated worldwide attention from late December 2024 to January by consecutively releasing two advanced open-source AI models, V3
Continue reading...

Open Source LLM

Mold 2.38 Linker Adds Support For LLVM's CREL Format

Mold 2.38 is out this weekend as the latest feature update to this open-source, high-speed linker.

The most notable feature of Mold 2.38 is introducing experimental support for CREL, the experimental relocation table format currently being developed within the LLVM tree. CREL was originally known as RELLEB and is a compact relocation format for ELF files. CREL is much more efficient than the likes of REL and RELA for ELF files. Mold 2.38 adds initial support for reading object files with
Continue reading...

Open Source LLM

Intel Enabling Ultra Low Latency Scheduling "ULLS" For Lunar Lake GPU Compute

While last week Intel released an update Compute Runtime for GPU compute with the OpenCL and Level Zero APIs on Windows and Linux, today they released a new preview version for readying a shiny new feature: Ultra Low Latency Scheduling "ULLS" for Lunar Lake Xe2 graphics.

Intel engineers have been working on Ultra Low Latency Scheduling "ULLS" as a feature for a while to allow direct submission of work to the GPU for bypassing some of the driver overhead and helping with lower latency for compute kernels. ULLS is also referred
Continue reading...

Open Source LLM

Baidu offers new AI models with enhanced features, lower cost than DeepSeek’s products

Baidu has made available to developers its latest artificial intelligence models, Ernie 4.5 Turbo and X1 Turbo. Photo: Shutterstock

in Wuhan, China

At a developer conference held in Wuhan, capital of central Hubei province, Baidu co-founder, chairman and chief executive Robin Li Yanhong unveiled the multimodal Ernie 4.5 Turbo and X1 Turbo reasoning models.
According to Li, Ernie 4.5 Turbo costs about 40 per cent less than DeepSeek’s namesake V3 large language model (LLM), while X1 Turbo is being offered at a quarter

Open Source LLM|Intel

Intel Updates Its PyTorch Extension With DeepSeek-R1 Support, New Optimizations

Intel today released a new version of the Intel Extension for PyTorch in order to apply optimizations to PyTorch for benefiting Intel's hardware. With the Intel Extension for PyTorch v2.7 release, there is support for new large language models (LLMs) as well as various performance optimizations and other enhancements.

The Intel Extension for PyTorch 2.7 release adds support for the popular DeepSeek-R1 model, including enabling INT8 precision on modern Intel Xeon hardware. The updated Intel extension also supports the recently released Microsoft
Continue reading...