Open Source LLM

Alibaba unleashes Qwen3 coding model for developers to push AI agent adoption

Alibaba’s Qwen3-Coder is expected to give developers a boost in creating AI agents. Photo: Shutterstock

Touted as the company’s “most advanced agentic AI coding model to date”, the Qwen3-Coder-480B-A35B-Instruct – built on a so-called Mixture of Experts (MoE) architecture – features a total of 480 billion parameters, 35 billion of which are active, and supports a 256,000-token context window, expandable to 1


Open Source LLM

Alibaba launches open-source AI coding model, touted as its most advanced to date - Reuters

Alibaba launches open-source AI coding model, touted as its most advanced to date Reuters

Open Source LLM

Alibaba upgrades Qwen3 model to outperform OpenAI, DeepSeek in maths, coding

Alibaba Group Holding unveiled an upgraded version of its third-generation Qwen3 family of large language models (LLMs), improving one of its members to score higher in maths and coding than products from rivals OpenAI and DeepSeek.

The new Qwen3-235B-A22B-Instruct-2507-FP8 is an open-source model that achieved “significant improvements in general capabilities, including instruction following, logical reasoning, text comprehension, mathematics, science, coding and tool usage


Open Source LLM

Moonshot’s Kimi K2 soars in popularity amid experts’ praise for Chinese AI

The launch of the Kimi K2 artificial intelligence model by Alibaba Group Holding-backed Moonshot AI has drawn rapid uptake amid praise from industry experts.

Downloads of Kimi K2, launched on July 11, doubled to 145,000 on Monday from 76,000 on Friday, according to AI and machine-learning developer platform Hugging Face.

The large language model (LLM) from the Beijing-based start-up uses a mixture-of-experts (MOE) architecture and


Open Source LLM

UK and ChatGPT maker OpenAI sign new strategic partnership - Reuters

UK and ChatGPT maker OpenAI sign new strategic partnership Reuters

Open Source LLM

Apple details how it trained its new AI models: 4 interesting highlights

During WWDC25, Apple announced new versions of its on-device and cloud-based foundation models. Now, they have published a tech report detailing how those models were trained, optimized, and evaluated. And the report includes some genuinely interesting under-the-hood tidbits.

In a comprehensive document called “ “, the company walks through multiple aspects of the new models, including their architecture, data sources, pre-training, post-training, tool use development, optimizations, and benchmarks.

Continue reading...

Open Source LLM

LLVM Begins Landing Distributed ThinLTO "DTLTO" Support

The LLVM compiler toolchain has begun upstreaming support for Distributed ThinLTO "DTLTO" as a new means of handling ThinLTO compilations for leveraging link-time optimizations.

ThinLTO is the more scalable and incremental approach for handling link-time optimizations by LLVM. With Distributed ThinLTO, the distribution of backend ThinLTO compilations can be done via external distribution systems.

The DTLTO Design Overview explains of Distributed ThinLTO:
"DTLTO enables the distribution of backend ThinLTO compilations via external distribution systems, such as Incredibuild. Existing support for distributing ThinLTO compilations typically involves separate

Open Source LLM

Chinese open-source AI models occupy top spots among global developers: ranking

China is home to the world’s top artificial intelligence (AI) models that are open-sourced, according to an American benchmarking platform created by researchers from the University of California, Berkeley.

Kimi K2, MiniMax M1, Qwen 3 and a variant of DeepSeek R1 were ranked as the world’s top open-sourced AI models, beating out offerings like Google’s Gemma 3-72B and Meta’s Llama 4-Maverick, LMArena said in a report on Friday.

The


Open Source LLM

NVIDIA Brings Reasoning Models to Consumers Ranging from 1.5B to 32B Parameters

Today, NVIDIA unveiled OpenReasoning-Nemotron, a quartet of distilled reasoning models with 1.5B, 7B, 14B, and 32B parameters, all derived from the 671B-parameter DeepSeek R1 0528. By compressing that massive teacher into four leaner Qwen‑2.5-based students, NVIDIA is making advanced reasoning experiments accessible even on standard gaming rigs, without the need to worry about hefty GPU bills and cloud usage. The key is not some elaborate
Continue reading...

Open Source LLM

Open-Source & Rust-Written Burn MATMUL Kernels Can Compete With NVIDIA's CUDA/cuBLAS

The open-source and Rust-based Burn deep learning framework developed by Tracel AI shared that their open-source matrix multiplication kernel performance can compete with and even outperform the NVIDIA CUDA cuBLAS performance. Plus Burn isn't limited to just NVIDIA GPUs but can work on most hardware/drivers, including a Vulkan back-end.

On Friday the Burn developers published a lengthy blog post going over their exciting MATMUL kernel performance relative to NVIDIA CUDA cuBLAS/CUTLASS and showing some really splendid results for this cross-platform,
Continue reading...
Menu