Open Source LLM

Russia's Sberbank plans to unveil LLM with reasoning capacity - Reuters

Russia's Sberbank plans to unveil LLM with reasoning capacity Reuters

Open Source LLM

Unpacking the bias of large language models

In a new study, researchers discover the root cause of a type of bias in LLMs, paving the way for more accurate and reliable AI systems.

Open Source LLM

DeepSeek rival MiniMax says its first AI reasoning model halves compute of R1

Shanghai-based artificial intelligence (AI) start-up MiniMax has launched an open-source reasoning model that it said requires just half the computing resources of rival DeepSeek ’s models for some tasks.

On Tuesday, the company announced the release of MiniMax-M1, its first reasoning model, on its official WeChat account. M1 consumes less than half the computing power of DeepSeek-R1 for reasoning tasks with a generation length of 64,000 tokens or fewer, according to a technical


Open Source LLM

Alibaba updates open-source Qwen3 models for AI deployment on Apple devices

In the MLX format, Qwen3 makes it possible to train and run a range of artificial intelligence models on the iPhone and other Apple devices. Photo: Shutterstock

In a Monday post on social media platform X , the Qwen team of Alibaba’s cloud computing unit said it launched open-source Qwen3 models optimised for Apple’s MLX framework for machine learning. Alibaba owns the South China Morning Post.

The MLX framework is intended to be user-friendly, but still efficient in training and deploying AI models on Apple’s silicon hardware. It has seen growing adoption among developers within the Apple


Open Source LLM

Chinese scientists find first evidence that AI could think like a human

Researchers from CAS and the South China University of Technology used behavioural experiments, computational modelling and neuroimaging analysis to investigate the relationship between object concept representations in LLMs and human cognition. Image: Shutterstock

Chinese researchers have confirmed for the first time that artificial intelligence large language models can spontaneously create a humanlike system to comprehend and sort natural objects, a process considered a pillar of human cognition.

It provides new evidence in a debate over the cognitive capacity of AI models, suggesting that artificial systems that reflect key aspects of human thinking may be possible.

“Understanding how humans conceptualise and categorise natural objects offers critical insights into perception and cognition,” the

Continue reading...

Open Source LLM

New paper pushes back on Apple’s LLM ‘reasoning collapse’ study

Apple’s recent AI research paper, “ ”, has been making waves for its blunt conclusion: even the most advanced Large Reasoning Models (LRMs) collapse on complex tasks. But not everyone agrees with that framing.

Today, Alex Lawsen, a researcher at Open Philanthropy, published a detailed rebuttal arguing that many of Apple’s most headline-grabbing findings boil down to experimental design flaws, not fundamental reasoning limits. The paper also credits Anthropic’s Claude Opus model as its co-author.

Lawsen’


Open Source LLM

AMD ROCm 7 Announced: MI350 Support, New Algorithms, Models & Advanced Features For AI Added, Focus on Inference With 3.5x Uplfit

AMD goes official with its next version of open software stack technologies in the form of , which further accelerates AI & developer productivity.

AMD Unveils ROCm 7: The Next-Generation of Open Stack Software Innovations With Focus on AI Inferencing

With the announcement of ROCm 7, AMD is finally moving forward from its ROCm 6 software stack, which itself has seen various updates over the last few years and since the advent of AI computing. The following are some of the main features that AMD is focusing on with ROCm 7:

  • Latest

Open Source LLM

OpenAI開源權重的模型要延到夏天

OpenAI推出開源權重模型的承諾,基於一些技術問題要延後到夏天了。

有鑒於DeepSeek帶動的開源AI模型風潮,也讓封閉模型業者如OpenAI開始思考靠攏開放路線。 OpenAI執行長Sam Altman 3月底宣佈 ,公司正在推進開源權重模型的專案,預計未來幾個月


Open Source LLM

Apple just gave developers access to its new local AI models, here’s how they perform

One of the very first announcements on this year’s WWDC was that for the first time, third‑party developers will get to tap directly into Apple’s on‑device AI with the new Foundation Models framework. But how do these models actually compare against what’s already out there?

With the new Foundation Models framework, third-party developers can now build on the same on-device AI stack used by Apple’s native apps.

In other words, this means that developers will now be able

Continue reading...

Open Source LLM

Joe Tsai says open-source AI will boost Alibaba’s cloud business

Alibaba chairman Joe Tsai speaks on stage during the VivaTech trade show in Paris on Wednesday. Photo: Xinmei Shen

in Paris, France

Alibaba Group Holding chairman Joe Tsai said that open-sourcing large language models (LLMs) would spur a surge in artificial intelligence (AI) applications and boost demand for cloud computing, as the company refines the focus of its sprawling business empire after “a period of huge ordeal”.

One reason Alibaba has opted to open-source its Qwen models is that it “democratises the usage of AI” and “proliferates

Continue reading...
Menu